Scaling Down to Scale Up: A Cost-Benefit Analysis of Replacing OpenAI's LLM with Open Source SLMs in Production
Fuente:
arXiv
Salvato in:
| Autori principali: | Irugalbandara, Chandra, Mahendra, Ashish, Daynauth, Roland, Arachchige, Tharuka Kasthuri, Dantanarayana, Jayanaka, Flautner, Krisztian, Tang, Lingjia, Kang, Yiping, Mars, Jason |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2023
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
TOBUGraph: Knowledge Graph-Based Retrieval for Enhanced LLM Performance Beyond RAG
di: Kashmira, Savini, et al.
Pubblicazione: (2024)
di: Kashmira, Savini, et al.
Pubblicazione: (2024)
GraphRunner: A Multi-Stage Framework for Efficient and Accurate Graph-Based Retrieval
di: Kashmira, Savini, et al.
Pubblicazione: (2025)
di: Kashmira, Savini, et al.
Pubblicazione: (2025)
GraphMend: Code Transformations for Fixing Graph Breaks in PyTorch 2
di: Kashmira, Savini, et al.
Pubblicazione: (2025)
di: Kashmira, Savini, et al.
Pubblicazione: (2025)
SLMEval: Entropy-Based Calibration for Human-Aligned Evaluation of Large Language Models
di: Daynauth, Roland, et al.
Pubblicazione: (2025)
di: Daynauth, Roland, et al.
Pubblicazione: (2025)
Ranking Unraveled: Recipes for LLM Rankings in Head-to-Head AI Combat
di: Daynauth, Roland, et al.
Pubblicazione: (2024)
di: Daynauth, Roland, et al.
Pubblicazione: (2024)
Prompt Less, Smile More: MTP with Semantic Engineering in Lieu of Prompt Engineering
di: Dantanarayana, Jayanaka L., et al.
Pubblicazione: (2025)
di: Dantanarayana, Jayanaka L., et al.
Pubblicazione: (2025)
MTP: A Meaning-Typed Language Abstraction for AI-Integrated Programming
di: Dantanarayana, Jayanaka L., et al.
Pubblicazione: (2024)
di: Dantanarayana, Jayanaka L., et al.
Pubblicazione: (2024)
FeDABoost: Fairness Aware Federated Learning with Adaptive Boosting
di: Arachchige, Tharuka Kasthuri, et al.
Pubblicazione: (2025)
di: Arachchige, Tharuka Kasthuri, et al.
Pubblicazione: (2025)
Aligning Model Evaluations with Human Preferences: Mitigating Token Count Bias in Language Model Assessments
di: Daynauth, Roland, et al.
Pubblicazione: (2024)
di: Daynauth, Roland, et al.
Pubblicazione: (2024)
Meaning Typed Prompting: A Technique for Efficient, Reliable Structured Output Generation
di: Irugalbandara, Chandra
Pubblicazione: (2024)
di: Irugalbandara, Chandra
Pubblicazione: (2024)
Guylingo: The Republic of Guyana Creole Corpora
di: Clarke, Christopher, et al.
Pubblicazione: (2024)
di: Clarke, Christopher, et al.
Pubblicazione: (2024)
A Large-Scale Empirical Analysis of Custom GPTs' Vulnerabilities in the OpenAI Ecosystem
di: Ogundoyin, Sunday Oyinlola, et al.
Pubblicazione: (2025)
di: Ogundoyin, Sunday Oyinlola, et al.
Pubblicazione: (2025)
Elevating intelligent voice assistant chatbots with natural language processing, and OpenAI technologies
di: Nilesh B. Korade, Mahendra B. Salunke, Amol A. Bhosle, Gayatri G. Asalkar, Bechoo Lal, Prashant B. Kumbharkar
Pubblicazione: (2025)
di: Nilesh B. Korade, Mahendra B. Salunke, Amol A. Bhosle, Gayatri G. Asalkar, Bechoo Lal, Prashant B. Kumbharkar
Pubblicazione: (2025)
Sora OpenAI's Prelude: Social Media Perspectives on Sora OpenAI and the Future of AI Video Generation
di: Mogavi, Reza Hadi, et al.
Pubblicazione: (2024)
di: Mogavi, Reza Hadi, et al.
Pubblicazione: (2024)
OpenAI GPT-5 System Card
di: Singh, Aaditya, et al.
Pubblicazione: (2025)
di: Singh, Aaditya, et al.
Pubblicazione: (2025)
Privacy and Security Threat for OpenAI GPTs
di: Wenying, Wei, et al.
Pubblicazione: (2025)
di: Wenying, Wei, et al.
Pubblicazione: (2025)
OpenAI o1 System Card
di: OpenAI, et al.
Pubblicazione: (2024)
di: OpenAI, et al.
Pubblicazione: (2024)
OpenAI Partners with Learning Management System
Pubblicazione: (2025)
Pubblicazione: (2025)
Ozone O3 A Dynamic Neuromorphic Intelligence Architecture for Adaptive Intelligences
di: Balasooriya, Tharuka
Pubblicazione: (2024)
di: Balasooriya, Tharuka
Pubblicazione: (2024)
BlazingAML: High-Throughput Anti-Money Laundering (AML) via Multi-Stage Graph Mining
di: Ye, Haojie, et al.
Pubblicazione: (2026)
di: Ye, Haojie, et al.
Pubblicazione: (2026)
On Sarcasm Detection with OpenAI GPT-based Models
di: Gole, Montgomery, et al.
Pubblicazione: (2023)
di: Gole, Montgomery, et al.
Pubblicazione: (2023)
Evaluating the Effectiveness of OpenAI's Parental Control System
di: Ersoz, Kerem, et al.
Pubblicazione: (2026)
di: Ersoz, Kerem, et al.
Pubblicazione: (2026)
Chapter 16 Scaling Up, Down, and Across
di: Druckman, Daniel
Pubblicazione: (2024)
di: Druckman, Daniel
Pubblicazione: (2024)
Open-Reasoner-Zero: An Open Source Approach to Scaling Up Reinforcement Learning on the Base Model
di: Hu, Jingcheng, et al.
Pubblicazione: (2025)
di: Hu, Jingcheng, et al.
Pubblicazione: (2025)
One Agent Too Many: User Perspectives on Approaches to Multi-agent Conversational AI
di: Clarke, Christopher, et al.
Pubblicazione: (2024)
di: Clarke, Christopher, et al.
Pubblicazione: (2024)
OpenAI for OpenAPI: Automated generation of REST API specification via LLMs
di: Chen, Hao, et al.
Pubblicazione: (2026)
di: Chen, Hao, et al.
Pubblicazione: (2026)
Evaluating Test-Time Scaling LLMs for Legal Reasoning: OpenAI o1, DeepSeek-R1, and Beyond
di: Hu, Yinghao, et al.
Pubblicazione: (2025)
di: Hu, Yinghao, et al.
Pubblicazione: (2025)
Evaluation of OpenAI o1: Opportunities and Challenges of AGI
di: Zhong, Tianyang, et al.
Pubblicazione: (2024)
di: Zhong, Tianyang, et al.
Pubblicazione: (2024)
Quantization for OpenAI's Whisper Models: A Comparative Analysis
di: Andreyev, Allison
Pubblicazione: (2025)
di: Andreyev, Allison
Pubblicazione: (2025)
An Empirical Study of OpenAI API Discussions on Stack Overflow
di: Chen, Xiang, et al.
Pubblicazione: (2025)
di: Chen, Xiang, et al.
Pubblicazione: (2025)
Using OpenAI GPT to Generate Reading Comprehension Items
di: Ayfer Sayin, et al.
Pubblicazione: (2024)
di: Ayfer Sayin, et al.
Pubblicazione: (2024)
Position: AI Scaling: From Up to Down and Out
di: Wang, Yunke, et al.
Pubblicazione: (2025)
di: Wang, Yunke, et al.
Pubblicazione: (2025)
Scaling Down to Scale Up: A Guide to Parameter-Efficient Fine-Tuning
di: Lialin, Vladislav, et al.
Pubblicazione: (2023)
di: Lialin, Vladislav, et al.
Pubblicazione: (2023)
Cost-Benefit of Connected Milk Tank Scales
di: DEPUILLE, Laurence
Pubblicazione: (2026)
di: DEPUILLE, Laurence
Pubblicazione: (2026)
Is GPT-OSS Good? A Comprehensive Evaluation of OpenAI's Latest Open Source Models
di: Bi, Ziqian, et al.
Pubblicazione: (2025)
di: Bi, Ziqian, et al.
Pubblicazione: (2025)
Health Benefits and Therapeutic Potential of Quercetin
di: Mahendra Aryal
Pubblicazione: (2026)
di: Mahendra Aryal
Pubblicazione: (2026)
OpenAI ChatGPT interprets Radiological Images: GPT-4 as a Medical Doctor for a Fast Check-Up
di: Aydin, Omer, et al.
Pubblicazione: (2025)
di: Aydin, Omer, et al.
Pubblicazione: (2025)
Unlocking NACE Classification Embeddings with OpenAI for Enhanced Analysis and Processing
di: Vidali, Andrea, et al.
Pubblicazione: (2024)
di: Vidali, Andrea, et al.
Pubblicazione: (2024)
Understanding and Benchmarking Artificial Intelligence: OpenAI's o3 Is Not AGI
di: Pfister, Rolf, et al.
Pubblicazione: (2025)
di: Pfister, Rolf, et al.
Pubblicazione: (2025)
OpenAI's Approach to External Red Teaming for AI Models and Systems
di: Ahmad, Lama, et al.
Pubblicazione: (2025)
di: Ahmad, Lama, et al.
Pubblicazione: (2025)
Documenti analoghi
-
TOBUGraph: Knowledge Graph-Based Retrieval for Enhanced LLM Performance Beyond RAG
di: Kashmira, Savini, et al.
Pubblicazione: (2024) -
GraphRunner: A Multi-Stage Framework for Efficient and Accurate Graph-Based Retrieval
di: Kashmira, Savini, et al.
Pubblicazione: (2025) -
GraphMend: Code Transformations for Fixing Graph Breaks in PyTorch 2
di: Kashmira, Savini, et al.
Pubblicazione: (2025) -
SLMEval: Entropy-Based Calibration for Human-Aligned Evaluation of Large Language Models
di: Daynauth, Roland, et al.
Pubblicazione: (2025) -
Ranking Unraveled: Recipes for LLM Rankings in Head-to-Head AI Combat
di: Daynauth, Roland, et al.
Pubblicazione: (2024)