BAR Conjecture: the Feasibility of Inference Budget-Constrained LLM Services with Authenticity and Reasoning
Fuente:
arXiv
Saved in:
| Main Authors: | Zhou, Jinan, Ghosh, Rajat, Bhargava, Vaishnavi, Dutta, Debojyoti, Singhal, Aryan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
CPP-UT-Bench: Can LLMs Write Complex Unit Tests in C++?
by: Bhargava, Vaishnavi, et al.
Published: (2024)
by: Bhargava, Vaishnavi, et al.
Published: (2024)
Predictive Scaling Laws for Efficient GRPO Training of Large Reasoning Models
by: Nimmaturi, Datta, et al.
Published: (2025)
by: Nimmaturi, Datta, et al.
Published: (2025)
Go-UT-Bench: A Fine-Tuning Dataset for LLM-Based Unit Test Generation in Go
by: Pipalani, Yashshi, et al.
Published: (2025)
by: Pipalani, Yashshi, et al.
Published: (2025)
RANGER -- Repository-Level Agent for Graph-Enhanced Retrieval
by: Shah, Pratik, et al.
Published: (2025)
by: Shah, Pratik, et al.
Published: (2025)
SWE-Tester: Training Open-Source LLMs for Issue Reproduction in Real-World Repositories
by: Soni, Aditya Bharat, et al.
Published: (2026)
by: Soni, Aditya Bharat, et al.
Published: (2026)
Action Shapley: A Training Data Selection Metric for World Model in Reinforcement Learning
by: Ghosh, Rajat, et al.
Published: (2026)
by: Ghosh, Rajat, et al.
Published: (2026)
Efficient Alignment of Large Language Models via Data Sampling
by: Khera, Amrit, et al.
Published: (2024)
by: Khera, Amrit, et al.
Published: (2024)
A Multi-Agent Framework for Stateful Inference-Time Search
by: Lalan, Arshika, et al.
Published: (2025)
by: Lalan, Arshika, et al.
Published: (2025)
XAI-Driven Deep Learning for Protein Sequence Functional Group Classification
by: Chakraborty, Pratik, et al.
Published: (2025)
by: Chakraborty, Pratik, et al.
Published: (2025)
The ALCHEmist: Automated Labeling 500x CHEaper Than LLM Data Annotators
by: Huang, Tzu-Heng, et al.
Published: (2024)
by: Huang, Tzu-Heng, et al.
Published: (2024)
BudgetThinker: Empowering Budget-aware LLM Reasoning with Control Tokens
by: Wen, Hao, et al.
Published: (2025)
by: Wen, Hao, et al.
Published: (2025)
Token-Budget-Aware LLM Reasoning
by: Han, Tingxu, et al.
Published: (2024)
by: Han, Tingxu, et al.
Published: (2024)
Semantic Search and Recommendation Algorithm
by: Duhan, Aryan, et al.
Published: (2024)
by: Duhan, Aryan, et al.
Published: (2024)
Budget-Constrained Online Retrieval-Augmented Generation: The Chunk-as-a-Service Model
by: Al-Maliki, Shawqi, et al.
Published: (2026)
by: Al-Maliki, Shawqi, et al.
Published: (2026)
Input Guided Multiple Deconstruction Single Reconstruction neural network models for Matrix Factorization
by: Dutta, Prasun, et al.
Published: (2024)
by: Dutta, Prasun, et al.
Published: (2024)
Forward-Cooperation-Backward (FCB) learning in a Multi-Encoding Uni-Decoding neural network architecture
by: Dutta, Prasun, et al.
Published: (2025)
by: Dutta, Prasun, et al.
Published: (2025)
A review on Machine Learning based User-Centric Multimedia Streaming Techniques
by: Ghosh, Monalisa, et al.
Published: (2024)
by: Ghosh, Monalisa, et al.
Published: (2024)
Efficient Contextual LLM Cascades through Budget-Constrained Policy Learning
by: Zhang, Xuechen, et al.
Published: (2024)
by: Zhang, Xuechen, et al.
Published: (2024)
Learning to Bid with Unknown Private Values in Budget-Constrained First-Price Auctions
by: Hu, Zihao, et al.
Published: (2026)
by: Hu, Zihao, et al.
Published: (2026)
Targeted Tests for LLM Reasoning: An Audit-Constrained Protocol
by: Li, Hongmin
Published: (2026)
by: Li, Hongmin
Published: (2026)
Dependent Randomized Rounding for Budget Constrained Experimental Design
by: Yamin, Khurram, et al.
Published: (2025)
by: Yamin, Khurram, et al.
Published: (2025)
LLM-as-Judge on a Budget
by: Saha, Aadirupa, et al.
Published: (2026)
by: Saha, Aadirupa, et al.
Published: (2026)
ConceptSearch: Towards Efficient Program Search Using LLMs for Abstraction and Reasoning Corpus (ARC)
by: Singhal, Kartik, et al.
Published: (2024)
by: Singhal, Kartik, et al.
Published: (2024)
CR-Bench: Evaluating the Real-World Utility of AI Code Review Agents
by: Pereira, Kristen, et al.
Published: (2026)
by: Pereira, Kristen, et al.
Published: (2026)
MCAP: Deployment-Time Layer Profiling for Memory-Constrained LLM Inference
by: Das, Anurita
Published: (2026)
by: Das, Anurita
Published: (2026)
Fixed-Budget Constrained Best Arm Identification in Grouped Bandits
by: Mukherjee, Raunak, et al.
Published: (2026)
by: Mukherjee, Raunak, et al.
Published: (2026)
Elastic Spectral State Space Models for Budgeted Inference
by: Song, Dachuan, et al.
Published: (2026)
by: Song, Dachuan, et al.
Published: (2026)
UCB-type Algorithm for Budget-Constrained Expert Learning
by: Latypov, Ilgam, et al.
Published: (2025)
by: Latypov, Ilgam, et al.
Published: (2025)
Spend Less, Reason Better: Budget-Aware Value Tree Search for LLM Agents
by: Li, Yushu, et al.
Published: (2026)
by: Li, Yushu, et al.
Published: (2026)
DARTS: Targeting Prognostic Covariates in Budget-Constrained Sequential Experiments
by: Husar, Kateryna, et al.
Published: (2026)
by: Husar, Kateryna, et al.
Published: (2026)
An Adaptable Budget Planner for Enhancing Budget-Constrained Auto-Bidding in Online Advertising
by: Duan, Zhijian, et al.
Published: (2025)
by: Duan, Zhijian, et al.
Published: (2025)
Leveraging a Simulator for Learning Causal Representations from Post-Treatment Covariates for CATE
by: Nagalapatti, Lokesh, et al.
Published: (2025)
by: Nagalapatti, Lokesh, et al.
Published: (2025)
PairNet: Training with Observed Pairs to Estimate Individual Treatment Effect
by: Nagalapatti, Lokesh, et al.
Published: (2024)
by: Nagalapatti, Lokesh, et al.
Published: (2024)
Adaptive LLM Routing under Budget Constraints
by: Panda, Pranoy, et al.
Published: (2025)
by: Panda, Pranoy, et al.
Published: (2025)
Adaptive Budget Allocation in LLM-Augmented Surveys
by: Ye, Zikun, et al.
Published: (2026)
by: Ye, Zikun, et al.
Published: (2026)
Improving LLM Reasoning through Scaling Inference Computation with Collaborative Verification
by: Liang, Zhenwen, et al.
Published: (2024)
by: Liang, Zhenwen, et al.
Published: (2024)
Partial Causal Structure Learning for Valid Selective Conformal Inference under Interventions
by: Asiaee, Amir, et al.
Published: (2026)
by: Asiaee, Amir, et al.
Published: (2026)
Enhancing Speech Emotion Recognition via Fine-Tuning Pre-Trained Models and Hyper-Parameter Optimisation
by: Golbaghi, Aryan, et al.
Published: (2025)
by: Golbaghi, Aryan, et al.
Published: (2025)
Low-Budget Simulation-Based Inference with Bayesian Neural Networks
by: Delaunoy, Arnaud, et al.
Published: (2024)
by: Delaunoy, Arnaud, et al.
Published: (2024)
SpecOffload: Unlocking Latent GPU Capacity for LLM Inference on Resource-Constrained Devices
by: Zhuge, Xiangwen, et al.
Published: (2025)
by: Zhuge, Xiangwen, et al.
Published: (2025)
Similar Items
-
CPP-UT-Bench: Can LLMs Write Complex Unit Tests in C++?
by: Bhargava, Vaishnavi, et al.
Published: (2024) -
Predictive Scaling Laws for Efficient GRPO Training of Large Reasoning Models
by: Nimmaturi, Datta, et al.
Published: (2025) -
Go-UT-Bench: A Fine-Tuning Dataset for LLM-Based Unit Test Generation in Go
by: Pipalani, Yashshi, et al.
Published: (2025) -
RANGER -- Repository-Level Agent for Graph-Enhanced Retrieval
by: Shah, Pratik, et al.
Published: (2025) -
SWE-Tester: Training Open-Source LLMs for Issue Reproduction in Real-World Repositories
by: Soni, Aditya Bharat, et al.
Published: (2026)