When Do We Need LLMs? A Diagnostic for Language-Driven Bandits
Fuente:
arXiv
Saved in:
| Main Authors: | Berdica, Uljad, Acero, Fernando, Ipsen, Anton, Zehtabi, Parisa, Cashmore, Michael, Veloso, Manuela |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Beyond Manual Planning: Seating Allocation for Large Organizations
by: Ipsen, Anton, et al.
Published: (2026)
by: Ipsen, Anton, et al.
Published: (2026)
Deep Reinforcement Learning and Mean-Variance Strategies for Responsible Portfolio Optimization
by: Acero, Fernando, et al.
Published: (2024)
by: Acero, Fernando, et al.
Published: (2024)
Accelerating Cutting-Plane Algorithms via Reinforcement Learning Surrogates
by: Mana, Kyle, et al.
Published: (2023)
by: Mana, Kyle, et al.
Published: (2023)
Temporal Fairness in Decision Making Problems
by: Torres, Manuel R., et al.
Published: (2024)
by: Torres, Manuel R., et al.
Published: (2024)
Surrogate Assisted Monte Carlo Tree Search in Combinatorial Optimization
by: Amiri, Saeid, et al.
Published: (2024)
by: Amiri, Saeid, et al.
Published: (2024)
The Subset Sum Matching Problem
by: Wu, Yufei, et al.
Published: (2025)
by: Wu, Yufei, et al.
Published: (2025)
Evolving Many Worlds: Towards Open-Ended Discovery in Petri Dish NCA via Population-Based Training
by: Berdica, Uljad, et al.
Published: (2026)
by: Berdica, Uljad, et al.
Published: (2026)
Intent Factored Generation: Unleashing the Diversity in Your Language Model
by: Ahmed, Eltayeb, et al.
Published: (2025)
by: Ahmed, Eltayeb, et al.
Published: (2025)
A Clean Slate for Offline Reinforcement Learning
by: Jackson, Matthew Thomas, et al.
Published: (2025)
by: Jackson, Matthew Thomas, et al.
Published: (2025)
SOReL and TOReL: Two Methods for Fully Offline Reinforcement Learning
by: Fellows, Mattie, et al.
Published: (2025)
by: Fellows, Mattie, et al.
Published: (2025)
Capacity Planning and Scheduling for Jobs with Uncertainty in Resource Usage and Duration
by: Patra, Sunandita, et al.
Published: (2025)
by: Patra, Sunandita, et al.
Published: (2025)
Contrastive Explanations of Centralized Multi-agent Optimization Solutions
by: Zehtabi, Parisa, et al.
Published: (2023)
by: Zehtabi, Parisa, et al.
Published: (2023)
ELATE: Evolutionary Language model for Automated Time-series Engineering
by: Murray, Andrew, et al.
Published: (2025)
by: Murray, Andrew, et al.
Published: (2025)
GenePlan: Evolving Better Generalized PDDL Plans using Large Language Models
by: Murray, Andrew, et al.
Published: (2026)
by: Murray, Andrew, et al.
Published: (2026)
Do We Need Adam? Surprisingly Strong and Sparse Reinforcement Learning with SGD in LLMs
by: Mukherjee, Sagnik, et al.
Published: (2026)
by: Mukherjee, Sagnik, et al.
Published: (2026)
No One Size Fits All: QueryBandits for Hallucination Mitigation
by: Cho, Nicole, et al.
Published: (2026)
by: Cho, Nicole, et al.
Published: (2026)
QueryBandits for Hallucination Mitigation: Exploiting Semantic Features for No-Regret Rewriting
by: Cho, Nicole, et al.
Published: (2025)
by: Cho, Nicole, et al.
Published: (2025)
Intelligent Execution through Plan Analysis
by: Borrajo, Daniel, et al.
Published: (2024)
by: Borrajo, Daniel, et al.
Published: (2024)
Do We Need Bigger Models for Science? Task-Aware Retrieval with Small Language Models
by: Kelber, Florian, et al.
Published: (2026)
by: Kelber, Florian, et al.
Published: (2026)
Counterfactual Reasoning in Automated Planning
by: Pozanco, Alberto, et al.
Published: (2026)
by: Pozanco, Alberto, et al.
Published: (2026)
On Computing Plans with Uniform Action Costs
by: Pozanco, Alberto, et al.
Published: (2024)
by: Pozanco, Alberto, et al.
Published: (2024)
Rating Multi-Modal Time-Series Forecasting Models (MM-TSFM) for Robustness Through a Causal Lens
by: Lakkaraju, Kausik, et al.
Published: (2024)
by: Lakkaraju, Kausik, et al.
Published: (2024)
Do We Need Frontier Models to Verify Mathematical Proofs?
by: Naik, Aaditya, et al.
Published: (2026)
by: Naik, Aaditya, et al.
Published: (2026)
Do We Really Need to Approach the Entire Pareto Front in Many-Objective Bayesian Optimisation?
by: Jiang, Chao, et al.
Published: (2026)
by: Jiang, Chao, et al.
Published: (2026)
How Far Are We From AGI: Are LLMs All We Need?
by: Feng, Tao, et al.
Published: (2024)
by: Feng, Tao, et al.
Published: (2024)
When to ASK: Uncertainty-Gated Language Assistance for Reinforcement Learning
by: Monteiro, Juarez, et al.
Published: (2026)
by: Monteiro, Juarez, et al.
Published: (2026)
Do We Need Distinct Representations for Every Speech Token? Unveiling and Exploiting Redundancy in Large Speech Language Models
by: Xiang, Bajian, et al.
Published: (2026)
by: Xiang, Bajian, et al.
Published: (2026)
Do We Really Need a Large Number of Visual Prompts?
by: Kim, Youngeun, et al.
Published: (2023)
by: Kim, Youngeun, et al.
Published: (2023)
A Planning Compilation to Reason about Goal Achievement at Planning Time
by: Pozanco, Alberto, et al.
Published: (2025)
by: Pozanco, Alberto, et al.
Published: (2025)
Pharmacology Knowledge Graphs: Do We Need Chemical Structure for Drug Repurposing?
by: Abo-Dahab, Youssef, et al.
Published: (2026)
by: Abo-Dahab, Youssef, et al.
Published: (2026)
Do LLMs Build World Models From Text? A Multilingual Diagnostic of Spatial Reasoning
by: Pan, Zhikai, et al.
Published: (2026)
by: Pan, Zhikai, et al.
Published: (2026)
Learning When to Trust in Contextual Bandits
by: Ghasemi, Majid, et al.
Published: (2026)
by: Ghasemi, Majid, et al.
Published: (2026)
MambaOut: Do We Really Need Mamba for Vision?
by: Yu, Weihao, et al.
Published: (2024)
by: Yu, Weihao, et al.
Published: (2024)
Do We Need Large VLMs for Spotting Soccer Actions?
by: Chakraborty, Ritabrata, et al.
Published: (2025)
by: Chakraborty, Ritabrata, et al.
Published: (2025)
Do Large Language Models Mentalize When They Teach?
by: Harootonian, Sevan K., et al.
Published: (2026)
by: Harootonian, Sevan K., et al.
Published: (2026)
Creating a Causally Grounded Rating Method for Assessing the Robustness of AI Models for Time-Series Forecasting
by: Lakkaraju, Kausik, et al.
Published: (2025)
by: Lakkaraju, Kausik, et al.
Published: (2025)
Interpreting Language Reward Models via Contrastive Explanations
by: Jiang, Junqi, et al.
Published: (2024)
by: Jiang, Junqi, et al.
Published: (2024)
FlowMind: Automatic Workflow Generation with LLMs
by: Zeng, Zhen, et al.
Published: (2024)
by: Zeng, Zhen, et al.
Published: (2024)
We Urgently Need Intrinsically Kind Machines
by: Hewson, Joshua T. S.
Published: (2024)
by: Hewson, Joshua T. S.
Published: (2024)
Planning with Minimal Disruption
by: Pozanco, Alberto, et al.
Published: (2025)
by: Pozanco, Alberto, et al.
Published: (2025)
Similar Items
-
Beyond Manual Planning: Seating Allocation for Large Organizations
by: Ipsen, Anton, et al.
Published: (2026) -
Deep Reinforcement Learning and Mean-Variance Strategies for Responsible Portfolio Optimization
by: Acero, Fernando, et al.
Published: (2024) -
Accelerating Cutting-Plane Algorithms via Reinforcement Learning Surrogates
by: Mana, Kyle, et al.
Published: (2023) -
Temporal Fairness in Decision Making Problems
by: Torres, Manuel R., et al.
Published: (2024) -
Surrogate Assisted Monte Carlo Tree Search in Combinatorial Optimization
by: Amiri, Saeid, et al.
Published: (2024)