R2V Agent: Teaching SLMs When to Ask for Help
Fuente:
arXiv
Saved in:
| Main Authors: | Hemadri, Raghu Vamshi, Mohammed, Humaira Firdowse, Maheshwary, Rishabh, Daruru, Srivatsava, Davasam, Sagar, Yadav, Vikas, Sunkara, Srinivas, Rajeswar, Sai |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Terminal Agents Suffice for Enterprise Automation
by: Bechard, Patrice, et al.
Published: (2026)
by: Bechard, Patrice, et al.
Published: (2026)
Optimizing What Matters: AUC-Driven Learning for Robust Neural Retrieval
by: Sheikholeslami, Nima, et al.
Published: (2025)
by: Sheikholeslami, Nima, et al.
Published: (2025)
EnterpriseOps-Gym: Environments and Evaluations for Stateful Agentic Planning and Tool Use in Enterprise Settings
by: Malay, Shiva Krishna Reddy, et al.
Published: (2026)
by: Malay, Shiva Krishna Reddy, et al.
Published: (2026)
Augmenting LLM Reasoning with Dynamic Notes Writing for Complex QA
by: Maheshwary, Rishabh, et al.
Published: (2025)
by: Maheshwary, Rishabh, et al.
Published: (2025)
PrefixLLM: LLM-aided Prefix Circuit Design
by: Xiao, Weihua, et al.
Published: (2024)
by: Xiao, Weihua, et al.
Published: (2024)
OncoReason: Structuring Clinical Reasoning in LLMs for Robust and Interpretable Survival Prediction
by: Hemadri, Raghu Vamshi, et al.
Published: (2025)
by: Hemadri, Raghu Vamshi, et al.
Published: (2025)
Do Enterprise Systems Need Learned World Models? The Importance of Context to Infer Dynamics
by: Nair, Jishnu Sethumadhavan, et al.
Published: (2026)
by: Nair, Jishnu Sethumadhavan, et al.
Published: (2026)
Curry-DPO: Enhancing Alignment using Curriculum Learning & Ranked Preferences
by: Pattnaik, Pulkit, et al.
Published: (2024)
by: Pattnaik, Pulkit, et al.
Published: (2024)
M2Lingual: Enhancing Multilingual, Multi-Turn Instruction Alignment in Large Language Models
by: Maheshwary, Rishabh, et al.
Published: (2024)
by: Maheshwary, Rishabh, et al.
Published: (2024)
Apriel-1.5-15b-Thinker
by: Radhakrishna, Shruthan, et al.
Published: (2025)
by: Radhakrishna, Shruthan, et al.
Published: (2025)
ColMate: Contrastive Late Interaction and Masked Text for Multimodal Document Retrieval
by: Masry, Ahmed, et al.
Published: (2025)
by: Masry, Ahmed, et al.
Published: (2025)
Revitalizing Saturated Benchmarks: A Weighted Metric Approach for Differentiating Large Language Model Performance
by: Etzine, Bryan, et al.
Published: (2025)
by: Etzine, Bryan, et al.
Published: (2025)
VeriDispatcher: Multi-Model Dispatching through Pre-Inference Difficulty Prediction for RTL Generation Optimization
by: Wang, Zeng, et al.
Published: (2025)
by: Wang, Zeng, et al.
Published: (2025)
Layer-Wise Quantization: A Pragmatic and Effective Method for Quantizing LLMs Beyond Integer Bit-Levels
by: Dumitru, Razvan-Gabriel, et al.
Published: (2024)
by: Dumitru, Razvan-Gabriel, et al.
Published: (2024)
Prompting with Phonemes: Enhancing LLMs' Multilinguality for Non-Latin Script Languages
by: Nguyen, Hoang H, et al.
Published: (2024)
by: Nguyen, Hoang H, et al.
Published: (2024)
Framework for Co-distillation Driven Federated Learning to Address Class Imbalance in Healthcare
by: Racha, Suraj, et al.
Published: (2024)
by: Racha, Suraj, et al.
Published: (2024)
VeriLoC: Line-of-Code Level Prediction of Hardware Design Quality from Verilog Code
by: Hemadri, Raghu Vamshi, et al.
Published: (2025)
by: Hemadri, Raghu Vamshi, et al.
Published: (2025)
TrojanLoC: LLM-based Framework for RTL Trojan Localization
by: Xiao, Weihua, et al.
Published: (2025)
by: Xiao, Weihua, et al.
Published: (2025)
LLM-Assisted Relevance Assessments: When Should We Ask LLMs for Help?
by: Takehi, Rikiya, et al.
Published: (2024)
by: Takehi, Rikiya, et al.
Published: (2024)
From crisis to transformation: Evaluating the implementation of DIKSHA in India's EdTech landscape
by: Rahul Pachori, et al.
Published: (2026)
by: Rahul Pachori, et al.
Published: (2026)
Apriel-Nemotron-15B-Thinker
by: Radhakrishna, Shruthan, et al.
Published: (2025)
by: Radhakrishna, Shruthan, et al.
Published: (2025)
Limits of Difficulty Scaling: Hard Samples Yield Diminishing Returns in GRPO-Tuned SLMs
by: Yadav, Suraj, et al.
Published: (2026)
by: Yadav, Suraj, et al.
Published: (2026)
Learning When to Ask for Help: Efficient Interactive Navigation via Implicit Uncertainty Estimation
by: Igbinedion, Ifueko, et al.
Published: (2023)
by: Igbinedion, Ifueko, et al.
Published: (2023)
Grammar Search for Multi-Agent Systems
by: Singh, Mayank, et al.
Published: (2025)
by: Singh, Mayank, et al.
Published: (2025)
Need Help?...Ask Your Mentor.
by: Logsdon, Janis
Published: (1992)
by: Logsdon, Janis
Published: (1992)
Analyzing Reluctance to Ask for Help When Cooperating With Robots: Insights to Integrate Artificial Agents in HRC
by: Martin, Ane San, et al.
Published: (2025)
by: Martin, Ane San, et al.
Published: (2025)
HiL-Bench (Human-in-Loop Benchmark): Do Agents Know When to Ask for Help?
by: Trinh, Tu, et al.
Published: (2026)
by: Trinh, Tu, et al.
Published: (2026)
Emotions in the Loop: A Survey of Affective Computing for Emotional Support
by: Hegde, Karishma, et al.
Published: (2025)
by: Hegde, Karishma, et al.
Published: (2025)
Knowing When Not to Answer: Evaluating Abstention in Multimodal Reasoning Systems
by: Madhusudhan, Nishanth, et al.
Published: (2026)
by: Madhusudhan, Nishanth, et al.
Published: (2026)
Performance Analysis and Evaluation of Cognitive Radio-Based WRAN Systems Using Adaptive CFAR Thresholding
by: Jenigala, Vamshi, et al.
Published: (2026)
by: Jenigala, Vamshi, et al.
Published: (2026)
STaR-GATE: Teaching Language Models to Ask Clarifying Questions
by: Andukuri, Chinmaya, et al.
Published: (2024)
by: Andukuri, Chinmaya, et al.
Published: (2024)
Shot Noise near Quantum-Criticality
by: Raghu, Srinivas, et al.
Published: (2024)
by: Raghu, Srinivas, et al.
Published: (2024)
Ask Early, Ask Late, Ask Right: When Does Clarification Timing Matter for Long-Horizon Agents?
by: Gulati, Anmol, et al.
Published: (2026)
by: Gulati, Anmol, et al.
Published: (2026)
Pulsatile hydromagnetic Carreau–Yasuda nanofluid flow in a nonlinearly radiated inclined porous channel with nonuniform heat source/sink: Entropy analysis
by: Joseph Josuva, et al.
Published: (2025)
by: Joseph Josuva, et al.
Published: (2025)
When and What to Ask: AskBench and Rubric-Guided RLVR for LLM Clarification
by: Zhao, Jiale, et al.
Published: (2026)
by: Zhao, Jiale, et al.
Published: (2026)
On ${\cal M}$-Theory Dual of Large-$N$ Thermal QCD-Like Theories up to ${\cal O}(R^4)$ and $G$-Structure Classification of Underlying Non-Supersymmetric Geometries
by: Yadav, Vikas, et al.
Published: (2020)
by: Yadav, Vikas, et al.
Published: (2020)
Therefore I am. I Think
by: Esakkiraja, Esakkivel, et al.
Published: (2026)
by: Esakkiraja, Esakkivel, et al.
Published: (2026)
May We Help You Find Something? AskNSDL!
by: Silverstein, Joanne
Published: (2003)
by: Silverstein, Joanne
Published: (2003)
On metacyclic p-group codes
by: Chahal, Seema, et al.
Published: (2025)
by: Chahal, Seema, et al.
Published: (2025)
Cut Groups: The Progress And The Problems
by: Chahal, Seema, et al.
Published: (2025)
by: Chahal, Seema, et al.
Published: (2025)
Similar Items
-
Terminal Agents Suffice for Enterprise Automation
by: Bechard, Patrice, et al.
Published: (2026) -
Optimizing What Matters: AUC-Driven Learning for Robust Neural Retrieval
by: Sheikholeslami, Nima, et al.
Published: (2025) -
EnterpriseOps-Gym: Environments and Evaluations for Stateful Agentic Planning and Tool Use in Enterprise Settings
by: Malay, Shiva Krishna Reddy, et al.
Published: (2026) -
Augmenting LLM Reasoning with Dynamic Notes Writing for Complex QA
by: Maheshwary, Rishabh, et al.
Published: (2025) -
PrefixLLM: LLM-aided Prefix Circuit Design
by: Xiao, Weihua, et al.
Published: (2024)