One Supervisor, Many Modalities: Adaptive Tool Orchestration for Autonomous Queries
Fuente:
arXiv
Salvato in:
| Autori principali: | Saini, Mayank, Bishwas, Arit Kumar |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Towards Resource-Efficient Multimodal Intelligence: Learned Routing among Specialized Expert Models
di: Saini, Mayank, et al.
Pubblicazione: (2025)
di: Saini, Mayank, et al.
Pubblicazione: (2025)
Evaluating the Efficacy of Hybrid Deep Learning Models in Distinguishing AI-Generated Text
di: Oketunji, Abiodun Finbarrs
Pubblicazione: (2023)
di: Oketunji, Abiodun Finbarrs
Pubblicazione: (2023)
How Many Parameters Does Your Task Really Need? Task Specific Pruning with LLM-Sieve
di: Reda, Waleed, et al.
Pubblicazione: (2025)
di: Reda, Waleed, et al.
Pubblicazione: (2025)
Large Language Model (LLM) Bias Index -- LLMBI
di: Oketunji, Abiodun Finbarrs, et al.
Pubblicazione: (2023)
di: Oketunji, Abiodun Finbarrs, et al.
Pubblicazione: (2023)
KV-Fold: One-Step KV-Cache Recurrence for Long-Context Inference
di: Nadali, Alireza, et al.
Pubblicazione: (2026)
di: Nadali, Alireza, et al.
Pubblicazione: (2026)
Adaptive Circuit Behavior and Generalization in Mechanistic Interpretability
di: Nainani, Jatin, et al.
Pubblicazione: (2024)
di: Nainani, Jatin, et al.
Pubblicazione: (2024)
ObfusQAte: A Proposed Framework to Evaluate LLM Robustness on Obfuscated Factual Question Answering
di: Ghosh, Shubhra, et al.
Pubblicazione: (2025)
di: Ghosh, Shubhra, et al.
Pubblicazione: (2025)
The Deterministic Horizon: When Extended Reasoning Fails and Tool Delegation Becomes Necessary
di: Guo, Dongxin, et al.
Pubblicazione: (2026)
di: Guo, Dongxin, et al.
Pubblicazione: (2026)
Adapting While Learning: Grounding LLMs for Scientific Problems with Intelligent Tool Usage Adaptation
di: Lyu, Bohan, et al.
Pubblicazione: (2024)
di: Lyu, Bohan, et al.
Pubblicazione: (2024)
Pressure-Testing Deception Probes in LLMs: Scaling, Robustness, and the Geometry of Deceptive Representations
di: Kumar, Sachin
Pubblicazione: (2026)
di: Kumar, Sachin
Pubblicazione: (2026)
LayerRoute: Input-Conditioned Adaptive Layer Skipping via LoRA Fine-Tuning for Agentic Language Models
di: Sikdar, Prateek Kumar
Pubblicazione: (2026)
di: Sikdar, Prateek Kumar
Pubblicazione: (2026)
Entropy-Based Measurement of Value Drift and Alignment Work in Large Language Models
di: Fadli, Samih
Pubblicazione: (2025)
di: Fadli, Samih
Pubblicazione: (2025)
Many LLMs Are More Utilitarian Than One
di: Keshmirian, Anita, et al.
Pubblicazione: (2025)
di: Keshmirian, Anita, et al.
Pubblicazione: (2025)
Beyond Hallucinations: A Composite Score for Measuring Reliability in Open-Source Large Language Models
di: Salla, Rohit Kumar, et al.
Pubblicazione: (2025)
di: Salla, Rohit Kumar, et al.
Pubblicazione: (2025)
Sleepless Nights, Sugary Days: Creating Synthetic Users with Health Conditions for Realistic Coaching Agent Interactions
di: Yun, Taedong, et al.
Pubblicazione: (2025)
di: Yun, Taedong, et al.
Pubblicazione: (2025)
Dealing with Annotator Disagreement in Hate Speech Classification
di: Dehghan, Somaiyeh, et al.
Pubblicazione: (2025)
di: Dehghan, Somaiyeh, et al.
Pubblicazione: (2025)
Improving Discrete Diffusion Unmasking Policies Beyond Explicit Reference Policies
di: Hong, Chunsan, et al.
Pubblicazione: (2025)
di: Hong, Chunsan, et al.
Pubblicazione: (2025)
The Metacognitive Probe: Five Behavioural Calibration Diagnostics for LLMs
di: Oliveira, Rafael C. T.
Pubblicazione: (2026)
di: Oliveira, Rafael C. T.
Pubblicazione: (2026)
Towards Intrinsic Interpretability of Large Language Models:A Survey of Design Principles and Architectures
di: Gao, Yutong, et al.
Pubblicazione: (2026)
di: Gao, Yutong, et al.
Pubblicazione: (2026)
IntentGrasp: A Comprehensive Benchmark for Intent Understanding
di: Yin, Yuwei, et al.
Pubblicazione: (2026)
di: Yin, Yuwei, et al.
Pubblicazione: (2026)
ALBA: A European Portuguese Benchmark for Evaluating Language and Linguistic Dimensions in Generative LLMs
di: Vieira, Inês, et al.
Pubblicazione: (2026)
di: Vieira, Inês, et al.
Pubblicazione: (2026)
Positional Failures in Long-Context LLMs: A Blind Spot in Reasoning Benchmarks
di: Zhang, Chuyifei, et al.
Pubblicazione: (2026)
di: Zhang, Chuyifei, et al.
Pubblicazione: (2026)
Do Personality Traits Interfere? Geometric Limitations of Steering in Large Language Models
di: Bhandari, Pranav, et al.
Pubblicazione: (2026)
di: Bhandari, Pranav, et al.
Pubblicazione: (2026)
Hista and Numca: Estimate State Value Effectively for LLM Reinforcement Learning
di: Chen, Zizhe, et al.
Pubblicazione: (2026)
di: Chen, Zizhe, et al.
Pubblicazione: (2026)
Path-Lock Expert: Separating Reasoning Mode in Hybrid Thinking via Architecture-Level Separation
di: Wang, Shouren, et al.
Pubblicazione: (2026)
di: Wang, Shouren, et al.
Pubblicazione: (2026)
PSK at SemEval-2026 Task 9: Multilingual Polarization Detection Using Ensemble Gemma Models with Synthetic Data Augmentation
di: Pulipaka, Srikar Kashyap
Pubblicazione: (2026)
di: Pulipaka, Srikar Kashyap
Pubblicazione: (2026)
Harmful Intent as a Geometrically Recoverable Feature of LLM Residual Streams
di: Llorente-Saguer, Isaac
Pubblicazione: (2026)
di: Llorente-Saguer, Isaac
Pubblicazione: (2026)
Skill Availability and Presentation Granularity in Large-Language-Model Agents: A Controlled SkillsBench Study
di: Xu, Xiaonan, et al.
Pubblicazione: (2026)
di: Xu, Xiaonan, et al.
Pubblicazione: (2026)
The Expert Strikes Back: Interpreting Mixture-of-Experts Language Models at Expert Level
di: Herbst, Jeremy, et al.
Pubblicazione: (2026)
di: Herbst, Jeremy, et al.
Pubblicazione: (2026)
Disentangling Direction and Magnitude in Transformer Representations: A Double Dissociation Through L2-Matched Perturbation Analysis
di: Vardhan, Mangadoddi Srikar, et al.
Pubblicazione: (2026)
di: Vardhan, Mangadoddi Srikar, et al.
Pubblicazione: (2026)
Three Regimes of Context-Parametric Conflict: A Predictive Framework and Empirical Validation
di: Venkata, Pruthvinath Jeripity
Pubblicazione: (2026)
di: Venkata, Pruthvinath Jeripity
Pubblicazione: (2026)
Large Language Models Generate Harmful Content Using a Distinct, Unified Mechanism
di: Orgad, Hadas, et al.
Pubblicazione: (2026)
di: Orgad, Hadas, et al.
Pubblicazione: (2026)
Self-Consistency from Only Two Samples: CoT-PoT Ensembling for Efficient LLM Reasoning
di: Saparkhan, Raman, et al.
Pubblicazione: (2026)
di: Saparkhan, Raman, et al.
Pubblicazione: (2026)
Robust Explanations for User Trust in Enterprise NLP Systems
di: Zhang, Guilin, et al.
Pubblicazione: (2026)
di: Zhang, Guilin, et al.
Pubblicazione: (2026)
AMALIA Technical Report: A Fully Open Source Large Language Model for European Portuguese
di: Simplício, Afonso, et al.
Pubblicazione: (2026)
di: Simplício, Afonso, et al.
Pubblicazione: (2026)
The Geometry of Harmful Intent: Training-Free Anomaly Detection via Angular Deviation in LLM Residual Streams
di: Llorente-Saguer, Isaac
Pubblicazione: (2026)
di: Llorente-Saguer, Isaac
Pubblicazione: (2026)
Metaphors are a Source of Cross-Domain Misalignment of Large Reasoning Models
di: Hu, Zhibo, et al.
Pubblicazione: (2026)
di: Hu, Zhibo, et al.
Pubblicazione: (2026)
Do Models Know Why They Changed Their Mind? Interpretability and Faithfulness of Chain-of-Thought Under Knowledge Conflict
di: Venkata, Pruthvinath Jeripity
Pubblicazione: (2026)
di: Venkata, Pruthvinath Jeripity
Pubblicazione: (2026)
Does LLM Alignment Really Need Diversity? An Empirical Study of Adapting RLVR Methods for Moral Reasoning
di: Zhang, Zhaowei, et al.
Pubblicazione: (2026)
di: Zhang, Zhaowei, et al.
Pubblicazione: (2026)
Extracting Small Translation Specialists from LLMs by Aggressively Pruning Experts
di: Martin, Liu O., et al.
Pubblicazione: (2026)
di: Martin, Liu O., et al.
Pubblicazione: (2026)
Documenti analoghi
-
Towards Resource-Efficient Multimodal Intelligence: Learned Routing among Specialized Expert Models
di: Saini, Mayank, et al.
Pubblicazione: (2025) -
Evaluating the Efficacy of Hybrid Deep Learning Models in Distinguishing AI-Generated Text
di: Oketunji, Abiodun Finbarrs
Pubblicazione: (2023) -
How Many Parameters Does Your Task Really Need? Task Specific Pruning with LLM-Sieve
di: Reda, Waleed, et al.
Pubblicazione: (2025) -
Large Language Model (LLM) Bias Index -- LLMBI
di: Oketunji, Abiodun Finbarrs, et al.
Pubblicazione: (2023) -
KV-Fold: One-Step KV-Cache Recurrence for Long-Context Inference
di: Nadali, Alireza, et al.
Pubblicazione: (2026)