Gespeichert in:
| Hauptverfasser: | Ding, Wenxuan, Tomlin, Nicholas, Durrett, Greg |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2602.16699 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Molecular Facts: Desiderata for Decontextualization in LLM Fact Verification
von: Gunjal, Anisha, et al.
Veröffentlicht: (2024)
von: Gunjal, Anisha, et al.
Veröffentlicht: (2024)
SynthesizRR: Generating Diverse Datasets with Retrieval Augmentation
von: Divekar, Abhishek, et al.
Veröffentlicht: (2024)
von: Divekar, Abhishek, et al.
Veröffentlicht: (2024)
MiniCheck: Efficient Fact-Checking of LLMs on Grounding Documents
von: Tang, Liyan, et al.
Veröffentlicht: (2024)
von: Tang, Liyan, et al.
Veröffentlicht: (2024)
PropMEND: Hypernetworks for Knowledge Propagation in LLMs
von: Liu, Zeyu Leo, et al.
Veröffentlicht: (2025)
von: Liu, Zeyu Leo, et al.
Veröffentlicht: (2025)
Adaptive Margin RLHF via Preference over Preferences
von: Chittepu, Yaswanth, et al.
Veröffentlicht: (2025)
von: Chittepu, Yaswanth, et al.
Veröffentlicht: (2025)
Contrastive Learning to Improve Retrieval for Real-world Fact Checking
von: Sriram, Aniruddh, et al.
Veröffentlicht: (2024)
von: Sriram, Aniruddh, et al.
Veröffentlicht: (2024)
Learning Composable Chains-of-Thought
von: Yin, Fangcong, et al.
Veröffentlicht: (2025)
von: Yin, Fangcong, et al.
Veröffentlicht: (2025)
SkillFactory: Self-Distillation For Learning Cognitive Behaviors
von: Sprague, Zayne, et al.
Veröffentlicht: (2025)
von: Sprague, Zayne, et al.
Veröffentlicht: (2025)
Ghostbuster: Detecting Text Ghostwritten by Large Language Models
von: Verma, Vivek, et al.
Veröffentlicht: (2023)
von: Verma, Vivek, et al.
Veröffentlicht: (2023)
Decision-Oriented Dialogue for Human-AI Collaboration
von: Lin, Jessy, et al.
Veröffentlicht: (2023)
von: Lin, Jessy, et al.
Veröffentlicht: (2023)
CascadeDebate: Multi-Agent Deliberation for Cost-Aware LLM Cascades
von: Chang, Raeyoung, et al.
Veröffentlicht: (2026)
von: Chang, Raeyoung, et al.
Veröffentlicht: (2026)
Is the Top Still Spinning? Evaluating Subjectivity in Narrative Understanding
von: Subbiah, Melanie, et al.
Veröffentlicht: (2025)
von: Subbiah, Melanie, et al.
Veröffentlicht: (2025)
ProofWala: A Framework for Multilingual Proof Data Synthesis and Theorem-Proving
von: Thakur, Amitayush, et al.
Veröffentlicht: (2025)
von: Thakur, Amitayush, et al.
Veröffentlicht: (2025)
GroundAct: Can LLM Agents Ground Actions in Environmental States?
von: Wang, Zixuan, et al.
Veröffentlicht: (2025)
von: Wang, Zixuan, et al.
Veröffentlicht: (2025)
RankAlign: A Ranking View of the Generator-Validator Gap in Large Language Models
von: Rodriguez, Juan Diego, et al.
Veröffentlicht: (2025)
von: Rodriguez, Juan Diego, et al.
Veröffentlicht: (2025)
ARUQULA -- An LLM based Text2SPARQL Approach using ReAct and Knowledge Graph Exploration Utilities
von: Brei, Felix, et al.
Veröffentlicht: (2025)
von: Brei, Felix, et al.
Veröffentlicht: (2025)
Bayesian Orchestration of Multi-LLM Agents for Cost-Aware Sequential Decision-Making
von: Amin, Danial
Veröffentlicht: (2026)
von: Amin, Danial
Veröffentlicht: (2026)
Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing
von: Ding, Dujian, et al.
Veröffentlicht: (2024)
von: Ding, Dujian, et al.
Veröffentlicht: (2024)
Look Before You Leap: Autonomous Exploration for LLM Agents
von: Ye, Ziang, et al.
Veröffentlicht: (2026)
von: Ye, Ziang, et al.
Veröffentlicht: (2026)
Act-Adaptive Margin: Dynamically Calibrating Reward Models for Subjective Ambiguity
von: Fang, Feiteng, et al.
Veröffentlicht: (2025)
von: Fang, Feiteng, et al.
Veröffentlicht: (2025)
AgentHER: Hindsight Experience Replay for LLM Agent Trajectory Relabeling
von: Ding, Liang
Veröffentlicht: (2026)
von: Ding, Liang
Veröffentlicht: (2026)
EcoAct: Economic Agent Determines When to Register What Action
von: Zhang, Shaokun, et al.
Veröffentlicht: (2024)
von: Zhang, Shaokun, et al.
Veröffentlicht: (2024)
SPARTA ALIGNMENT: Collectively Aligning Multiple Language Models through Combat
von: Jiang, Yuru, et al.
Veröffentlicht: (2025)
von: Jiang, Yuru, et al.
Veröffentlicht: (2025)
SDE-SQL: Enhancing Text-to-SQL Generation in Large Language Models via Self-Driven Exploration with SQL Probes
von: Xie, Wenxuan, et al.
Veröffentlicht: (2025)
von: Xie, Wenxuan, et al.
Veröffentlicht: (2025)
ActMem: Bridging the Gap Between Memory Retrieval and Reasoning in LLM Agents
von: Zhang, Xiaohui, et al.
Veröffentlicht: (2026)
von: Zhang, Xiaohui, et al.
Veröffentlicht: (2026)
A Communication-Theoretic Framework for LLM Agents: Cost-Aware Adaptive Reliability
von: Omidvar, Hamed, et al.
Veröffentlicht: (2026)
von: Omidvar, Hamed, et al.
Veröffentlicht: (2026)
Polyrating: A Cost-Effective and Bias-Aware Rating System for LLM Evaluation
von: Dekoninck, Jasper, et al.
Veröffentlicht: (2024)
von: Dekoninck, Jasper, et al.
Veröffentlicht: (2024)
Code Execution as Grounded Supervision for LLM Reasoning
von: Jung, Dongwon, et al.
Veröffentlicht: (2025)
von: Jung, Dongwon, et al.
Veröffentlicht: (2025)
PreAct: Prediction Enhances Agent's Planning Ability
von: Fu, Dayuan, et al.
Veröffentlicht: (2024)
von: Fu, Dayuan, et al.
Veröffentlicht: (2024)
TREX: Automating LLM Fine-tuning via Agent-Driven Tree-based Exploration
von: Ma, Zerun, et al.
Veröffentlicht: (2026)
von: Ma, Zerun, et al.
Veröffentlicht: (2026)
StateAct: Enhancing LLM Base Agents via Self-prompting and State-tracking
von: Rozanov, Nikolai, et al.
Veröffentlicht: (2024)
von: Rozanov, Nikolai, et al.
Veröffentlicht: (2024)
Steering Evaluation-Aware Language Models to Act Like They Are Deployed
von: Hua, Tim Tian, et al.
Veröffentlicht: (2025)
von: Hua, Tim Tian, et al.
Veröffentlicht: (2025)
CostBench: Evaluating Multi-Turn Cost-Optimal Planning and Adaptation in Dynamic Environments for LLM Tool-Use Agents
von: Liu, Jiayu, et al.
Veröffentlicht: (2025)
von: Liu, Jiayu, et al.
Veröffentlicht: (2025)
MTRouter: Cost-Aware Multi-Turn LLM Routing with History-Model Joint Embeddings
von: Zhang, Yiqun, et al.
Veröffentlicht: (2026)
von: Zhang, Yiqun, et al.
Veröffentlicht: (2026)
Automated Survey Collection with LLM-based Conversational Agents
von: Kaiyrbekov, Kurmanbek, et al.
Veröffentlicht: (2025)
von: Kaiyrbekov, Kurmanbek, et al.
Veröffentlicht: (2025)
Trial and Error: Exploration-Based Trajectory Optimization for LLM Agents
von: Song, Yifan, et al.
Veröffentlicht: (2024)
von: Song, Yifan, et al.
Veröffentlicht: (2024)
AdaRubric: Task-Adaptive Rubrics for Reliable LLM Agent Evaluation and Reward Learning
von: Ding, Liang
Veröffentlicht: (2026)
von: Ding, Liang
Veröffentlicht: (2026)
PABU: Progress-Aware Belief Update for Efficient LLM Agents
von: Jiang, Haitao, et al.
Veröffentlicht: (2026)
von: Jiang, Haitao, et al.
Veröffentlicht: (2026)
R-Judge: Benchmarking Safety Risk Awareness for LLM Agents
von: Yuan, Tongxin, et al.
Veröffentlicht: (2024)
von: Yuan, Tongxin, et al.
Veröffentlicht: (2024)
Preference-Aware Memory Update for Long-Term LLM Agents
von: Sun, Haoran, et al.
Veröffentlicht: (2025)
von: Sun, Haoran, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Molecular Facts: Desiderata for Decontextualization in LLM Fact Verification
von: Gunjal, Anisha, et al.
Veröffentlicht: (2024) -
SynthesizRR: Generating Diverse Datasets with Retrieval Augmentation
von: Divekar, Abhishek, et al.
Veröffentlicht: (2024) -
MiniCheck: Efficient Fact-Checking of LLMs on Grounding Documents
von: Tang, Liyan, et al.
Veröffentlicht: (2024) -
PropMEND: Hypernetworks for Knowledge Propagation in LLMs
von: Liu, Zeyu Leo, et al.
Veröffentlicht: (2025) -
Adaptive Margin RLHF via Preference over Preferences
von: Chittepu, Yaswanth, et al.
Veröffentlicht: (2025)