Gespeichert in:
| Hauptverfasser: | Lattimer, Barrett Martin, Gangal, Varun, McDonald, Ryan, Yang, Yi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2409.04617 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Enhancing Hallucination Detection through Perturbation-Based Synthetic Data Generation in System Responses
von: Zhang, Dongxu, et al.
Veröffentlicht: (2024)
von: Zhang, Dongxu, et al.
Veröffentlicht: (2024)
Fast and Accurate Factual Inconsistency Detection Over Long Documents
von: Lattimer, Barrett Martin, et al.
Veröffentlicht: (2023)
von: Lattimer, Barrett Martin, et al.
Veröffentlicht: (2023)
Human Inspired Progressive Alignment and Comparative Learning for Grounded Word Acquisition
von: Bao, Yuwei, et al.
Veröffentlicht: (2023)
von: Bao, Yuwei, et al.
Veröffentlicht: (2023)
Multi-Step Dialogue Workflow Action Prediction
von: Ramakrishnan, Ramya, et al.
Veröffentlicht: (2023)
von: Ramakrishnan, Ramya, et al.
Veröffentlicht: (2023)
MEMTRACK: Evaluating Long-Term Memory and State Tracking in Multi-Platform Dynamic Agent Environments
von: Deshpande, Darshan, et al.
Veröffentlicht: (2025)
von: Deshpande, Darshan, et al.
Veröffentlicht: (2025)
Can We Afford The Perfect Prompt? Balancing Cost and Accuracy with the Economical Prompting Index
von: McDonald, Tyler, et al.
Veröffentlicht: (2024)
von: McDonald, Tyler, et al.
Veröffentlicht: (2024)
Defenses & Enablers For Skill Injection Attacks on Terminal Based Agents
von: Fujinuma, Yoshinari, et al.
Veröffentlicht: (2026)
von: Fujinuma, Yoshinari, et al.
Veröffentlicht: (2026)
TRAIL: Trace Reasoning and Agentic Issue Localization
von: Deshpande, Darshan, et al.
Veröffentlicht: (2025)
von: Deshpande, Darshan, et al.
Veröffentlicht: (2025)
Trace-of-Thought Prompting: Investigating Prompt-Based Knowledge Distillation Through Question Decomposition
von: McDonald, Tyler, et al.
Veröffentlicht: (2025)
von: McDonald, Tyler, et al.
Veröffentlicht: (2025)
ARIA: Training Language Agents with Intention-Driven Reward Aggregation
von: Yang, Ruihan, et al.
Veröffentlicht: (2025)
von: Yang, Ruihan, et al.
Veröffentlicht: (2025)
Aligning Dialogue Agents with Global Feedback via Large Language Model Multimodal Reward Decomposition
von: Lee, Dong Won, et al.
Veröffentlicht: (2025)
von: Lee, Dong Won, et al.
Veröffentlicht: (2025)
Retrieval, Reward, and Training Protocols: What Matters in Training Search Agents?
von: Zhao, Yibo, et al.
Veröffentlicht: (2026)
von: Zhao, Yibo, et al.
Veröffentlicht: (2026)
Challenges and opportunities in portraying emotion in generated sign language
von: McDonald, John C., et al.
Veröffentlicht: (2025)
von: McDonald, John C., et al.
Veröffentlicht: (2025)
Bootstrapping LLM-based Task-Oriented Dialogue Agents via Self-Talk
von: Ulmer, Dennis, et al.
Veröffentlicht: (2024)
von: Ulmer, Dennis, et al.
Veröffentlicht: (2024)
HalluWorld: A Controlled Benchmark for Hallucination via Reference World Models
von: Liu, Emmy, et al.
Veröffentlicht: (2026)
von: Liu, Emmy, et al.
Veröffentlicht: (2026)
To Memorize or to Retrieve: Scaling Laws for RAG-Considerate Pretraining
von: Singh, Karan, et al.
Veröffentlicht: (2026)
von: Singh, Karan, et al.
Veröffentlicht: (2026)
Zero-shot and Few-shot Generation Strategies for Artificial Clinical Records
von: Frayling, Erlend, et al.
Veröffentlicht: (2024)
von: Frayling, Erlend, et al.
Veröffentlicht: (2024)
SEAD: Self-Evolving Agent for Multi-Turn Service Dialogue
von: Dai, Yuqin, et al.
Veröffentlicht: (2026)
von: Dai, Yuqin, et al.
Veröffentlicht: (2026)
From Self-Evolving Synthetic Data to Verifiable-Reward RL: Post-Training Multi-turn Interactive Tool-Using Agents
von: Gao, Jiaxuan, et al.
Veröffentlicht: (2026)
von: Gao, Jiaxuan, et al.
Veröffentlicht: (2026)
GRAM-R$^2$: Self-Training Generative Foundation Reward Models for Reward Reasoning
von: Wang, Chenglong, et al.
Veröffentlicht: (2025)
von: Wang, Chenglong, et al.
Veröffentlicht: (2025)
State Value Generation with Prompt Learning and Self-Training for Low-Resource Dialogue State Tracking
von: Gu, Ming, et al.
Veröffentlicht: (2024)
von: Gu, Ming, et al.
Veröffentlicht: (2024)
Re-ReST: Reflection-Reinforced Self-Training for Language Agents
von: Dou, Zi-Yi, et al.
Veröffentlicht: (2024)
von: Dou, Zi-Yi, et al.
Veröffentlicht: (2024)
A Comparative Benchmark of Large Language Models for Labelling Wind Turbine Maintenance Logs
von: Malyi, Max, et al.
Veröffentlicht: (2025)
von: Malyi, Max, et al.
Veröffentlicht: (2025)
ReSeek: A Self-Correcting Framework for Search Agents with Instructive Rewards
von: Li, Shiyu, et al.
Veröffentlicht: (2025)
von: Li, Shiyu, et al.
Veröffentlicht: (2025)
Cohesive Conversations: Enhancing Authenticity in Multi-Agent Simulated Dialogues
von: Chu, KuanChao, et al.
Veröffentlicht: (2024)
von: Chu, KuanChao, et al.
Veröffentlicht: (2024)
STOP! Benchmarking Large Language Models with Sensitivity Testing on Offensive Progressions
von: Morabito, Robert, et al.
Veröffentlicht: (2024)
von: Morabito, Robert, et al.
Veröffentlicht: (2024)
Why Synthetic Isn't Real Yet: A Diagnostic Framework for Contact Center Dialogue Generation
von: Devanathan, Rishikesh, et al.
Veröffentlicht: (2025)
von: Devanathan, Rishikesh, et al.
Veröffentlicht: (2025)
Enhancing Personalized Multi-Turn Dialogue with Curiosity Reward
von: Wan, Yanming, et al.
Veröffentlicht: (2025)
von: Wan, Yanming, et al.
Veröffentlicht: (2025)
Training Dialogue Systems by AI Feedback for Improving Overall Dialogue Impression
von: Yoshida, Kai, et al.
Veröffentlicht: (2025)
von: Yoshida, Kai, et al.
Veröffentlicht: (2025)
Agent-RLVR: Training Software Engineering Agents via Guidance and Environment Rewards
von: Da, Jeff, et al.
Veröffentlicht: (2025)
von: Da, Jeff, et al.
Veröffentlicht: (2025)
NYT-Connections: A Deceptively Simple Text Classification Task that Stumps System-1 Thinkers
von: Lopez, Angel Yahir Loredo, et al.
Veröffentlicht: (2024)
von: Lopez, Angel Yahir Loredo, et al.
Veröffentlicht: (2024)
Temporal Fact Conflicts in LLMs: Reproducibility Insights from Unifying DYNAMICQA and MULAN
von: Dey, Ritajit, et al.
Veröffentlicht: (2026)
von: Dey, Ritajit, et al.
Veröffentlicht: (2026)
RRM: Robust Reward Model Training Mitigates Reward Hacking
von: Liu, Tianqi, et al.
Veröffentlicht: (2024)
von: Liu, Tianqi, et al.
Veröffentlicht: (2024)
Educational-Psychological Dialogue Robot Based on Multi-Agent Collaboration
von: Ni, Shiwen, et al.
Veröffentlicht: (2024)
von: Ni, Shiwen, et al.
Veröffentlicht: (2024)
SDiaReward: Modeling and Benchmarking Spoken Dialogue Rewards with Modality and Colloquialness
von: Lu, Jingyu, et al.
Veröffentlicht: (2026)
von: Lu, Jingyu, et al.
Veröffentlicht: (2026)
Sparse Reward Subsystem in Large Language Models
von: Xu, Guowei, et al.
Veröffentlicht: (2026)
von: Xu, Guowei, et al.
Veröffentlicht: (2026)
Can LLMs Understand the Implication of Emphasized Sentences in Dialogue?
von: Lin, Guan-Ting, et al.
Veröffentlicht: (2024)
von: Lin, Guan-Ting, et al.
Veröffentlicht: (2024)
MMWOZ: Building Multimodal Agent for Task-oriented Dialogue
von: Yang, Pu-Hai, et al.
Veröffentlicht: (2025)
von: Yang, Pu-Hai, et al.
Veröffentlicht: (2025)
R1-Reward: Training Multimodal Reward Model Through Stable Reinforcement Learning
von: Zhang, Yi-Fan, et al.
Veröffentlicht: (2025)
von: Zhang, Yi-Fan, et al.
Veröffentlicht: (2025)
SE-Search: Self-Evolving Search Agent via Memory and Dense Reward
von: Li, Jian, et al.
Veröffentlicht: (2026)
von: Li, Jian, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Enhancing Hallucination Detection through Perturbation-Based Synthetic Data Generation in System Responses
von: Zhang, Dongxu, et al.
Veröffentlicht: (2024) -
Fast and Accurate Factual Inconsistency Detection Over Long Documents
von: Lattimer, Barrett Martin, et al.
Veröffentlicht: (2023) -
Human Inspired Progressive Alignment and Comparative Learning for Grounded Word Acquisition
von: Bao, Yuwei, et al.
Veröffentlicht: (2023) -
Multi-Step Dialogue Workflow Action Prediction
von: Ramakrishnan, Ramya, et al.
Veröffentlicht: (2023) -
MEMTRACK: Evaluating Long-Term Memory and State Tracking in Multi-Platform Dynamic Agent Environments
von: Deshpande, Darshan, et al.
Veröffentlicht: (2025)