Direct Advantage Regression: Aligning LLMs with Online AI Reward
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | He, Li, Zhao, He, Wan, Stephen, Wang, Dadong, Yao, Lina, Liu, Tongliang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Direct Language Model Alignment from Online AI Feedback
von: Guo, Shangmin, et al.
Veröffentlicht: (2024)
von: Guo, Shangmin, et al.
Veröffentlicht: (2024)
The Alignment Floor: How Persona Customization Breaks Safety in Weakly-Aligned LLMs
von: Zhang, Xing, et al.
Veröffentlicht: (2026)
von: Zhang, Xing, et al.
Veröffentlicht: (2026)
Aligning LLMs with Individual Preferences via Interaction
von: Wu, Shujin, et al.
Veröffentlicht: (2024)
von: Wu, Shujin, et al.
Veröffentlicht: (2024)
SARGes: Semantically Aligned Reliable Gesture Generation via Intent Chain
von: Gao, Nan, et al.
Veröffentlicht: (2025)
von: Gao, Nan, et al.
Veröffentlicht: (2025)
Effects of Theory of Mind and Prosocial Beliefs on Steering Human-Aligned Behaviors of LLMs in Ultimatum Games
von: Yadav, Neemesh, et al.
Veröffentlicht: (2025)
von: Yadav, Neemesh, et al.
Veröffentlicht: (2025)
Your Co-Workers Matter: Evaluating Collaborative Capabilities of Language Models in Blocks World
von: Wu, Guande, et al.
Veröffentlicht: (2024)
von: Wu, Guande, et al.
Veröffentlicht: (2024)
Are Today's LLMs Ready to Explain Well-Being Concepts?
von: Jiang, Bohan, et al.
Veröffentlicht: (2025)
von: Jiang, Bohan, et al.
Veröffentlicht: (2025)
Human-AI Interaction and User Satisfaction: Empirical Evidence from Online Reviews of AI Products
von: Pasch, Stefan, et al.
Veröffentlicht: (2025)
von: Pasch, Stefan, et al.
Veröffentlicht: (2025)
Strategic Chain-of-Thought: Guiding Accurate Reasoning in LLMs through Strategy Elicitation
von: Wang, Yu, et al.
Veröffentlicht: (2024)
von: Wang, Yu, et al.
Veröffentlicht: (2024)
Interaction Dynamics as a Reward Signal for LLMs
von: Gooding, Sian, et al.
Veröffentlicht: (2025)
von: Gooding, Sian, et al.
Veröffentlicht: (2025)
AI Conversational Interviewing: Transforming Surveys with LLMs as Adaptive Interviewers
von: Wuttke, Alexander, et al.
Veröffentlicht: (2024)
von: Wuttke, Alexander, et al.
Veröffentlicht: (2024)
CulturalTeaming: AI-Assisted Interactive Red-Teaming for Challenging LLMs' (Lack of) Multicultural Knowledge
von: Chiu, Yu Ying, et al.
Veröffentlicht: (2024)
von: Chiu, Yu Ying, et al.
Veröffentlicht: (2024)
Copiloting Diagnosis of Autism in Real Clinical Scenarios via LLMs
von: Jiang, Yi, et al.
Veröffentlicht: (2024)
von: Jiang, Yi, et al.
Veröffentlicht: (2024)
Inclusion Arena: An Open Platform for Evaluating Large Foundation Models with Real-World Apps
von: Wang, Kangyu, et al.
Veröffentlicht: (2025)
von: Wang, Kangyu, et al.
Veröffentlicht: (2025)
Human Empathy as Encoder: AI-Assisted Depression Assessment in Special Education
von: Zhao, Boning, et al.
Veröffentlicht: (2025)
von: Zhao, Boning, et al.
Veröffentlicht: (2025)
CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs
von: Liu, Hongtao, et al.
Veröffentlicht: (2025)
von: Liu, Hongtao, et al.
Veröffentlicht: (2025)
Talking to Machines: do you read me?
von: Rojas-Barahona, Lina M.
Veröffentlicht: (2024)
von: Rojas-Barahona, Lina M.
Veröffentlicht: (2024)
ConsistencyAI: A Benchmark to Assess LLMs' Factual Consistency When Responding to Different Demographic Groups
von: Banyas, Peter, et al.
Veröffentlicht: (2025)
von: Banyas, Peter, et al.
Veröffentlicht: (2025)
Olapa-MCoT: Enhancing the Chinese Mathematical Reasoning Capability of LLMs
von: Zhu, Shaojie, et al.
Veröffentlicht: (2023)
von: Zhu, Shaojie, et al.
Veröffentlicht: (2023)
HICode: Hierarchical Inductive Coding with LLMs
von: Zhong, Mian, et al.
Veröffentlicht: (2025)
von: Zhong, Mian, et al.
Veröffentlicht: (2025)
Analyzing Cognitive Differences Among Large Language Models through the Lens of Social Worldview
von: Li, Jiatao, et al.
Veröffentlicht: (2025)
von: Li, Jiatao, et al.
Veröffentlicht: (2025)
An AI-Powered Research Assistant in the Lab: A Practical Guide for Text Analysis Through Iterative Collaboration with LLMs
von: Carmona-Díaz, Gino, et al.
Veröffentlicht: (2025)
von: Carmona-Díaz, Gino, et al.
Veröffentlicht: (2025)
A Comprehensive Survey of Bias in LLMs: Current Landscape and Future Directions
von: Ranjan, Rajesh, et al.
Veröffentlicht: (2024)
von: Ranjan, Rajesh, et al.
Veröffentlicht: (2024)
Flowco: Rethinking Data Analysis in the Age of LLMs
von: Freund, Stephen N., et al.
Veröffentlicht: (2025)
von: Freund, Stephen N., et al.
Veröffentlicht: (2025)
From Control to Foresight: Simulation as a New Paradigm for Human-Agent Collaboration
von: He, Gaole, et al.
Veröffentlicht: (2026)
von: He, Gaole, et al.
Veröffentlicht: (2026)
LLMs for XAI: Future Directions for Explaining Explanations
von: Zytek, Alexandra, et al.
Veröffentlicht: (2024)
von: Zytek, Alexandra, et al.
Veröffentlicht: (2024)
EUDAIMONIA: Evaluating Undesirable Dynamics in AI
von: Huang, Jun Rui, et al.
Veröffentlicht: (2026)
von: Huang, Jun Rui, et al.
Veröffentlicht: (2026)
StressPrompt: Does Stress Impact Large Language Models and Human Performance Similarly?
von: Shen, Guobin, et al.
Veröffentlicht: (2024)
von: Shen, Guobin, et al.
Veröffentlicht: (2024)
SymbolicThought: Integrating Language Models and Symbolic Reasoning for Consistent and Interpretable Human Relationship Understanding
von: Zhao, Runcong, et al.
Veröffentlicht: (2025)
von: Zhao, Runcong, et al.
Veröffentlicht: (2025)
AutoLife: Automatic Life Journaling with Smartphones and LLMs
von: Xu, Huatao, et al.
Veröffentlicht: (2024)
von: Xu, Huatao, et al.
Veröffentlicht: (2024)
Awaking the Slides: A Tuning-free and Knowledge-regulated AI Tutoring System via Language Model Coordination
von: Zhang-Li, Daniel, et al.
Veröffentlicht: (2024)
von: Zhang-Li, Daniel, et al.
Veröffentlicht: (2024)
Performance Gains of LLMs With Humans in a World of LLMs Versus Humans
von: McCullum, Lucas, et al.
Veröffentlicht: (2025)
von: McCullum, Lucas, et al.
Veröffentlicht: (2025)
Model-in-the-Loop (MILO): Accelerating Multimodal AI Data Annotation with LLMs
von: Wang, Yifan, et al.
Veröffentlicht: (2024)
von: Wang, Yifan, et al.
Veröffentlicht: (2024)
Advancing Speech Language Models by Scaling Supervised Fine-Tuning with Over 60,000 Hours of Synthetic Speech Dialogue Data
von: Zhao, Shuaijiang, et al.
Veröffentlicht: (2024)
von: Zhao, Shuaijiang, et al.
Veröffentlicht: (2024)
Aligning Model Evaluations with Human Preferences: Mitigating Token Count Bias in Language Model Assessments
von: Daynauth, Roland, et al.
Veröffentlicht: (2024)
von: Daynauth, Roland, et al.
Veröffentlicht: (2024)
The Good, The Bad, and Why: Unveiling Emotions in Generative AI
von: Li, Cheng, et al.
Veröffentlicht: (2023)
von: Li, Cheng, et al.
Veröffentlicht: (2023)
From Voices to Validity: Leveraging Large Language Models (LLMs) for Textual Analysis of Policy Stakeholder Interviews
von: Liu, Alex, et al.
Veröffentlicht: (2023)
von: Liu, Alex, et al.
Veröffentlicht: (2023)
ProactiveEval: A Unified Evaluation Framework for Proactive Dialogue Agents
von: Liu, Tianjian, et al.
Veröffentlicht: (2025)
von: Liu, Tianjian, et al.
Veröffentlicht: (2025)
When Models Know More Than They Can Explain: Quantifying Knowledge Transfer in Human-AI Collaboration
von: Shi, Quan, et al.
Veröffentlicht: (2025)
von: Shi, Quan, et al.
Veröffentlicht: (2025)
User-Assistant Bias in LLMs
von: Pan, Xu, et al.
Veröffentlicht: (2025)
von: Pan, Xu, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Direct Language Model Alignment from Online AI Feedback
von: Guo, Shangmin, et al.
Veröffentlicht: (2024) -
The Alignment Floor: How Persona Customization Breaks Safety in Weakly-Aligned LLMs
von: Zhang, Xing, et al.
Veröffentlicht: (2026) -
Aligning LLMs with Individual Preferences via Interaction
von: Wu, Shujin, et al.
Veröffentlicht: (2024) -
SARGes: Semantically Aligned Reliable Gesture Generation via Intent Chain
von: Gao, Nan, et al.
Veröffentlicht: (2025) -
Effects of Theory of Mind and Prosocial Beliefs on Steering Human-Aligned Behaviors of LLMs in Ultimatum Games
von: Yadav, Neemesh, et al.
Veröffentlicht: (2025)