Direct Language Model Alignment from Online AI Feedback
Fuente:
arXiv
Saved in:
| Main Authors: | Guo, Shangmin, Zhang, Biao, Liu, Tianlin, Liu, Tianqi, Khalman, Misha, Llinares, Felipe, Rame, Alexandre, Mesnard, Thomas, Zhao, Yao, Piot, Bilal, Ferret, Johan, Blondel, Mathieu |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
On Teacher Hacking in Language Model Distillation
by: Tiapkin, Daniil, et al.
Published: (2025)
by: Tiapkin, Daniil, et al.
Published: (2025)
Decoding-time Realignment of Language Models
by: Liu, Tianlin, et al.
Published: (2024)
by: Liu, Tianlin, et al.
Published: (2024)
Exploring AI-assisted Ideation and Prototyping for Choreography
by: Liu, Yimeng, et al.
Published: (2024)
by: Liu, Yimeng, et al.
Published: (2024)
DanceGen: Supporting Choreography Ideation and Prototyping with Generative AI
by: Liu, Yimeng, et al.
Published: (2024)
by: Liu, Yimeng, et al.
Published: (2024)
TaskLens: Generating Task-Conditioned Scaffolded Interfaces for Learning Professional Creative Software
by: Liu, Yimeng, et al.
Published: (2025)
by: Liu, Yimeng, et al.
Published: (2025)
Designing Scaffolded Interfaces for Enhanced Learning and Performance in Professional Software
by: Liu, Yimeng, et al.
Published: (2025)
by: Liu, Yimeng, et al.
Published: (2025)
Embedded vs. Situated: An Evaluation of AR Facial Training Feedback
by: Nargund, Avinash Ajit, et al.
Published: (2026)
by: Nargund, Avinash Ajit, et al.
Published: (2026)
CrowdGenUI: Aligning LLM-Based UI Generation with Crowdsourced User Preferences
by: Liu, Yimeng, et al.
Published: (2024)
by: Liu, Yimeng, et al.
Published: (2024)
TR-LLM: Integrating Trajectory Data for Scene-Aware LLM-Based Human Action Prediction
by: Takeyama, Kojiro, et al.
Published: (2024)
by: Takeyama, Kojiro, et al.
Published: (2024)
AlignUI: A Method for Designing LLM-Generated UIs Aligned with User Preferences
by: Liu, Yimeng, et al.
Published: (2026)
by: Liu, Yimeng, et al.
Published: (2026)
Direct Advantage Regression: Aligning LLMs with Online AI Reward
by: He, Li, et al.
Published: (2025)
by: He, Li, et al.
Published: (2025)
Statistical Rejection Sampling Improves Preference Optimization
by: Liu, Tianqi, et al.
Published: (2023)
by: Liu, Tianqi, et al.
Published: (2023)
ARLang: An Outdoor Augmented Reality Application for Portuguese Vocabulary Learning
by: Caetano, Arthur, et al.
Published: (2024)
by: Caetano, Arthur, et al.
Published: (2024)
ARfy: A Pipeline for Adapting 3D Scenes to Augmented Reality
by: Caetano, Arthur, et al.
Published: (2024)
by: Caetano, Arthur, et al.
Published: (2024)
The Language of Approval: Identifying the Drivers of Positive Feedback Online
by: Goyal, Agam, et al.
Published: (2025)
by: Goyal, Agam, et al.
Published: (2025)
ReUseIt: Synthesizing Reusable AI Agent Workflows for Web Automation
by: Liu, Yimeng, et al.
Published: (2025)
by: Liu, Yimeng, et al.
Published: (2025)
LocoVR: Multiuser Indoor Locomotion Dataset in Virtual Reality
by: Takeyama, Kojiro, et al.
Published: (2024)
by: Takeyama, Kojiro, et al.
Published: (2024)
Feedback to the European Data Protection Board's Guidelines 2/2023 on Technical Scope of Art. 5(3) of ePrivacy Directive
by: Santos, Cristiana, et al.
Published: (2024)
by: Santos, Cristiana, et al.
Published: (2024)
Constrained Online Recursive Source Separation Framework for Real-time Electrophysiological Signal Processing
by: Li, Yao, et al.
Published: (2024)
by: Li, Yao, et al.
Published: (2024)
Exploring and Analyzing the Effect of Avatar's Visual Style on Anxiety of English as Second Language (ESL) Speakers
by: Liu, Tianqi, et al.
Published: (2023)
by: Liu, Tianqi, et al.
Published: (2023)
Routers in Vision Mixture of Experts: An Empirical Study
by: Liu, Tianlin, et al.
Published: (2024)
by: Liu, Tianlin, et al.
Published: (2024)
Directional Alignment and Narrative Agency in Human-LLM Co-Writing
by: Fundal, Halfdan Nordahl, et al.
Published: (2026)
by: Fundal, Halfdan Nordahl, et al.
Published: (2026)
Dynamic Personalization Through Continuous Feedback Loops in Interactive AI Systems
by: He, Liu
Published: (2026)
by: He, Liu
Published: (2026)
Virtual Steps: The Experience of Walking for a Lifelong Wheelchair User in Virtual Reality
by: Taheri, Atieh, et al.
Published: (2024)
by: Taheri, Atieh, et al.
Published: (2024)
Feeds Don't Tell the Whole Story: Measuring Online-Offline Emotion Alignment
by: Elahimanesh, Sina, et al.
Published: (2026)
by: Elahimanesh, Sina, et al.
Published: (2026)
Learning Spatial Awareness for Laparoscopic Surgery with AI Assisted Visual Feedback
by: Liu, Songyang, et al.
Published: (2025)
by: Liu, Songyang, et al.
Published: (2025)
GraV: Grasp Volume Data for the Design of One-Handed XR Interfaces
by: Aponte, Alejandro, et al.
Published: (2024)
by: Aponte, Alejandro, et al.
Published: (2024)
Semantic Direct Modeling
by: Zou, Qiang, et al.
Published: (2025)
by: Zou, Qiang, et al.
Published: (2025)
DECAN: A Denoising Encoder via Contrastive Alignment Network for Dry Electrode EEG Emotion Recognition
by: Zhang, Meihong, et al.
Published: (2024)
by: Zhang, Meihong, et al.
Published: (2024)
MYCloth: Towards Intelligent and Interactive Online T-Shirt Customization based on User's Preference
by: Liu, Yexin, et al.
Published: (2024)
by: Liu, Yexin, et al.
Published: (2024)
AutoLegend: A User Feedback-Driven Adaptive Legend Generator for Visualizations
by: Liu, Can, et al.
Published: (2024)
by: Liu, Can, et al.
Published: (2024)
Virtual Buddy: Redefining Conversational AI Interactions for Individuals with Hand Motor Disabilities
by: Taheri, Atieh, et al.
Published: (2024)
by: Taheri, Atieh, et al.
Published: (2024)
Exploration of Radar-based Obstacle Visualizations to Support Safety and Presence in Camera-Free Outdoor VR
by: Nargund, Avinash Ajit, et al.
Published: (2026)
by: Nargund, Avinash Ajit, et al.
Published: (2026)
Audience in the Loop: Viewer Feedback-Driven Content Creation in Micro-drama Production on Social Media
by: Cao, Gengchen, et al.
Published: (2026)
by: Cao, Gengchen, et al.
Published: (2026)
Sketch Then Generate: Providing Incremental User Feedback and Guiding LLM Code Generation through Language-Oriented Code Sketches
by: Zhu-Tian, Chen, et al.
Published: (2024)
by: Zhu-Tian, Chen, et al.
Published: (2024)
A Multi-Label EEG Dataset for Mental Attention State Classification in Online Learning
by: Liu, Huan, et al.
Published: (2024)
by: Liu, Huan, et al.
Published: (2024)
Engaging with AI: How Interface Design Shapes Human-AI Collaboration in High-Stakes Decision-Making
by: Chen, Zichen, et al.
Published: (2025)
by: Chen, Zichen, et al.
Published: (2025)
An Interaction Design Toolkit for Physical Task Guidance with Artificial Intelligence and Mixed Reality
by: Caetano, Arthur, et al.
Published: (2024)
by: Caetano, Arthur, et al.
Published: (2024)
Alignment Drift in Long-Term Human-LLM Interaction: A Mechanism-Oriented Framework
by: Yao, Xintong
Published: (2026)
by: Yao, Xintong
Published: (2026)
Modeling the Impact of Visual Stimuli on Redirection Noticeability with Gaze Behavior in Virtual Reality
by: Li, Zhipeng, et al.
Published: (2025)
by: Li, Zhipeng, et al.
Published: (2025)
Similar Items
-
On Teacher Hacking in Language Model Distillation
by: Tiapkin, Daniil, et al.
Published: (2025) -
Decoding-time Realignment of Language Models
by: Liu, Tianlin, et al.
Published: (2024) -
Exploring AI-assisted Ideation and Prototyping for Choreography
by: Liu, Yimeng, et al.
Published: (2024) -
DanceGen: Supporting Choreography Ideation and Prototyping with Generative AI
by: Liu, Yimeng, et al.
Published: (2024) -
TaskLens: Generating Task-Conditioned Scaffolded Interfaces for Learning Professional Creative Software
by: Liu, Yimeng, et al.
Published: (2025)