Improving Human-AI Coordination through Online Adversarial Training and Generative Models
Fuente:
arXiv
Saved in:
| Main Authors: | Chaudhary, Paresh, Liang, Yancheng, Chen, Daphne, Du, Simon S., Jaques, Natasha |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Learning to Cooperate with Humans using Generative Agents
by: Liang, Yancheng, et al.
Published: (2024)
by: Liang, Yancheng, et al.
Published: (2024)
Cross-environment Cooperation Enables Zero-shot Multi-agent Coordination
by: Jha, Kunal, et al.
Published: (2025)
by: Jha, Kunal, et al.
Published: (2025)
Maximizing Mutual Information Between Prompt and Response Improves LLM Performance With No Additional Data
by: Nam, Hyunji, et al.
Published: (2026)
by: Nam, Hyunji, et al.
Published: (2026)
Generative Modeling for Robust Deep Reinforcement Learning on the Traveling Salesman Problem
by: Li, Michael, et al.
Published: (2025)
by: Li, Michael, et al.
Published: (2025)
ReaLJam: Real-Time Human-AI Music Jamming with Reinforcement Learning-Tuned Transformers
by: Scarlatos, Alexander, et al.
Published: (2025)
by: Scarlatos, Alexander, et al.
Published: (2025)
Infer Human's Intentions Before Following Natural Language Instructions
by: Wan, Yanming, et al.
Published: (2024)
by: Wan, Yanming, et al.
Published: (2024)
Consistently Simulating Human Personas with Multi-Turn Reinforcement Learning
by: Abdulhai, Marwa, et al.
Published: (2025)
by: Abdulhai, Marwa, et al.
Published: (2025)
Personalizing Reinforcement Learning from Human Feedback with Variational Preference Learning
by: Poddar, Sriyash, et al.
Published: (2024)
by: Poddar, Sriyash, et al.
Published: (2024)
Improving Interactive In-Context Learning from Natural Language Feedback
by: Klissarov, Martin, et al.
Published: (2026)
by: Klissarov, Martin, et al.
Published: (2026)
AgenticRed: Evolving Agentic Systems for Red-Teaming
by: Yuan, Jiayi, et al.
Published: (2026)
by: Yuan, Jiayi, et al.
Published: (2026)
Beyond Cooperative Simulators: Generating Realistic User Personas for Robust Evaluation of LLM Agents
by: Chopra, Harshita, et al.
Published: (2026)
by: Chopra, Harshita, et al.
Published: (2026)
Modeling Others' Minds as Code
by: Jha, Kunal, et al.
Published: (2025)
by: Jha, Kunal, et al.
Published: (2025)
Impossibility Theorems for Feature Attribution
by: Bilodeau, Blair, et al.
Published: (2022)
by: Bilodeau, Blair, et al.
Published: (2022)
Human-compatible driving partners through data-regularized self-play reinforcement learning
by: Cornelisse, Daphne, et al.
Published: (2024)
by: Cornelisse, Daphne, et al.
Published: (2024)
Annealing Self-Distillation Rectification Improves Adversarial Training
by: Wu, Yu-Yu, et al.
Published: (2023)
by: Wu, Yu-Yu, et al.
Published: (2023)
PRM-Free Security Alignment of Large Models via Red Teaming and Adversarial Training
by: Du, Pengfei
Published: (2025)
by: Du, Pengfei
Published: (2025)
Enhancing Personalized Multi-Turn Dialogue with Curiosity Reward
by: Wan, Yanming, et al.
Published: (2025)
by: Wan, Yanming, et al.
Published: (2025)
Tackling Cooperative Incompatibility for Zero-Shot Human-AI Coordination
by: Li, Yang, et al.
Published: (2023)
by: Li, Yang, et al.
Published: (2023)
Chasing Moving Targets with Online Self-Play Reinforcement Learning for Safer Language Models
by: Liu, Mickel, et al.
Published: (2025)
by: Liu, Mickel, et al.
Published: (2025)
Triadic-OCD: Asynchronous Online Change Detection with Provable Robustness, Optimality, and Convergence
by: Huang, Yancheng, et al.
Published: (2024)
by: Huang, Yancheng, et al.
Published: (2024)
General Information Metrics for Improving AI Model Training Efficiency
by: Xu, Jianfeng, et al.
Published: (2025)
by: Xu, Jianfeng, et al.
Published: (2025)
Online Handbook of Argumentation for AI: Volume 4
by: Bengel, Lars, et al.
Published: (2023)
by: Bengel, Lars, et al.
Published: (2023)
Automatic Curriculum Design for Zero-Shot Human-AI Coordination
by: You, Won-Sang, et al.
Published: (2025)
by: You, Won-Sang, et al.
Published: (2025)
Adaptive Human-AI Coordination via Hierarchical Action Disentanglement
by: Ahmad, Adnan, et al.
Published: (2026)
by: Ahmad, Adnan, et al.
Published: (2026)
Evaluating & Reducing Deceptive Dialogue From Language Models with Multi-turn RL
by: Abdulhai, Marwa, et al.
Published: (2025)
by: Abdulhai, Marwa, et al.
Published: (2025)
RoS-Guard: Robust and Scalable Online Change Detection with Delay-Optimal Guarantees
by: Zhu, Zelin, et al.
Published: (2025)
by: Zhu, Zelin, et al.
Published: (2025)
Improved Training Mechanism for Reinforcement Learning via Online Model Selection
by: Afshar, Aida, et al.
Published: (2025)
by: Afshar, Aida, et al.
Published: (2025)
A Quantum-Inspired Analysis of Human Disambiguation Processes
by: Wang, Daphne
Published: (2024)
by: Wang, Daphne
Published: (2024)
Learning to summarize user information for personalized reinforcement learning from human feedback
by: Nam, Hyunji, et al.
Published: (2025)
by: Nam, Hyunji, et al.
Published: (2025)
Cyborg Data: Merging Human with AI Generated Training Data
by: North, Kai, et al.
Published: (2025)
by: North, Kai, et al.
Published: (2025)
Explanation-Guided Adversarial Training for Robust and Interpretable Models
by: Chen, Chao, et al.
Published: (2026)
by: Chen, Chao, et al.
Published: (2026)
Model Spec Midtraining: Improving How Alignment Training Generalizes
by: Li, Chloe, et al.
Published: (2026)
by: Li, Chloe, et al.
Published: (2026)
Training Stratigraphy: Persistent Behavioral Artifacts in Large Language Models Observed Through Longitudinal AI-Human Interaction
by: Claude, Chen Ying, et al.
Published: (2026)
by: Claude, Chen Ying, et al.
Published: (2026)
When AI Persuades: Adversarial Explanation Attacks on Human Trust in AI-Assisted Decision Making
by: Fan, Shutong, et al.
Published: (2026)
by: Fan, Shutong, et al.
Published: (2026)
Band Together: Untargeted Adversarial Training with Multimodal Coordination against Evasion-based Promotion Attacks
by: Xian, Guanmeng, et al.
Published: (2026)
by: Xian, Guanmeng, et al.
Published: (2026)
MuSC: Improving Complex Instruction Following with Multi-granularity Self-Contrastive Training
by: Huang, Hui, et al.
Published: (2025)
by: Huang, Hui, et al.
Published: (2025)
MESA: Cooperative Meta-Exploration in Multi-Agent Learning through Exploiting State-Action Space Structure
by: Zhang, Zhicheng, et al.
Published: (2024)
by: Zhang, Zhicheng, et al.
Published: (2024)
Human Alignment of Large Language Models through Online Preference Optimisation
by: Calandriello, Daniele, et al.
Published: (2024)
by: Calandriello, Daniele, et al.
Published: (2024)
Chasing Random: Instruction Selection Strategies Fail to Generalize
by: Diddee, Harshita, et al.
Published: (2024)
by: Diddee, Harshita, et al.
Published: (2024)
Governing What You Cannot Observe: Adaptive Runtime Governance for Autonomous AI Agents
by: Marin, German, et al.
Published: (2026)
by: Marin, German, et al.
Published: (2026)
Similar Items
-
Learning to Cooperate with Humans using Generative Agents
by: Liang, Yancheng, et al.
Published: (2024) -
Cross-environment Cooperation Enables Zero-shot Multi-agent Coordination
by: Jha, Kunal, et al.
Published: (2025) -
Maximizing Mutual Information Between Prompt and Response Improves LLM Performance With No Additional Data
by: Nam, Hyunji, et al.
Published: (2026) -
Generative Modeling for Robust Deep Reinforcement Learning on the Traveling Salesman Problem
by: Li, Michael, et al.
Published: (2025) -
ReaLJam: Real-Time Human-AI Music Jamming with Reinforcement Learning-Tuned Transformers
by: Scarlatos, Alexander, et al.
Published: (2025)