Aligning LLM agents with human learning and adjustment behavior: a dual agent approach
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Liu, Tianming, Yang, Jirong, Yin, Yafeng, Li, Manzi, Wang, Linghao, Zhu, Zheng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Aligning LLM with human travel choices: a persona-based embedding learning approach
von: Liu, Tianming, et al.
Veröffentlicht: (2025)
von: Liu, Tianming, et al.
Veröffentlicht: (2025)
Toward LLM-Agent-Based Modeling of Transportation Systems: A Conceptual Framework
von: Liu, Tianming, et al.
Veröffentlicht: (2024)
von: Liu, Tianming, et al.
Veröffentlicht: (2024)
AIonopedia: an LLM agent orchestrating multimodal learning for ionic liquid discovery
von: Yin, Yuqi, et al.
Veröffentlicht: (2025)
von: Yin, Yuqi, et al.
Veröffentlicht: (2025)
The impact of behavioral diversity in multi-agent reinforcement learning
von: Bettini, Matteo, et al.
Veröffentlicht: (2024)
von: Bettini, Matteo, et al.
Veröffentlicht: (2024)
Understanding the planning of LLM agents: A survey
von: Huang, Xu, et al.
Veröffentlicht: (2024)
von: Huang, Xu, et al.
Veröffentlicht: (2024)
Robust agents learn causal world models
von: Richens, Jonathan, et al.
Veröffentlicht: (2024)
von: Richens, Jonathan, et al.
Veröffentlicht: (2024)
Ethics2vec: aligning automatic agents and human preferences
von: Bontempi, Gianluca
Veröffentlicht: (2025)
von: Bontempi, Gianluca
Veröffentlicht: (2025)
Single-agent or Multi-agent Systems? Why Not Both?
von: Gao, Mingyan, et al.
Veröffentlicht: (2025)
von: Gao, Mingyan, et al.
Veröffentlicht: (2025)
ACEGEN: Reinforcement learning of generative chemical agents for drug discovery
von: Bou, Albert, et al.
Veröffentlicht: (2024)
von: Bou, Albert, et al.
Veröffentlicht: (2024)
MADiff: Offline Multi-agent Learning with Diffusion Models
von: Zhu, Zhengbang, et al.
Veröffentlicht: (2023)
von: Zhu, Zhengbang, et al.
Veröffentlicht: (2023)
The challenge of hidden gifts in multi-agent reinforcement learning
von: Malenfant, Dane, et al.
Veröffentlicht: (2025)
von: Malenfant, Dane, et al.
Veröffentlicht: (2025)
A representational framework for learning and encoding structurally enriched trajectories in complex agent environments
von: Catarau-Cotutiu, Corina, et al.
Veröffentlicht: (2025)
von: Catarau-Cotutiu, Corina, et al.
Veröffentlicht: (2025)
Breaking $\textit{Winner-Takes-All}$: Cooperative Policy Optimization Improves Diverse LLM Reasoning
von: Chen, Haoxuan, et al.
Veröffentlicht: (2026)
von: Chen, Haoxuan, et al.
Veröffentlicht: (2026)
AOAD-MAT: Transformer-based multi-agent deep reinforcement learning model considering agents' order of action decisions
von: Takayama, Shota, et al.
Veröffentlicht: (2025)
von: Takayama, Shota, et al.
Veröffentlicht: (2025)
RE-Bench: Evaluating frontier AI R&D capabilities of language model agents against human experts
von: Wijk, Hjalmar, et al.
Veröffentlicht: (2024)
von: Wijk, Hjalmar, et al.
Veröffentlicht: (2024)
Time Series Forecasting via Direct Per-Step Probability Distribution Modeling
von: Kong, Linghao, et al.
Veröffentlicht: (2025)
von: Kong, Linghao, et al.
Veröffentlicht: (2025)
Select to Perfect: Imitating desired behavior from large multi-agent data
von: Franzmeyer, Tim, et al.
Veröffentlicht: (2024)
von: Franzmeyer, Tim, et al.
Veröffentlicht: (2024)
SAC-GLAM: Improving Online RL for LLM agents with Soft Actor-Critic and Hindsight Relabeling
von: Gaven, Loris, et al.
Veröffentlicht: (2024)
von: Gaven, Loris, et al.
Veröffentlicht: (2024)
LLM-ABM for Transportation: Assessing the Potential of LLM Agents in System Analysis
von: Liu, Tianming, et al.
Veröffentlicht: (2025)
von: Liu, Tianming, et al.
Veröffentlicht: (2025)
Taxon: Hierarchical Tax Code Prediction with Semantically Aligned LLM Expert Guidance
von: Li, Jihang, et al.
Veröffentlicht: (2026)
von: Li, Jihang, et al.
Veröffentlicht: (2026)
ELLMob: Event-Driven Human Mobility Generation with Self-Aligned LLM Framework
von: Wang, Yusong, et al.
Veröffentlicht: (2026)
von: Wang, Yusong, et al.
Veröffentlicht: (2026)
HINTBench: Horizon-agent Intrinsic Non-attack Trajectory Benchmark
von: Wang, Jiacheng, et al.
Veröffentlicht: (2026)
von: Wang, Jiacheng, et al.
Veröffentlicht: (2026)
Efficient Multi-agent Reinforcement Learning by Planning
von: Liu, Qihan, et al.
Veröffentlicht: (2024)
von: Liu, Qihan, et al.
Veröffentlicht: (2024)
LearnAlign: Data Selection for LLM Reinforcement Learning with Improved Gradient Alignment
von: Li, Shipeng, et al.
Veröffentlicht: (2025)
von: Li, Shipeng, et al.
Veröffentlicht: (2025)
STAS: Spatial-Temporal Return Decomposition for Multi-agent Reinforcement Learning
von: Chen, Sirui, et al.
Veröffentlicht: (2023)
von: Chen, Sirui, et al.
Veröffentlicht: (2023)
Position: agentic AI orchestration should be Bayes-consistent
von: Papamarkou, Theodore, et al.
Veröffentlicht: (2026)
von: Papamarkou, Theodore, et al.
Veröffentlicht: (2026)
OKG-LLM: Aligning Ocean Knowledge Graph with Observation Data via LLMs for Global Sea Surface Temperature Prediction
von: Yang, Hanchen, et al.
Veröffentlicht: (2025)
von: Yang, Hanchen, et al.
Veröffentlicht: (2025)
Engineering LLM Powered Multi-agent Framework for Autonomous CloudOps
von: Parthasarathy, Kannan, et al.
Veröffentlicht: (2025)
von: Parthasarathy, Kannan, et al.
Veröffentlicht: (2025)
MAFA: A multi-agent framework for annotation
von: Hegazy, Mahmood, et al.
Veröffentlicht: (2025)
von: Hegazy, Mahmood, et al.
Veröffentlicht: (2025)
Variational Offline Multi-agent Skill Discovery
von: Chen, Jiayu, et al.
Veröffentlicht: (2024)
von: Chen, Jiayu, et al.
Veröffentlicht: (2024)
Cooperative Multi-agent RL with Communication Constraints
von: Xiong, Nuoya, et al.
Veröffentlicht: (2026)
von: Xiong, Nuoya, et al.
Veröffentlicht: (2026)
A modular framework for automated evaluation of procedural content generation in serious games with deep reinforcement learning agents
von: Kalafatis, Eleftherios, et al.
Veröffentlicht: (2025)
von: Kalafatis, Eleftherios, et al.
Veröffentlicht: (2025)
Concept frustration: Aligning human concepts and machine representations
von: Parisini, Enrico, et al.
Veröffentlicht: (2026)
von: Parisini, Enrico, et al.
Veröffentlicht: (2026)
Neuromorphic dreaming: A pathway to efficient learning in artificial agents
von: Blakowski, Ingo, et al.
Veröffentlicht: (2024)
von: Blakowski, Ingo, et al.
Veröffentlicht: (2024)
The cognitive companion: a lightweight parallel monitoring architecture for detecting and recovering from reasoning degradation in LLM agents
von: Khan, Rafflesia, et al.
Veröffentlicht: (2026)
von: Khan, Rafflesia, et al.
Veröffentlicht: (2026)
HELENE: Hessian Layer-wise Clipping and Gradient Annealing for Accelerating Fine-tuning LLM with Zeroth-order Optimization
von: Zhao, Huaqin, et al.
Veröffentlicht: (2024)
von: Zhao, Huaqin, et al.
Veröffentlicht: (2024)
Data-driven inventory management for new products: An adjusted Dyna-$Q$ approach with transfer learning
von: Qu, Xinye, et al.
Veröffentlicht: (2025)
von: Qu, Xinye, et al.
Veröffentlicht: (2025)
GradAlign: Gradient-Aligned Data Selection for LLM Reinforcement Learning
von: Yang, Ningyuan, et al.
Veröffentlicht: (2026)
von: Yang, Ningyuan, et al.
Veröffentlicht: (2026)
KVCOMM: Online Cross-context KV-cache Communication for Efficient LLM-based Multi-agent Systems
von: Ye, Hancheng, et al.
Veröffentlicht: (2025)
von: Ye, Hancheng, et al.
Veröffentlicht: (2025)
Single-agent Reinforcement Learning Model for Regional Adaptive Traffic Signal Control
von: Li, Qiang, et al.
Veröffentlicht: (2025)
von: Li, Qiang, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Aligning LLM with human travel choices: a persona-based embedding learning approach
von: Liu, Tianming, et al.
Veröffentlicht: (2025) -
Toward LLM-Agent-Based Modeling of Transportation Systems: A Conceptual Framework
von: Liu, Tianming, et al.
Veröffentlicht: (2024) -
AIonopedia: an LLM agent orchestrating multimodal learning for ionic liquid discovery
von: Yin, Yuqi, et al.
Veröffentlicht: (2025) -
The impact of behavioral diversity in multi-agent reinforcement learning
von: Bettini, Matteo, et al.
Veröffentlicht: (2024) -
Understanding the planning of LLM agents: A survey
von: Huang, Xu, et al.
Veröffentlicht: (2024)