WALL-E: World Alignment by Rule Learning Improves World Model-based LLM Agents
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhou, Siyu, Zhou, Tianyi, Yang, Yijun, Long, Guodong, Ye, Deheng, Jiang, Jing, Zhang, Chengqi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
WALL-E 2.0: World Alignment by NeuroSymbolic Learning improves World Model-based LLM Agents
von: Zhou, Siyu, et al.
Veröffentlicht: (2025)
von: Zhou, Siyu, et al.
Veröffentlicht: (2025)
Personalized Federated Collaborative Filtering: A Variational AutoEncoder Approach
von: Li, Zhiwei, et al.
Veröffentlicht: (2024)
von: Li, Zhiwei, et al.
Veröffentlicht: (2024)
What Hides behind Unfairness? Exploring Dynamics Fairness in Reinforcement Learning
von: Deng, Zhihong, et al.
Veröffentlicht: (2024)
von: Deng, Zhihong, et al.
Veröffentlicht: (2024)
Influence-oriented Personalized Federated Learning
von: Tan, Yue, et al.
Veröffentlicht: (2024)
von: Tan, Yue, et al.
Veröffentlicht: (2024)
RuleR: Improving LLM Controllability by Rule-based Data Recycling
von: Li, Ming, et al.
Veröffentlicht: (2024)
von: Li, Ming, et al.
Veröffentlicht: (2024)
Retrieval-Feedback-Driven Distillation and Preference Alignment for Efficient LLM-based Query Expansion
von: Li, Minghan, et al.
Veröffentlicht: (2026)
von: Li, Minghan, et al.
Veröffentlicht: (2026)
Federated Vision-Language-Recommendation with Personalized Fusion
von: Li, Zhiwei, et al.
Veröffentlicht: (2024)
von: Li, Zhiwei, et al.
Veröffentlicht: (2024)
Vul-R2: A Reasoning LLM for Automated Vulnerability Repair
von: Wen, Xin-Cheng, et al.
Veröffentlicht: (2025)
von: Wen, Xin-Cheng, et al.
Veröffentlicht: (2025)
Multiple LLM Agents Debate for Equitable Cultural Alignment
von: Ki, Dayeon, et al.
Veröffentlicht: (2025)
von: Ki, Dayeon, et al.
Veröffentlicht: (2025)
Multi-objective Large Language Model Alignment with Hierarchical Experts
von: Li, Zhuo, et al.
Veröffentlicht: (2025)
von: Li, Zhuo, et al.
Veröffentlicht: (2025)
GTR: Guided Thought Reinforcement Prevents Thought Collapse in RL-based VLM Agent Training
von: Wei, Tong, et al.
Veröffentlicht: (2025)
von: Wei, Tong, et al.
Veröffentlicht: (2025)
A Survey of Personalized Federated Foundation Models for Privacy-Preserving Recommendation
von: Li, Zhiwei, et al.
Veröffentlicht: (2025)
von: Li, Zhiwei, et al.
Veröffentlicht: (2025)
Navigating the Future of Federated Recommendation Systems with Foundation Models
von: Li, Zhiwei, et al.
Veröffentlicht: (2024)
von: Li, Zhiwei, et al.
Veröffentlicht: (2024)
HiRO-Nav: Hybrid ReasOning Enables Efficient Embodied Navigation
von: Zhao, He, et al.
Veröffentlicht: (2026)
von: Zhao, He, et al.
Veröffentlicht: (2026)
Exploring Expert Failures Improves LLM Agent Tuning
von: Lan, Li-Cheng, et al.
Veröffentlicht: (2025)
von: Lan, Li-Cheng, et al.
Veröffentlicht: (2025)
WorldGPT: Empowering LLM as Multimodal World Model
von: Ge, Zhiqi, et al.
Veröffentlicht: (2024)
von: Ge, Zhiqi, et al.
Veröffentlicht: (2024)
World Models as an Intermediary between Agents and the Real World
von: Yang, Sherry
Veröffentlicht: (2026)
von: Yang, Sherry
Veröffentlicht: (2026)
TradeTrap: Are LLM-based Trading Agents Truly Reliable and Faithful?
von: Yan, Lewen, et al.
Veröffentlicht: (2025)
von: Yan, Lewen, et al.
Veröffentlicht: (2025)
AgriWorld:A World Tools Protocol Framework for Verifiable Agricultural Reasoning with Code-Executing LLM Agents
von: Zhang, Zhixing, et al.
Veröffentlicht: (2026)
von: Zhang, Zhixing, et al.
Veröffentlicht: (2026)
Boosting Vulnerability Detection of LLMs via Curriculum Preference Optimization with Synthetic Reasoning Data
von: Wen, Xin-Cheng, et al.
Veröffentlicht: (2025)
von: Wen, Xin-Cheng, et al.
Veröffentlicht: (2025)
World Modelling Improves Language Model Agents
von: Guo, Shangmin, et al.
Veröffentlicht: (2025)
von: Guo, Shangmin, et al.
Veröffentlicht: (2025)
Learning to Wait: Synchronizing Agents with the Physical World
von: She, Yifei, et al.
Veröffentlicht: (2025)
von: She, Yifei, et al.
Veröffentlicht: (2025)
RuleArena: A Benchmark for Rule-Guided Reasoning with LLMs in Real-World Scenarios
von: Zhou, Ruiwen, et al.
Veröffentlicht: (2024)
von: Zhou, Ruiwen, et al.
Veröffentlicht: (2024)
TerminalWorld: Benchmarking Agents on Real-World Terminal Tasks
von: Chu, Zhaoyang, et al.
Veröffentlicht: (2026)
von: Chu, Zhaoyang, et al.
Veröffentlicht: (2026)
PV-SQL: Synergizing Database Probing and Rule-based Verification for Text-to-SQL Agents
von: Tian, Yuan, et al.
Veröffentlicht: (2026)
von: Tian, Yuan, et al.
Veröffentlicht: (2026)
Genesis: Evolving Attack Strategies for LLM Web Agent Red-Teaming
von: Zhang, Zheng, et al.
Veröffentlicht: (2025)
von: Zhang, Zheng, et al.
Veröffentlicht: (2025)
AriGraph: Learning Knowledge Graph World Models with Episodic Memory for LLM Agents
von: Anokhin, Petr, et al.
Veröffentlicht: (2024)
von: Anokhin, Petr, et al.
Veröffentlicht: (2024)
RLVR-World: Training World Models with Reinforcement Learning
von: Wu, Jialong, et al.
Veröffentlicht: (2025)
von: Wu, Jialong, et al.
Veröffentlicht: (2025)
Learning Multiple Probabilistic Decisions from Latent World Model in Autonomous Driving
von: Xiao, Lingyu, et al.
Veröffentlicht: (2024)
von: Xiao, Lingyu, et al.
Veröffentlicht: (2024)
Transferable Expertise for Autonomous Agents via Real-World Case-Based Learning
von: Ma, Zhenyu, et al.
Veröffentlicht: (2026)
von: Ma, Zhenyu, et al.
Veröffentlicht: (2026)
Coalitions of AI-based Methods Predict 15-Year Risks of Breast Cancer Metastasis Using Real-World Clinical Data with AUC up to 0.9
von: Jiang, Xia, et al.
Veröffentlicht: (2024)
von: Jiang, Xia, et al.
Veröffentlicht: (2024)
World4RL: Diffusion World Models for Policy Refinement with Reinforcement Learning for Robotic Manipulation
von: Jiang, Zhennan, et al.
Veröffentlicht: (2025)
von: Jiang, Zhennan, et al.
Veröffentlicht: (2025)
MIPO: Mutual Integration of Patient Journey and Medical Ontology for Healthcare Representation Learning
von: Peng, Xueping, et al.
Veröffentlicht: (2021)
von: Peng, Xueping, et al.
Veröffentlicht: (2021)
MAPF-World: Action World Model for Multi-Agent Path Finding
von: Yang, Zhanjiang, et al.
Veröffentlicht: (2025)
von: Yang, Zhanjiang, et al.
Veröffentlicht: (2025)
COMAP: Co-Evolving World Models and Agent Policies for LLM Agents
von: Liu, Youwei, et al.
Veröffentlicht: (2026)
von: Liu, Youwei, et al.
Veröffentlicht: (2026)
Explainable Reinforcement Learning Agents Using World Models
von: Singh, Madhuri, et al.
Veröffentlicht: (2025)
von: Singh, Madhuri, et al.
Veröffentlicht: (2025)
Learning to Be A Doctor: Searching for Effective Medical Agent Architectures
von: Zhuang, Yangyang, et al.
Veröffentlicht: (2025)
von: Zhuang, Yangyang, et al.
Veröffentlicht: (2025)
PROF: An LLM-based Reward Code Preference Optimization Framework for Offline Imitation Learning
von: Sun, Shengjie, et al.
Veröffentlicht: (2025)
von: Sun, Shengjie, et al.
Veröffentlicht: (2025)
Social World Models
von: Zhou, Xuhui, et al.
Veröffentlicht: (2025)
von: Zhou, Xuhui, et al.
Veröffentlicht: (2025)
Toward Real-World Table Agents: Capabilities, Workflows, and Design Principles for LLM-based Table Intelligence
von: Tian, Jiaming, et al.
Veröffentlicht: (2025)
von: Tian, Jiaming, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
WALL-E 2.0: World Alignment by NeuroSymbolic Learning improves World Model-based LLM Agents
von: Zhou, Siyu, et al.
Veröffentlicht: (2025) -
Personalized Federated Collaborative Filtering: A Variational AutoEncoder Approach
von: Li, Zhiwei, et al.
Veröffentlicht: (2024) -
What Hides behind Unfairness? Exploring Dynamics Fairness in Reinforcement Learning
von: Deng, Zhihong, et al.
Veröffentlicht: (2024) -
Influence-oriented Personalized Federated Learning
von: Tan, Yue, et al.
Veröffentlicht: (2024) -
RuleR: Improving LLM Controllability by Rule-based Data Recycling
von: Li, Ming, et al.
Veröffentlicht: (2024)