Towards Unraveling and Improving Generalization in World Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Fang, Qiaoyi, Du, Weiyu, Wang, Hang, Zhang, Junshan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Heterogeneous Decision Making in Mixed Traffic: Uncertainty-aware Planning and Bounded Rationality
von: Wang, Hang, et al.
Veröffentlicht: (2025)
von: Wang, Hang, et al.
Veröffentlicht: (2025)
Towards Causal Relationship in Indefinite Data: Baseline Model and New Datasets
von: Chen, Hang, et al.
Veröffentlicht: (2024)
von: Chen, Hang, et al.
Veröffentlicht: (2024)
Adversarial Reward Auditing for Active Detection and Mitigation of Reward Hacking
von: Beigi, Mohammad, et al.
Veröffentlicht: (2026)
von: Beigi, Mohammad, et al.
Veröffentlicht: (2026)
Towards Improved Preference Optimization Pipeline: from Data Generation to Budget-Controlled Regularization
von: Chen, Zhuotong, et al.
Veröffentlicht: (2024)
von: Chen, Zhuotong, et al.
Veröffentlicht: (2024)
Masked Generative Priors Improve World Models Sequence Modelling Capabilities
von: Meo, Cristian, et al.
Veröffentlicht: (2024)
von: Meo, Cristian, et al.
Veröffentlicht: (2024)
Representational Homomorphism Predicts and Improves Compositional Generalization In Transformer Language Model
von: An, Zhiyu, et al.
Veröffentlicht: (2026)
von: An, Zhiyu, et al.
Veröffentlicht: (2026)
Ego-centric Learning of Communicative World Models for Autonomous Driving
von: Wang, Hang, et al.
Veröffentlicht: (2025)
von: Wang, Hang, et al.
Veröffentlicht: (2025)
World Action Verifier: Self-Improving World Models via Forward-Inverse Asymmetry
von: Liu, Yuejiang, et al.
Veröffentlicht: (2026)
von: Liu, Yuejiang, et al.
Veröffentlicht: (2026)
Unraveling the Potential of Diffusion Models in Small Molecule Generation
von: Zhang, Peining, et al.
Veröffentlicht: (2025)
von: Zhang, Peining, et al.
Veröffentlicht: (2025)
IR$^3$: Contrastive Inverse Reinforcement Learning for Interpretable Detection and Mitigation of Reward Hacking
von: Beigi, Mohammad, et al.
Veröffentlicht: (2026)
von: Beigi, Mohammad, et al.
Veröffentlicht: (2026)
Towards Generalization-Oriented Models for Vehicle Routing Problems with Mixture-of-Experts
von: Miao, Changhao, et al.
Veröffentlicht: (2026)
von: Miao, Changhao, et al.
Veröffentlicht: (2026)
Tackling the Non-IID Issue in Heterogeneous Federated Learning by Gradient Harmonization
von: Zhang, Xinyu, et al.
Veröffentlicht: (2023)
von: Zhang, Xinyu, et al.
Veröffentlicht: (2023)
World Modelling Improves Language Model Agents
von: Guo, Shangmin, et al.
Veröffentlicht: (2025)
von: Guo, Shangmin, et al.
Veröffentlicht: (2025)
Improving Token-Based World Models with Parallel Observation Prediction
von: Cohen, Lior, et al.
Veröffentlicht: (2024)
von: Cohen, Lior, et al.
Veröffentlicht: (2024)
OLLIE: Imitation Learning from Offline Pretraining to Online Finetuning
von: Yue, Sheng, et al.
Veröffentlicht: (2024)
von: Yue, Sheng, et al.
Veröffentlicht: (2024)
Improving Generalization of Neural Vehicle Routing Problem Solvers Through the Lens of Model Architecture
von: Xiao, Yubin, et al.
Veröffentlicht: (2024)
von: Xiao, Yubin, et al.
Veröffentlicht: (2024)
Improving Transformer World Models for Data-Efficient RL
von: Dedieu, Antoine, et al.
Veröffentlicht: (2025)
von: Dedieu, Antoine, et al.
Veröffentlicht: (2025)
Towards Theoretical Understandings of Self-Consuming Generative Models
von: Fu, Shi, et al.
Veröffentlicht: (2024)
von: Fu, Shi, et al.
Veröffentlicht: (2024)
How to Leverage Diverse Demonstrations in Offline Imitation Learning
von: Yue, Sheng, et al.
Veröffentlicht: (2024)
von: Yue, Sheng, et al.
Veröffentlicht: (2024)
Emergent Semantic Representations in World Models through Physical Interaction without Linguistic Supervision
von: Fang, Jiayi
Veröffentlicht: (2026)
von: Fang, Jiayi
Veröffentlicht: (2026)
MolTC: Towards Molecular Relational Modeling In Language Models
von: Fang, Junfeng, et al.
Veröffentlicht: (2024)
von: Fang, Junfeng, et al.
Veröffentlicht: (2024)
Provably Invincible Adversarial Attacks on Reinforcement Learning Systems: A Rate-Distortion Information-Theoretic Approach
von: Lu, Ziqing, et al.
Veröffentlicht: (2025)
von: Lu, Ziqing, et al.
Veröffentlicht: (2025)
Improving World Models using Deep Supervision with Linear Probes
von: Zahorodnii, Andrii
Veröffentlicht: (2025)
von: Zahorodnii, Andrii
Veröffentlicht: (2025)
Spatiotemporal Forecasting as Planning: A Model-Based Reinforcement Learning Approach with Generative World Models
von: Wu, Hao, et al.
Veröffentlicht: (2025)
von: Wu, Hao, et al.
Veröffentlicht: (2025)
Neuro-Symbolic Artificial Intelligence: Towards Improving the Reasoning Abilities of Large Language Models
von: Yang, Xiao-Wen, et al.
Veröffentlicht: (2025)
von: Yang, Xiao-Wen, et al.
Veröffentlicht: (2025)
Unraveling Spatio-Temporal Foundation Models via the Pipeline Lens: A Comprehensive Review
von: Fang, Yuchen, et al.
Veröffentlicht: (2025)
von: Fang, Yuchen, et al.
Veröffentlicht: (2025)
Towards Urban General Intelligence: A Review and Outlook of Urban Foundation Models
von: Zhang, Weijia, et al.
Veröffentlicht: (2024)
von: Zhang, Weijia, et al.
Veröffentlicht: (2024)
Towards an Improved Metric for Evaluating Disentangled Representations
von: Julka, Sahib, et al.
Veröffentlicht: (2024)
von: Julka, Sahib, et al.
Veröffentlicht: (2024)
Towards Large-Scale In-Context Reinforcement Learning by Meta-Training in Randomized Worlds
von: Wang, Fan, et al.
Veröffentlicht: (2025)
von: Wang, Fan, et al.
Veröffentlicht: (2025)
Locality Sensitive Sparse Encoding for Learning World Models Online
von: Liu, Zichen, et al.
Veröffentlicht: (2024)
von: Liu, Zichen, et al.
Veröffentlicht: (2024)
Toward Privileged Foundation Models:LUPI for Accelerated and Improved Learning
von: Ding, Xueying, et al.
Veröffentlicht: (2026)
von: Ding, Xueying, et al.
Veröffentlicht: (2026)
Towards General-Purpose Model-Free Reinforcement Learning
von: Fujimoto, Scott, et al.
Veröffentlicht: (2025)
von: Fujimoto, Scott, et al.
Veröffentlicht: (2025)
Continual Reinforcement Learning by Planning with Online World Models
von: Liu, Zichen, et al.
Veröffentlicht: (2025)
von: Liu, Zichen, et al.
Veröffentlicht: (2025)
Dreaming of Many Worlds: Learning Contextual World Models Aids Zero-Shot Generalization
von: Prasanna, Sai, et al.
Veröffentlicht: (2024)
von: Prasanna, Sai, et al.
Veröffentlicht: (2024)
VAM: Verbalized Action Masking for Controllable Exploration in RL Post-Training -- A Chess Case Study
von: Zhang, Zhicheng, et al.
Veröffentlicht: (2026)
von: Zhang, Zhicheng, et al.
Veröffentlicht: (2026)
Topology-Independent Robustness of the Weighted Mean under Label Poisoning Attacks in Heterogeneous Decentralized Learning
von: Peng, Jie, et al.
Veröffentlicht: (2026)
von: Peng, Jie, et al.
Veröffentlicht: (2026)
Meta-World+: An Improved, Standardized, RL Benchmark
von: McLean, Reginald, et al.
Veröffentlicht: (2025)
von: McLean, Reginald, et al.
Veröffentlicht: (2025)
Towards Flash Thinking via Decoupled Advantage Policy Optimization
von: Tan, Zezhong, et al.
Veröffentlicht: (2025)
von: Tan, Zezhong, et al.
Veröffentlicht: (2025)
Towards a Scalable Reference-Free Evaluation of Generative Models
von: Ospanov, Azim, et al.
Veröffentlicht: (2024)
von: Ospanov, Azim, et al.
Veröffentlicht: (2024)
Unraveling Text Generation in LLMs: A Stochastic Differential Equation Approach
von: Zhang, Yukun
Veröffentlicht: (2024)
von: Zhang, Yukun
Veröffentlicht: (2024)
Ähnliche Einträge
-
Heterogeneous Decision Making in Mixed Traffic: Uncertainty-aware Planning and Bounded Rationality
von: Wang, Hang, et al.
Veröffentlicht: (2025) -
Towards Causal Relationship in Indefinite Data: Baseline Model and New Datasets
von: Chen, Hang, et al.
Veröffentlicht: (2024) -
Adversarial Reward Auditing for Active Detection and Mitigation of Reward Hacking
von: Beigi, Mohammad, et al.
Veröffentlicht: (2026) -
Towards Improved Preference Optimization Pipeline: from Data Generation to Budget-Controlled Regularization
von: Chen, Zhuotong, et al.
Veröffentlicht: (2024) -
Masked Generative Priors Improve World Models Sequence Modelling Capabilities
von: Meo, Cristian, et al.
Veröffentlicht: (2024)