DreamSAC: Learning Hamiltonian World Models via Symmetry Exploration
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Tang, Jinzhou, Feng, Fan, Fu, Minghao, Lin, Wenjun, Huang, Biwei, Wang, Keze |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
STORM: Search-Guided Generative World Models for Robotic Manipulation
von: Lin, Wenjun, et al.
Veröffentlicht: (2025)
von: Lin, Wenjun, et al.
Veröffentlicht: (2025)
Beyond Pixels: Introducing Geometric-Semantic World Priors for Video-based Embodied Models via Spatio-temporal Alignment
von: Tang, Jinzhou, et al.
Veröffentlicht: (2025)
von: Tang, Jinzhou, et al.
Veröffentlicht: (2025)
SCAR: Self-Supervised Continuous Action Representation Learning
von: Liu, Hongjia, et al.
Veröffentlicht: (2026)
von: Liu, Hongjia, et al.
Veröffentlicht: (2026)
LASAR: Towards Spatio-temporal Reasoning with Latent Cognitive Map
von: Tang, Jinzhou, et al.
Veröffentlicht: (2026)
von: Tang, Jinzhou, et al.
Veröffentlicht: (2026)
CURLing the Dream: Contrastive Representations for World Modeling in Reinforcement Learning
von: Kich, Victor Augusto, et al.
Veröffentlicht: (2024)
von: Kich, Victor Augusto, et al.
Veröffentlicht: (2024)
MicroVerse: A Preliminary Exploration Toward a Micro-World Simulation
von: Wang, Rongsheng, et al.
Veröffentlicht: (2026)
von: Wang, Rongsheng, et al.
Veröffentlicht: (2026)
GTMA: Dynamic Representation Optimization for OOD Vision-Language Models
von: Zhang, Jensen, et al.
Veröffentlicht: (2025)
von: Zhang, Jensen, et al.
Veröffentlicht: (2025)
From Motion to Behavior: Hierarchical Modeling of Humanoid Generative Behavior Control
von: Zhang, Jusheng, et al.
Veröffentlicht: (2025)
von: Zhang, Jusheng, et al.
Veröffentlicht: (2025)
PTTA: A Pure Text-to-Animation Framework for High-Quality Creation
von: Chen, Ruiqi, et al.
Veröffentlicht: (2025)
von: Chen, Ruiqi, et al.
Veröffentlicht: (2025)
DreamVLA: A Vision-Language-Action Model Dreamed with Comprehensive World Knowledge
von: Zhang, Wenyao, et al.
Veröffentlicht: (2025)
von: Zhang, Wenyao, et al.
Veröffentlicht: (2025)
DTL: Disentangled Transfer Learning for Visual Recognition
von: Fu, Minghao, et al.
Veröffentlicht: (2023)
von: Fu, Minghao, et al.
Veröffentlicht: (2023)
Enhancing targeted transferability via feature space fine-tuning
von: Zeng, Hui, et al.
Veröffentlicht: (2024)
von: Zeng, Hui, et al.
Veröffentlicht: (2024)
DreamWorld: Unified World Modeling in Video Generation
von: Tan, Boming, et al.
Veröffentlicht: (2026)
von: Tan, Boming, et al.
Veröffentlicht: (2026)
Learning Plug-and-play Memory for Guiding Video Diffusion Models
von: Song, Selena, et al.
Veröffentlicht: (2025)
von: Song, Selena, et al.
Veröffentlicht: (2025)
Low-rank Attention Side-Tuning for Parameter-Efficient Fine-Tuning
von: Tang, Ningyuan, et al.
Veröffentlicht: (2024)
von: Tang, Ningyuan, et al.
Veröffentlicht: (2024)
A Stepwise Distillation Learning Strategy for Non-differentiable Visual Programming Frameworks on Visual Reasoning Tasks
von: Wan, Wentao, et al.
Veröffentlicht: (2023)
von: Wan, Wentao, et al.
Veröffentlicht: (2023)
3D-Agent:Tri-Modal Multi-Agent Collaboration for Scalable 3D Object Annotation
von: Zhang, Jusheng, et al.
Veröffentlicht: (2026)
von: Zhang, Jusheng, et al.
Veröffentlicht: (2026)
ResAgent: Entropy-based Prior Point Discovery and Visual Reasoning for Referring Expression Segmentation
von: Wang, Yihao, et al.
Veröffentlicht: (2026)
von: Wang, Yihao, et al.
Veröffentlicht: (2026)
Learning Vision-Language-Action World Models for Autonomous Driving
von: Wang, Guoqing, et al.
Veröffentlicht: (2026)
von: Wang, Guoqing, et al.
Veröffentlicht: (2026)
DreamPolish: Domain Score Distillation With Progressive Geometry Generation
von: Cheng, Yean, et al.
Veröffentlicht: (2024)
von: Cheng, Yean, et al.
Veröffentlicht: (2024)
Sekai: A Video Dataset towards World Exploration
von: Li, Zhen, et al.
Veröffentlicht: (2025)
von: Li, Zhen, et al.
Veröffentlicht: (2025)
Lattice Boltzmann Model for Learning Real-World Pixel Dynamicity
von: Zheng, Guangze, et al.
Veröffentlicht: (2025)
von: Zheng, Guangze, et al.
Veröffentlicht: (2025)
Dream2Flow: Bridging Video Generation and Open-World Manipulation with 3D Object Flow
von: Dharmarajan, Karthik, et al.
Veröffentlicht: (2025)
von: Dharmarajan, Karthik, et al.
Veröffentlicht: (2025)
HybridToken-VLM: Hybrid Token Compression for Vision-Language Models
von: Zhang, Jusheng, et al.
Veröffentlicht: (2025)
von: Zhang, Jusheng, et al.
Veröffentlicht: (2025)
Top-Down Semantic Refinement for Image Captioning
von: Zhang, Jusheng, et al.
Veröffentlicht: (2025)
von: Zhang, Jusheng, et al.
Veröffentlicht: (2025)
DreamActor-M2: Universal Character Image Animation via Spatiotemporal In-Context Learning
von: Luo, Mingshuang, et al.
Veröffentlicht: (2026)
von: Luo, Mingshuang, et al.
Veröffentlicht: (2026)
Dream to Generalize: Zero-Shot Model-Based Reinforcement Learning for Unseen Visual Distractions
von: Ha, Jeongsoo, et al.
Veröffentlicht: (2025)
von: Ha, Jeongsoo, et al.
Veröffentlicht: (2025)
MapDream: Task-Driven Map Learning for Vision-Language Navigation
von: Lian, Guoxin, et al.
Veröffentlicht: (2026)
von: Lian, Guoxin, et al.
Veröffentlicht: (2026)
DreamLoop: Controllable Cinemagraph Generation from a Single Photograph
von: Mahapatra, Aniruddha, et al.
Veröffentlicht: (2026)
von: Mahapatra, Aniruddha, et al.
Veröffentlicht: (2026)
DreamDojo: A Generalist Robot World Model from Large-Scale Human Videos
von: Gao, Shenyuan, et al.
Veröffentlicht: (2026)
von: Gao, Shenyuan, et al.
Veröffentlicht: (2026)
DreamPolisher: Towards High-Quality Text-to-3D Generation via Geometric Diffusion
von: Lin, Yuanze, et al.
Veröffentlicht: (2024)
von: Lin, Yuanze, et al.
Veröffentlicht: (2024)
StressDream: Steering Video World Models for Robust Policy Evaluation and Improvement
von: Seo, Junwon, et al.
Veröffentlicht: (2026)
von: Seo, Junwon, et al.
Veröffentlicht: (2026)
DreamID: High-Fidelity and Fast diffusion-based Face Swapping via Triplet ID Group Learning
von: Ye, Fulong, et al.
Veröffentlicht: (2025)
von: Ye, Fulong, et al.
Veröffentlicht: (2025)
MAT-Agent: Adaptive Multi-Agent Training Optimization
von: Zhang, Jusheng, et al.
Veröffentlicht: (2025)
von: Zhang, Jusheng, et al.
Veröffentlicht: (2025)
Back to Parsimonious Latents: Learning Task-Centric World Models from Visual Foundations
von: Fu, Minghao, et al.
Veröffentlicht: (2026)
von: Fu, Minghao, et al.
Veröffentlicht: (2026)
Supervised Learning Model for Key Frame Identification from Cow Teat Videos
von: Wang, Minghao, et al.
Veröffentlicht: (2024)
von: Wang, Minghao, et al.
Veröffentlicht: (2024)
Chain of World: World Model Thinking in Latent Motion
von: Yang, Fuxiang, et al.
Veröffentlicht: (2026)
von: Yang, Fuxiang, et al.
Veröffentlicht: (2026)
DreamDance: Animating Character Art via Inpainting Stable Gaussian Worlds
von: Zhang, Jiaxu, et al.
Veröffentlicht: (2025)
von: Zhang, Jiaxu, et al.
Veröffentlicht: (2025)
How Far is Video Generation from World Model: A Physical Law Perspective
von: Kang, Bingyi, et al.
Veröffentlicht: (2024)
von: Kang, Bingyi, et al.
Veröffentlicht: (2024)
Adaptive-VoCo: Complexity-Aware Visual Token Compression for Vision-Language Models
von: Guo, Xiaoyang, et al.
Veröffentlicht: (2025)
von: Guo, Xiaoyang, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
STORM: Search-Guided Generative World Models for Robotic Manipulation
von: Lin, Wenjun, et al.
Veröffentlicht: (2025) -
Beyond Pixels: Introducing Geometric-Semantic World Priors for Video-based Embodied Models via Spatio-temporal Alignment
von: Tang, Jinzhou, et al.
Veröffentlicht: (2025) -
SCAR: Self-Supervised Continuous Action Representation Learning
von: Liu, Hongjia, et al.
Veröffentlicht: (2026) -
LASAR: Towards Spatio-temporal Reasoning with Latent Cognitive Map
von: Tang, Jinzhou, et al.
Veröffentlicht: (2026) -
CURLing the Dream: Contrastive Representations for World Modeling in Reinforcement Learning
von: Kich, Victor Augusto, et al.
Veröffentlicht: (2024)