Diffusion Forcing: Next-token Prediction Meets Full-Sequence Diffusion
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chen, Boyuan, Monso, Diego Marti, Du, Yilun, Simchowitz, Max, Tedrake, Russ, Sitzmann, Vincent |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
History-Guided Video Diffusion
von: Song, Kiwhan, et al.
Veröffentlicht: (2025)
von: Song, Kiwhan, et al.
Veröffentlicht: (2025)
Large Video Planner Enables Generalizable Robot Control
von: Chen, Boyuan, et al.
Veröffentlicht: (2025)
von: Chen, Boyuan, et al.
Veröffentlicht: (2025)
Turning Video Models into Generalist Robot Policies
von: Li, Sizhe Lester, et al.
Veröffentlicht: (2026)
von: Li, Sizhe Lester, et al.
Veröffentlicht: (2026)
Diffusion Forcing for Multi-Agent Interaction Sequence Modeling
von: Maluleke, Vongani H., et al.
Veröffentlicht: (2025)
von: Maluleke, Vongani H., et al.
Veröffentlicht: (2025)
SMART: Scalable Multi-agent Real-time Motion Generation via Next-token Prediction
von: Wu, Wei, et al.
Veröffentlicht: (2024)
von: Wu, Wei, et al.
Veröffentlicht: (2024)
Potential Based Diffusion Motion Planning
von: Luo, Yunhao, et al.
Veröffentlicht: (2024)
von: Luo, Yunhao, et al.
Veröffentlicht: (2024)
Controlling diverse robots by inferring Jacobian fields with deep networks
von: Li, Sizhe Lester, et al.
Veröffentlicht: (2024)
von: Li, Sizhe Lester, et al.
Veröffentlicht: (2024)
SceneSmith: Agentic Generation of Simulation-Ready Indoor Scenes
von: Pfaff, Nicholas, et al.
Veröffentlicht: (2026)
von: Pfaff, Nicholas, et al.
Veröffentlicht: (2026)
Diffusion-FS: Multimodal Free-Space Prediction via Diffusion for Autonomous Driving
von: Gupta, Keshav, et al.
Veröffentlicht: (2025)
von: Gupta, Keshav, et al.
Veröffentlicht: (2025)
Constrained Bimanual Planning with Analytic Inverse Kinematics
von: Cohn, Thomas, et al.
Veröffentlicht: (2023)
von: Cohn, Thomas, et al.
Veröffentlicht: (2023)
Anomalies by Synthesis: Anomaly Detection using Generative Diffusion Models for Off-Road Navigation
von: Ancha, Siddharth, et al.
Veröffentlicht: (2025)
von: Ancha, Siddharth, et al.
Veröffentlicht: (2025)
Locality in Image Diffusion Models Emerges from Data Statistics
von: Lukoianov, Artem, et al.
Veröffentlicht: (2025)
von: Lukoianov, Artem, et al.
Veröffentlicht: (2025)
NextBestPath: Efficient 3D Mapping of Unseen Environments
von: Li, Shiyao, et al.
Veröffentlicht: (2025)
von: Li, Shiyao, et al.
Veröffentlicht: (2025)
AMP: Autoregressive Motion Prediction Revisited with Next Token Prediction for Autonomous Driving
von: Jia, Xiaosong, et al.
Veröffentlicht: (2024)
von: Jia, Xiaosong, et al.
Veröffentlicht: (2024)
DAP: Diffusion-based Affordance Prediction for Multi-modality Storage
von: Chang, Haonan, et al.
Veröffentlicht: (2024)
von: Chang, Haonan, et al.
Veröffentlicht: (2024)
DAgger Diffusion Navigation: DAgger Boosted Diffusion Policy for Vision-Language Navigation
von: Shi, Haoxiang, et al.
Veröffentlicht: (2025)
von: Shi, Haoxiang, et al.
Veröffentlicht: (2025)
DiffusionDrive: Truncated Diffusion Model for End-to-End Autonomous Driving
von: Liao, Bencheng, et al.
Veröffentlicht: (2024)
von: Liao, Bencheng, et al.
Veröffentlicht: (2024)
Diffusion Transformer Policy
von: Hou, Zhi, et al.
Veröffentlicht: (2024)
von: Hou, Zhi, et al.
Veröffentlicht: (2024)
DINOv3-Diffusion Policy: Self-Supervised Large Visual Model for Visuomotor Diffusion Policy Learning
von: Egbe, ThankGod, et al.
Veröffentlicht: (2025)
von: Egbe, ThankGod, et al.
Veröffentlicht: (2025)
SuperPC: A Single Diffusion Model for Point Cloud Completion, Upsampling, Denoising, and Colorization
von: Du, Yi, et al.
Veröffentlicht: (2025)
von: Du, Yi, et al.
Veröffentlicht: (2025)
Inference-Time Enhancement of Generative Robot Policies via Predictive World Modeling
von: Qi, Han, et al.
Veröffentlicht: (2025)
von: Qi, Han, et al.
Veröffentlicht: (2025)
Unified Diffusion VLA: Vision-Language-Action Model via Joint Discrete Denoising Diffusion Process
von: Chen, Jiayi, et al.
Veröffentlicht: (2025)
von: Chen, Jiayi, et al.
Veröffentlicht: (2025)
DATR: Diffusion-based 3D Apple Tree Reconstruction Framework with Sparse-View
von: Qiu, Tian, et al.
Veröffentlicht: (2025)
von: Qiu, Tian, et al.
Veröffentlicht: (2025)
ADM: Accelerated Diffusion Model via Estimated Priors for Robust Motion Prediction under Uncertainties
von: Li, Jiahui, et al.
Veröffentlicht: (2024)
von: Li, Jiahui, et al.
Veröffentlicht: (2024)
SceneStreamer: Continuous Scenario Generation as Next Token Group Prediction
von: Peng, Zhenghao, et al.
Veröffentlicht: (2025)
von: Peng, Zhenghao, et al.
Veröffentlicht: (2025)
Diffusion-Based Environment-Aware Trajectory Prediction
von: Westny, Theodor, et al.
Veröffentlicht: (2024)
von: Westny, Theodor, et al.
Veröffentlicht: (2024)
EP-Diffuser: An Efficient Diffusion Model for Traffic Scene Generation and Prediction via Polynomial Representations
von: Yao, Yue, et al.
Veröffentlicht: (2025)
von: Yao, Yue, et al.
Veröffentlicht: (2025)
X-Diffusion: Training Diffusion Policies on Cross-Embodiment Human Demonstrations
von: Pace, Maximus A., et al.
Veröffentlicht: (2025)
von: Pace, Maximus A., et al.
Veröffentlicht: (2025)
3D-Mem: 3D Scene Memory for Embodied Exploration and Reasoning
von: Yang, Yuncong, et al.
Veröffentlicht: (2024)
von: Yang, Yuncong, et al.
Veröffentlicht: (2024)
Generative View Stitching
von: Song, Chonghyuk, et al.
Veröffentlicht: (2025)
von: Song, Chonghyuk, et al.
Veröffentlicht: (2025)
FusionForce: End-to-end Differentiable Neural-Symbolic Layer for Trajectory Prediction
von: Agishev, Ruslan, et al.
Veröffentlicht: (2025)
von: Agishev, Ruslan, et al.
Veröffentlicht: (2025)
Learning 3D Persistent Embodied World Models
von: Zhou, Siyuan, et al.
Veröffentlicht: (2025)
von: Zhou, Siyuan, et al.
Veröffentlicht: (2025)
SIMPACT: Simulation-Enabled Action Planning using Vision-Language Models
von: Liu, Haowen, et al.
Veröffentlicht: (2025)
von: Liu, Haowen, et al.
Veröffentlicht: (2025)
Diffusion Meets DAgger: Supercharging Eye-in-hand Imitation Learning
von: Zhang, Xiaoyu, et al.
Veröffentlicht: (2024)
von: Zhang, Xiaoyu, et al.
Veröffentlicht: (2024)
Spatial Calibration of Diffuse LiDARs
von: Behari, Nikhil, et al.
Veröffentlicht: (2026)
von: Behari, Nikhil, et al.
Veröffentlicht: (2026)
Compositional Generative Modeling: A Single Model is Not All You Need
von: Du, Yilun, et al.
Veröffentlicht: (2024)
von: Du, Yilun, et al.
Veröffentlicht: (2024)
Grounding Video Models to Actions through Goal Conditioned Exploration
von: Luo, Yunhao, et al.
Veröffentlicht: (2024)
von: Luo, Yunhao, et al.
Veröffentlicht: (2024)
Bayesian-Optimized One-Step Diffusion Model with Knowledge Distillation for Real-Time 3D Human Motion Prediction
von: Tian, Sibo, et al.
Veröffentlicht: (2024)
von: Tian, Sibo, et al.
Veröffentlicht: (2024)
Humanoid Locomotion as Next Token Prediction
von: Radosavovic, Ilija, et al.
Veröffentlicht: (2024)
von: Radosavovic, Ilija, et al.
Veröffentlicht: (2024)
Task-Oriented Human Grasp Synthesis via Context- and Task-Aware Diffusers
von: Liu, An-Lun, et al.
Veröffentlicht: (2025)
von: Liu, An-Lun, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
History-Guided Video Diffusion
von: Song, Kiwhan, et al.
Veröffentlicht: (2025) -
Large Video Planner Enables Generalizable Robot Control
von: Chen, Boyuan, et al.
Veröffentlicht: (2025) -
Turning Video Models into Generalist Robot Policies
von: Li, Sizhe Lester, et al.
Veröffentlicht: (2026) -
Diffusion Forcing for Multi-Agent Interaction Sequence Modeling
von: Maluleke, Vongani H., et al.
Veröffentlicht: (2025) -
SMART: Scalable Multi-agent Real-time Motion Generation via Next-token Prediction
von: Wu, Wei, et al.
Veröffentlicht: (2024)