MAD: Motion Appearance Decoupling for efficient Driving World Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Rahimi, Ahmad, Gerard, Valentin, Zablocki, Eloi, Cord, Matthieu, Alahi, Alexandre |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Towards Motion Forecasting with Real-World Perception Inputs: Are End-to-End Approaches Competitive?
von: Xu, Yihong, et al.
Veröffentlicht: (2023)
von: Xu, Yihong, et al.
Veröffentlicht: (2023)
ReGentS: Real-World Safety-Critical Driving Scenario Generation Made Stable
von: Yin, Yuan, et al.
Veröffentlicht: (2024)
von: Yin, Yuan, et al.
Veröffentlicht: (2024)
GIFT: A Framework Towards Global Interpretable Faithful Textual Explanations of Vision Classifiers
von: Zablocki, Éloi, et al.
Veröffentlicht: (2024)
von: Zablocki, Éloi, et al.
Veröffentlicht: (2024)
UniTraj: A Unified Framework for Scalable Vehicle Trajectory Prediction
von: Feng, Lan, et al.
Veröffentlicht: (2024)
von: Feng, Lan, et al.
Veröffentlicht: (2024)
Annealed Winner-Takes-All for Motion Forecasting
von: Xu, Yihong, et al.
Veröffentlicht: (2024)
von: Xu, Yihong, et al.
Veröffentlicht: (2024)
GaussRender: Learning 3D Occupancy with Gaussian Rendering
von: Chambon, Loïck, et al.
Veröffentlicht: (2025)
von: Chambon, Loïck, et al.
Veröffentlicht: (2025)
PPT: Pretraining with Pseudo-Labeled Trajectories for Motion Forecasting
von: Xu, Yihong, et al.
Veröffentlicht: (2024)
von: Xu, Yihong, et al.
Veröffentlicht: (2024)
RAP: 3D Rasterization Augmented End-to-End Planning
von: Feng, Lan, et al.
Veröffentlicht: (2025)
von: Feng, Lan, et al.
Veröffentlicht: (2025)
LLM-wrapper: Black-Box Semantic-Aware Adaptation of Vision-Language Models for Referring Expression Comprehension
von: Cardiel, Amaia, et al.
Veröffentlicht: (2024)
von: Cardiel, Amaia, et al.
Veröffentlicht: (2024)
NAF: Zero-Shot Feature Upsampling via Neighborhood Attention Filtering
von: Chambon, Loick, et al.
Veröffentlicht: (2025)
von: Chambon, Loick, et al.
Veröffentlicht: (2025)
Towards Generalizable Trajectory Prediction Using Dual-Level Representation Learning And Adaptive Prompting
von: Messaoud, Kaouther, et al.
Veröffentlicht: (2025)
von: Messaoud, Kaouther, et al.
Veröffentlicht: (2025)
PointBeV: A Sparse Approach to BeV Predictions
von: Chambon, Loick, et al.
Veröffentlicht: (2023)
von: Chambon, Loick, et al.
Veröffentlicht: (2023)
A Multi-Loss Strategy for Vehicle Trajectory Prediction: Combining Off-Road, Diversity, and Directional Consistency Losses
von: Rahimi, Ahmad, et al.
Veröffentlicht: (2024)
von: Rahimi, Ahmad, et al.
Veröffentlicht: (2024)
Driving on Registers
von: Kirby, Ellington, et al.
Veröffentlicht: (2026)
von: Kirby, Ellington, et al.
Veröffentlicht: (2026)
GA-Drive: Geometry-Appearance Decoupled Modeling for Free-viewpoint Driving Scene Generation
von: Zhang, Hao, et al.
Veröffentlicht: (2026)
von: Zhang, Hao, et al.
Veröffentlicht: (2026)
Multi-Transmotion: Pre-trained Model for Human Motion Prediction
von: Gao, Yang, et al.
Veröffentlicht: (2024)
von: Gao, Yang, et al.
Veröffentlicht: (2024)
Anchored Video Generation: Decoupling Scene Construction and Temporal Synthesis in Text-to-Video Diffusion Models
von: Hassan, Mariam, et al.
Veröffentlicht: (2025)
von: Hassan, Mariam, et al.
Veröffentlicht: (2025)
Improved Baselines for Data-efficient Perceptual Augmentation of LLMs
von: Vallaeys, Théophane, et al.
Veröffentlicht: (2024)
von: Vallaeys, Théophane, et al.
Veröffentlicht: (2024)
Unsupervised Object Localization in the Era of Self-Supervised ViTs: A Survey
von: Siméoni, Oriane, et al.
Veröffentlicht: (2023)
von: Siméoni, Oriane, et al.
Veröffentlicht: (2023)
VaViM and VaVAM: Autonomous Driving through Video Generative Modeling
von: Bartoccioni, Florent, et al.
Veröffentlicht: (2025)
von: Bartoccioni, Florent, et al.
Veröffentlicht: (2025)
Retrieval-Based Interleaved Visual Chain-of-Thought in Real-World Driving Scenarios
von: Corbière, Charles, et al.
Veröffentlicht: (2025)
von: Corbière, Charles, et al.
Veröffentlicht: (2025)
Beyond Task Performance: Evaluating and Reducing the Flaws of Large Multimodal Models with In-Context Learning
von: Shukor, Mustafa, et al.
Veröffentlicht: (2023)
von: Shukor, Mustafa, et al.
Veröffentlicht: (2023)
From Generation to Generalization: Emergent Few-Shot Learning in Video Diffusion Models
von: Acuaviva, Pablo, et al.
Veröffentlicht: (2025)
von: Acuaviva, Pablo, et al.
Veröffentlicht: (2025)
MoSA: Motion-Coherent Human Video Generation via Structure-Appearance Decoupling
von: Wang, Haoyu, et al.
Veröffentlicht: (2025)
von: Wang, Haoyu, et al.
Veröffentlicht: (2025)
FG$^2$: Fine-Grained Cross-View Localization by Fine-Grained Feature Matching
von: Xia, Zimin, et al.
Veröffentlicht: (2025)
von: Xia, Zimin, et al.
Veröffentlicht: (2025)
DeMo++: Motion Decoupling for Autonomous Driving
von: Zhang, Bozhou, et al.
Veröffentlicht: (2025)
von: Zhang, Bozhou, et al.
Veröffentlicht: (2025)
Valeo4Cast: A Modular Approach to End-to-End Forecasting
von: Xu, Yihong, et al.
Veröffentlicht: (2024)
von: Xu, Yihong, et al.
Veröffentlicht: (2024)
SenCache: Accelerating Diffusion Model Inference via Sensitivity-Aware Caching
von: Haghighi, Yasaman, et al.
Veröffentlicht: (2026)
von: Haghighi, Yasaman, et al.
Veröffentlicht: (2026)
NAT: Learning to Attack Neurons for Enhanced Adversarial Transferability
von: Nakka, Krishna Kanth, et al.
Veröffentlicht: (2025)
von: Nakka, Krishna Kanth, et al.
Veröffentlicht: (2025)
Drift-Resistant Navigation World Model with Anchored Epipolar Guidance
von: Luan, Po-Chien, et al.
Veröffentlicht: (2026)
von: Luan, Po-Chien, et al.
Veröffentlicht: (2026)
GEM: A Generalizable Ego-Vision Multimodal World Model for Fine-Grained Ego-Motion, Object Dynamics, and Scene Composition Control
von: Hassan, Mariam, et al.
Veröffentlicht: (2024)
von: Hassan, Mariam, et al.
Veröffentlicht: (2024)
Skipping Computations in Multimodal LLMs
von: Shukor, Mustafa, et al.
Veröffentlicht: (2024)
von: Shukor, Mustafa, et al.
Veröffentlicht: (2024)
SSDD: Single-Step Diffusion Decoder for Efficient Image Tokenization
von: Vallaeys, Théophane, et al.
Veröffentlicht: (2025)
von: Vallaeys, Théophane, et al.
Veröffentlicht: (2025)
Motion Consistency Model: Accelerating Video Diffusion with Disentangled Motion-Appearance Distillation
von: Zhai, Yuanhao, et al.
Veröffentlicht: (2024)
von: Zhai, Yuanhao, et al.
Veröffentlicht: (2024)
An evaluation of SVBRDF Prediction from Generative Image Models for Appearance Modeling of 3D Scenes
von: Gauthier, Alban, et al.
Veröffentlicht: (2025)
von: Gauthier, Alban, et al.
Veröffentlicht: (2025)
Loc$^2$: Interpretable Cross-View Localization via Depth-Lifted Local Feature Matching
von: Xia, Zimin, et al.
Veröffentlicht: (2025)
von: Xia, Zimin, et al.
Veröffentlicht: (2025)
Social-Pose: Enhancing Trajectory Prediction with Human Body Pose
von: Gao, Yang, et al.
Veröffentlicht: (2025)
von: Gao, Yang, et al.
Veröffentlicht: (2025)
Implicit Multimodal Alignment: On the Generalization of Frozen LLMs to Multimodal Inputs
von: Shukor, Mustafa, et al.
Veröffentlicht: (2024)
von: Shukor, Mustafa, et al.
Veröffentlicht: (2024)
Forecast-PEFT: Parameter-Efficient Fine-Tuning for Pre-trained Motion Forecasting Models
von: Wang, Jifeng, et al.
Veröffentlicht: (2024)
von: Wang, Jifeng, et al.
Veröffentlicht: (2024)
VideoJAM: Joint Appearance-Motion Representations for Enhanced Motion Generation in Video Models
von: Chefer, Hila, et al.
Veröffentlicht: (2025)
von: Chefer, Hila, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Towards Motion Forecasting with Real-World Perception Inputs: Are End-to-End Approaches Competitive?
von: Xu, Yihong, et al.
Veröffentlicht: (2023) -
ReGentS: Real-World Safety-Critical Driving Scenario Generation Made Stable
von: Yin, Yuan, et al.
Veröffentlicht: (2024) -
GIFT: A Framework Towards Global Interpretable Faithful Textual Explanations of Vision Classifiers
von: Zablocki, Éloi, et al.
Veröffentlicht: (2024) -
UniTraj: A Unified Framework for Scalable Vehicle Trajectory Prediction
von: Feng, Lan, et al.
Veröffentlicht: (2024) -
Annealed Winner-Takes-All for Motion Forecasting
von: Xu, Yihong, et al.
Veröffentlicht: (2024)