Diff-2-in-1: Bridging Generation and Dense Perception with Diffusion Models
Fuente:
arXiv
Guardado en:
| Autores principales: | Zheng, Shuhong, Bao, Zhipeng, Zhao, Ruoyu, Hebert, Martial, Wang, Yu-Xiong |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Lexicon3D: Probing Visual Foundation Models for Complex 3D Scene Understanding
por: Man, Yunze, et al.
Publicado: (2024)
por: Man, Yunze, et al.
Publicado: (2024)
DexHandDiff: Interaction-aware Diffusion Planning for Adaptive Dexterous Manipulation
por: Liang, Zhixuan, et al.
Publicado: (2024)
por: Liang, Zhixuan, et al.
Publicado: (2024)
Track, Inpaint, Resplat: Subject-driven 3D and 4D Generation with Progressive Texture Infilling
por: Zheng, Shuhong, et al.
Publicado: (2025)
por: Zheng, Shuhong, et al.
Publicado: (2025)
Fractional Diffusion Bridge Models
por: Nobis, Gabriel, et al.
Publicado: (2025)
por: Nobis, Gabriel, et al.
Publicado: (2025)
EMPERROR: A Flexible Generative Perception Error Model for Probing Self-Driving Planners
por: Hanselmann, Niklas, et al.
Publicado: (2024)
por: Hanselmann, Niklas, et al.
Publicado: (2024)
EP-Diffuser: An Efficient Diffusion Model for Traffic Scene Generation and Prediction via Polynomial Representations
por: Yao, Yue, et al.
Publicado: (2025)
por: Yao, Yue, et al.
Publicado: (2025)
MapDiffusion: Generative Diffusion for Vectorized Online HD Map Construction and Uncertainty Estimation in Autonomous Driving
por: Monninger, Thomas, et al.
Publicado: (2025)
por: Monninger, Thomas, et al.
Publicado: (2025)
Learning Part-Aware Dense 3D Feature Field for Generalizable Articulated Object Manipulation
por: Chen, Yue, et al.
Publicado: (2026)
por: Chen, Yue, et al.
Publicado: (2026)
Separate-and-Enhance: Compositional Finetuning for Text2Image Diffusion Models
por: Bao, Zhipeng, et al.
Publicado: (2023)
por: Bao, Zhipeng, et al.
Publicado: (2023)
Dense Policy: Bidirectional Autoregressive Learning of Actions
por: Su, Yue, et al.
Publicado: (2025)
por: Su, Yue, et al.
Publicado: (2025)
SyncDiff: Synchronized Motion Diffusion for Multi-Body Human-Object Interaction Synthesis
por: He, Wenkun, et al.
Publicado: (2024)
por: He, Wenkun, et al.
Publicado: (2024)
Distributed NeRF Learning for Collaborative Multi-Robot Perception
por: Zhao, Hongrui, et al.
Publicado: (2024)
por: Zhao, Hongrui, et al.
Publicado: (2024)
Multi-Step Guided Diffusion for Image Restoration on Edge Devices: Toward Lightweight Perception in Embodied AI
por: Chakravarty, Aditya
Publicado: (2025)
por: Chakravarty, Aditya
Publicado: (2025)
Generalizable Dense Reward for Long-Horizon Robotic Tasks
por: Yong, Silong, et al.
Publicado: (2026)
por: Yong, Silong, et al.
Publicado: (2026)
Anomalies by Synthesis: Anomaly Detection using Generative Diffusion Models for Off-Road Navigation
por: Ancha, Siddharth, et al.
Publicado: (2025)
por: Ancha, Siddharth, et al.
Publicado: (2025)
VO-DP: Semantic-Geometric Adaptive Diffusion Policy for Vision-Only Robotic Manipulation
por: Ni, Zehao, et al.
Publicado: (2025)
por: Ni, Zehao, et al.
Publicado: (2025)
Non-rigid Relative Placement through 3D Dense Diffusion
por: Cai, Eric, et al.
Publicado: (2024)
por: Cai, Eric, et al.
Publicado: (2024)
DiffCloud: Real-to-Sim from Point Clouds with Differentiable Simulation and Rendering of Deformable Objects
por: Sundaresan, Priya, et al.
Publicado: (2022)
por: Sundaresan, Priya, et al.
Publicado: (2022)
DiffGen: Robot Demonstration Generation via Differentiable Physics Simulation, Differentiable Rendering, and Vision-Language Model
por: Jin, Yang, et al.
Publicado: (2024)
por: Jin, Yang, et al.
Publicado: (2024)
Consistency Diffusion Bridge Models
por: He, Guande, et al.
Publicado: (2024)
por: He, Guande, et al.
Publicado: (2024)
UFM: A Simple Path towards Unified Dense Correspondence with Flow
por: Zhang, Yuchen, et al.
Publicado: (2025)
por: Zhang, Yuchen, et al.
Publicado: (2025)
4D Contrastive Superflows are Dense 3D Representation Learners
por: Xu, Xiang, et al.
Publicado: (2024)
por: Xu, Xiang, et al.
Publicado: (2024)
Motion-prior Contrast Maximization for Dense Continuous-Time Motion Estimation
por: Hamann, Friedhelm, et al.
Publicado: (2024)
por: Hamann, Friedhelm, et al.
Publicado: (2024)
NeurAll: Towards a Unified Visual Perception Model for Automated Driving
por: Sistu, Ganesh, et al.
Publicado: (2019)
por: Sistu, Ganesh, et al.
Publicado: (2019)
NavigScene: Bridging Local Perception and Global Navigation for Beyond-Visual-Range Autonomous Driving
por: Peng, Qucheng, et al.
Publicado: (2025)
por: Peng, Qucheng, et al.
Publicado: (2025)
Grasp2Grasp: Vision-Based Dexterous Grasp Translation via Schrödinger Bridges
por: Zhong, Tao, et al.
Publicado: (2025)
por: Zhong, Tao, et al.
Publicado: (2025)
ImitDiff: Transferring Foundation-Model Priors for Distraction Robust Visuomotor Policy
por: Dong, Yuhang, et al.
Publicado: (2025)
por: Dong, Yuhang, et al.
Publicado: (2025)
Model-Agnostic Open-Set Air-to-Air Visual Object Detection for Reliable UAV Perception
por: Loukovitis, Spyridon, et al.
Publicado: (2025)
por: Loukovitis, Spyridon, et al.
Publicado: (2025)
NeuralRemaster: Phase-Preserving Diffusion for Structure-Aligned Generation
por: Zeng, Yu, et al.
Publicado: (2025)
por: Zeng, Yu, et al.
Publicado: (2025)
Dense-depth map guided deep Lidar-Visual Odometry with Sparse Point Clouds and Images
por: Huang, JunYing, et al.
Publicado: (2025)
por: Huang, JunYing, et al.
Publicado: (2025)
The GOOSE Dataset for Perception in Unstructured Environments
por: Mortimer, Peter, et al.
Publicado: (2023)
por: Mortimer, Peter, et al.
Publicado: (2023)
DiWA: Diffusion Policy Adaptation with World Models
por: Chandra, Akshay L, et al.
Publicado: (2025)
por: Chandra, Akshay L, et al.
Publicado: (2025)
ESPIRE: A Diagnostic Benchmark for Embodied Spatial Reasoning of Vision-Language Models
por: Zhao, Yanpeng, et al.
Publicado: (2026)
por: Zhao, Yanpeng, et al.
Publicado: (2026)
R3D2: Realistic 3D Asset Insertion via Diffusion for Autonomous Driving Simulation
por: Ljungbergh, William, et al.
Publicado: (2025)
por: Ljungbergh, William, et al.
Publicado: (2025)
Panoptic Perception for Autonomous Driving: A Survey
por: Li, Yunge, et al.
Publicado: (2024)
por: Li, Yunge, et al.
Publicado: (2024)
SingularTrajectory: Universal Trajectory Predictor Using Diffusion Model
por: Bae, Inhwan, et al.
Publicado: (2024)
por: Bae, Inhwan, et al.
Publicado: (2024)
Bridging Human Oversight and Black-box Driver Assistance: Vision-Language Models for Predictive Alerting in Lane Keeping Assist Systems
por: Wang, Yuhang, et al.
Publicado: (2025)
por: Wang, Yuhang, et al.
Publicado: (2025)
M2H: Multi-Task Learning with Efficient Window-Based Cross-Task Attention for Monocular Spatial Perception
por: Udugama, U. V. B. L, et al.
Publicado: (2025)
por: Udugama, U. V. B. L, et al.
Publicado: (2025)
Planning-Guided Diffusion Policy Learning for Generalizable Contact-Rich Bimanual Manipulation
por: Li, Xuanlin, et al.
Publicado: (2024)
por: Li, Xuanlin, et al.
Publicado: (2024)
TransDiffuser: Diverse Trajectory Generation with Decorrelated Multi-modal Representation for End-to-end Autonomous Driving
por: Jiang, Xuefeng, et al.
Publicado: (2025)
por: Jiang, Xuefeng, et al.
Publicado: (2025)
Ejemplares similares
-
Lexicon3D: Probing Visual Foundation Models for Complex 3D Scene Understanding
por: Man, Yunze, et al.
Publicado: (2024) -
DexHandDiff: Interaction-aware Diffusion Planning for Adaptive Dexterous Manipulation
por: Liang, Zhixuan, et al.
Publicado: (2024) -
Track, Inpaint, Resplat: Subject-driven 3D and 4D Generation with Progressive Texture Infilling
por: Zheng, Shuhong, et al.
Publicado: (2025) -
Fractional Diffusion Bridge Models
por: Nobis, Gabriel, et al.
Publicado: (2025) -
EMPERROR: A Flexible Generative Perception Error Model for Probing Self-Driving Planners
por: Hanselmann, Niklas, et al.
Publicado: (2024)