DiffuBox: Refining 3D Object Detection with Point Diffusion
Fuente:
arXiv
Saved in:
| Main Authors: | Chen, Xiangyu, Liu, Zhenzhen, Luo, Katie Z, Datta, Siddhartha, Polavaram, Adhitya, Wang, Yan, You, Yurong, Li, Boyi, Pavone, Marco, Chao, Wei-Lun, Campbell, Mark, Hariharan, Bharath, Weinberger, Kilian Q. |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
On the Feasibility and Opportunity of Autoregressive 3D Object Detection
by: Huang, Zanming, et al.
Published: (2026)
by: Huang, Zanming, et al.
Published: (2026)
Pre-Training LiDAR-Based 3D Object Detectors Through Colorization
by: Pan, Tai-Yu, et al.
Published: (2023)
by: Pan, Tai-Yu, et al.
Published: (2023)
Better Monocular 3D Detectors with LiDAR from the Past
by: You, Yurong, et al.
Published: (2024)
by: You, Yurong, et al.
Published: (2024)
Stop-Think-AutoRegress: Language Modeling with Latent Diffusion Planning
by: Lovelace, Justin, et al.
Published: (2026)
by: Lovelace, Justin, et al.
Published: (2026)
Mixed Signals: A Diverse Point Cloud Dataset for Heterogeneous LiDAR V2X Collaboration
by: Luo, Katie Z, et al.
Published: (2025)
by: Luo, Katie Z, et al.
Published: (2025)
Learning 3D Perception from Others' Predictions
by: Yoo, Jinsu, et al.
Published: (2024)
by: Yoo, Jinsu, et al.
Published: (2024)
Transfer Your Perspective: Controllable 3D Generation from Any Viewpoint in a Driving Scene
by: Pan, Tai-Yu, et al.
Published: (2025)
by: Pan, Tai-Yu, et al.
Published: (2025)
MOD-UV: Learning Mobile Object Detectors from Unlabeled Videos
by: Sun, Yihong, et al.
Published: (2024)
by: Sun, Yihong, et al.
Published: (2024)
Online Feature Updates Improve Online (Generalized) Label Shift Adaptation
by: Wu, Ruihan, et al.
Published: (2024)
by: Wu, Ruihan, et al.
Published: (2024)
When the City Teaches the Car: Label-Free 3D Perception from Infrastructure
by: Xu, Zhen, et al.
Published: (2026)
by: Xu, Zhen, et al.
Published: (2026)
Detecting Out-of-Distribution Objects through Class-Conditioned Inpainting
by: Nguyen, Quang-Huy, et al.
Published: (2024)
by: Nguyen, Quang-Huy, et al.
Published: (2024)
Tracking and Understanding Object Transformations
by: Sun, Yihong, et al.
Published: (2025)
by: Sun, Yihong, et al.
Published: (2025)
Generate, Transfer, Adapt: Learning Functional Dexterous Grasping from a Single Human Demonstration
by: He, Xingyi, et al.
Published: (2026)
by: He, Xingyi, et al.
Published: (2026)
Orchestrating LLMs with Different Personalizations
by: Zhou, Jin Peng, et al.
Published: (2024)
by: Zhou, Jin Peng, et al.
Published: (2024)
Numerical Words and Linguistic Loops: The Perpetual Four-Letter Routine
by: Polavaram, Krishna Chaitanya
Published: (2025)
by: Polavaram, Krishna Chaitanya
Published: (2025)
Correction with Backtracking Reduces Hallucination in Summarization
by: Liu, Zhenzhen, et al.
Published: (2023)
by: Liu, Zhenzhen, et al.
Published: (2023)
Counter-Current Learning: A Biologically Plausible Dual Network Approach for Deep Learning
by: Kao, Chia-Hsiang, et al.
Published: (2024)
by: Kao, Chia-Hsiang, et al.
Published: (2024)
Extrapolated Urban View Synthesis Benchmark
by: Han, Xiangyu, et al.
Published: (2024)
by: Han, Xiangyu, et al.
Published: (2024)
DiffuReason: Bridging Latent Reasoning and Generative Refinement for Sequential Recommendation
by: Jiang, Jie, et al.
Published: (2026)
by: Jiang, Jie, et al.
Published: (2026)
Musical Cities
by: Adhitya, Sara
Published: (2018)
by: Adhitya, Sara
Published: (2018)
geoyanzhan3/ThermoDiffuDA: ThermoDiffuDA
by: Yan Zhan
Published: (2026)
by: Yan Zhan
Published: (2026)
Music Transcription with (Almost) No Supervision
by: Shin, Saebyeol, et al.
Published: (2026)
by: Shin, Saebyeol, et al.
Published: (2026)
DiffuMatting: Synthesizing Arbitrary Objects with Matting-level Annotation
by: Hu, Xiaobin, et al.
Published: (2024)
by: Hu, Xiaobin, et al.
Published: (2024)
Diffusion Guided Language Modeling
by: Lovelace, Justin, et al.
Published: (2024)
by: Lovelace, Justin, et al.
Published: (2024)
Benchmark Datasets for Lead-Lag Forecasting on Social Platforms
by: Kazemian, Kimia, et al.
Published: (2025)
by: Kazemian, Kimia, et al.
Published: (2025)
Towards Efficient and Effective Multi-Camera Encoding for End-to-End Driving
by: Yang, Jiawei, et al.
Published: (2025)
by: Yang, Jiawei, et al.
Published: (2025)
DreamDrive: Generative 4D Scene Modeling from Street View Images
by: Mao, Jiageng, et al.
Published: (2024)
by: Mao, Jiageng, et al.
Published: (2024)
Composing People Together: Iterative Pose-Image Generation for Multi-Person Interaction Scenes
by: Peng, Wenxuan, et al.
Published: (2026)
by: Peng, Wenxuan, et al.
Published: (2026)
Adaptive Moments are Surprisingly Effective for Plug-and-Play Diffusion Sampling
by: Belardi, Christian, et al.
Published: (2026)
by: Belardi, Christian, et al.
Published: (2026)
Denoising Vision Transformers
by: Yang, Jiawei, et al.
Published: (2024)
by: Yang, Jiawei, et al.
Published: (2024)
Efficient Multi-Camera Tokenization with Triplanes for End-to-End Driving
by: Ivanovic, Boris, et al.
Published: (2025)
by: Ivanovic, Boris, et al.
Published: (2025)
Accelerating Structured Chain-of-Thought in Autonomous Vehicles
by: Gu, Yi, et al.
Published: (2026)
by: Gu, Yi, et al.
Published: (2026)
Learning Feature Descriptors using Camera Pose Supervision
by: Wang, Qianqian, et al.
Published: (2020)
by: Wang, Qianqian, et al.
Published: (2020)
RanDeS: Randomized Delta Superposition for Multi-Model Compression
by: Zhou, Hangyu, et al.
Published: (2025)
by: Zhou, Hangyu, et al.
Published: (2025)
ObjectCarver: Semi-automatic segmentation, reconstruction and separation of 3D objects
by: Hassena, Gemmechu, et al.
Published: (2024)
by: Hassena, Gemmechu, et al.
Published: (2024)
Learning from Synthetic Data Improves Multi-hop Reasoning
by: Kabra, Anmol, et al.
Published: (2026)
by: Kabra, Anmol, et al.
Published: (2026)
Prescriptive Scaling Laws for Data Constrained Training
by: Lovelace, Justin, et al.
Published: (2026)
by: Lovelace, Justin, et al.
Published: (2026)
Sample-Efficient Diffusion for Text-To-Speech Synthesis
by: Lovelace, Justin, et al.
Published: (2024)
by: Lovelace, Justin, et al.
Published: (2024)
IncDSI: Incrementally Updatable Document Retrieval
by: Kishore, Varsha, et al.
Published: (2023)
by: Kishore, Varsha, et al.
Published: (2023)
INPROVF: Leveraging Large Language Models to Repair High-level Robot Controllers from Assumption Violations
by: Meng, Qian, et al.
Published: (2025)
by: Meng, Qian, et al.
Published: (2025)
Similar Items
-
On the Feasibility and Opportunity of Autoregressive 3D Object Detection
by: Huang, Zanming, et al.
Published: (2026) -
Pre-Training LiDAR-Based 3D Object Detectors Through Colorization
by: Pan, Tai-Yu, et al.
Published: (2023) -
Better Monocular 3D Detectors with LiDAR from the Past
by: You, Yurong, et al.
Published: (2024) -
Stop-Think-AutoRegress: Language Modeling with Latent Diffusion Planning
by: Lovelace, Justin, et al.
Published: (2026) -
Mixed Signals: A Diverse Point Cloud Dataset for Heterogeneous LiDAR V2X Collaboration
by: Luo, Katie Z, et al.
Published: (2025)