On the Feasibility and Opportunity of Autoregressive 3D Object Detection
Fuente:
arXiv
Saved in:
| Main Authors: | Huang, Zanming, Yoo, Jinsu, Jeon, Sooyoung, Liu, Zhenzhen, Campbell, Mark, Weinberger, Kilian Q, Hariharan, Bharath, Chao, Wei-Lun, Luo, Katie Z |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Transfer Your Perspective: Controllable 3D Generation from Any Viewpoint in a Driving Scene
by: Pan, Tai-Yu, et al.
Published: (2025)
by: Pan, Tai-Yu, et al.
Published: (2025)
When the City Teaches the Car: Label-Free 3D Perception from Infrastructure
by: Xu, Zhen, et al.
Published: (2026)
by: Xu, Zhen, et al.
Published: (2026)
Leveraging Sparse LiDAR for RAFT-Stereo: A Depth Pre-Fill Perspective
by: Yoo, Jinsu, et al.
Published: (2025)
by: Yoo, Jinsu, et al.
Published: (2025)
DiffuBox: Refining 3D Object Detection with Point Diffusion
by: Chen, Xiangyu, et al.
Published: (2024)
by: Chen, Xiangyu, et al.
Published: (2024)
Pre-Training LiDAR-Based 3D Object Detectors Through Colorization
by: Pan, Tai-Yu, et al.
Published: (2023)
by: Pan, Tai-Yu, et al.
Published: (2023)
Better Monocular 3D Detectors with LiDAR from the Past
by: You, Yurong, et al.
Published: (2024)
by: You, Yurong, et al.
Published: (2024)
Learning 3D Perception from Others' Predictions
by: Yoo, Jinsu, et al.
Published: (2024)
by: Yoo, Jinsu, et al.
Published: (2024)
Mixed Signals: A Diverse Point Cloud Dataset for Heterogeneous LiDAR V2X Collaboration
by: Luo, Katie Z, et al.
Published: (2025)
by: Luo, Katie Z, et al.
Published: (2025)
Detecting Out-of-Distribution Objects through Class-Conditioned Inpainting
by: Nguyen, Quang-Huy, et al.
Published: (2024)
by: Nguyen, Quang-Huy, et al.
Published: (2024)
MOD-UV: Learning Mobile Object Detectors from Unlabeled Videos
by: Sun, Yihong, et al.
Published: (2024)
by: Sun, Yihong, et al.
Published: (2024)
Orchestrating LLMs with Different Personalizations
by: Zhou, Jin Peng, et al.
Published: (2024)
by: Zhou, Jin Peng, et al.
Published: (2024)
Correction with Backtracking Reduces Hallucination in Summarization
by: Liu, Zhenzhen, et al.
Published: (2023)
by: Liu, Zhenzhen, et al.
Published: (2023)
Tracking and Understanding Object Transformations
by: Sun, Yihong, et al.
Published: (2025)
by: Sun, Yihong, et al.
Published: (2025)
Music Transcription with (Almost) No Supervision
by: Shin, Saebyeol, et al.
Published: (2026)
by: Shin, Saebyeol, et al.
Published: (2026)
FRED: Towards a Full Rotation-Equivariance in Aerial Image Object Detection
by: Lee, Chanho, et al.
Published: (2023)
by: Lee, Chanho, et al.
Published: (2023)
Counter-Current Learning: A Biologically Plausible Dual Network Approach for Deep Learning
by: Kao, Chia-Hsiang, et al.
Published: (2024)
by: Kao, Chia-Hsiang, et al.
Published: (2024)
Benchmark Datasets for Lead-Lag Forecasting on Social Platforms
by: Kazemian, Kimia, et al.
Published: (2025)
by: Kazemian, Kimia, et al.
Published: (2025)
IncDSI: Incrementally Updatable Document Retrieval
by: Kishore, Varsha, et al.
Published: (2023)
by: Kishore, Varsha, et al.
Published: (2023)
Continual Unlearning for Text-to-Image Diffusion Models: A Regularization Perspective
by: Lee, Justin, et al.
Published: (2025)
by: Lee, Justin, et al.
Published: (2025)
ObjectCarver: Semi-automatic segmentation, reconstruction and separation of 3D objects
by: Hassena, Gemmechu, et al.
Published: (2024)
by: Hassena, Gemmechu, et al.
Published: (2024)
Denoising Vision Transformers
by: Yang, Jiawei, et al.
Published: (2024)
by: Yang, Jiawei, et al.
Published: (2024)
Diffusion Guided Language Modeling
by: Lovelace, Justin, et al.
Published: (2024)
by: Lovelace, Justin, et al.
Published: (2024)
Learning from Synthetic Data Improves Multi-hop Reasoning
by: Kabra, Anmol, et al.
Published: (2026)
by: Kabra, Anmol, et al.
Published: (2026)
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models
by: Mai, Zheda, et al.
Published: (2025)
by: Mai, Zheda, et al.
Published: (2025)
C3Po: Cross-View Cross-Modality Correspondence by Pointmap Prediction
by: Huang, Kuan Wei, et al.
Published: (2025)
by: Huang, Kuan Wei, et al.
Published: (2025)
Composing People Together: Iterative Pose-Image Generation for Multi-Person Interaction Scenes
by: Peng, Wenxuan, et al.
Published: (2026)
by: Peng, Wenxuan, et al.
Published: (2026)
Adaptive Moments are Surprisingly Effective for Plug-and-Play Diffusion Sampling
by: Belardi, Christian, et al.
Published: (2026)
by: Belardi, Christian, et al.
Published: (2026)
Marriage as a legal gateway to care: Legal entitlement and gender inequality in the 2020 and 2025 Korean time use survey
by: Sooyoung Kim
Published: (2026)
by: Sooyoung Kim
Published: (2026)
$π$-CoT: Prolog-Initialized Chain-of-Thought Prompting for Multi-Hop Question-Answering
by: Wan, Chao, et al.
Published: (2025)
by: Wan, Chao, et al.
Published: (2025)
Learning Feature Descriptors using Camera Pose Supervision
by: Wang, Qianqian, et al.
Published: (2020)
by: Wang, Qianqian, et al.
Published: (2020)
RanDeS: Randomized Delta Superposition for Multi-Model Compression
by: Zhou, Hangyu, et al.
Published: (2025)
by: Zhou, Hangyu, et al.
Published: (2025)
Prescriptive Scaling Laws for Data Constrained Training
by: Lovelace, Justin, et al.
Published: (2026)
by: Lovelace, Justin, et al.
Published: (2026)
Sample-Efficient Diffusion for Text-To-Speech Synthesis
by: Lovelace, Justin, et al.
Published: (2024)
by: Lovelace, Justin, et al.
Published: (2024)
INPROVF: Leveraging Large Language Models to Repair High-level Robot Controllers from Assumption Violations
by: Meng, Qian, et al.
Published: (2025)
by: Meng, Qian, et al.
Published: (2025)
Learning to decode logical circuits
by: Zhou, Yiqing, et al.
Published: (2025)
by: Zhou, Yiqing, et al.
Published: (2025)
Probabilistic Uncertainty Quantification of Prediction Models with Application to Visual Localization
by: Chen, Junan, et al.
Published: (2023)
by: Chen, Junan, et al.
Published: (2023)
Lessons and Open Questions from a Unified Study of Camera-Trap Species Recognition Over Time
by: Jeon, Sooyoung, et al.
Published: (2026)
by: Jeon, Sooyoung, et al.
Published: (2026)
AesFA: An Aesthetic Feature-Aware Arbitrary Neural Style Transfer
by: Kwon, Joonwoo, et al.
Published: (2023)
by: Kwon, Joonwoo, et al.
Published: (2023)
SpeechOp: Inference-Time Task Composition for Generative Speech Processing
by: Lovelace, Justin, et al.
Published: (2025)
by: Lovelace, Justin, et al.
Published: (2025)
Stop-Think-AutoRegress: Language Modeling with Latent Diffusion Planning
by: Lovelace, Justin, et al.
Published: (2026)
by: Lovelace, Justin, et al.
Published: (2026)
Similar Items
-
Transfer Your Perspective: Controllable 3D Generation from Any Viewpoint in a Driving Scene
by: Pan, Tai-Yu, et al.
Published: (2025) -
When the City Teaches the Car: Label-Free 3D Perception from Infrastructure
by: Xu, Zhen, et al.
Published: (2026) -
Leveraging Sparse LiDAR for RAFT-Stereo: A Depth Pre-Fill Perspective
by: Yoo, Jinsu, et al.
Published: (2025) -
DiffuBox: Refining 3D Object Detection with Point Diffusion
by: Chen, Xiangyu, et al.
Published: (2024) -
Pre-Training LiDAR-Based 3D Object Detectors Through Colorization
by: Pan, Tai-Yu, et al.
Published: (2023)