Towards Stable 3D Object Detection
Fuente:
arXiv
Salvato in:
| Autori principali: | Wang, Jiabao, Meng, Qiang, Liu, Guochao, Yan, Liujiang, Wang, Ke, Cheng, Ming-Ming, Hou, Qibin |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
OPUS: Occupancy Prediction Using a Sparse Set
di: Wang, Jiabao, et al.
Pubblicazione: (2024)
di: Wang, Jiabao, et al.
Pubblicazione: (2024)
CrossKD: Cross-Head Knowledge Distillation for Object Detection
di: Wang, Jiabao, et al.
Pubblicazione: (2023)
di: Wang, Jiabao, et al.
Pubblicazione: (2023)
YOLO-MS: Rethinking Multi-Scale Representation Learning for Real-time Object Detection
di: Chen, Yuming, et al.
Pubblicazione: (2023)
di: Chen, Yuming, et al.
Pubblicazione: (2023)
Zone Evaluation: Revealing Spatial Bias in Object Detection
di: Zheng, Zhaohui, et al.
Pubblicazione: (2023)
di: Zheng, Zhaohui, et al.
Pubblicazione: (2023)
Referring Camouflaged Object Detection
di: Zhang, Xuying, et al.
Pubblicazione: (2023)
di: Zhang, Xuying, et al.
Pubblicazione: (2023)
Small, Versatile and Mighty: A Range-View Perception Framework
di: Meng, Qiang, et al.
Pubblicazione: (2024)
di: Meng, Qiang, et al.
Pubblicazione: (2024)
AR-1-to-3: Single Image to Consistent 3D Object Generation via Next-View Prediction
di: Zhang, Xuying, et al.
Pubblicazione: (2025)
di: Zhang, Xuying, et al.
Pubblicazione: (2025)
Rethinking RGB-D Salient Object Detection: Models, Data Sets, and Large-Scale Benchmarks
di: Fan, Deng-Ping, et al.
Pubblicazione: (2019)
di: Fan, Deng-Ping, et al.
Pubblicazione: (2019)
Strip R-CNN: Large Strip Convolution for Remote Sensing Object Detection
di: Yuan, Xinbin, et al.
Pubblicazione: (2025)
di: Yuan, Xinbin, et al.
Pubblicazione: (2025)
SM3Det: A Unified Model for Multi-Modal Remote Sensing Object Detection
di: Li, Yuxuan, et al.
Pubblicazione: (2024)
di: Li, Yuxuan, et al.
Pubblicazione: (2024)
SARDet-100K: Towards Open-Source Benchmark and ToolKit for Large-Scale SAR Object Detection
di: Li, Yuxuan, et al.
Pubblicazione: (2024)
di: Li, Yuxuan, et al.
Pubblicazione: (2024)
Towards Universal Video MLLMs with Attribute-Structured and Quality-Verified Instructions
di: Li, Yunheng, et al.
Pubblicazione: (2026)
di: Li, Yunheng, et al.
Pubblicazione: (2026)
TAR3D: Creating High-Quality 3D Assets via Next-Part Prediction
di: Zhang, Xuying, et al.
Pubblicazione: (2024)
di: Zhang, Xuying, et al.
Pubblicazione: (2024)
Towards RAW Object Detection in Diverse Conditions
di: Li, Zhong-Yu, et al.
Pubblicazione: (2024)
di: Li, Zhong-Yu, et al.
Pubblicazione: (2024)
Unbiased Region-Language Alignment for Open-Vocabulary Dense Prediction
di: Li, Yunheng, et al.
Pubblicazione: (2024)
di: Li, Yunheng, et al.
Pubblicazione: (2024)
StoryDiffusion: Consistent Self-Attention for Long-Range Image and Video Generation
di: Zhou, Yupeng, et al.
Pubblicazione: (2024)
di: Zhou, Yupeng, et al.
Pubblicazione: (2024)
DFormerv2: Geometry Self-Attention for RGBD Semantic Segmentation
di: Yin, Bo-Wen, et al.
Pubblicazione: (2025)
di: Yin, Bo-Wen, et al.
Pubblicazione: (2025)
Mutual Forcing: Dual-Mode Self-Evolution for Fast Autoregressive Audio-Video Character Generation
di: Zhou, Yupeng, et al.
Pubblicazione: (2026)
di: Zhou, Yupeng, et al.
Pubblicazione: (2026)
Re-Aligning Language to Visual Objects with an Agentic Workflow
di: Chen, Yuming, et al.
Pubblicazione: (2025)
di: Chen, Yuming, et al.
Pubblicazione: (2025)
DFormer: Rethinking RGBD Representation Learning for Semantic Segmentation
di: Yin, Bowen, et al.
Pubblicazione: (2023)
di: Yin, Bowen, et al.
Pubblicazione: (2023)
GeoWorld: Unlocking the Potential of Geometry Models to Facilitate High-Fidelity 3D Scene Generation
di: Wan, Yuhao, et al.
Pubblicazione: (2025)
di: Wan, Yuhao, et al.
Pubblicazione: (2025)
Cascade-CLIP: Cascaded Vision-Language Embeddings Alignment for Zero-Shot Semantic Segmentation
di: Li, Yunheng, et al.
Pubblicazione: (2024)
di: Li, Yunheng, et al.
Pubblicazione: (2024)
SRFormerV2: Taking a Closer Look at Permuted Self-Attention for Image Super-Resolution
di: Zhou, Yupeng, et al.
Pubblicazione: (2023)
di: Zhou, Yupeng, et al.
Pubblicazione: (2023)
Concealed Object Detection
di: Fan, Deng-Ping, et al.
Pubblicazione: (2021)
di: Fan, Deng-Ping, et al.
Pubblicazione: (2021)
Sora Generates Videos with Stunning Geometrical Consistency
di: Li, Xuanyi, et al.
Pubblicazione: (2024)
di: Li, Xuanyi, et al.
Pubblicazione: (2024)
Revisiting Efficient Semantic Segmentation: Learning Offsets for Better Spatial and Class Feature Alignment
di: Zhang, Shi-Chen, et al.
Pubblicazione: (2025)
di: Zhang, Shi-Chen, et al.
Pubblicazione: (2025)
OpenM3D: Open Vocabulary Multi-view Indoor 3D Object Detection without Human Annotations
di: Hsu, Peng-Hao, et al.
Pubblicazione: (2025)
di: Hsu, Peng-Hao, et al.
Pubblicazione: (2025)
Rethinking Token-Level Policy Optimization for Multimodal Chain-of-Thought
di: Li, Yunheng, et al.
Pubblicazione: (2026)
di: Li, Yunheng, et al.
Pubblicazione: (2026)
TempSamp-R1: Effective Temporal Sampling with Reinforcement Fine-Tuning for Video LLMs
di: Li, Yunheng, et al.
Pubblicazione: (2025)
di: Li, Yunheng, et al.
Pubblicazione: (2025)
RGB-D Indiscernible Object Counting in Underwater Scenes
di: Sun, Guolei, et al.
Pubblicazione: (2023)
di: Sun, Guolei, et al.
Pubblicazione: (2023)
HazyDet: Open-Source Benchmark for Drone-View Object Detection with Depth-Cues in Hazy Scenes
di: Feng, Changfeng, et al.
Pubblicazione: (2024)
di: Feng, Changfeng, et al.
Pubblicazione: (2024)
High-Quality Mask Tuning Matters for Open-Vocabulary Segmentation
di: Zeng, Quan-Sheng, et al.
Pubblicazione: (2024)
di: Zeng, Quan-Sheng, et al.
Pubblicazione: (2024)
Traffic Scene Parsing through the TSP6K Dataset
di: Jiang, Peng-Tao, et al.
Pubblicazione: (2023)
di: Jiang, Peng-Tao, et al.
Pubblicazione: (2023)
Future Does Matter: Boosting 3D Object Detection with Temporal Motion Estimation in Point Cloud Sequences
di: Yu, Rui, et al.
Pubblicazione: (2024)
di: Yu, Rui, et al.
Pubblicazione: (2024)
OpenAD: Open-World Autonomous Driving Benchmark for 3D Object Detection
di: Xia, Zhongyu, et al.
Pubblicazione: (2024)
di: Xia, Zhongyu, et al.
Pubblicazione: (2024)
A Glimpse to Compress: Dynamic Visual Token Pruning for Large Vision-Language Models
di: Zeng, Quan-Sheng, et al.
Pubblicazione: (2025)
di: Zeng, Quan-Sheng, et al.
Pubblicazione: (2025)
Mixture of Style Experts for Diverse Image Stylization
di: Zhu, Shihao, et al.
Pubblicazione: (2026)
di: Zhu, Shihao, et al.
Pubblicazione: (2026)
The Consistency Critic: Correcting Inconsistencies in Generated Images via Reference-Guided Attentive Alignment
di: Ouyang, Ziheng, et al.
Pubblicazione: (2025)
di: Ouyang, Ziheng, et al.
Pubblicazione: (2025)
KAC: Kolmogorov-Arnold Classifier for Continual Learning
di: Hu, Yusong, et al.
Pubblicazione: (2025)
di: Hu, Yusong, et al.
Pubblicazione: (2025)
Every Dataset Counts: Scaling up Monocular 3D Object Detection with Joint Datasets Training
di: Ma, Fulong, et al.
Pubblicazione: (2023)
di: Ma, Fulong, et al.
Pubblicazione: (2023)
Documenti analoghi
-
OPUS: Occupancy Prediction Using a Sparse Set
di: Wang, Jiabao, et al.
Pubblicazione: (2024) -
CrossKD: Cross-Head Knowledge Distillation for Object Detection
di: Wang, Jiabao, et al.
Pubblicazione: (2023) -
YOLO-MS: Rethinking Multi-Scale Representation Learning for Real-time Object Detection
di: Chen, Yuming, et al.
Pubblicazione: (2023) -
Zone Evaluation: Revealing Spatial Bias in Object Detection
di: Zheng, Zhaohui, et al.
Pubblicazione: (2023) -
Referring Camouflaged Object Detection
di: Zhang, Xuying, et al.
Pubblicazione: (2023)