Weakly Supervised Monocular 3D Detection with a Single-View Image
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Jiang, Xueying, Jin, Sheng, Lu, Lewei, Zhang, Xiaoqin, Lu, Shijian |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
MonoMAE: Enhancing Monocular 3D Detection through Depth-Aware Masked Autoencoders
von: Jiang, Xueying, et al.
Veröffentlicht: (2024)
von: Jiang, Xueying, et al.
Veröffentlicht: (2024)
LLMs Meet VLMs: Boost Open Vocabulary Object Detection with Fine-grained Descriptors
von: Jin, Sheng, et al.
Veröffentlicht: (2024)
von: Jin, Sheng, et al.
Veröffentlicht: (2024)
Multimodal 3D Reasoning Segmentation with Complex Scenes
von: Jiang, Xueying, et al.
Veröffentlicht: (2024)
von: Jiang, Xueying, et al.
Veröffentlicht: (2024)
Exploring 3D Reasoning-Driven Planning: From Implicit Human Intentions to Route-Aware Activity Planning
von: Jiang, Xueying, et al.
Veröffentlicht: (2025)
von: Jiang, Xueying, et al.
Veröffentlicht: (2025)
STS-Mixer: Spatio-Temporal-Spectral Mixer for 4D Point Cloud Video Understanding
von: Li, Wenhao, et al.
Veröffentlicht: (2026)
von: Li, Wenhao, et al.
Veröffentlicht: (2026)
Masked AutoDecoder is Effective Multi-Task Vision Generalist
von: Qiu, Han, et al.
Veröffentlicht: (2024)
von: Qiu, Han, et al.
Veröffentlicht: (2024)
Spatial Preference Rewarding for MLLMs Spatial Understanding
von: Qiu, Han, et al.
Veröffentlicht: (2025)
von: Qiu, Han, et al.
Veröffentlicht: (2025)
MuSASplat: Efficient Sparse-View 3D Gaussian Splats via Lightweight Multi-Scale Adaptation
von: Xu, Muyu, et al.
Veröffentlicht: (2025)
von: Xu, Muyu, et al.
Veröffentlicht: (2025)
L3DR: 3D-aware LiDAR Diffusion and Rectification
von: Liu, Quan, et al.
Veröffentlicht: (2026)
von: Liu, Quan, et al.
Veröffentlicht: (2026)
A Survey of Label-Efficient Deep Learning for 3D Point Clouds
von: Xiao, Aoran, et al.
Veröffentlicht: (2023)
von: Xiao, Aoran, et al.
Veröffentlicht: (2023)
Weakly Supervised 3D Open-vocabulary Segmentation
von: Liu, Kunhao, et al.
Veröffentlicht: (2023)
von: Liu, Kunhao, et al.
Veröffentlicht: (2023)
PCR-GS: COLMAP-Free 3D Gaussian Splatting via Pose Co-Regularizations
von: Wei, Yu, et al.
Veröffentlicht: (2025)
von: Wei, Yu, et al.
Veröffentlicht: (2025)
Modeling Continuous Motion for 3D Point Cloud Object Tracking
von: Luo, Zhipeng, et al.
Veröffentlicht: (2023)
von: Luo, Zhipeng, et al.
Veröffentlicht: (2023)
Data-Efficient Generalization for Zero-shot Composed Image Retrieval
von: Chen, Zining, et al.
Veröffentlicht: (2025)
von: Chen, Zining, et al.
Veröffentlicht: (2025)
Towards Camera-Robust 3D Localization: Equation-Anchored Tool-Use for MLLMs
von: Jiang, Xueying, et al.
Veröffentlicht: (2026)
von: Jiang, Xueying, et al.
Veröffentlicht: (2026)
Learning to Prompt Segment Anything Models
von: Huang, Jiaxing, et al.
Veröffentlicht: (2024)
von: Huang, Jiaxing, et al.
Veröffentlicht: (2024)
Multi-modality Affinity Inference for Weakly Supervised 3D Semantic Segmentation
von: Li, Xiawei, et al.
Veröffentlicht: (2023)
von: Li, Xiawei, et al.
Veröffentlicht: (2023)
MVAT: Multi-View Aware Teacher for Weakly Supervised 3D Object Detection
von: Lahlali, Saad, et al.
Veröffentlicht: (2025)
von: Lahlali, Saad, et al.
Veröffentlicht: (2025)
Vision-Language Models for Vision Tasks: A Survey
von: Zhang, Jingyi, et al.
Veröffentlicht: (2023)
von: Zhang, Jingyi, et al.
Veröffentlicht: (2023)
Novel View Extrapolation with Video Diffusion Priors
von: Liu, Kunhao, et al.
Veröffentlicht: (2024)
von: Liu, Kunhao, et al.
Veröffentlicht: (2024)
PacGDC: Label-Efficient Generalizable Depth Completion with Projection Ambiguity and Consistency
von: Wang, Haotian, et al.
Veröffentlicht: (2025)
von: Wang, Haotian, et al.
Veröffentlicht: (2025)
ToDRE: Effective Visual Token Pruning via Token Diversity and Task Relevance
von: Li, Duo, et al.
Veröffentlicht: (2025)
von: Li, Duo, et al.
Veröffentlicht: (2025)
A Comprehensive Study on Visual Token Redundancy for Discrete Diffusion-based Multimodal Large Language Models
von: Li, Duo, et al.
Veröffentlicht: (2025)
von: Li, Duo, et al.
Veröffentlicht: (2025)
Self-Supervised Monocular Depth Estimation in the Dark: Towards Data Distribution Compensation
von: Yang, Haolin, et al.
Veröffentlicht: (2024)
von: Yang, Haolin, et al.
Veröffentlicht: (2024)
CA-W3D: Leveraging Context-Aware Knowledge for Weakly Supervised Monocular 3D Detection
von: Liu, Chupeng, et al.
Veröffentlicht: (2025)
von: Liu, Chupeng, et al.
Veröffentlicht: (2025)
Decoupled Pseudo-labeling for Semi-Supervised Monocular 3D Object Detection
von: Zhang, Jiacheng, et al.
Veröffentlicht: (2024)
von: Zhang, Jiacheng, et al.
Veröffentlicht: (2024)
WARM-3D: A Weakly-Supervised Sim2Real Domain Adaptation Framework for Roadside Monocular 3D Object Detection
von: Zhou, Xingcheng, et al.
Veröffentlicht: (2024)
von: Zhou, Xingcheng, et al.
Veröffentlicht: (2024)
From 2D Images to 3D Model:Weakly Supervised Multi-View Face Reconstruction with Deep Fusion
von: Zhao, Weiguang, et al.
Veröffentlicht: (2022)
von: Zhao, Weiguang, et al.
Veröffentlicht: (2022)
VirPro: Visual-referred Probabilistic Prompt Learning for Weakly-Supervised Monocular 3D Detection
von: Liu, Chupeng, et al.
Veröffentlicht: (2026)
von: Liu, Chupeng, et al.
Veröffentlicht: (2026)
Occlusion-Aware Self-Supervised Monocular Depth Estimation for Weak-Texture Endoscopic Images
von: Huang, Zebo, et al.
Veröffentlicht: (2025)
von: Huang, Zebo, et al.
Veröffentlicht: (2025)
General Geometry-aware Weakly Supervised 3D Object Detection
von: Zhang, Guowen, et al.
Veröffentlicht: (2024)
von: Zhang, Guowen, et al.
Veröffentlicht: (2024)
Stylos: Multi-View 3D Stylization with Single-Forward Gaussian Splatting
von: Liu, Hanzhou, et al.
Veröffentlicht: (2025)
von: Liu, Hanzhou, et al.
Veröffentlicht: (2025)
SOGS: Second-Order Anchor for Advanced 3D Gaussian Splatting
von: Zhang, Jiahui, et al.
Veröffentlicht: (2025)
von: Zhang, Jiahui, et al.
Veröffentlicht: (2025)
Weakly Supervised 3D Object Detection with Multi-Stage Generalization
von: He, Jiawei, et al.
Veröffentlicht: (2023)
von: He, Jiawei, et al.
Veröffentlicht: (2023)
On Robust Cross-View Consistency in Self-Supervised Monocular Depth Estimation
von: Zhao, Haimei, et al.
Veröffentlicht: (2022)
von: Zhao, Haimei, et al.
Veröffentlicht: (2022)
DivAvatar: Diverse 3D Avatar Generation with a Single Prompt
von: Tao, Weijing, et al.
Veröffentlicht: (2024)
von: Tao, Weijing, et al.
Veröffentlicht: (2024)
Weakly-Supervised Image Forgery Localization via Vision-Language Collaborative Reasoning Framework
von: Sheng, Ziqi, et al.
Veröffentlicht: (2025)
von: Sheng, Ziqi, et al.
Veröffentlicht: (2025)
Depth3DLane: Fusing Monocular 3D Lane Detection with Self-Supervised Monocular Depth Estimation
von: Hoven, Max van den, et al.
Veröffentlicht: (2025)
von: Hoven, Max van den, et al.
Veröffentlicht: (2025)
Open Vocabulary Monocular 3D Object Detection
von: Yao, Jin, et al.
Veröffentlicht: (2024)
von: Yao, Jin, et al.
Veröffentlicht: (2024)
VFMM3D: Releasing the Potential of Image by Vision Foundation Model for Monocular 3D Object Detection
von: Ding, Bonan, et al.
Veröffentlicht: (2024)
von: Ding, Bonan, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
MonoMAE: Enhancing Monocular 3D Detection through Depth-Aware Masked Autoencoders
von: Jiang, Xueying, et al.
Veröffentlicht: (2024) -
LLMs Meet VLMs: Boost Open Vocabulary Object Detection with Fine-grained Descriptors
von: Jin, Sheng, et al.
Veröffentlicht: (2024) -
Multimodal 3D Reasoning Segmentation with Complex Scenes
von: Jiang, Xueying, et al.
Veröffentlicht: (2024) -
Exploring 3D Reasoning-Driven Planning: From Implicit Human Intentions to Route-Aware Activity Planning
von: Jiang, Xueying, et al.
Veröffentlicht: (2025) -
STS-Mixer: Spatio-Temporal-Spectral Mixer for 4D Point Cloud Video Understanding
von: Li, Wenhao, et al.
Veröffentlicht: (2026)