Detect Anything 3D in the Wild
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Zhang, Hanxue, Jiang, Haoran, Yao, Qingsong, Sun, Yanan, Zhang, Renrui, Zhao, Hao, Li, Hongyang, Zhu, Hongzi, Yang, Zetong |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Test-time Correction: An Online 3D Detection System via Visual Prompting
par: Zhang, Hanxue, et autres
Publié: (2024)
par: Zhang, Hanxue, et autres
Publié: (2024)
Visual Point Cloud Forecasting enables Scalable Autonomous Driving
par: Yang, Zetong, et autres
Publié: (2023)
par: Yang, Zetong, et autres
Publié: (2023)
Decoupled Diffusion Sparks Adaptive Scene Generation
par: Zhou, Yunsong, et autres
Publié: (2025)
par: Zhou, Yunsong, et autres
Publié: (2025)
NTO3D: Neural Target Object 3D Reconstruction with Segment Anything
par: Wei, Xiaobao, et autres
Publié: (2023)
par: Wei, Xiaobao, et autres
Publié: (2023)
Perceive Anything: Recognize, Explain, Caption, and Segment Anything in Images and Videos
par: Lin, Weifeng, et autres
Publié: (2025)
par: Lin, Weifeng, et autres
Publié: (2025)
How to Efficiently Annotate Images for Best-Performing Deep Learning Based Segmentation Models: An Empirical Study with Weak and Noisy Annotations and Segment Anything Model
par: Zhang, Yixin, et autres
Publié: (2023)
par: Zhang, Yixin, et autres
Publié: (2023)
Fully Sparse 3D Occupancy Prediction
par: Liu, Haisong, et autres
Publié: (2023)
par: Liu, Haisong, et autres
Publié: (2023)
Coca-Splat: Collaborative Optimization for Camera Parameters and 3D Gaussians
par: Wu, Jiamin, et autres
Publié: (2025)
par: Wu, Jiamin, et autres
Publié: (2025)
WildDet3D: Scaling Promptable 3D Detection in the Wild
par: Huang, Weikai, et autres
Publié: (2026)
par: Huang, Weikai, et autres
Publié: (2026)
Depth Anything in $360^\circ$: Towards Scale Invariance in the Wild
par: Jiang, Hualie, et autres
Publié: (2025)
par: Jiang, Hualie, et autres
Publié: (2025)
LocateAnything3D: Vision-Language 3D Detection with Chain-of-Sight
par: Man, Yunze, et autres
Publié: (2025)
par: Man, Yunze, et autres
Publié: (2025)
CAMotion: A High-Quality Benchmark for Camouflaged Moving Object Detection in the Wild
par: Yao, Siyuan, et autres
Publié: (2026)
par: Yao, Siyuan, et autres
Publié: (2026)
Improving Distant 3D Object Detection Using 2D Box Supervision
par: Yang, Zetong, et autres
Publié: (2024)
par: Yang, Zetong, et autres
Publié: (2024)
Language-Assisted 3D Scene Understanding
par: Wu, Yanmin, et autres
Publié: (2023)
par: Wu, Yanmin, et autres
Publié: (2023)
Amodal Depth Anything: Amodal Depth Estimation in the Wild
par: Li, Zhenyu, et autres
Publié: (2024)
par: Li, Zhenyu, et autres
Publié: (2024)
Med3D-R1: Incentivizing Clinical Reasoning in 3D Medical Vision-Language Models for Abnormality Diagnosis
par: Lai, Haoran, et autres
Publié: (2026)
par: Lai, Haoran, et autres
Publié: (2026)
ETA: Efficiency through Thinking Ahead, A Dual Approach to Self-Driving with Large Models
par: Hamdan, Shadi, et autres
Publié: (2025)
par: Hamdan, Shadi, et autres
Publié: (2025)
Segment Anything in 3D with Radiance Fields
par: Cen, Jiazhong, et autres
Publié: (2023)
par: Cen, Jiazhong, et autres
Publié: (2023)
How to build the best medical image segmentation algorithm using foundation models: a comprehensive empirical study with Segment Anything Model
par: Gu, Hanxue, et autres
Publié: (2024)
par: Gu, Hanxue, et autres
Publié: (2024)
OVSeg3R: Learn Open-vocabulary Instance Segmentation from 2D via 3D Reconstruction
par: Li, Hongyang, et autres
Publié: (2025)
par: Li, Hongyang, et autres
Publié: (2025)
H3DE-Net: Efficient and Accurate 3D Landmark Detection in Medical Imaging
par: Huang, Zhen, et autres
Publié: (2025)
par: Huang, Zhen, et autres
Publié: (2025)
SAMamba3D: adapting Segment Anything for generalizable 3D segmentation of multiphase pore-scale images
par: Zhang, Rui, et autres
Publié: (2026)
par: Zhang, Rui, et autres
Publié: (2026)
STAL3D: Unsupervised Domain Adaptation for 3D Object Detection via Collaborating Self-Training and Adversarial Learning
par: Zhang, Yanan, et autres
Publié: (2024)
par: Zhang, Yanan, et autres
Publié: (2024)
MedLSAM: Localize and Segment Anything Model for 3D CT Images
par: Lei, Wenhui, et autres
Publié: (2023)
par: Lei, Wenhui, et autres
Publié: (2023)
BlinkBud: Detecting Hazards from Behind via Sampled Monocular 3D Detection on a Single Earbud
par: Li, Yunzhe, et autres
Publié: (2025)
par: Li, Yunzhe, et autres
Publié: (2025)
Judge Anything: MLLM as a Judge Across Any Modality
par: Pu, Shu, et autres
Publié: (2025)
par: Pu, Shu, et autres
Publié: (2025)
Roadside Monocular 3D Detection Prompted by 2D Detection
par: Ma, Yechi, et autres
Publié: (2024)
par: Ma, Yechi, et autres
Publié: (2024)
Adapting Segment Anything Model for Change Detection in HR Remote Sensing Images
par: Ding, Lei, et autres
Publié: (2023)
par: Ding, Lei, et autres
Publié: (2023)
Cubify Anything: Scaling Indoor 3D Object Detection
par: Lazarow, Justin, et autres
Publié: (2024)
par: Lazarow, Justin, et autres
Publié: (2024)
Multimodal OCR: Parse Anything from Documents
par: Zheng, Handong, et autres
Publié: (2026)
par: Zheng, Handong, et autres
Publié: (2026)
DragAnything: Motion Control for Anything using Entity Representation
par: Wu, Weijia, et autres
Publié: (2024)
par: Wu, Weijia, et autres
Publié: (2024)
Detect Anything via Next Point Prediction
par: Jiang, Qing, et autres
Publié: (2025)
par: Jiang, Qing, et autres
Publié: (2025)
Exploring Deeper! Segment Anything Model with Depth Perception for Camouflaged Object Detection
par: Yu, Zhenni, et autres
Publié: (2024)
par: Yu, Zhenni, et autres
Publié: (2024)
SegVGGT: Joint 3D Reconstruction and Instance Segmentation from Multi-View Images
par: Qu, Jinyuan, et autres
Publié: (2026)
par: Qu, Jinyuan, et autres
Publié: (2026)
FSD-BEV: Foreground Self-Distillation for Multi-view 3D Object Detection
par: Jiang, Zheng, et autres
Publié: (2024)
par: Jiang, Zheng, et autres
Publié: (2024)
Depth Anything with Any Prior
par: Wang, Zehan, et autres
Publié: (2025)
par: Wang, Zehan, et autres
Publié: (2025)
Personalize Anything for Free with Diffusion Transformer
par: Feng, Haoran, et autres
Publié: (2025)
par: Feng, Haoran, et autres
Publié: (2025)
Bridged Semantic Alignment for Zero-shot 3D Medical Image Diagnosis
par: Lai, Haoran, et autres
Publié: (2025)
par: Lai, Haoran, et autres
Publié: (2025)
Orient Anything: Learning Robust Object Orientation Estimation from Rendering 3D Models
par: Wang, Zehan, et autres
Publié: (2024)
par: Wang, Zehan, et autres
Publié: (2024)
GSOT3D: Towards Generic 3D Single Object Tracking in the Wild
par: Jiao, Yifan, et autres
Publié: (2024)
par: Jiao, Yifan, et autres
Publié: (2024)
Documents similaires
-
Test-time Correction: An Online 3D Detection System via Visual Prompting
par: Zhang, Hanxue, et autres
Publié: (2024) -
Visual Point Cloud Forecasting enables Scalable Autonomous Driving
par: Yang, Zetong, et autres
Publié: (2023) -
Decoupled Diffusion Sparks Adaptive Scene Generation
par: Zhou, Yunsong, et autres
Publié: (2025) -
NTO3D: Neural Target Object 3D Reconstruction with Segment Anything
par: Wei, Xiaobao, et autres
Publié: (2023) -
Perceive Anything: Recognize, Explain, Caption, and Segment Anything in Images and Videos
par: Lin, Weifeng, et autres
Publié: (2025)