Depth Anything at Any Condition
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Sun, Boyuan, Jin, Modi, Yin, Bowen, Hou, Qibin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
See What I Mean: Aligning Vision and Language Representations for Video Fine-grained Object Understanding
von: Sun, Boyuan, et al.
Veröffentlicht: (2026)
von: Sun, Boyuan, et al.
Veröffentlicht: (2026)
LLaVA-Octopus: Unlocking Instruction-Driven Adaptive Projector Fusion for Video Understanding
von: Sun, Boyuan, et al.
Veröffentlicht: (2025)
von: Sun, Boyuan, et al.
Veröffentlicht: (2025)
LLaVA-Scissor: Token Compression with Semantic Connected Components for Video LLMs
von: Sun, Boyuan, et al.
Veröffentlicht: (2025)
von: Sun, Boyuan, et al.
Veröffentlicht: (2025)
Compress Any Segment Anything Model (SAM)
von: Fan, Juntong, et al.
Veröffentlicht: (2025)
von: Fan, Juntong, et al.
Veröffentlicht: (2025)
Describe Anything Anywhere At Any Moment
von: Gorlo, Nicolas, et al.
Veröffentlicht: (2025)
von: Gorlo, Nicolas, et al.
Veröffentlicht: (2025)
Tracking and Segmenting Anything in Any Modality
von: Zhang, Tianlu, et al.
Veröffentlicht: (2025)
von: Zhang, Tianlu, et al.
Veröffentlicht: (2025)
X-SAM: From Segment Anything to Any Segmentation
von: Wang, Hao, et al.
Veröffentlicht: (2025)
von: Wang, Hao, et al.
Veröffentlicht: (2025)
Depth Any Video with Scalable Synthetic Data
von: Yang, Honghui, et al.
Veröffentlicht: (2024)
von: Yang, Honghui, et al.
Veröffentlicht: (2024)
AnyMo: Scaling Any-Modality Conditional Motion Generation with Masked Modeling
von: Li, Yiheng, et al.
Veröffentlicht: (2026)
von: Li, Yiheng, et al.
Veröffentlicht: (2026)
Video Depth Anything: Consistent Depth Estimation for Super-Long Videos
von: Chen, Sili, et al.
Veröffentlicht: (2025)
von: Chen, Sili, et al.
Veröffentlicht: (2025)
Advancing Depth Anything Model for Unsupervised Monocular Depth Estimation in Endoscopy
von: Li, Bojian, et al.
Veröffentlicht: (2024)
von: Li, Bojian, et al.
Veröffentlicht: (2024)
Depth Any Camera: Zero-Shot Metric Depth Estimation from Any Camera
von: Guo, Yuliang, et al.
Veröffentlicht: (2025)
von: Guo, Yuliang, et al.
Veröffentlicht: (2025)
Depth Anything with Any Prior
von: Wang, Zehan, et al.
Veröffentlicht: (2025)
von: Wang, Zehan, et al.
Veröffentlicht: (2025)
Any2Caption:Interpreting Any Condition to Caption for Controllable Video Generation
von: Wu, Shengqiong, et al.
Veröffentlicht: (2025)
von: Wu, Shengqiong, et al.
Veröffentlicht: (2025)
MetricAnything: Scaling Metric Depth Pretraining with Noisy Heterogeneous Sources
von: Ma, Baorui, et al.
Veröffentlicht: (2026)
von: Ma, Baorui, et al.
Veröffentlicht: (2026)
FindAnything: Open-Vocabulary and Object-Centric Mapping for Robot Exploration in Any Environment
von: Laina, Sebastián Barbas, et al.
Veröffentlicht: (2025)
von: Laina, Sebastián Barbas, et al.
Veröffentlicht: (2025)
Any-to-Any Learning in Computational Pathology via Triplet Multimodal Pretraining
von: Sun, Qichen, et al.
Veröffentlicht: (2025)
von: Sun, Qichen, et al.
Veröffentlicht: (2025)
UnSAMv2: Self-Supervised Learning Enables Segment Anything at Any Granularity
von: Yu, Junwei, et al.
Veröffentlicht: (2025)
von: Yu, Junwei, et al.
Veröffentlicht: (2025)
DA$^{2}$: Depth Anything in Any Direction
von: Li, Haodong, et al.
Veröffentlicht: (2025)
von: Li, Haodong, et al.
Veröffentlicht: (2025)
AnyTrans: Translate AnyText in the Image with Large Scale Models
von: Qian, Zhipeng, et al.
Veröffentlicht: (2024)
von: Qian, Zhipeng, et al.
Veröffentlicht: (2024)
SAM 3D: 3Dfy Anything in Images
von: SAM 3D Team, et al.
Veröffentlicht: (2025)
von: SAM 3D Team, et al.
Veröffentlicht: (2025)
Any6D: Model-free 6D Pose Estimation of Novel Objects
von: Lee, Taeyeop, et al.
Veröffentlicht: (2025)
von: Lee, Taeyeop, et al.
Veröffentlicht: (2025)
Segment Anything in Pathology Images with Natural Language
von: Chen, Zhixuan, et al.
Veröffentlicht: (2025)
von: Chen, Zhixuan, et al.
Veröffentlicht: (2025)
ReDepth Anything: Test-Time Depth Refinement via Self-Supervised Re-lighting
von: Bhattarai, Ananta R., et al.
Veröffentlicht: (2025)
von: Bhattarai, Ananta R., et al.
Veröffentlicht: (2025)
Animate Any Character in Any World
von: Wang, Yitong, et al.
Veröffentlicht: (2025)
von: Wang, Yitong, et al.
Veröffentlicht: (2025)
AnySR: Realizing Image Super-Resolution as Any-Scale, Any-Resource
von: Zhan, Wengyi, et al.
Veröffentlicht: (2024)
von: Zhan, Wengyi, et al.
Veröffentlicht: (2024)
Composition Vision-Language Understanding via Segment and Depth Anything Model
von: Huo, Mingxiao, et al.
Veröffentlicht: (2024)
von: Huo, Mingxiao, et al.
Veröffentlicht: (2024)
Neural Gaffer: Relighting Any Object via Diffusion
von: Jin, Haian, et al.
Veröffentlicht: (2024)
von: Jin, Haian, et al.
Veröffentlicht: (2024)
Any to Full: Prompting Depth Anything for Depth Completion in One Stage
von: Zhou, Zhiyuan, et al.
Veröffentlicht: (2026)
von: Zhou, Zhiyuan, et al.
Veröffentlicht: (2026)
Segment-Anything Models Achieve Zero-shot Robustness in Autonomous Driving
von: Yan, Jun, et al.
Veröffentlicht: (2024)
von: Yan, Jun, et al.
Veröffentlicht: (2024)
Task Me Anything
von: Zhang, Jieyu, et al.
Veröffentlicht: (2024)
von: Zhang, Jieyu, et al.
Veröffentlicht: (2024)
PhyT2V: LLM-Guided Iterative Self-Refinement for Physics-Grounded Text-to-Video Generation
von: Xue, Qiyao, et al.
Veröffentlicht: (2024)
von: Xue, Qiyao, et al.
Veröffentlicht: (2024)
Describe Anything: Detailed Localized Image and Video Captioning
von: Lian, Long, et al.
Veröffentlicht: (2025)
von: Lian, Long, et al.
Veröffentlicht: (2025)
AnyTSR: Any-Scale Thermal Super-Resolution for UAV
von: Li, Mengyuan, et al.
Veröffentlicht: (2025)
von: Li, Mengyuan, et al.
Veröffentlicht: (2025)
AnyPattern: Towards In-context Image Copy Detection
von: Wang, Wenhao, et al.
Veröffentlicht: (2024)
von: Wang, Wenhao, et al.
Veröffentlicht: (2024)
Event-VStream: Event-Driven Real-Time Understanding for Long Video Streams
von: Guo, Zhenghui, et al.
Veröffentlicht: (2026)
von: Guo, Zhenghui, et al.
Veröffentlicht: (2026)
CamSAM2: Segment Anything Accurately in Camouflaged Videos
von: Zhou, Yuli, et al.
Veröffentlicht: (2025)
von: Zhou, Yuli, et al.
Veröffentlicht: (2025)
WeatherDepth: Curriculum Contrastive Learning for Self-Supervised Depth Estimation under Adverse Weather Conditions
von: Wang, Jiyuan, et al.
Veröffentlicht: (2023)
von: Wang, Jiyuan, et al.
Veröffentlicht: (2023)
SAM 3: Segment Anything with Concepts
von: Carion, Nicolas, et al.
Veröffentlicht: (2025)
von: Carion, Nicolas, et al.
Veröffentlicht: (2025)
Asynchronous Perception Machine For Efficient Test-Time-Training
von: Modi, Rajat, et al.
Veröffentlicht: (2024)
von: Modi, Rajat, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
See What I Mean: Aligning Vision and Language Representations for Video Fine-grained Object Understanding
von: Sun, Boyuan, et al.
Veröffentlicht: (2026) -
LLaVA-Octopus: Unlocking Instruction-Driven Adaptive Projector Fusion for Video Understanding
von: Sun, Boyuan, et al.
Veröffentlicht: (2025) -
LLaVA-Scissor: Token Compression with Semantic Connected Components for Video LLMs
von: Sun, Boyuan, et al.
Veröffentlicht: (2025) -
Compress Any Segment Anything Model (SAM)
von: Fan, Juntong, et al.
Veröffentlicht: (2025) -
Describe Anything Anywhere At Any Moment
von: Gorlo, Nicolas, et al.
Veröffentlicht: (2025)