Describe Anything Anywhere At Any Moment
Fuente:
arXiv
Saved in:
| Main Authors: | Gorlo, Nicolas, Schmid, Lukas, Carlone, Luca |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
FindAnything: Open-Vocabulary and Object-Centric Mapping for Robot Exploration in Any Environment
by: Laina, Sebastián Barbas, et al.
Published: (2025)
by: Laina, Sebastián Barbas, et al.
Published: (2025)
Long-Term Human Trajectory Prediction using 3D Dynamic Scene Graphs
by: Gorlo, Nicolas, et al.
Published: (2024)
by: Gorlo, Nicolas, et al.
Published: (2024)
Learning to Manipulate Anywhere: A Visual Generalizable Framework For Reinforcement Learning
by: Yuan, Zhecheng, et al.
Published: (2024)
by: Yuan, Zhecheng, et al.
Published: (2024)
Picasso: Holistic Scene Reconstruction with Physics-Constrained Sampling
by: Yu, Xihang, et al.
Published: (2026)
by: Yu, Xihang, et al.
Published: (2026)
Track Anything Rapter(TAR)
by: Puthanveettil, Tharun V., et al.
Published: (2024)
by: Puthanveettil, Tharun V., et al.
Published: (2024)
The RoboDrive Challenge: Drive Anytime Anywhere in Any Condition
by: Kong, Lingdong, et al.
Published: (2024)
by: Kong, Lingdong, et al.
Published: (2024)
Depth Any Camera: Zero-Shot Metric Depth Estimation from Any Camera
by: Guo, Yuliang, et al.
Published: (2025)
by: Guo, Yuliang, et al.
Published: (2025)
Depth Anything at Any Condition
by: Sun, Boyuan, et al.
Published: (2025)
by: Sun, Boyuan, et al.
Published: (2025)
UnSAMFlow: Unsupervised Optical Flow Guided by Segment Anything Model
by: Yuan, Shuai, et al.
Published: (2024)
by: Yuan, Shuai, et al.
Published: (2024)
AnyScene: Towards Highly Controllable Driving Scene Generation at Anywhere and Beyond
by: Zhang, Haiming, et al.
Published: (2026)
by: Zhang, Haiming, et al.
Published: (2026)
AnyPlace: Learning Generalized Object Placement for Robot Manipulation
by: Zhao, Yuchi, et al.
Published: (2025)
by: Zhao, Yuchi, et al.
Published: (2025)
DexGrasp Anything: Towards Universal Robotic Dexterous Grasping with Physics Awareness
by: Zhong, Yiming, et al.
Published: (2025)
by: Zhong, Yiming, et al.
Published: (2025)
VGGT-SLAM 2.0: Real-time Dense Feed-forward Scene Reconstruction
by: Maggio, Dominic, et al.
Published: (2026)
by: Maggio, Dominic, et al.
Published: (2026)
CUPS: Improving Human Pose-Shape Estimators with Conformalized Deep Uncertainty
by: Zhang, Harry, et al.
Published: (2024)
by: Zhang, Harry, et al.
Published: (2024)
Describe Anything: Detailed Localized Image and Video Captioning
by: Lian, Long, et al.
Published: (2025)
by: Lian, Long, et al.
Published: (2025)
AnyTraverse: An off-road traversability framework with VLM and human operator in the loop
by: Sahu, Sattwik, et al.
Published: (2025)
by: Sahu, Sattwik, et al.
Published: (2025)
Any6D: Model-free 6D Pose Estimation of Novel Objects
by: Lee, Taeyeop, et al.
Published: (2025)
by: Lee, Taeyeop, et al.
Published: (2025)
AnyTouch 2: General Optical Tactile Representation Learning For Dynamic Tactile Perception
by: Feng, Ruoxuan, et al.
Published: (2026)
by: Feng, Ruoxuan, et al.
Published: (2026)
Towards Predicting Any Human Trajectory In Context
by: Fujii, Ryo, et al.
Published: (2025)
by: Fujii, Ryo, et al.
Published: (2025)
Compress Any Segment Anything Model (SAM)
by: Fan, Juntong, et al.
Published: (2025)
by: Fan, Juntong, et al.
Published: (2025)
Gondola: Grounded Vision Language Planning for Generalizable Robotic Manipulation
by: Chen, Shizhe, et al.
Published: (2025)
by: Chen, Shizhe, et al.
Published: (2025)
Hearing Anywhere in Any Environment
by: Liu, Xiulong, et al.
Published: (2025)
by: Liu, Xiulong, et al.
Published: (2025)
Tracking and Segmenting Anything in Any Modality
by: Zhang, Tianlu, et al.
Published: (2025)
by: Zhang, Tianlu, et al.
Published: (2025)
X-SAM: From Segment Anything to Any Segmentation
by: Wang, Hao, et al.
Published: (2025)
by: Wang, Hao, et al.
Published: (2025)
Uncertainty Quantification for Visual Object Pose Estimation
by: Shaikewitz, Lorenzo, et al.
Published: (2025)
by: Shaikewitz, Lorenzo, et al.
Published: (2025)
Category-Level Object Shape and Pose Estimation in Less Than a Millisecond
by: Shaikewitz, Lorenzo, et al.
Published: (2025)
by: Shaikewitz, Lorenzo, et al.
Published: (2025)
A Deep Learning-based Pest Insect Monitoring System for Ultra-low Power Pocket-sized Drones
by: Crupi, Luca, et al.
Published: (2024)
by: Crupi, Luca, et al.
Published: (2024)
MapAnything: Universal Feed-Forward Metric 3D Reconstruction
by: Keetha, Nikhil, et al.
Published: (2025)
by: Keetha, Nikhil, et al.
Published: (2025)
AnyThermal: Towards Learning Universal Representations for Thermal Perception
by: Maheshwari, Parv, et al.
Published: (2026)
by: Maheshwari, Parv, et al.
Published: (2026)
LocateAnything: Fast and High-Quality Vision-Language Grounding with Parallel Box Decoding
by: Wang, Shihao, et al.
Published: (2026)
by: Wang, Shihao, et al.
Published: (2026)
Any4D: Unified Feed-Forward Metric 4D Reconstruction
by: Karhade, Jay, et al.
Published: (2025)
by: Karhade, Jay, et al.
Published: (2025)
Flow-guided Motion Prediction with Semantics and Dynamic Occupancy Grid Maps
by: Asghar, Rabbia, et al.
Published: (2024)
by: Asghar, Rabbia, et al.
Published: (2024)
Towards Zero-Shot Point Cloud Registration Across Diverse Scales, Scenes, and Sensor Setups
by: Lim, Hyungtae, et al.
Published: (2026)
by: Lim, Hyungtae, et al.
Published: (2026)
Multi-Model 3D Registration: Finding Multiple Moving Objects in Cluttered Point Clouds
by: Jin, David, et al.
Published: (2024)
by: Jin, David, et al.
Published: (2024)
Pandora: Articulated 3D Scene Graphs from Egocentric Vision
by: Yu, Alan, et al.
Published: (2026)
by: Yu, Alan, et al.
Published: (2026)
Anytime, Anywhere, Anyone: Investigating the Feasibility of Segment Anything Model for Crowd-Sourcing Medical Image Annotations
by: Kulkarni, Pranav, et al.
Published: (2024)
by: Kulkarni, Pranav, et al.
Published: (2024)
FlexCap: Describe Anything in Images in Controllable Detail
by: Dwibedi, Debidatta, et al.
Published: (2024)
by: Dwibedi, Debidatta, et al.
Published: (2024)
ContactHandover: Contact-Guided Robot-to-Human Object Handover
by: Wang, Zixi, et al.
Published: (2024)
by: Wang, Zixi, et al.
Published: (2024)
Segment Anything Model for automated image data annotation: empirical studies using text prompts from Grounding DINO
by: Mumuni, Fuseini, et al.
Published: (2024)
by: Mumuni, Fuseini, et al.
Published: (2024)
S.T.A.R.-Track: Latent Motion Models for End-to-End 3D Object Tracking with Adaptive Spatio-Temporal Appearance Representations
by: Doll, Simon, et al.
Published: (2023)
by: Doll, Simon, et al.
Published: (2023)
Similar Items
-
FindAnything: Open-Vocabulary and Object-Centric Mapping for Robot Exploration in Any Environment
by: Laina, Sebastián Barbas, et al.
Published: (2025) -
Long-Term Human Trajectory Prediction using 3D Dynamic Scene Graphs
by: Gorlo, Nicolas, et al.
Published: (2024) -
Learning to Manipulate Anywhere: A Visual Generalizable Framework For Reinforcement Learning
by: Yuan, Zhecheng, et al.
Published: (2024) -
Picasso: Holistic Scene Reconstruction with Physics-Constrained Sampling
by: Yu, Xihang, et al.
Published: (2026) -
Track Anything Rapter(TAR)
by: Puthanveettil, Tharun V., et al.
Published: (2024)