BYE: Build Your Encoder with One Sequence of Exploration Data for Long-Term Dynamic Scene Understanding
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Huang, Chenguang, Yan, Shengchao, Burgard, Wolfram |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Hierarchical Open-Vocabulary 3D Scene Graphs for Language-Grounded Robot Navigation
von: Werby, Abdelrhman, et al.
Veröffentlicht: (2024)
von: Werby, Abdelrhman, et al.
Veröffentlicht: (2024)
Multimodal Spatial Language Maps for Robot Navigation and Manipulation
von: Huang, Chenguang, et al.
Veröffentlicht: (2025)
von: Huang, Chenguang, et al.
Veröffentlicht: (2025)
LetsMap: Unsupervised Representation Learning for Semantic BEV Mapping
von: Gosala, Nikhil, et al.
Veröffentlicht: (2024)
von: Gosala, Nikhil, et al.
Veröffentlicht: (2024)
Planning with the Views via Scene Self-Exploration
von: Wang, Kangrui, et al.
Veröffentlicht: (2026)
von: Wang, Kangrui, et al.
Veröffentlicht: (2026)
Pixels-to-Graph: Real-time Integration of Building Information Models and Scene Graphs for Semantic-Geometric Human-Robot Understanding
von: Longo, Antonello, et al.
Veröffentlicht: (2025)
von: Longo, Antonello, et al.
Veröffentlicht: (2025)
Text-Scene: A Scene-to-Language Parsing Framework for 3D Scene Understanding
von: Li, Haoyuan, et al.
Veröffentlicht: (2025)
von: Li, Haoyuan, et al.
Veröffentlicht: (2025)
Articulated Object Estimation in the Wild
von: Werby, Abdelrhman, et al.
Veröffentlicht: (2025)
von: Werby, Abdelrhman, et al.
Veröffentlicht: (2025)
GLEAM: Learning Generalizable Exploration Policy for Active Mapping in Complex 3D Indoor Scenes
von: Chen, Xiao, et al.
Veröffentlicht: (2025)
von: Chen, Xiao, et al.
Veröffentlicht: (2025)
RayFronts: Open-Set Semantic Ray Frontiers for Online Scene Understanding and Exploration
von: Alama, Omar, et al.
Veröffentlicht: (2025)
von: Alama, Omar, et al.
Veröffentlicht: (2025)
SpaceTimePilot: Generative Rendering of Dynamic Scenes Across Space and Time
von: Huang, Zhening, et al.
Veröffentlicht: (2025)
von: Huang, Zhening, et al.
Veröffentlicht: (2025)
Explainable Scene Understanding with Qualitative Representations and Graph Neural Networks
von: Belmecheri, Nassim, et al.
Veröffentlicht: (2025)
von: Belmecheri, Nassim, et al.
Veröffentlicht: (2025)
Enhancing Vision-Language Models with Scene Graphs for Traffic Accident Understanding
von: Lohner, Aaron, et al.
Veröffentlicht: (2024)
von: Lohner, Aaron, et al.
Veröffentlicht: (2024)
Pixel-level Scene Understanding in One Token: Visual States Need What-is-Where Composition
von: Lee, Seokmin, et al.
Veröffentlicht: (2026)
von: Lee, Seokmin, et al.
Veröffentlicht: (2026)
DGSG-Mind: Dynamic 3D Gaussian Scene Graphs for Long-Term Scene Understanding and Grounding
von: Ge, Luzhou, et al.
Veröffentlicht: (2026)
von: Ge, Luzhou, et al.
Veröffentlicht: (2026)
SPGrasp: Spatiotemporal Prompt-driven Grasp Synthesis in Dynamic Scenes
von: Mei, Yunpeng, et al.
Veröffentlicht: (2025)
von: Mei, Yunpeng, et al.
Veröffentlicht: (2025)
Embodied Agents for Efficient Exploration and Smart Scene Description
von: Bigazzi, Roberto, et al.
Veröffentlicht: (2023)
von: Bigazzi, Roberto, et al.
Veröffentlicht: (2023)
On Deep Learning for Geometric and Semantic Scene Understanding Using On-Vehicle 3D LiDAR
von: Li, Li
Veröffentlicht: (2024)
von: Li, Li
Veröffentlicht: (2024)
Multi-modal Situated Reasoning in 3D Scenes
von: Linghu, Xiongkun, et al.
Veröffentlicht: (2024)
von: Linghu, Xiongkun, et al.
Veröffentlicht: (2024)
RoboEXP: Action-Conditioned Scene Graph via Interactive Exploration for Robotic Manipulation
von: Jiang, Hanxiao, et al.
Veröffentlicht: (2024)
von: Jiang, Hanxiao, et al.
Veröffentlicht: (2024)
CrashSight: A Phase-Aware, Infrastructure-Centric Video Benchmark for Traffic Crash Scene Understanding and Reasoning
von: Gan, Rui, et al.
Veröffentlicht: (2026)
von: Gan, Rui, et al.
Veröffentlicht: (2026)
DiWA: Diffusion Policy Adaptation with World Models
von: Chandra, Akshay L, et al.
Veröffentlicht: (2025)
von: Chandra, Akshay L, et al.
Veröffentlicht: (2025)
SceneVerse: Scaling 3D Vision-Language Learning for Grounded Scene Understanding
von: Jia, Baoxiong, et al.
Veröffentlicht: (2024)
von: Jia, Baoxiong, et al.
Veröffentlicht: (2024)
Genie 4D: Semantic-Prior-Guided 4D Dynamic Scene Reconstruction
von: Yang, Yiru, et al.
Veröffentlicht: (2026)
von: Yang, Yiru, et al.
Veröffentlicht: (2026)
EmbodiedVSR: Dynamic Scene Graph-Guided Chain-of-Thought Reasoning for Visual Spatial Tasks
von: Zhang, Yi, et al.
Veröffentlicht: (2025)
von: Zhang, Yi, et al.
Veröffentlicht: (2025)
From Scene to Object: Text-Guided Dual-Gaze Prediction
von: Ke, Zehong, et al.
Veröffentlicht: (2026)
von: Ke, Zehong, et al.
Veröffentlicht: (2026)
NuPlanQA: A Large-Scale Dataset and Benchmark for Multi-View Driving Scene Understanding in Multi-Modal Large Language Models
von: Park, Sung-Yeon, et al.
Veröffentlicht: (2025)
von: Park, Sung-Yeon, et al.
Veröffentlicht: (2025)
FunGraph: Functionality Aware 3D Scene Graphs for Language-Prompted Scene Interaction
von: Rotondi, Dennis, et al.
Veröffentlicht: (2025)
von: Rotondi, Dennis, et al.
Veröffentlicht: (2025)
SD-OVON: A Semantics-aware Dataset and Benchmark Generation Pipeline for Open-Vocabulary Object Navigation in Dynamic Scenes
von: Qiu, Dicong, et al.
Veröffentlicht: (2025)
von: Qiu, Dicong, et al.
Veröffentlicht: (2025)
Distilling Knowledge for Short-to-Long Term Trajectory Prediction
von: Das, Sourav, et al.
Veröffentlicht: (2023)
von: Das, Sourav, et al.
Veröffentlicht: (2023)
Vision-Based Autonomous UAV Navigation and Landing for Urban Search and Rescue
von: Mittal, Mayank, et al.
Veröffentlicht: (2019)
von: Mittal, Mayank, et al.
Veröffentlicht: (2019)
CoDEPS: Online Continual Learning for Depth Estimation and Panoptic Segmentation
von: Vödisch, Niclas, et al.
Veröffentlicht: (2023)
von: Vödisch, Niclas, et al.
Veröffentlicht: (2023)
VideoArtGS: Building Digital Twins of Articulated Objects from Monocular Video
von: Liu, Yu, et al.
Veröffentlicht: (2025)
von: Liu, Yu, et al.
Veröffentlicht: (2025)
Teaching Robots to Build Simulations of Themselves
von: Hu, Yuhang, et al.
Veröffentlicht: (2023)
von: Hu, Yuhang, et al.
Veröffentlicht: (2023)
MMScan: A Multi-Modal 3D Scene Dataset with Hierarchical Grounded Language Annotations
von: Lyu, Ruiyuan, et al.
Veröffentlicht: (2024)
von: Lyu, Ruiyuan, et al.
Veröffentlicht: (2024)
Application of Vision-Language Model to Pedestrians Behavior and Scene Understanding in Autonomous Driving
von: Gao, Haoxiang, et al.
Veröffentlicht: (2025)
von: Gao, Haoxiang, et al.
Veröffentlicht: (2025)
Your Robot Will Feel You Now: Empathy in Robots and Embodied Agents
von: Lim, Angelica, et al.
Veröffentlicht: (2026)
von: Lim, Angelica, et al.
Veröffentlicht: (2026)
MANSION: Multi-floor lANguage-to-3D Scene generatIOn for loNg-horizon tasks
von: Che, Lirong, et al.
Veröffentlicht: (2026)
von: Che, Lirong, et al.
Veröffentlicht: (2026)
PhyScene: Physically Interactable 3D Scene Synthesis for Embodied AI
von: Yang, Yandan, et al.
Veröffentlicht: (2024)
von: Yang, Yandan, et al.
Veröffentlicht: (2024)
SEM: Enhancing Spatial Understanding for Robust Robot Manipulation
von: Lin, Xuewu, et al.
Veröffentlicht: (2025)
von: Lin, Xuewu, et al.
Veröffentlicht: (2025)
Diffusion-guided Generalizable Enhancer for Urban Scene Reconstruction
von: Che, Henry, et al.
Veröffentlicht: (2026)
von: Che, Henry, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Hierarchical Open-Vocabulary 3D Scene Graphs for Language-Grounded Robot Navigation
von: Werby, Abdelrhman, et al.
Veröffentlicht: (2024) -
Multimodal Spatial Language Maps for Robot Navigation and Manipulation
von: Huang, Chenguang, et al.
Veröffentlicht: (2025) -
LetsMap: Unsupervised Representation Learning for Semantic BEV Mapping
von: Gosala, Nikhil, et al.
Veröffentlicht: (2024) -
Planning with the Views via Scene Self-Exploration
von: Wang, Kangrui, et al.
Veröffentlicht: (2026) -
Pixels-to-Graph: Real-time Integration of Building Information Models and Scene Graphs for Semantic-Geometric Human-Robot Understanding
von: Longo, Antonello, et al.
Veröffentlicht: (2025)