Salvato in:
| Autori principali: | Guo, Xiaodong, Lin, Zi'ang, Hu, Luwen, Deng, Zhihong, Liu, Tong, Zhou, Wujie |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2506.17869 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
TUNI: Real-time RGB-T Semantic Segmentation with Unified Multi-Modal Feature Extraction and Cross-Modal Feature Fusion
di: Guo, Xiaodong, et al.
Pubblicazione: (2025)
di: Guo, Xiaodong, et al.
Pubblicazione: (2025)
Semantic Scene Segmentation for Robotics
di: Hurtado, Juana Valeria, et al.
Pubblicazione: (2024)
di: Hurtado, Juana Valeria, et al.
Pubblicazione: (2024)
Semantic Segmentation and Scene Reconstruction of RGB-D Image Frames: An End-to-End Modular Pipeline for Robotic Applications
di: Zheng, Zhiwu, et al.
Pubblicazione: (2024)
di: Zheng, Zhiwu, et al.
Pubblicazione: (2024)
TOSS: Real-time Tracking and Moving Object Segmentation for Static Scene Mapping
di: Jang, Seoyeon, et al.
Pubblicazione: (2024)
di: Jang, Seoyeon, et al.
Pubblicazione: (2024)
WildScenes: A Benchmark for 2D and 3D Semantic Segmentation in Large-scale Natural Environments
di: Vidanapathirana, Kavisha, et al.
Pubblicazione: (2023)
di: Vidanapathirana, Kavisha, et al.
Pubblicazione: (2023)
S3M: Semantic Segmentation Sparse Mapping for UAVs with RGB-D Camera
di: Canh, Thanh Nguyen, et al.
Pubblicazione: (2024)
di: Canh, Thanh Nguyen, et al.
Pubblicazione: (2024)
FeasibleCap: Real-Time Embodiment Constraint Guidance for In-the-Wild Robot Demonstration Collection
di: Yin, Zi, et al.
Pubblicazione: (2026)
di: Yin, Zi, et al.
Pubblicazione: (2026)
Real-time 3D Semantic Scene Perception for Egocentric Robots with Binocular Vision
di: Nguyen, K., et al.
Pubblicazione: (2024)
di: Nguyen, K., et al.
Pubblicazione: (2024)
Excavating in the Wild: The GOOSE-Ex Dataset for Semantic Segmentation
di: Hagmanns, Raphael, et al.
Pubblicazione: (2024)
di: Hagmanns, Raphael, et al.
Pubblicazione: (2024)
RoadFormer: Duplex Transformer for RGB-Normal Semantic Road Scene Parsing
di: Li, Jiahang, et al.
Pubblicazione: (2023)
di: Li, Jiahang, et al.
Pubblicazione: (2023)
Complementary Random Masking for RGB-Thermal Semantic Segmentation
di: Shin, Ukcheol, et al.
Pubblicazione: (2023)
di: Shin, Ukcheol, et al.
Pubblicazione: (2023)
METDrive: Multi-modal End-to-end Autonomous Driving with Temporal Guidance
di: Guo, Ziang, et al.
Pubblicazione: (2024)
di: Guo, Ziang, et al.
Pubblicazione: (2024)
VLM-Auto: VLM-based Autonomous Driving Assistant with Human-like Behavior and Understanding for Complex Road Scenes
di: Guo, Ziang, et al.
Pubblicazione: (2024)
di: Guo, Ziang, et al.
Pubblicazione: (2024)
Caltech Aerial RGB-Thermal Dataset in the Wild
di: Lee, Connor, et al.
Pubblicazione: (2024)
di: Lee, Connor, et al.
Pubblicazione: (2024)
EquiBot: SIM(3)-Equivariant Diffusion Policy for Generalizable and Data Efficient Learning
di: Yang, Jingyun, et al.
Pubblicazione: (2024)
di: Yang, Jingyun, et al.
Pubblicazione: (2024)
HawkDrive: A Transformer-driven Visual Perception System for Autonomous Driving in Night Scene
di: Guo, Ziang, et al.
Pubblicazione: (2024)
di: Guo, Ziang, et al.
Pubblicazione: (2024)
VDRive: Leveraging Reinforced VLA and Diffusion Policy for End-to-end Autonomous Driving
di: Guo, Ziang, et al.
Pubblicazione: (2025)
di: Guo, Ziang, et al.
Pubblicazione: (2025)
FEAST: A Flexible Mealtime-Assistance System Towards In-the-Wild Personalization
di: Jenamani, Rajat Kumar, et al.
Pubblicazione: (2025)
di: Jenamani, Rajat Kumar, et al.
Pubblicazione: (2025)
Privacy-Preserving Semantic Segmentation from Ultra-Low-Resolution RGB Inputs
di: Huang, Xuying, et al.
Pubblicazione: (2025)
di: Huang, Xuying, et al.
Pubblicazione: (2025)
DiffPixelFormer: Differential Pixel-Aware Transformer for RGB-D Indoor Scene Segmentation
di: Gong, Yan, et al.
Pubblicazione: (2025)
di: Gong, Yan, et al.
Pubblicazione: (2025)
SGFormer: Satellite-Ground Fusion for 3D Semantic Scene Completion
di: Guo, Xiyue, et al.
Pubblicazione: (2025)
di: Guo, Xiyue, et al.
Pubblicazione: (2025)
Wild-Drive: Off-Road Scene Captioning and Path Planning via Robust Multi-modal Routing and Efficient Large Language Model
di: Wang, Zihang, et al.
Pubblicazione: (2026)
di: Wang, Zihang, et al.
Pubblicazione: (2026)
BikeScenes: Online LiDAR Semantic Segmentation for Bicycles
di: Goren, Denniz, et al.
Pubblicazione: (2025)
di: Goren, Denniz, et al.
Pubblicazione: (2025)
Unveiling the Potential of Segment Anything Model 2 for RGB-Thermal Semantic Segmentation with Language Guidance
di: Zhao, Jiayi, et al.
Pubblicazione: (2025)
di: Zhao, Jiayi, et al.
Pubblicazione: (2025)
TA-VLA: Elucidating the Design Space of Torque-aware Vision-Language-Action Models
di: Zhang, Zongzheng, et al.
Pubblicazione: (2025)
di: Zhang, Zongzheng, et al.
Pubblicazione: (2025)
SR-SLAM: Scene-reliability Based RGB-D SLAM in Diverse Environments
di: Zhang, Haolan, et al.
Pubblicazione: (2025)
di: Zhang, Haolan, et al.
Pubblicazione: (2025)
Multi-modal NeRF Self-Supervision for LiDAR Semantic Segmentation
di: Timoneda, Xavier, et al.
Pubblicazione: (2024)
di: Timoneda, Xavier, et al.
Pubblicazione: (2024)
A Semantic Communication System for Real-time 3D Reconstruction Tasks
di: Zhang, Jiaxing, et al.
Pubblicazione: (2024)
di: Zhang, Jiaxing, et al.
Pubblicazione: (2024)
Real-time Monocular 2D and 3D Perception of Endoluminal Scenes for Controlling Flexible Robotic Endoscopic Instruments
di: Wei, Ruofeng, et al.
Pubblicazione: (2026)
di: Wei, Ruofeng, et al.
Pubblicazione: (2026)
GeomPrompt: Geometric Prompt Learning for RGB-D Semantic Segmentation Under Missing and Degraded Depth
di: Jaganathan, Krishna, et al.
Pubblicazione: (2026)
di: Jaganathan, Krishna, et al.
Pubblicazione: (2026)
Agentic Scene Policies: Unifying Space, Semantics, and Affordances for Robot Action
di: Morin, Sacha, et al.
Pubblicazione: (2025)
di: Morin, Sacha, et al.
Pubblicazione: (2025)
CSCPR: Cross-Source-Context Indoor RGB-D Place Recognition
di: Liang, Jing, et al.
Pubblicazione: (2024)
di: Liang, Jing, et al.
Pubblicazione: (2024)
CEI: A Unified Interface for Cross-Embodiment Visuomotor Policy Learning in 3D Space
di: Wu, Tong, et al.
Pubblicazione: (2026)
di: Wu, Tong, et al.
Pubblicazione: (2026)
RoboOcc: Enhancing the Geometric and Semantic Scene Understanding for Robots
di: Zhang, Zhang, et al.
Pubblicazione: (2025)
di: Zhang, Zhang, et al.
Pubblicazione: (2025)
Topological Analysis and Structural Determination of 3D Covalent Organic Frameworks
di: Zi'ang Guo, et al.
Pubblicazione: (2024)
di: Zi'ang Guo, et al.
Pubblicazione: (2024)
Efficient Multi-Task Scene Analysis with RGB-D Transformers
di: Fischedick, Söhnke Benedikt, et al.
Pubblicazione: (2023)
di: Fischedick, Söhnke Benedikt, et al.
Pubblicazione: (2023)
EWMBench: Evaluating Scene, Motion, and Semantic Quality in Embodied World Models
di: Yue, Hu, et al.
Pubblicazione: (2025)
di: Yue, Hu, et al.
Pubblicazione: (2025)
Real-time Recognition of Human Interactions from a Single RGB-D Camera for Socially-Aware Robot Navigation
di: Nguyen, Thanh Long, et al.
Pubblicazione: (2025)
di: Nguyen, Thanh Long, et al.
Pubblicazione: (2025)
Clio: Real-time Task-Driven Open-Set 3D Scene Graphs
di: Maggio, Dominic, et al.
Pubblicazione: (2024)
di: Maggio, Dominic, et al.
Pubblicazione: (2024)
MOSU: Autonomous Long-range Robot Navigation with Multi-modal Scene Understanding
di: Liang, Jing, et al.
Pubblicazione: (2025)
di: Liang, Jing, et al.
Pubblicazione: (2025)
Documenti analoghi
-
TUNI: Real-time RGB-T Semantic Segmentation with Unified Multi-Modal Feature Extraction and Cross-Modal Feature Fusion
di: Guo, Xiaodong, et al.
Pubblicazione: (2025) -
Semantic Scene Segmentation for Robotics
di: Hurtado, Juana Valeria, et al.
Pubblicazione: (2024) -
Semantic Segmentation and Scene Reconstruction of RGB-D Image Frames: An End-to-End Modular Pipeline for Robotic Applications
di: Zheng, Zhiwu, et al.
Pubblicazione: (2024) -
TOSS: Real-time Tracking and Moving Object Segmentation for Static Scene Mapping
di: Jang, Seoyeon, et al.
Pubblicazione: (2024) -
WildScenes: A Benchmark for 2D and 3D Semantic Segmentation in Large-scale Natural Environments
di: Vidanapathirana, Kavisha, et al.
Pubblicazione: (2023)