Resource-Efficient Affordance Grounding with Complementary Depth and Semantic Prompts
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Huang, Yizhou, Yang, Fan, Zhu, Guoliang, Li, Gen, Shi, Hao, Zuo, Yukun, Chen, Wenrui, Li, Zhiyong, Yang, Kailun |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
PanoAffordanceNet: Towards Holistic Affordance Grounding in 360° Indoor Environments
von: Zhu, Guoliang, et al.
Veröffentlicht: (2026)
von: Zhu, Guoliang, et al.
Veröffentlicht: (2026)
One-Shot Affordance Grounding of Deformable Objects in Egocentric Organizing Scenes
von: Jia, Wanjun, et al.
Veröffentlicht: (2025)
von: Jia, Wanjun, et al.
Veröffentlicht: (2025)
Multi-Keypoint Affordance Representation for Functional Dexterous Grasping
von: Yang, Fan, et al.
Veröffentlicht: (2025)
von: Yang, Fan, et al.
Veröffentlicht: (2025)
Learning Granularity-Aware Affordances from Human-Object Interaction for Tool-Based Functional Dexterous Grasping
von: Yang, Fan, et al.
Veröffentlicht: (2024)
von: Yang, Fan, et al.
Veröffentlicht: (2024)
Learning Fine-Grained Correspondence with Cross-Perspective Perception for Open-Vocabulary 6D Object Pose Estimation
von: Qin, Yu, et al.
Veröffentlicht: (2026)
von: Qin, Yu, et al.
Veröffentlicht: (2026)
PVPUFormer: Probabilistic Visual Prompt Unified Transformer for Interactive Image Segmentation
von: Zhang, Xu, et al.
Veröffentlicht: (2023)
von: Zhang, Xu, et al.
Veröffentlicht: (2023)
GenMapping: Unleashing the Potential of Inverse Perspective Mapping for Robust Online HD Map Construction
von: Li, Siyu, et al.
Veröffentlicht: (2024)
von: Li, Siyu, et al.
Veröffentlicht: (2024)
NRSeg: Noise-Resilient Learning for BEV Semantic Segmentation via Driving World Models
von: Li, Siyu, et al.
Veröffentlicht: (2025)
von: Li, Siyu, et al.
Veröffentlicht: (2025)
TS-CGNet: Temporal-Spatial Fusion Meets Centerline-Guided Diffusion for BEV Mapping
von: Hong, Xinying, et al.
Veröffentlicht: (2025)
von: Hong, Xinying, et al.
Veröffentlicht: (2025)
HierDAMap: Towards Universal Domain Adaptive BEV Mapping via Hierarchical Perspective Priors
von: Li, Siyu, et al.
Veröffentlicht: (2025)
von: Li, Siyu, et al.
Veröffentlicht: (2025)
O3N: Omnidirectional Open-Vocabulary Occupancy Prediction
von: Duan, Mengfei, et al.
Veröffentlicht: (2026)
von: Duan, Mengfei, et al.
Veröffentlicht: (2026)
Event-aided Semantic Scene Completion
von: Guo, Shangwei, et al.
Veröffentlicht: (2025)
von: Guo, Shangwei, et al.
Veröffentlicht: (2025)
Unveiling the Potential of Segment Anything Model 2 for RGB-Thermal Semantic Segmentation with Language Guidance
von: Zhao, Jiayi, et al.
Veröffentlicht: (2025)
von: Zhao, Jiayi, et al.
Veröffentlicht: (2025)
DTCLMapper: Dual Temporal Consistent Learning for Vectorized HD Map Construction
von: Li, Siyu, et al.
Veröffentlicht: (2024)
von: Li, Siyu, et al.
Veröffentlicht: (2024)
S$^3$-MonoDETR: Supervised Shape&Scale-perceptive Deformable Transformer for Monocular 3D Object Detection
von: He, Xuan, et al.
Veröffentlicht: (2023)
von: He, Xuan, et al.
Veröffentlicht: (2023)
Out-of-Distribution Semantic Occupancy Prediction
von: Zhang, Yuheng, et al.
Veröffentlicht: (2025)
von: Zhang, Yuheng, et al.
Veröffentlicht: (2025)
CoBEVMoE: Heterogeneity-aware Feature Fusion with Dynamic Mixture-of-Experts for Collaborative Perception
von: Kong, Lingzhao, et al.
Veröffentlicht: (2025)
von: Kong, Lingzhao, et al.
Veröffentlicht: (2025)
MambaMOS: LiDAR-based 3D Moving Object Segmentation with Motion-aware State Space Model
von: Zeng, Kang, et al.
Veröffentlicht: (2024)
von: Zeng, Kang, et al.
Veröffentlicht: (2024)
UniFucGrasp: Human-Hand-Inspired Unified Functional Grasp Annotation Strategy and Dataset for Diverse Dexterous Hands
von: Lin, Haoran, et al.
Veröffentlicht: (2025)
von: Lin, Haoran, et al.
Veröffentlicht: (2025)
CoBEV: Elevating Roadside 3D Object Detection with Depth and Height Complementarity
von: Shi, Hao, et al.
Veröffentlicht: (2023)
von: Shi, Hao, et al.
Veröffentlicht: (2023)
TransKD: Transformer Knowledge Distillation for Efficient Semantic Segmentation
von: Liu, Ruiping, et al.
Veröffentlicht: (2022)
von: Liu, Ruiping, et al.
Veröffentlicht: (2022)
DepTR-MOT: Unveiling the Potential of Depth-Informed Trajectory Refinement for Multi-Object Tracking
von: Deng, Buyin, et al.
Veröffentlicht: (2025)
von: Deng, Buyin, et al.
Veröffentlicht: (2025)
EgoEV-HandPose: Egocentric 3D Hand Pose Estimation and Gesture Recognition with Stereo Event Cameras
von: Wang, Luming, et al.
Veröffentlicht: (2026)
von: Wang, Luming, et al.
Veröffentlicht: (2026)
LF Tracy: A Unified Single-Pipeline Approach for Salient Object Detection in Light Field Cameras
von: Teng, Fei, et al.
Veröffentlicht: (2024)
von: Teng, Fei, et al.
Veröffentlicht: (2024)
OneOcc: Semantic Occupancy Prediction for Legged Robots with a Single Panoramic Camera
von: Shi, Hao, et al.
Veröffentlicht: (2025)
von: Shi, Hao, et al.
Veröffentlicht: (2025)
NOVA: Next-step Open-Vocabulary Autoregression for 3D Multi-Object Tracking in Autonomous Driving
von: Luo, Kai, et al.
Veröffentlicht: (2026)
von: Luo, Kai, et al.
Veröffentlicht: (2026)
Language-Driven Dual Style Mixing for Single-Domain Generalized Object Detection
von: Qin, Hongda, et al.
Veröffentlicht: (2025)
von: Qin, Hongda, et al.
Veröffentlicht: (2025)
Can we Trust Unreliable Voxels? Exploring 3D Semantic Occupancy Prediction under Label Noise
von: Li, Wenxin, et al.
Veröffentlicht: (2026)
von: Li, Wenxin, et al.
Veröffentlicht: (2026)
E-VLA: Event-Augmented Vision-Language-Action Model for Dark and Blurred Scenes
von: Zhai, Jiajun, et al.
Veröffentlicht: (2026)
von: Zhai, Jiajun, et al.
Veröffentlicht: (2026)
OccTrack360: 4D Panoptic Occupancy Tracking from Surround-View Fisheye Cameras
von: Lin, Yongzhi, et al.
Veröffentlicht: (2026)
von: Lin, Yongzhi, et al.
Veröffentlicht: (2026)
Towards Precise 3D Human Pose Estimation with Multi-Perspective Spatial-Temporal Relational Transformers
von: Jiao, Jianbin, et al.
Veröffentlicht: (2024)
von: Jiao, Jianbin, et al.
Veröffentlicht: (2024)
OAFuser: Towards Omni-Aperture Fusion for Light Field Semantic Segmentation
von: Teng, Fei, et al.
Veröffentlicht: (2023)
von: Teng, Fei, et al.
Veröffentlicht: (2023)
Panoramic Multimodal Semantic Occupancy Prediction for Quadruped Robots
von: Zhao, Guoqiang, et al.
Veröffentlicht: (2026)
von: Zhao, Guoqiang, et al.
Veröffentlicht: (2026)
Panoramic Out-of-Distribution Segmentation
von: Duan, Mengfei, et al.
Veröffentlicht: (2025)
von: Duan, Mengfei, et al.
Veröffentlicht: (2025)
Behind Every Domain There is a Shift: Adapting Distortion-aware Vision Transformers for Panoramic Semantic Segmentation
von: Zhang, Jiaming, et al.
Veröffentlicht: (2022)
von: Zhang, Jiaming, et al.
Veröffentlicht: (2022)
Hallucinating 360°: Panoramic Street-View Generation via Local Scenes Diffusion and Probabilistic Prompting
von: Teng, Fei, et al.
Veröffentlicht: (2025)
von: Teng, Fei, et al.
Veröffentlicht: (2025)
Towards Source-free Domain Adaptive Semantic Segmentation via Importance-aware and Prototype-contrast Learning
von: Cao, Yihong, et al.
Veröffentlicht: (2023)
von: Cao, Yihong, et al.
Veröffentlicht: (2023)
Offboard Occupancy Refinement with Hybrid Propagation for Autonomous Driving
von: Shi, Hao, et al.
Veröffentlicht: (2024)
von: Shi, Hao, et al.
Veröffentlicht: (2024)
Towards Single-Lens Controllable Depth-of-Field Imaging via Depth-Aware Point Spread Functions
von: Qian, Xiaolong, et al.
Veröffentlicht: (2024)
von: Qian, Xiaolong, et al.
Veröffentlicht: (2024)
$M^2$-Occ: Resilient 3D Semantic Occupancy Prediction for Autonomous Driving with Incomplete Camera Inputs
von: Lin, Kaixin, et al.
Veröffentlicht: (2026)
von: Lin, Kaixin, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
PanoAffordanceNet: Towards Holistic Affordance Grounding in 360° Indoor Environments
von: Zhu, Guoliang, et al.
Veröffentlicht: (2026) -
One-Shot Affordance Grounding of Deformable Objects in Egocentric Organizing Scenes
von: Jia, Wanjun, et al.
Veröffentlicht: (2025) -
Multi-Keypoint Affordance Representation for Functional Dexterous Grasping
von: Yang, Fan, et al.
Veröffentlicht: (2025) -
Learning Granularity-Aware Affordances from Human-Object Interaction for Tool-Based Functional Dexterous Grasping
von: Yang, Fan, et al.
Veröffentlicht: (2024) -
Learning Fine-Grained Correspondence with Cross-Perspective Perception for Open-Vocabulary 6D Object Pose Estimation
von: Qin, Yu, et al.
Veröffentlicht: (2026)