RGB-D Video Object Segmentation via Enhanced Multi-store Feature Memory
Fuente:
arXiv
Saved in:
| Main Authors: | Xu, Boyue, Hou, Ruichao, Ren, Tongwei, Wu, Gangshan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
RGB-D Tracking via Hierarchical Modality Aggregation and Distribution Network
by: Xu, Boyue, et al.
Published: (2025)
by: Xu, Boyue, et al.
Published: (2025)
Learning Frequency and Memory-Aware Prompts for Multi-Modal Object Tracking
by: Xu, Boyue, et al.
Published: (2025)
by: Xu, Boyue, et al.
Published: (2025)
VL-UniTrack: A Unified Framework with Visual-Language Prompts for UAV-Ground Visual Tracking
by: Xu, Boyue, et al.
Published: (2026)
by: Xu, Boyue, et al.
Published: (2026)
MTNet: Learning modality-aware representation with transformer for RGBT tracking
by: Hou, Ruichao, et al.
Published: (2025)
by: Hou, Ruichao, et al.
Published: (2025)
KAN-SAM: Kolmogorov-Arnold Network Guided Segment Anything Model for RGB-T Salient Object Detection
by: Li, Xingyuan, et al.
Published: (2025)
by: Li, Xingyuan, et al.
Published: (2025)
SwiTrack: Tri-State Switch for Cross-Modal Object Tracking
by: Xu, Boyue, et al.
Published: (2025)
by: Xu, Boyue, et al.
Published: (2025)
HyPSAM: Hybrid Prompt-driven Segment Anything Model for RGB-Thermal Salient Object Detection
by: Hou, Ruichao, et al.
Published: (2025)
by: Hou, Ruichao, et al.
Published: (2025)
Spatial-Temporal Human-Object Interaction Detection
by: Sun, Xu, et al.
Published: (2025)
by: Sun, Xu, et al.
Published: (2025)
Joint Modeling of Feature, Correspondence, and a Compressed Memory for Video Object Segmentation
by: Zhang, Jiaming, et al.
Published: (2023)
by: Zhang, Jiaming, et al.
Published: (2023)
Visual Object Tracking on Multi-modal RGB-D Videos: A Review
by: Zhu, Xue-Feng, et al.
Published: (2022)
by: Zhu, Xue-Feng, et al.
Published: (2022)
Salient Object Detection in RGB-D Videos
by: Mou, Ao, et al.
Published: (2023)
by: Mou, Ao, et al.
Published: (2023)
A Saliency Enhanced Feature Fusion based multiscale RGB-D Salient Object Detection Network
by: Huang, Rui, et al.
Published: (2024)
by: Huang, Rui, et al.
Published: (2024)
M$^4$-SAM: Multi-Modal Mixture-of-Experts with Memory-Augmented SAM for RGB-D Video Salient Object Detection
by: Liu, Jiyuan, et al.
Published: (2026)
by: Liu, Jiyuan, et al.
Published: (2026)
Enhancing Video Object Segmentation in TrackRAD Using XMem Memory Network
by: Deng, Pengchao, et al.
Published: (2025)
by: Deng, Pengchao, et al.
Published: (2025)
Efficient Video Object Segmentation via Modulated Cross-Attention Memory
by: Shaker, Abdelrahman, et al.
Published: (2024)
by: Shaker, Abdelrahman, et al.
Published: (2024)
ObjFiller3D: Scaling 3D Object Inpainting to Dense Multi-View Consistency
by: Feng, Haitang, et al.
Published: (2025)
by: Feng, Haitang, et al.
Published: (2025)
Selective Complementary Feature Fusion and Modal Feature Compression Interaction for Brain Tumor Segmentation
by: Chen, Dong, et al.
Published: (2025)
by: Chen, Dong, et al.
Published: (2025)
Glass Surface Segmentation with an RGB-D Camera via Weighted Feature Fusion for Service Robots
by: Lin, Henghong, et al.
Published: (2025)
by: Lin, Henghong, et al.
Published: (2025)
SAM-DAQ: Segment Anything Model with Depth-guided Adaptive Queries for RGB-D Video Salient Object Detection
by: Lin, Jia, et al.
Published: (2025)
by: Lin, Jia, et al.
Published: (2025)
X Modality Assisting RGBT Object Tracking
by: Ding, Zhaisheng, et al.
Published: (2023)
by: Ding, Zhaisheng, et al.
Published: (2023)
SportsHHI: A Dataset for Human-Human Interaction Detection in Sports Videos
by: Wu, Tao, et al.
Published: (2024)
by: Wu, Tao, et al.
Published: (2024)
Multi-Granularity Video Object Segmentation
by: Lim, Sangbeom, et al.
Published: (2024)
by: Lim, Sangbeom, et al.
Published: (2024)
IAM: Enhancing RGB-D Instance Segmentation with New Benchmarks
by: Jung, Aecheon, et al.
Published: (2025)
by: Jung, Aecheon, et al.
Published: (2025)
One-Shot Medical Video Object Segmentation via Temporal Contrastive Memory Networks
by: Chen, Yaxiong, et al.
Published: (2025)
by: Chen, Yaxiong, et al.
Published: (2025)
Temporally Consistent Referring Video Object Segmentation with Hybrid Memory
by: Miao, Bo, et al.
Published: (2024)
by: Miao, Bo, et al.
Published: (2024)
Efficient RGB-D Scene Understanding via Multi-task Adaptive Learning and Cross-dimensional Feature Guidance
by: Sun, Guodong, et al.
Published: (2026)
by: Sun, Guodong, et al.
Published: (2026)
RGBD Objects in the Wild: Scaling Real-World 3D Object Learning from RGB-D Videos
by: Xia, Hongchi, et al.
Published: (2024)
by: Xia, Hongchi, et al.
Published: (2024)
Shallow Features Matter: Hierarchical Memory with Heterogeneous Interaction for Unsupervised Video Object Segmentation
by: Xiangyu, Zheng, et al.
Published: (2025)
by: Xiangyu, Zheng, et al.
Published: (2025)
Enhanced Automotive Object Detection via RGB-D Fusion in a DiffusionDet Framework
by: Orfaig, Eliraz, et al.
Published: (2024)
by: Orfaig, Eliraz, et al.
Published: (2024)
Latest Object Memory Management for Temporally Consistent Video Instance Segmentation
by: Lee, Seunghun, et al.
Published: (2025)
by: Lee, Seunghun, et al.
Published: (2025)
Dexterous Manipulation Policies from RGB Human Videos via 3D Hand-Object Trajectory Reconstruction
by: Chen, Hongyi, et al.
Published: (2026)
by: Chen, Hongyi, et al.
Published: (2026)
MotionRAG: Motion Retrieval-Augmented Image-to-Video Generation
by: Zhu, Chenhui, et al.
Published: (2025)
by: Zhu, Chenhui, et al.
Published: (2025)
MAMBA: Multi-level Aggregation via Memory Bank for Video Object Detection
by: Sun, Guanxiong, et al.
Published: (2024)
by: Sun, Guanxiong, et al.
Published: (2024)
Spatial-Temporal Graph Enhanced DETR Towards Multi-Frame 3D Object Detection
by: Zhang, Yifan, et al.
Published: (2023)
by: Zhang, Yifan, et al.
Published: (2023)
XTrack: Multimodal Training Boosts RGB-X Video Object Trackers
by: Tan, Yuedong, et al.
Published: (2024)
by: Tan, Yuedong, et al.
Published: (2024)
Studying Image Diffusion Features for Zero-Shot Video Object Segmentation
by: Delatolas, Thanos, et al.
Published: (2025)
by: Delatolas, Thanos, et al.
Published: (2025)
STENet: Superpixel Token Enhancing Network for RGB-D Salient Object Detection
by: Chen, Jianlin, et al.
Published: (2026)
by: Chen, Jianlin, et al.
Published: (2026)
Weakly Supervised Point Clouds Transformer for 3D Object Detection
by: Tang, Zuojin, et al.
Published: (2023)
by: Tang, Zuojin, et al.
Published: (2023)
VideoSeg-R1:Reasoning Video Object Segmentation via Reinforcement Learning
by: Xu, Zishan, et al.
Published: (2025)
by: Xu, Zishan, et al.
Published: (2025)
TSMS-SAM2: Multi-scale Temporal Sampling Augmentation and Memory-Splitting Pruning for Promptable Video Object Segmentation and Tracking in Surgical Scenarios
by: Xu, Guoping, et al.
Published: (2025)
by: Xu, Guoping, et al.
Published: (2025)
Similar Items
-
RGB-D Tracking via Hierarchical Modality Aggregation and Distribution Network
by: Xu, Boyue, et al.
Published: (2025) -
Learning Frequency and Memory-Aware Prompts for Multi-Modal Object Tracking
by: Xu, Boyue, et al.
Published: (2025) -
VL-UniTrack: A Unified Framework with Visual-Language Prompts for UAV-Ground Visual Tracking
by: Xu, Boyue, et al.
Published: (2026) -
MTNet: Learning modality-aware representation with transformer for RGBT tracking
by: Hou, Ruichao, et al.
Published: (2025) -
KAN-SAM: Kolmogorov-Arnold Network Guided Segment Anything Model for RGB-T Salient Object Detection
by: Li, Xingyuan, et al.
Published: (2025)