Mixture of Scale Experts for Alignment-free RGBT Video Object Detection and A Unified Benchmark
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Qishun, Tu, Zhengzheng, Wang, Kunpeng, Gu, Le, Guo, Chuanwang |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Multimodal Spatio-temporal Graph Learning for Alignment-free RGBT Video Object Detection
by: Wang, Qishun, et al.
Published: (2025)
by: Wang, Qishun, et al.
Published: (2025)
Alignment-Free RGBT Salient Object Detection: Semantics-guided Asymmetric Correlation Network and A Unified Benchmark
by: Wang, Kunpeng, et al.
Published: (2024)
by: Wang, Kunpeng, et al.
Published: (2024)
Unified-modal Salient Object Detection via Adaptive Prompt Learning
by: Wang, Kunpeng, et al.
Published: (2023)
by: Wang, Kunpeng, et al.
Published: (2023)
Alignment-Free RGB-T Salient Object Detection: A Large-scale Dataset and Progressive Correlation Network
by: Wang, Kunpeng, et al.
Published: (2024)
by: Wang, Kunpeng, et al.
Published: (2024)
Learning Adaptive Fusion Bank for Multi-modal Salient Object Detection
by: Wang, Kunpeng, et al.
Published: (2024)
by: Wang, Kunpeng, et al.
Published: (2024)
Adapting Segment Anything Model to Multi-modal Salient Object Detection with Semantic Feature Fusion Guidance
by: Wang, Kunpeng, et al.
Published: (2024)
by: Wang, Kunpeng, et al.
Published: (2024)
A Spatial-Temporal Progressive Fusion Network for Breast Lesion Segmentation in Ultrasound Videos
by: Tu, Zhengzheng, et al.
Published: (2024)
by: Tu, Zhengzheng, et al.
Published: (2024)
COXNet: Cross-Layer Fusion with Adaptive Alignment and Scale Integration for RGBT Tiny Object Detection
by: Peng, Peiran, et al.
Published: (2025)
by: Peng, Peiran, et al.
Published: (2025)
SAMSOD: Rethinking SAM Optimization for RGB-T Salient Object Detection
by: Liu, Zhengyi, et al.
Published: (2025)
by: Liu, Zhengyi, et al.
Published: (2025)
LFSamba: Marry SAM with Mamba for Light Field Salient Object Detection
by: Liu, Zhengyi, et al.
Published: (2024)
by: Liu, Zhengyi, et al.
Published: (2024)
CADTrack: Learning Contextual Aggregation with Deformable Alignment for Robust RGBT Tracking
by: Li, Hao, et al.
Published: (2025)
by: Li, Hao, et al.
Published: (2025)
X Modality Assisting RGBT Object Tracking
by: Ding, Zhaisheng, et al.
Published: (2023)
by: Ding, Zhaisheng, et al.
Published: (2023)
Breaking Shallow Limits: Task-Driven Pixel Fusion for Gap-free RGBT Tracking
by: Lu, Andong, et al.
Published: (2025)
by: Lu, Andong, et al.
Published: (2025)
CattleFace-RGBT: RGB-T Cattle Facial Landmark Benchmark
by: Coffman, Ethan, et al.
Published: (2024)
by: Coffman, Ethan, et al.
Published: (2024)
Ultrasound SAM Adapter: Adapting SAM for Breast Lesion Segmentation in Ultrasound Images
by: Tu, Zhengzheng, et al.
Published: (2024)
by: Tu, Zhengzheng, et al.
Published: (2024)
Mitigating the Impact of Prominent Position Shift in Drone-based RGBT Object Detection
by: Zhang, Yan, et al.
Published: (2025)
by: Zhang, Yan, et al.
Published: (2025)
Decoupled Cross-Modal Alignment Network for Text-RGBT Person Retrieval and A High-Quality Benchmark
by: Deng, Yifei, et al.
Published: (2025)
by: Deng, Yifei, et al.
Published: (2025)
M$^2$CD: A Unified MultiModal Framework for Optical-SAR Change Detection with Mixture of Experts and Self-Distillation
by: Liu, Ziyuan, et al.
Published: (2025)
by: Liu, Ziyuan, et al.
Published: (2025)
YOLOv11-RGBT: Towards a Comprehensive Single-Stage Multispectral Object Detection Framework
by: Wan, Dahang, et al.
Published: (2025)
by: Wan, Dahang, et al.
Published: (2025)
YOLO Meets Mixture-of-Experts: Adaptive Expert Routing for Robust Object Detection
by: Meiraz, Ori, et al.
Published: (2025)
by: Meiraz, Ori, et al.
Published: (2025)
Visual Self-Fulfilling Alignment: Shaping Safety-Oriented Personas via Threat-Related Images
by: Yang, Qishun, et al.
Published: (2026)
by: Yang, Qishun, et al.
Published: (2026)
Dynamic Disentangled Fusion Network for RGBT Tracking
by: Li, Chenglong, et al.
Published: (2024)
by: Li, Chenglong, et al.
Published: (2024)
UniPaint: Unified Space-time Video Inpainting via Mixture-of-Experts
by: Wan, Zhen, et al.
Published: (2024)
by: Wan, Zhen, et al.
Published: (2024)
Unified Multimodal Visual Tracking with Dual Mixture-of-Experts
by: Hong, Lingyi, et al.
Published: (2026)
by: Hong, Lingyi, et al.
Published: (2026)
Temporal Adaptive RGBT Tracking with Modality Prompt
by: Wang, Hongyu, et al.
Published: (2024)
by: Wang, Hongyu, et al.
Published: (2024)
Unity is Strength: Unifying Convolutional and Transformeral Features for Better Person Re-Identification
by: Wang, Yuhao, et al.
Published: (2024)
by: Wang, Yuhao, et al.
Published: (2024)
Revisiting RGBT Tracking Benchmarks from the Perspective of Modality Validity: A New Benchmark, Problem, and Solution
by: Tang, Zhangyong, et al.
Published: (2024)
by: Tang, Zhangyong, et al.
Published: (2024)
RAGTrack: Language-aware RGBT Tracking with Retrieval-Augmented Generation
by: Li, Hao, et al.
Published: (2026)
by: Li, Hao, et al.
Published: (2026)
AFter: Attention-based Fusion Router for RGBT Tracking
by: Lu, Andong, et al.
Published: (2024)
by: Lu, Andong, et al.
Published: (2024)
Modality-missing RGBT Tracking: Invertible Prompt Learning and High-quality Benchmarks
by: Lu, Andong, et al.
Published: (2023)
by: Lu, Andong, et al.
Published: (2023)
Exposing and Defending the Achilles' Heel of Video Mixture-of-Experts
by: Wang, Songping, et al.
Published: (2026)
by: Wang, Songping, et al.
Published: (2026)
Video Relationship Detection Using Mixture of Experts
by: Shaabana, Ala, et al.
Published: (2024)
by: Shaabana, Ala, et al.
Published: (2024)
GaTector+: A Unified Head-free Framework for Gaze Object and Gaze Following Prediction
by: Jin, Yang, et al.
Published: (2025)
by: Jin, Yang, et al.
Published: (2025)
MoCaE: Mixture of Calibrated Experts Significantly Improves Object Detection
by: Oksuz, Kemal, et al.
Published: (2023)
by: Oksuz, Kemal, et al.
Published: (2023)
DeMamba: AI-Generated Video Detection on Million-Scale GenVideo Benchmark
by: Chen, Haoxing, et al.
Published: (2024)
by: Chen, Haoxing, et al.
Published: (2024)
Fourier Angle Alignment for Oriented Object Detection in Remote Sensing
by: Gu, Changyu, et al.
Published: (2026)
by: Gu, Changyu, et al.
Published: (2026)
RGBT-Ground Benchmark: Visual Grounding Beyond RGB in Complex Real-World Scenarios
by: Zhao, Tianyi, et al.
Published: (2025)
by: Zhao, Tianyi, et al.
Published: (2025)
M$^4$-SAM: Multi-Modal Mixture-of-Experts with Memory-Augmented SAM for RGB-D Video Salient Object Detection
by: Liu, Jiyuan, et al.
Published: (2026)
by: Liu, Jiyuan, et al.
Published: (2026)
Uni-MoE: Scaling Unified Multimodal LLMs with Mixture of Experts
by: Li, Yunxin, et al.
Published: (2024)
by: Li, Yunxin, et al.
Published: (2024)
Mixture-of-Attack-Experts with Class Regularization for Unified Physical-Digital Face Attack Detection
by: Chen, Shunxin, et al.
Published: (2025)
by: Chen, Shunxin, et al.
Published: (2025)
Similar Items
-
Multimodal Spatio-temporal Graph Learning for Alignment-free RGBT Video Object Detection
by: Wang, Qishun, et al.
Published: (2025) -
Alignment-Free RGBT Salient Object Detection: Semantics-guided Asymmetric Correlation Network and A Unified Benchmark
by: Wang, Kunpeng, et al.
Published: (2024) -
Unified-modal Salient Object Detection via Adaptive Prompt Learning
by: Wang, Kunpeng, et al.
Published: (2023) -
Alignment-Free RGB-T Salient Object Detection: A Large-scale Dataset and Progressive Correlation Network
by: Wang, Kunpeng, et al.
Published: (2024) -
Learning Adaptive Fusion Bank for Multi-modal Salient Object Detection
by: Wang, Kunpeng, et al.
Published: (2024)