M-SpecGene: Generalized Foundation Model for RGBT Multispectral Vision
Fuente:
arXiv
Saved in:
| Main Authors: | Zhou, Kailai, Yang, Fuqiang, Wang, Shixian, Wen, Bihan, Zi, Chongde, Chen, Linsen, Shen, Qiu, Cao, Xun |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Joint RGB-Spectral Decomposition Model Guided Image Enhancement in Mobile Photography
by: Zhou, Kailai, et al.
Published: (2024)
by: Zhou, Kailai, et al.
Published: (2024)
Hierarchical Spatial-Frequency Aggregation for Spectral Deconvolution Imaging
by: Lv, Tao, et al.
Published: (2025)
by: Lv, Tao, et al.
Published: (2025)
SpecGen: Neural Spectral BRDF Generation via Spectral-Spatial Tri-plane Aggregation
by: Jin, Zhenyu, et al.
Published: (2025)
by: Jin, Zhenyu, et al.
Published: (2025)
YOLOv11-RGBT: Towards a Comprehensive Single-Stage Multispectral Object Detection Framework
by: Wan, Dahang, et al.
Published: (2025)
by: Wan, Dahang, et al.
Published: (2025)
Gaseous Object Detection
by: Zhou, Kailai, et al.
Published: (2025)
by: Zhou, Kailai, et al.
Published: (2025)
Exploring Spatiotemporal Feature Propagation for Video-Level Compressive Spectral Reconstruction: Dataset, Model and Benchmark
by: Cai, Lijing, et al.
Published: (2026)
by: Cai, Lijing, et al.
Published: (2026)
Split-Layer: Enhancing Implicit Neural Representation by Maximizing the Dimensionality of Feature Space
by: Cai, Zhicheng, et al.
Published: (2025)
by: Cai, Zhicheng, et al.
Published: (2025)
SpecSwin3D: Generating Hyperspectral Imagery from Multispectral Data via Transformer Networks
by: Sui, Tang, et al.
Published: (2025)
by: Sui, Tang, et al.
Published: (2025)
Oscillating Dispersion for Maximal Light-throughput Spectral Imaging
by: Zhang, Jiuyun, et al.
Published: (2026)
by: Zhang, Jiuyun, et al.
Published: (2026)
Frames2Residual: Spatiotemporal Decoupling for Self-Supervised Video Denoising
by: Ji, Mingjie, et al.
Published: (2026)
by: Ji, Mingjie, et al.
Published: (2026)
RAGTrack: Language-aware RGBT Tracking with Retrieval-Augmented Generation
by: Li, Hao, et al.
Published: (2026)
by: Li, Hao, et al.
Published: (2026)
Mitigating the Impact of Prominent Position Shift in Drone-based RGBT Object Detection
by: Zhang, Yan, et al.
Published: (2025)
by: Zhang, Yan, et al.
Published: (2025)
Temporal Adaptive RGBT Tracking with Modality Prompt
by: Wang, Hongyu, et al.
Published: (2024)
by: Wang, Hongyu, et al.
Published: (2024)
Dynamic Disentangled Fusion Network for RGBT Tracking
by: Li, Chenglong, et al.
Published: (2024)
by: Li, Chenglong, et al.
Published: (2024)
X Modality Assisting RGBT Object Tracking
by: Ding, Zhaisheng, et al.
Published: (2023)
by: Ding, Zhaisheng, et al.
Published: (2023)
Cross-modulated Attention Transformer for RGBT Tracking
by: Xiao, Yun, et al.
Published: (2024)
by: Xiao, Yun, et al.
Published: (2024)
Transformer RGBT Tracking with Spatio-Temporal Multimodal Tokens
by: Sun, Dengdi, et al.
Published: (2024)
by: Sun, Dengdi, et al.
Published: (2024)
AFter: Attention-based Fusion Router for RGBT Tracking
by: Lu, Andong, et al.
Published: (2024)
by: Lu, Andong, et al.
Published: (2024)
SIGMAE: A Spectral-Index-Guided Foundation Model for Multispectral Remote Sensing
by: Zhang, Xiaokang, et al.
Published: (2026)
by: Zhang, Xiaokang, et al.
Published: (2026)
A Two-Stage Masked Autoencoder Based Network for Indoor Depth Completion
by: Sun, Kailai, et al.
Published: (2024)
by: Sun, Kailai, et al.
Published: (2024)
CattleFace-RGBT: RGB-T Cattle Facial Landmark Benchmark
by: Coffman, Ethan, et al.
Published: (2024)
by: Coffman, Ethan, et al.
Published: (2024)
Breaking Modality Gap in RGBT Tracking: Coupled Knowledge Distillation
by: Lu, Andong, et al.
Published: (2024)
by: Lu, Andong, et al.
Published: (2024)
SpectraIrisPAD: Leveraging Vision Foundation Models for Spectrally Conditioned Multispectral Iris Presentation Attack Detection
by: Ramachandra, Raghavendra, et al.
Published: (2025)
by: Ramachandra, Raghavendra, et al.
Published: (2025)
Open-RGBT: Open-vocabulary RGB-T Zero-shot Semantic Segmentation in Open-world Environments
by: Yu, Meng, et al.
Published: (2024)
by: Yu, Meng, et al.
Published: (2024)
RGBT-Ground Benchmark: Visual Grounding Beyond RGB in Complex Real-World Scenarios
by: Zhao, Tianyi, et al.
Published: (2025)
by: Zhao, Tianyi, et al.
Published: (2025)
CADTrack: Learning Contextual Aggregation with Deformable Alignment for Robust RGBT Tracking
by: Li, Hao, et al.
Published: (2025)
by: Li, Hao, et al.
Published: (2025)
MTNet: Learning modality-aware representation with transformer for RGBT tracking
by: Hou, Ruichao, et al.
Published: (2025)
by: Hou, Ruichao, et al.
Published: (2025)
ZoRI: Towards Discriminative Zero-Shot Remote Sensing Instance Segmentation
by: Huang, Shiqi, et al.
Published: (2024)
by: Huang, Shiqi, et al.
Published: (2024)
RSGround-R1: Rethinking Remote Sensing Visual Grounding through Spatial Reasoning
by: Huang, Shiqi, et al.
Published: (2026)
by: Huang, Shiqi, et al.
Published: (2026)
SpecVLM: Fast Speculative Decoding in Vision-Language Models
by: Huang, Haiduo, et al.
Published: (2025)
by: Huang, Haiduo, et al.
Published: (2025)
RGBT Tracking via All-layer Multimodal Interactions with Progressive Fusion Mamba
by: Lu, Andong, et al.
Published: (2024)
by: Lu, Andong, et al.
Published: (2024)
Modality-missing RGBT Tracking: Invertible Prompt Learning and High-quality Benchmarks
by: Lu, Andong, et al.
Published: (2023)
by: Lu, Andong, et al.
Published: (2023)
Rethinking Early-Fusion Strategies for Improved Multispectral Object Detection
by: Zhang, Xue, et al.
Published: (2024)
by: Zhang, Xue, et al.
Published: (2024)
Parameter-Efficient Fine-Tuning of Multispectral Foundation Models for Hyperspectral Image Classification
by: Ligan, Bernardin, et al.
Published: (2025)
by: Ligan, Bernardin, et al.
Published: (2025)
FSSD: Feature Fusion Single Shot Multibox Detector
by: Li, Zuoxin, et al.
Published: (2017)
by: Li, Zuoxin, et al.
Published: (2017)
SFFR: Spatial-Frequency Feature Reconstruction for Multispectral Aerial Object Detection
by: Zuo, Xin, et al.
Published: (2025)
by: Zuo, Xin, et al.
Published: (2025)
SpecAware: A Spectral-Content Aware Foundation Model for Unifying Multi-Sensor Learning in Hyperspectral Remote Sensing Mapping
by: Ji, Renjie, et al.
Published: (2025)
by: Ji, Renjie, et al.
Published: (2025)
Vision Foundation Models as Generalist Tokenizers for Image Generation
by: Zheng, Anlin, et al.
Published: (2026)
by: Zheng, Anlin, et al.
Published: (2026)
Spec-Gaussian: Anisotropic View-Dependent Appearance for 3D Gaussian Splatting
by: Yang, Ziyi, et al.
Published: (2024)
by: Yang, Ziyi, et al.
Published: (2024)
Multimodal Spatio-temporal Graph Learning for Alignment-free RGBT Video Object Detection
by: Wang, Qishun, et al.
Published: (2025)
by: Wang, Qishun, et al.
Published: (2025)
Similar Items
-
Joint RGB-Spectral Decomposition Model Guided Image Enhancement in Mobile Photography
by: Zhou, Kailai, et al.
Published: (2024) -
Hierarchical Spatial-Frequency Aggregation for Spectral Deconvolution Imaging
by: Lv, Tao, et al.
Published: (2025) -
SpecGen: Neural Spectral BRDF Generation via Spectral-Spatial Tri-plane Aggregation
by: Jin, Zhenyu, et al.
Published: (2025) -
YOLOv11-RGBT: Towards a Comprehensive Single-Stage Multispectral Object Detection Framework
by: Wan, Dahang, et al.
Published: (2025) -
Gaseous Object Detection
by: Zhou, Kailai, et al.
Published: (2025)