Bridging the RGB-IR Gap: Consensus and Discrepancy Modeling for Text-Guided Multispectral Detection
Fuente:
arXiv
Saved in:
| Main Authors: | Wu, Jiaqi, Wang, Zhen, Huang, Enhao, Shen, Kangqing, Wang, Yulin, Yue, Yang, Pu, Yifan, Huang, Gao |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Few-Step Distillation for Text-to-Image Generation: A Practical Guide
by: Pu, Yifan, et al.
Published: (2025)
by: Pu, Yifan, et al.
Published: (2025)
UAV-CB: A Complex-Background RGB-T Dataset and Local Frequency Bridge Network for UAV Detection
by: Huang, Shenghui, et al.
Published: (2026)
by: Huang, Shenghui, et al.
Published: (2026)
Multispectral State-Space Feature Fusion: Bridging Shared and Cross-Parametric Interactions for Object Detection
by: Shen, Jifeng, et al.
Published: (2025)
by: Shen, Jifeng, et al.
Published: (2025)
Representation Discrepancy Bridging Method for Remote Sensing Image-Text Retrieval
by: Ning, Hailong, et al.
Published: (2025)
by: Ning, Hailong, et al.
Published: (2025)
Bridging the Gap Between End-to-End and Two-Step Text Spotting
by: Huang, Mingxin, et al.
Published: (2024)
by: Huang, Mingxin, et al.
Published: (2024)
Multispectral Texture Synthesis using RGB Convolutional Neural Networks
by: Ollivier, Sélim, et al.
Published: (2024)
by: Ollivier, Sélim, et al.
Published: (2024)
Mind the Gap: Confidence Discrepancy Can Guide Federated Semi-Supervised Learning Across Pseudo-Mismatch
by: Liu, Yijie, et al.
Published: (2025)
by: Liu, Yijie, et al.
Published: (2025)
Rethinking Early-Fusion Strategies for Improved Multispectral Object Detection
by: Zhang, Xue, et al.
Published: (2024)
by: Zhang, Xue, et al.
Published: (2024)
GRA: Detecting Oriented Objects through Group-wise Rotating and Attention
by: Wang, Jiangshan, et al.
Published: (2024)
by: Wang, Jiangshan, et al.
Published: (2024)
Multi-clue Consistency Learning to Bridge Gaps Between General and Oriented Object in Semi-supervised Detection
by: Wang, Chenxu, et al.
Published: (2024)
by: Wang, Chenxu, et al.
Published: (2024)
Enhancing Traffic Object Detection in Variable Illumination with RGB-Event Fusion
by: Liu, Zhanwen, et al.
Published: (2023)
by: Liu, Zhanwen, et al.
Published: (2023)
Contour-Native Bridge Defect Detection and Compact Digital Archiving with Frequency-Supervised Fourier Contours
by: Liu, Jin, et al.
Published: (2026)
by: Liu, Jin, et al.
Published: (2026)
Optimizing Multispectral Object Detection: A Bag of Tricks and Comprehensive Benchmarks
by: Zhou, Chen, et al.
Published: (2024)
by: Zhou, Chen, et al.
Published: (2024)
MO R-CNN: Multispectral Oriented R-CNN for Object Detection in Remote Sensing Image
by: Wang, Leiyu, et al.
Published: (2025)
by: Wang, Leiyu, et al.
Published: (2025)
Bridging the Indoor-Outdoor Gap: Vision-Centric Instruction-Guided Embodied Navigation for the Last Meters
by: Zhao, Yuxiang, et al.
Published: (2026)
by: Zhao, Yuxiang, et al.
Published: (2026)
Rethinking the Architecture Design for Efficient Generic Event Boundary Detection
by: Zheng, Ziwei, et al.
Published: (2024)
by: Zheng, Ziwei, et al.
Published: (2024)
AMFD: Distillation via Adaptive Multimodal Fusion for Multispectral Pedestrian Detection
by: Chen, Zizhao, et al.
Published: (2024)
by: Chen, Zizhao, et al.
Published: (2024)
Bridging the Gap: Aligning Text-to-Image Diffusion Models with Specific Feedback
by: Niu, Xuexiang, et al.
Published: (2024)
by: Niu, Xuexiang, et al.
Published: (2024)
SFFR: Spatial-Frequency Feature Reconstruction for Multispectral Aerial Object Detection
by: Zuo, Xin, et al.
Published: (2025)
by: Zuo, Xin, et al.
Published: (2025)
Membership Inference on Text-to-Image Diffusion Models via Conditional Likelihood Discrepancy
by: Zhai, Shengfang, et al.
Published: (2024)
by: Zhai, Shengfang, et al.
Published: (2024)
Bringing RGB and IR Together: Hierarchical Multi-Modal Enhancement for Robust Transmission Line Detection
by: Zhang, Shengdong, et al.
Published: (2025)
by: Zhang, Shengdong, et al.
Published: (2025)
EchoWorld: Learning Motion-Aware World Models for Echocardiography Probe Guidance
by: Yue, Yang, et al.
Published: (2025)
by: Yue, Yang, et al.
Published: (2025)
CheXWorld: Exploring Image World Modeling for Radiograph Representation Learning
by: Yue, Yang, et al.
Published: (2025)
by: Yue, Yang, et al.
Published: (2025)
Consensus-Driven Uncertainty for Robotic Grasping based on RGB Perception
by: Joyce, Eric C., et al.
Published: (2025)
by: Joyce, Eric C., et al.
Published: (2025)
Bridging the Skill Gap in Clinical CBCT Interpretation with CBCTRepD
by: Wu, Qinxin, et al.
Published: (2026)
by: Wu, Qinxin, et al.
Published: (2026)
Mind the Gap: Bridging Occlusion in Gait Recognition via Residual Gap Correction
by: Gupta, Ayush, et al.
Published: (2025)
by: Gupta, Ayush, et al.
Published: (2025)
Self-Learning Hyperspectral and Multispectral Image Fusion via Adaptive Residual Guided Subspace Diffusion Model
by: Zhu, Jian, et al.
Published: (2025)
by: Zhu, Jian, et al.
Published: (2025)
End-to-End RGB-IR Joint Image Compression With Channel-wise Cross-modality Entropy Model
by: Wang, Haofeng, et al.
Published: (2025)
by: Wang, Haofeng, et al.
Published: (2025)
Concept Guided Co-salient Object Detection
by: Zhu, Jiayi, et al.
Published: (2024)
by: Zhu, Jiayi, et al.
Published: (2024)
FreeInit: Bridging Initialization Gap in Video Diffusion Models
by: Wu, Tianxing, et al.
Published: (2023)
by: Wu, Tianxing, et al.
Published: (2023)
Guided MRI Reconstruction via Schrödinger Bridge
by: Wang, Yue, et al.
Published: (2024)
by: Wang, Yue, et al.
Published: (2024)
Bridging the Divide: Reconsidering Softmax and Linear Attention
by: Han, Dongchen, et al.
Published: (2024)
by: Han, Dongchen, et al.
Published: (2024)
Direct Discrepancy Replay: Distribution-Discrepancy Condensation and Manifold-Consistent Replay for Continual Face Forgery Detection
by: Zhang, Tianshuo, et al.
Published: (2026)
by: Zhang, Tianshuo, et al.
Published: (2026)
RASMD: RGB And SWIR Multispectral Driving Dataset for Robust Perception in Adverse Conditions
by: Jin, Youngwan, et al.
Published: (2025)
by: Jin, Youngwan, et al.
Published: (2025)
MADiff: Text-Guided Fashion Image Editing with Mask Prediction and Attention-Enhanced Diffusion
by: Zhan, Zechao, et al.
Published: (2024)
by: Zhan, Zechao, et al.
Published: (2024)
Bridging Visual Affective Gap: Borrowing Textual Knowledge by Learning from Noisy Image-Text Pairs
by: Wu, Daiqing, et al.
Published: (2025)
by: Wu, Daiqing, et al.
Published: (2025)
MonoDGP: Monocular 3D Object Detection with Decoupled-Query and Geometry-Error Priors
by: Pu, Fanqi, et al.
Published: (2024)
by: Pu, Fanqi, et al.
Published: (2024)
Seeing It Before It Happens: In-Generation NSFW Detection for Diffusion-Based Text-to-Image Models
by: Yang, Fan, et al.
Published: (2025)
by: Yang, Fan, et al.
Published: (2025)
Bridging the Scale Gap: Balanced Tiny and General Object Detection in Remote Sensing Imagery
by: Zhao, Zhicheng, et al.
Published: (2025)
by: Zhao, Zhicheng, et al.
Published: (2025)
Bridging the Vision-Brain Gap with an Uncertainty-Aware Blur Prior
by: Wu, Haitao, et al.
Published: (2025)
by: Wu, Haitao, et al.
Published: (2025)
Similar Items
-
Few-Step Distillation for Text-to-Image Generation: A Practical Guide
by: Pu, Yifan, et al.
Published: (2025) -
UAV-CB: A Complex-Background RGB-T Dataset and Local Frequency Bridge Network for UAV Detection
by: Huang, Shenghui, et al.
Published: (2026) -
Multispectral State-Space Feature Fusion: Bridging Shared and Cross-Parametric Interactions for Object Detection
by: Shen, Jifeng, et al.
Published: (2025) -
Representation Discrepancy Bridging Method for Remote Sensing Image-Text Retrieval
by: Ning, Hailong, et al.
Published: (2025) -
Bridging the Gap Between End-to-End and Two-Step Text Spotting
by: Huang, Mingxin, et al.
Published: (2024)