InfMAE: A Foundation Model in the Infrared Modality
Fuente:
arXiv
Saved in:
| Main Authors: | Liu, Fangcen, Gao, Chenqiang, Zhang, Yaming, Guo, Junjie, Wang, Jinhao, Meng, Deyu |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
DPDETR: Decoupled Position Detection Transformer for Infrared-Visible Object Detection
by: Guo, Junjie, et al.
Published: (2024)
by: Guo, Junjie, et al.
Published: (2024)
IV-tuning: Parameter-Efficient Transfer Learning for Infrared-Visible Tasks
by: Zhang, Yaming, et al.
Published: (2024)
by: Zhang, Yaming, et al.
Published: (2024)
IVGF: The Fusion-Guided Infrared and Visible General Framework
by: Liu, Fangcen, et al.
Published: (2024)
by: Liu, Fangcen, et al.
Published: (2024)
DAMSDet: Dynamic Adaptive Multispectral Detection Transformer with Competitive Query Selection and Adaptive Feature Fusion
by: Guo, Junjie, et al.
Published: (2024)
by: Guo, Junjie, et al.
Published: (2024)
CM-Diff: A Single Generative Network for Bidirectional Cross-Modality Translation Diffusion Model Between Infrared and Visible Images
by: Hu, Bin, et al.
Published: (2025)
by: Hu, Bin, et al.
Published: (2025)
DOD-SA: Infrared-Visible Decoupled Object Detection with Single-Modality Annotations
by: Jin, Hang, et al.
Published: (2025)
by: Jin, Hang, et al.
Published: (2025)
Are Dense Labels Always Necessary for 3D Object Detection from Point Cloud?
by: Gao, Chenqiang, et al.
Published: (2024)
by: Gao, Chenqiang, et al.
Published: (2024)
Diffusion-Guided Mask-Consistent Paired Mixing for Endoscopic Image Segmentation
by: Jie, Pengyu, et al.
Published: (2025)
by: Jie, Pengyu, et al.
Published: (2025)
UNIV: Unified Foundation Model for Infrared and Visible Modalities
by: Mao, Fangyuan, et al.
Published: (2025)
by: Mao, Fangyuan, et al.
Published: (2025)
Towards Student Actions in Classroom Scenes: New Dataset and Baseline
by: Tan, Zhuolin, et al.
Published: (2024)
by: Tan, Zhuolin, et al.
Published: (2024)
A Point-Neighborhood Learning Framework for Nasal Endoscope Image Segmentation
by: Jie, Pengyu, et al.
Published: (2024)
by: Jie, Pengyu, et al.
Published: (2024)
A New Learning Paradigm for Foundation Model-based Remote Sensing Change Detection
by: Li, Kaiyu, et al.
Published: (2023)
by: Li, Kaiyu, et al.
Published: (2023)
Dynamic Background Reconstruction via MAE for Infrared Small Target Detection
by: Peng, Jingchao, et al.
Published: (2023)
by: Peng, Jingchao, et al.
Published: (2023)
DuGI-MAE: Improving Infrared Mask Autoencoders via Dual-Domain Guidance
by: Xing, Yinghui, et al.
Published: (2025)
by: Xing, Yinghui, et al.
Published: (2025)
Missing No More: Dictionary-Guided Cross-Modal Image Fusion under Missing Infrared
by: Zhang, Yafei, et al.
Published: (2026)
by: Zhang, Yafei, et al.
Published: (2026)
OceanMAE: A Foundation Model for Ocean Remote Sensing
by: Stamer, Viola-Joanna, et al.
Published: (2026)
by: Stamer, Viola-Joanna, et al.
Published: (2026)
AD-DINOv3: Enhancing DINOv3 for Zero-Shot Anomaly Detection with Anomaly-Aware Calibration
by: Yuan, Jingyi, et al.
Published: (2025)
by: Yuan, Jingyi, et al.
Published: (2025)
SPIRIT: Adapting Vision Foundation Models for Unified Single- and Multi-Frame Infrared Small Target Detection
by: Xu, Qian, et al.
Published: (2026)
by: Xu, Qian, et al.
Published: (2026)
SOFTooth: Semantics-Enhanced Order-Aware Fusion for Tooth Instance Segmentation
by: Li, Xiaolan, et al.
Published: (2025)
by: Li, Xiaolan, et al.
Published: (2025)
InfGen: A Resolution-Agnostic Paradigm for Scalable Image Synthesis
by: Han, Tao, et al.
Published: (2025)
by: Han, Tao, et al.
Published: (2025)
InfVSR: Toward Consistency-Driven Streaming Generative Video Super-Resolution
by: Zhang, Ziqing, et al.
Published: (2025)
by: Zhang, Ziqing, et al.
Published: (2025)
Mogao: An Omni Foundation Model for Interleaved Multi-Modal Generation
by: Liao, Chao, et al.
Published: (2025)
by: Liao, Chao, et al.
Published: (2025)
InfLVG: Reinforce Inference-Time Consistent Long Video Generation with GRPO
by: Fang, Xueji, et al.
Published: (2025)
by: Fang, Xueji, et al.
Published: (2025)
IRSAM: Advancing Segment Anything Model for Infrared Small Target Detection
by: Zhang, Mingjin, et al.
Published: (2024)
by: Zhang, Mingjin, et al.
Published: (2024)
Continual-MAE: Adaptive Distribution Masked Autoencoders for Continual Test-Time Adaptation
by: Liu, Jiaming, et al.
Published: (2023)
by: Liu, Jiaming, et al.
Published: (2023)
STMI: Segmentation-Guided Token Modulation with Cross-Modal Hypergraph Interaction for Multi-Modal Object Re-Identification
by: Xu, Xingguo, et al.
Published: (2026)
by: Xu, Xingguo, et al.
Published: (2026)
AFR-CLIP: Enhancing Zero-Shot Industrial Anomaly Detection with Stateless-to-Stateful Anomaly Feature Rectification
by: Yuan, Jingyi, et al.
Published: (2025)
by: Yuan, Jingyi, et al.
Published: (2025)
Improved mmFormer for Liver Fibrosis Staging via Missing-Modality Compensation
by: Zhang, Zhejia, et al.
Published: (2025)
by: Zhang, Zhejia, et al.
Published: (2025)
MultiMAE for Brain MRIs: Robustness to Missing Inputs Using Multi-Modal Masked Autoencoder
by: Erdur, Ayhan Can, et al.
Published: (2025)
by: Erdur, Ayhan Can, et al.
Published: (2025)
Spatial-Mamba: Effective Visual State Space Models via Structure-aware State Fusion
by: Xiao, Chaodong, et al.
Published: (2024)
by: Xiao, Chaodong, et al.
Published: (2024)
InfScene-SR: Arbitrary-Size Image Super-Resolution via Iterative Joint-Denoising
by: Sun, Shoukun, et al.
Published: (2026)
by: Sun, Shoukun, et al.
Published: (2026)
Unsupervised Visible-Infrared ReID via Pseudo-label Correction and Modality-level Alignment
by: Liu, Yexin, et al.
Published: (2024)
by: Liu, Yexin, et al.
Published: (2024)
HSIGene: A Foundation Model For Hyperspectral Image Generation
by: Pang, Li, et al.
Published: (2024)
by: Pang, Li, et al.
Published: (2024)
Modality Dominance-Aware Optimization for Embodied RGB-Infrared Perception
by: Liu, Xianhui, et al.
Published: (2026)
by: Liu, Xianhui, et al.
Published: (2026)
Generative Latent Kernel Modeling for Blind Motion Deblurring
by: Ding, Chenhao, et al.
Published: (2025)
by: Ding, Chenhao, et al.
Published: (2025)
InfRS: Incremental Few-Shot Object Detection in Remote Sensing Images
by: Li, Wuzhou, et al.
Published: (2024)
by: Li, Wuzhou, et al.
Published: (2024)
Modality-Transition Representation Learning for Visible-Infrared Person Re-Identification
by: Yuan, Chao, et al.
Published: (2025)
by: Yuan, Chao, et al.
Published: (2025)
PolarMAE: Efficient Fetal Ultrasound Pre-training via Semantic Screening and Polar-Guided Masking
by: Lv, Meng, et al.
Published: (2026)
by: Lv, Meng, et al.
Published: (2026)
Continuous Representation Methods, Theories, and Applications: An Overview and Perspectives
by: Luo, Yisi, et al.
Published: (2025)
by: Luo, Yisi, et al.
Published: (2025)
Rotation-Equivariant Self-Supervised Method in Image Denoising
by: Liu, Hanze, et al.
Published: (2025)
by: Liu, Hanze, et al.
Published: (2025)
Similar Items
-
DPDETR: Decoupled Position Detection Transformer for Infrared-Visible Object Detection
by: Guo, Junjie, et al.
Published: (2024) -
IV-tuning: Parameter-Efficient Transfer Learning for Infrared-Visible Tasks
by: Zhang, Yaming, et al.
Published: (2024) -
IVGF: The Fusion-Guided Infrared and Visible General Framework
by: Liu, Fangcen, et al.
Published: (2024) -
DAMSDet: Dynamic Adaptive Multispectral Detection Transformer with Competitive Query Selection and Adaptive Feature Fusion
by: Guo, Junjie, et al.
Published: (2024) -
CM-Diff: A Single Generative Network for Bidirectional Cross-Modality Translation Diffusion Model Between Infrared and Visible Images
by: Hu, Bin, et al.
Published: (2025)