Modality Dominance-Aware Optimization for Embodied RGB-Infrared Perception
Fuente:
arXiv
Guardado en:
| Autores principales: | Liu, Xianhui, Jiang, Siqi, Xie, Yi, Lin, Yuqing, Liu, Siao |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Modification Takes Courage: Seamless Image Stitching via Reference-Driven Inpainting
por: Xie, Ziqi, et al.
Publicado: (2024)
por: Xie, Ziqi, et al.
Publicado: (2024)
MEIA: Multimodal Embodied Perception and Interaction in Unknown Environments
por: Liu, Yang, et al.
Publicado: (2024)
por: Liu, Yang, et al.
Publicado: (2024)
TUNI: Real-time RGB-T Semantic Segmentation with Unified Multi-Modal Feature Extraction and Cross-Modal Feature Fusion
por: Guo, Xiaodong, et al.
Publicado: (2025)
por: Guo, Xiaodong, et al.
Publicado: (2025)
RACANet: Reliability-Aware Crowd Anchor Network for RGB-T Crowd Counting
por: Shi, Jinghao, et al.
Publicado: (2026)
por: Shi, Jinghao, et al.
Publicado: (2026)
LER-YOLO: Reliability-Aware Expert Routing for Misaligned RGB-Infrared UAV Detection
por: Hou, Liming, et al.
Publicado: (2026)
por: Hou, Liming, et al.
Publicado: (2026)
Beyond the Destination: A Novel Benchmark for Exploration-Aware Embodied Question Answering
por: Jiang, Kaixuan, et al.
Publicado: (2025)
por: Jiang, Kaixuan, et al.
Publicado: (2025)
M$^4$-SAM: Multi-Modal Mixture-of-Experts with Memory-Augmented SAM for RGB-D Video Salient Object Detection
por: Liu, Jiyuan, et al.
Publicado: (2026)
por: Liu, Jiyuan, et al.
Publicado: (2026)
Adversarial Robustness in RGB-Skeleton Action Recognition: Leveraging Attention Modality Reweighter
por: Liu, Chao, et al.
Publicado: (2024)
por: Liu, Chao, et al.
Publicado: (2024)
TherA: Thermal-Aware Visual-Language Prompting for Controllable RGB-to-Thermal Infrared Translation
por: Lee, Dong-Guw, et al.
Publicado: (2026)
por: Lee, Dong-Guw, et al.
Publicado: (2026)
Context-Aware Interaction Network for RGB-T Semantic Segmentation
por: Lv, Ying, et al.
Publicado: (2024)
por: Lv, Ying, et al.
Publicado: (2024)
AsymFormer: Asymmetrical Cross-Modal Representation Learning for Mobile Platform Real-Time RGB-D Semantic Segmentation
por: Du, Siqi, et al.
Publicado: (2023)
por: Du, Siqi, et al.
Publicado: (2023)
Modality-Aware Feature Matching: A Comprehensive Review of Single- and Cross-Modality Techniques
por: Liu, Weide, et al.
Publicado: (2025)
por: Liu, Weide, et al.
Publicado: (2025)
Move What Matters: Parameter-Efficient Domain Adaptation via Optimal Transport Flow for Collaborative Perception
por: Jia, Zesheng, et al.
Publicado: (2026)
por: Jia, Zesheng, et al.
Publicado: (2026)
RGB-D Tracking via Hierarchical Modality Aggregation and Distribution Network
por: Xu, Boyue, et al.
Publicado: (2025)
por: Xu, Boyue, et al.
Publicado: (2025)
Perception Matters: Enhancing Embodied AI with Uncertainty-Aware Semantic Segmentation
por: Prasanna, Sai, et al.
Publicado: (2024)
por: Prasanna, Sai, et al.
Publicado: (2024)
Removal then Selection: A Coarse-to-Fine Fusion Perspective for RGB-Infrared Object Detection
por: Zhao, Tianyi, et al.
Publicado: (2024)
por: Zhao, Tianyi, et al.
Publicado: (2024)
Unsupervised Visible-Infrared ReID via Pseudo-label Correction and Modality-level Alignment
por: Liu, Yexin, et al.
Publicado: (2024)
por: Liu, Yexin, et al.
Publicado: (2024)
Coordinate-Aware Thermal Infrared Tracking Via Natural Language Modeling
por: Yan, Miao, et al.
Publicado: (2024)
por: Yan, Miao, et al.
Publicado: (2024)
Look, Zoom, Understand: The Robotic Eyeball for Embodied Perception
por: Yang, Jiashu, et al.
Publicado: (2025)
por: Yang, Jiashu, et al.
Publicado: (2025)
Linking Modality Isolation in Heterogeneous Collaborative Perception
por: Liu, Changxing, et al.
Publicado: (2026)
por: Liu, Changxing, et al.
Publicado: (2026)
ISCS: Parameter-Guided Feature Pruning for Resource-Constrained Embodied Perception
por: Wang, Jinhao, et al.
Publicado: (2025)
por: Wang, Jinhao, et al.
Publicado: (2025)
Improving Domain Generalization in Self-supervised Monocular Depth Estimation via Stabilized Adversarial Training
por: Yao, Yuanqi, et al.
Publicado: (2024)
por: Yao, Yuanqi, et al.
Publicado: (2024)
Spectral-Aware Global Fusion for RGB-Thermal Semantic Segmentation
por: Zhang, Ce, et al.
Publicado: (2025)
por: Zhang, Ce, et al.
Publicado: (2025)
InfMAE: A Foundation Model in the Infrared Modality
por: Liu, Fangcen, et al.
Publicado: (2024)
por: Liu, Fangcen, et al.
Publicado: (2024)
REALM: An RGB and Event Aligned Latent Manifold for Cross-Modal Perception
por: Polizzi, Vincenzo, et al.
Publicado: (2026)
por: Polizzi, Vincenzo, et al.
Publicado: (2026)
SmartControl: Enhancing ControlNet for Handling Rough Visual Conditions
por: Liu, Xiaoyu, et al.
Publicado: (2024)
por: Liu, Xiaoyu, et al.
Publicado: (2024)
AD-DINO: Attention-Dynamic DINO for Distance-Aware Embodied Reference Understanding
por: Guo, Hao, et al.
Publicado: (2024)
por: Guo, Hao, et al.
Publicado: (2024)
Infrared and Visible Image Fusion with Hierarchical Human Perception
por: Yang, Guang, et al.
Publicado: (2024)
por: Yang, Guang, et al.
Publicado: (2024)
CalFuse: Multi-Modal Continual Learning via Feature Calibration and Parameter Fusion
por: Guo, Juncen, et al.
Publicado: (2025)
por: Guo, Juncen, et al.
Publicado: (2025)
ISPDiffuser: Learning RAW-to-sRGB Mappings with Texture-Aware Diffusion Models and Histogram-Guided Color Consistency
por: Ren, Yang, et al.
Publicado: (2025)
por: Ren, Yang, et al.
Publicado: (2025)
Heatmap Pooling Network for Action Recognition from RGB Videos
por: Liu, Mengyuan, et al.
Publicado: (2025)
por: Liu, Mengyuan, et al.
Publicado: (2025)
Memory Regulation and Alignment toward Generalizer RGB-Infrared Person
por: Chen, Feng, et al.
Publicado: (2021)
por: Chen, Feng, et al.
Publicado: (2021)
Modality-Aware Infrared and Visible Image Fusion with Target-Aware Supervision
por: Sun, Tianyao, et al.
Publicado: (2025)
por: Sun, Tianyao, et al.
Publicado: (2025)
Bridging the Gap: Multi-Level Cross-Modality Joint Alignment for Visible-Infrared Person Re-Identification
por: Liang, Tengfei, et al.
Publicado: (2023)
por: Liang, Tengfei, et al.
Publicado: (2023)
UNIV: Unified Foundation Model for Infrared and Visible Modalities
por: Mao, Fangyuan, et al.
Publicado: (2025)
por: Mao, Fangyuan, et al.
Publicado: (2025)
Reconstructing the Image Stitching Pipeline: Integrating Fusion and Rectangling into a Unified Inpainting Model
por: Xie, Ziqi, et al.
Publicado: (2024)
por: Xie, Ziqi, et al.
Publicado: (2024)
HATIR: Heat-Aware Diffusion for Turbulent Infrared Video Super-Resolution
por: Zou, Yang, et al.
Publicado: (2026)
por: Zou, Yang, et al.
Publicado: (2026)
Modality-Guided Dynamic Graph Fusion and Temporal Diffusion for Self-Supervised RGB-T Tracking
por: Li, Shenglan, et al.
Publicado: (2025)
por: Li, Shenglan, et al.
Publicado: (2025)
NavBench: Probing Multimodal Large Language Models for Embodied Navigation
por: Qiao, Yanyuan, et al.
Publicado: (2025)
por: Qiao, Yanyuan, et al.
Publicado: (2025)
Leveraging RGB-D Data with Cross-Modal Context Mining for Glass Surface Detection
por: Lin, Jiaying, et al.
Publicado: (2022)
por: Lin, Jiaying, et al.
Publicado: (2022)
Ejemplares similares
-
Modification Takes Courage: Seamless Image Stitching via Reference-Driven Inpainting
por: Xie, Ziqi, et al.
Publicado: (2024) -
MEIA: Multimodal Embodied Perception and Interaction in Unknown Environments
por: Liu, Yang, et al.
Publicado: (2024) -
TUNI: Real-time RGB-T Semantic Segmentation with Unified Multi-Modal Feature Extraction and Cross-Modal Feature Fusion
por: Guo, Xiaodong, et al.
Publicado: (2025) -
RACANet: Reliability-Aware Crowd Anchor Network for RGB-T Crowd Counting
por: Shi, Jinghao, et al.
Publicado: (2026) -
LER-YOLO: Reliability-Aware Expert Routing for Misaligned RGB-Infrared UAV Detection
por: Hou, Liming, et al.
Publicado: (2026)