Saved in:
| Main Authors: | Nguyen, Huy Hoang, Jung, Cédric, Salehi, Shirin, Glück, Tobias, Schmeink, Anke, Kugi, Andreas |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2603.23159 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Active Learning Using Aggregated Acquisition Functions: Accuracy and Sustainability Analysis
by: Jung, Cédric, et al.
Published: (2026)
by: Jung, Cédric, et al.
Published: (2026)
Lang2Lift: A Language-Guided Autonomous Forklift System for Outdoor Industrial Pallet Handling
by: Nguyen, Huy Hoang, et al.
Published: (2025)
by: Nguyen, Huy Hoang, et al.
Published: (2025)
A Full Compression Pipeline for Green Federated Learning in Communication-Constrained Environments
by: Colybes, Elouan, et al.
Published: (2026)
by: Colybes, Elouan, et al.
Published: (2026)
LoG-VMamba: Local-Global Vision Mamba for Medical Image Segmentation
by: Dang, Trung Dinh Quoc, et al.
Published: (2024)
by: Dang, Trung Dinh Quoc, et al.
Published: (2024)
MolFM-Lite: Multi-Modal Molecular Property Prediction with Conformer Ensemble Attention and Cross-Modal Fusion
by: Shah, Syed Omer, et al.
Published: (2026)
by: Shah, Syed Omer, et al.
Published: (2026)
ActiveAnno3D -- An Active Learning Framework for Multi-Modal 3D Object Detection
by: Ghita, Ahmed, et al.
Published: (2024)
by: Ghita, Ahmed, et al.
Published: (2024)
Cross-Modal Emotion Transfer for Emotion Editing in Talking Face Video
by: Choi, Chanhyuk, et al.
Published: (2026)
by: Choi, Chanhyuk, et al.
Published: (2026)
Mind the Gap: Learning Modality-Agnostic Representations with a Cross-Modality UNet
by: Niu, Xin, et al.
Published: (2026)
by: Niu, Xin, et al.
Published: (2026)
On the Value of Cross-Modal Misalignment in Multimodal Representation Learning
by: Cai, Yichao, et al.
Published: (2025)
by: Cai, Yichao, et al.
Published: (2025)
HiWave: Training-Free High-Resolution Image Generation via Wavelet-Based Diffusion Sampling
by: Vontobel, Tobias, et al.
Published: (2025)
by: Vontobel, Tobias, et al.
Published: (2025)
PROMISE: Prompt-Attentive Hierarchical Contrastive Learning for Robust Cross-Modal Representation with Missing Modalities
by: Chen, Jiajun, et al.
Published: (2025)
by: Chen, Jiajun, et al.
Published: (2025)
Improved Segmentation of Polyps and Visual Explainability Analysis
by: Asare, Akwasi, et al.
Published: (2025)
by: Asare, Akwasi, et al.
Published: (2025)
A Generalization Theory of Cross-Modality Distillation with Contrastive Learning
by: Lin, Hangyu, et al.
Published: (2024)
by: Lin, Hangyu, et al.
Published: (2024)
EVL-ECG: Efficient ECG Interpretation With Multi-Aspect Heterogeneous Knowledge Distillation
by: Hong, Dang Nguyen, et al.
Published: (2026)
by: Hong, Dang Nguyen, et al.
Published: (2026)
Enhancing Vietnamese VQA through Curriculum Learning on Raw and Augmented Text Representations
by: Nguyen, Khoi Anh, et al.
Published: (2025)
by: Nguyen, Khoi Anh, et al.
Published: (2025)
Retrospective Feature Estimation for Continual Learning
by: Nguyen, Nghia D., et al.
Published: (2024)
by: Nguyen, Nghia D., et al.
Published: (2024)
Revisit Visual Prompt Tuning: The Expressiveness of Prompt Experts
by: Le, Minh, et al.
Published: (2025)
by: Le, Minh, et al.
Published: (2025)
Cross-Modal Few-Shot Learning: a Generative Transfer Learning Framework
by: Yang, Zhengwei, et al.
Published: (2024)
by: Yang, Zhengwei, et al.
Published: (2024)
Data Selection for Fine-tuning Vision Language Models via Cross Modal Alignment Trajectories
by: Naharas, Nilay, et al.
Published: (2025)
by: Naharas, Nilay, et al.
Published: (2025)
Is Hierarchical Quantization Essential for Optimal Reconstruction?
by: Reyhanian, Shirin, et al.
Published: (2026)
by: Reyhanian, Shirin, et al.
Published: (2026)
Cross-modal Active Complementary Learning with Self-refining Correspondence
by: Qin, Yang, et al.
Published: (2023)
by: Qin, Yang, et al.
Published: (2023)
Enhancing Cross-Modal Fine-Tuning with Gradually Intermediate Modality Generation
by: Cai, Lincan, et al.
Published: (2024)
by: Cai, Lincan, et al.
Published: (2024)
Bayesian Cross-Modal Alignment Learning for Few-Shot Out-of-Distribution Generalization
by: Zhu, Lin, et al.
Published: (2025)
by: Zhu, Lin, et al.
Published: (2025)
Learnable Cross-modal Knowledge Distillation for Multi-modal Learning with Missing Modality
by: Wang, Hu, et al.
Published: (2023)
by: Wang, Hu, et al.
Published: (2023)
Ensemble Learning for Vietnamese Scene Text Spotting in Urban Environments
by: Nguyen, Hieu, et al.
Published: (2024)
by: Nguyen, Hieu, et al.
Published: (2024)
Enhancing CTC-Based Visual Speech Recognition
by: Laux, Hendrik, et al.
Published: (2024)
by: Laux, Hendrik, et al.
Published: (2024)
A Controllable 3D Deepfake Generation Framework with Gaussian Splatting
by: Liu, Wending, et al.
Published: (2025)
by: Liu, Wending, et al.
Published: (2025)
Cross-Modal Adapter: Parameter-Efficient Transfer Learning Approach for Vision-Language Models
by: Yang, Juncheng, et al.
Published: (2024)
by: Yang, Juncheng, et al.
Published: (2024)
QLAM: A Quantum Long-Attention Memory Approach to Long-Sequence Token Modeling
by: Nguyen, Hoang-Quan, et al.
Published: (2026)
by: Nguyen, Hoang-Quan, et al.
Published: (2026)
Generalized Contrastive Learning for Multi-Modal Retrieval and Ranking
by: Zhu, Tianyu, et al.
Published: (2024)
by: Zhu, Tianyu, et al.
Published: (2024)
$φ$-Adapt: A Physics-Informed Adaptation Learning Approach to 2D Quantum Material Discovery
by: Nguyen, Hoang-Quan, et al.
Published: (2025)
by: Nguyen, Hoang-Quan, et al.
Published: (2025)
Cross-Modal Redundancy and the Geometry of Vision-Language Embeddings
by: Dhimoïla, Grégoire, et al.
Published: (2026)
by: Dhimoïla, Grégoire, et al.
Published: (2026)
Multimodal Structure Learning: Disentangling Shared and Specific Topology via Cross-Modal Graphical Lasso
by: Wang, Fei, et al.
Published: (2026)
by: Wang, Fei, et al.
Published: (2026)
Quantifying Cross-Modality Memorization in Vision-Language Models
by: Wen, Yuxin, et al.
Published: (2025)
by: Wen, Yuxin, et al.
Published: (2025)
GNSP: Gradient Null Space Projection for Preserving Cross-Modal Alignment in VLMs Continual Learning
by: Peng, Tiantian, et al.
Published: (2025)
by: Peng, Tiantian, et al.
Published: (2025)
Balanced Multi-modal Federated Learning via Cross-Modal Infiltration
by: Fan, Yunfeng, et al.
Published: (2023)
by: Fan, Yunfeng, et al.
Published: (2023)
Cross-Modal Coordination Across a Diverse Set of Input Modalities
by: Sánchez, Jorge, et al.
Published: (2024)
by: Sánchez, Jorge, et al.
Published: (2024)
Learning Conformal Explainers for Image Classifiers
by: Alkhatib, Amr, et al.
Published: (2025)
by: Alkhatib, Amr, et al.
Published: (2025)
Unveiling Concept Attribution in Diffusion Models
by: Nguyen, Quang H., et al.
Published: (2024)
by: Nguyen, Quang H., et al.
Published: (2024)
Troika: Multi-Path Cross-Modal Traction for Compositional Zero-Shot Learning
by: Huang, Siteng, et al.
Published: (2023)
by: Huang, Siteng, et al.
Published: (2023)
Similar Items
-
Active Learning Using Aggregated Acquisition Functions: Accuracy and Sustainability Analysis
by: Jung, Cédric, et al.
Published: (2026) -
Lang2Lift: A Language-Guided Autonomous Forklift System for Outdoor Industrial Pallet Handling
by: Nguyen, Huy Hoang, et al.
Published: (2025) -
A Full Compression Pipeline for Green Federated Learning in Communication-Constrained Environments
by: Colybes, Elouan, et al.
Published: (2026) -
LoG-VMamba: Local-Global Vision Mamba for Medical Image Segmentation
by: Dang, Trung Dinh Quoc, et al.
Published: (2024) -
MolFM-Lite: Multi-Modal Molecular Property Prediction with Conformer Ensemble Attention and Cross-Modal Fusion
by: Shah, Syed Omer, et al.
Published: (2026)