Gespeichert in:
| Hauptverfasser: | Suharitdamrong, Wish, Alex, Tony, Awais, Muhammad, Ahmed, Sara |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2604.03314 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
PAL: Probing Audio Encoders via LLMs -- Audio Information Transfer into LLMs
von: Alex, Tony, et al.
Veröffentlicht: (2025)
von: Alex, Tony, et al.
Veröffentlicht: (2025)
Domain Adaptation Without the Compute Burden for Efficient Whole Slide Image Analysis
von: Marikkar, Umar, et al.
Veröffentlicht: (2026)
von: Marikkar, Umar, et al.
Veröffentlicht: (2026)
Omni-SMoLA: Boosting Generalist Multimodal Models with Soft Mixture of Low-rank Experts
von: Wu, Jialin, et al.
Veröffentlicht: (2023)
von: Wu, Jialin, et al.
Veröffentlicht: (2023)
Expressive and Generalizable Low-rank Adaptation for Large Models via Slow Cascaded Learning
von: Li, Siwei, et al.
Veröffentlicht: (2024)
von: Li, Siwei, et al.
Veröffentlicht: (2024)
CoLA: Collaborative Low-Rank Adaptation
von: Zhou, Yiyun, et al.
Veröffentlicht: (2025)
von: Zhou, Yiyun, et al.
Veröffentlicht: (2025)
CoLA: Conditional Dropout and Language-driven Robust Dual-modal Salient Object Detection
von: Hao, Shuang, et al.
Veröffentlicht: (2024)
von: Hao, Shuang, et al.
Veröffentlicht: (2024)
DeLoRA: Decoupling Angles and Strength in Low-rank Adaptation
von: Bini, Massimo, et al.
Veröffentlicht: (2025)
von: Bini, Massimo, et al.
Veröffentlicht: (2025)
Dynamic Context-oriented Decomposition for Task-aware Low-rank Adaptation with Less Forgetting and Faster Convergence
von: Yang, Yibo, et al.
Veröffentlicht: (2025)
von: Yang, Yibo, et al.
Veröffentlicht: (2025)
Leveraging Entity Information for Cross-Modality Correlation Learning: The Entity-Guided Multimodal Summarization
von: Zhang, Yanghai, et al.
Veröffentlicht: (2024)
von: Zhang, Yanghai, et al.
Veröffentlicht: (2024)
Low-Rank Adaptation with Task-Relevant Feature Enhancement for Fine-tuning Language Models
von: Li, Changqun, et al.
Veröffentlicht: (2024)
von: Li, Changqun, et al.
Veröffentlicht: (2024)
CROME: Cross-Modal Adapters for Efficient Multimodal LLM
von: Ebrahimi, Sayna, et al.
Veröffentlicht: (2024)
von: Ebrahimi, Sayna, et al.
Veröffentlicht: (2024)
Vision-Language Models Create Cross-Modal Task Representations
von: Luo, Grace, et al.
Veröffentlicht: (2024)
von: Luo, Grace, et al.
Veröffentlicht: (2024)
LowCLIP: Adapting the CLIP Model Architecture for Low-Resource Languages in Multimodal Image Retrieval Task
von: Asgarov, Ali, et al.
Veröffentlicht: (2024)
von: Asgarov, Ali, et al.
Veröffentlicht: (2024)
Shared and Private Information Learning in Multimodal Sentiment Analysis with Deep Modal Alignment and Self-supervised Multi-Task Learning
von: Lai, Songning, et al.
Veröffentlicht: (2023)
von: Lai, Songning, et al.
Veröffentlicht: (2023)
$\mathcal{V}isi\mathcal{P}runer$: Decoding Discontinuous Cross-Modal Dynamics for Efficient Multimodal LLMs
von: Fan, Yingqi, et al.
Veröffentlicht: (2025)
von: Fan, Yingqi, et al.
Veröffentlicht: (2025)
CoLA: A Choice Leakage Attack Framework to Expose Privacy Risks in Subset Training
von: Li, Qi, et al.
Veröffentlicht: (2026)
von: Li, Qi, et al.
Veröffentlicht: (2026)
CMAP: Cross-Modal Adaptive Prompting for Multi-Domain Task-Incremental Learning
von: Mandalika, Sriram
Veröffentlicht: (2026)
von: Mandalika, Sriram
Veröffentlicht: (2026)
Connect, Collapse, Corrupt: Learning Cross-Modal Tasks with Uni-Modal Data
von: Zhang, Yuhui, et al.
Veröffentlicht: (2024)
von: Zhang, Yuhui, et al.
Veröffentlicht: (2024)
Analyzing Reasoning Consistency in Large Multimodal Models under Cross-Modal Conflicts
von: Zhu, Zhihao, et al.
Veröffentlicht: (2026)
von: Zhu, Zhihao, et al.
Veröffentlicht: (2026)
Deciphering Cross-Modal Alignment in Large Vision-Language Models with Modality Integration Rate
von: Huang, Qidong, et al.
Veröffentlicht: (2024)
von: Huang, Qidong, et al.
Veröffentlicht: (2024)
MoExtend: Tuning New Experts for Modality and Task Extension
von: Zhong, Shanshan, et al.
Veröffentlicht: (2024)
von: Zhong, Shanshan, et al.
Veröffentlicht: (2024)
Agent-X: Evaluating Deep Multimodal Reasoning in Vision-Centric Agentic Tasks
von: Ashraf, Tajamul, et al.
Veröffentlicht: (2025)
von: Ashraf, Tajamul, et al.
Veröffentlicht: (2025)
CoTasks: Chain-of-Thought based Video Instruction Tuning Tasks
von: Wang, Yanan, et al.
Veröffentlicht: (2025)
von: Wang, Yanan, et al.
Veröffentlicht: (2025)
One Model for ALL: Low-Level Task Interaction Is a Key to Task-Agnostic Image Fusion
von: Cheng, Chunyang, et al.
Veröffentlicht: (2025)
von: Cheng, Chunyang, et al.
Veröffentlicht: (2025)
Efficient Stitchable Task Adaptation
von: He, Haoyu, et al.
Veröffentlicht: (2023)
von: He, Haoyu, et al.
Veröffentlicht: (2023)
Multimodal LLM Enhanced Cross-lingual Cross-modal Retrieval
von: Wang, Yabing, et al.
Veröffentlicht: (2024)
von: Wang, Yabing, et al.
Veröffentlicht: (2024)
MANTA: Cross-Modal Semantic Alignment and Information-Theoretic Optimization for Long-form Multimodal Understanding
von: Zhong, Ziqi, et al.
Veröffentlicht: (2025)
von: Zhong, Ziqi, et al.
Veröffentlicht: (2025)
Cross-Modal Projection in Multimodal LLMs Doesn't Really Project Visual Attributes to Textual Space
von: Verma, Gaurav, et al.
Veröffentlicht: (2024)
von: Verma, Gaurav, et al.
Veröffentlicht: (2024)
DoRA: Weight-Decomposed Low-Rank Adaptation
von: Liu, Shih-Yang, et al.
Veröffentlicht: (2024)
von: Liu, Shih-Yang, et al.
Veröffentlicht: (2024)
TASO: Task-Aligned Sparse Optimization for Parameter-Efficient Model Adaptation
von: Miao, Daiye, et al.
Veröffentlicht: (2025)
von: Miao, Daiye, et al.
Veröffentlicht: (2025)
Multimodal Prompt Learning with Missing Modalities for Sentiment Analysis and Emotion Recognition
von: Guo, Zirun, et al.
Veröffentlicht: (2024)
von: Guo, Zirun, et al.
Veröffentlicht: (2024)
Anthropogenic Regional Adaptation in Multimodal Vision-Language Model
von: Cahyawijaya, Samuel, et al.
Veröffentlicht: (2026)
von: Cahyawijaya, Samuel, et al.
Veröffentlicht: (2026)
Cross-Modal Rationale Transfer for Explainable Humanitarian Classification on Social Media
von: Nguyen, Thi Huyen, et al.
Veröffentlicht: (2026)
von: Nguyen, Thi Huyen, et al.
Veröffentlicht: (2026)
Cross-Modal Obfuscation for Jailbreak Attacks on Large Vision-Language Models
von: Jiang, Lei, et al.
Veröffentlicht: (2025)
von: Jiang, Lei, et al.
Veröffentlicht: (2025)
Cross-Modal Retrieval for Motion and Text via DropTriple Loss
von: Yan, Sheng, et al.
Veröffentlicht: (2023)
von: Yan, Sheng, et al.
Veröffentlicht: (2023)
Unraveling Cross-Modality Knowledge Conflicts in Large Vision-Language Models
von: Zhu, Tinghui, et al.
Veröffentlicht: (2024)
von: Zhu, Tinghui, et al.
Veröffentlicht: (2024)
Evaluating Cross-Modal Reasoning Ability and Problem Characteristics with Multimodal Item Response Theory
von: Uebayashi, Shunki, et al.
Veröffentlicht: (2026)
von: Uebayashi, Shunki, et al.
Veröffentlicht: (2026)
Summarization of Multimodal Presentations with Vision-Language Models: Study of the Effect of Modalities and Structure
von: Gigant, Théo, et al.
Veröffentlicht: (2025)
von: Gigant, Théo, et al.
Veröffentlicht: (2025)
Co-AttenDWG: Co-Attentive Dimension-Wise Gating and Expert Fusion for Multi-Modal Offensive Content Detection
von: Hossain, Md. Mithun, et al.
Veröffentlicht: (2025)
von: Hossain, Md. Mithun, et al.
Veröffentlicht: (2025)
Understanding Multimodal Procedural Knowledge by Sequencing Multimodal Instructional Manuals
von: Wu, Te-Lin, et al.
Veröffentlicht: (2021)
von: Wu, Te-Lin, et al.
Veröffentlicht: (2021)
Ähnliche Einträge
-
PAL: Probing Audio Encoders via LLMs -- Audio Information Transfer into LLMs
von: Alex, Tony, et al.
Veröffentlicht: (2025) -
Domain Adaptation Without the Compute Burden for Efficient Whole Slide Image Analysis
von: Marikkar, Umar, et al.
Veröffentlicht: (2026) -
Omni-SMoLA: Boosting Generalist Multimodal Models with Soft Mixture of Low-rank Experts
von: Wu, Jialin, et al.
Veröffentlicht: (2023) -
Expressive and Generalizable Low-rank Adaptation for Large Models via Slow Cascaded Learning
von: Li, Siwei, et al.
Veröffentlicht: (2024) -
CoLA: Collaborative Low-Rank Adaptation
von: Zhou, Yiyun, et al.
Veröffentlicht: (2025)