CoLA: Cross-Modal Low-rank Adaptation for Multimodal Downstream Tasks
Fuente:
arXiv
Salvato in:
| Autori principali: | Suharitdamrong, Wish, Alex, Tony, Awais, Muhammad, Ahmed, Sara |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
PAL: Probing Audio Encoders via LLMs -- Audio Information Transfer into LLMs
di: Alex, Tony, et al.
Pubblicazione: (2025)
di: Alex, Tony, et al.
Pubblicazione: (2025)
Omni-SMoLA: Boosting Generalist Multimodal Models with Soft Mixture of Low-rank Experts
di: Wu, Jialin, et al.
Pubblicazione: (2023)
di: Wu, Jialin, et al.
Pubblicazione: (2023)
Expressive and Generalizable Low-rank Adaptation for Large Models via Slow Cascaded Learning
di: Li, Siwei, et al.
Pubblicazione: (2024)
di: Li, Siwei, et al.
Pubblicazione: (2024)
Domain Adaptation Without the Compute Burden for Efficient Whole Slide Image Analysis
di: Marikkar, Umar, et al.
Pubblicazione: (2026)
di: Marikkar, Umar, et al.
Pubblicazione: (2026)
CoLA: Collaborative Low-Rank Adaptation
di: Zhou, Yiyun, et al.
Pubblicazione: (2025)
di: Zhou, Yiyun, et al.
Pubblicazione: (2025)
CoLA: Conditional Dropout and Language-driven Robust Dual-modal Salient Object Detection
di: Hao, Shuang, et al.
Pubblicazione: (2024)
di: Hao, Shuang, et al.
Pubblicazione: (2024)
DeLoRA: Decoupling Angles and Strength in Low-rank Adaptation
di: Bini, Massimo, et al.
Pubblicazione: (2025)
di: Bini, Massimo, et al.
Pubblicazione: (2025)
Leveraging Entity Information for Cross-Modality Correlation Learning: The Entity-Guided Multimodal Summarization
di: Zhang, Yanghai, et al.
Pubblicazione: (2024)
di: Zhang, Yanghai, et al.
Pubblicazione: (2024)
Low-Rank Adaptation with Task-Relevant Feature Enhancement for Fine-tuning Language Models
di: Li, Changqun, et al.
Pubblicazione: (2024)
di: Li, Changqun, et al.
Pubblicazione: (2024)
LowCLIP: Adapting the CLIP Model Architecture for Low-Resource Languages in Multimodal Image Retrieval Task
di: Asgarov, Ali, et al.
Pubblicazione: (2024)
di: Asgarov, Ali, et al.
Pubblicazione: (2024)
Dynamic Context-oriented Decomposition for Task-aware Low-rank Adaptation with Less Forgetting and Faster Convergence
di: Yang, Yibo, et al.
Pubblicazione: (2025)
di: Yang, Yibo, et al.
Pubblicazione: (2025)
CROME: Cross-Modal Adapters for Efficient Multimodal LLM
di: Ebrahimi, Sayna, et al.
Pubblicazione: (2024)
di: Ebrahimi, Sayna, et al.
Pubblicazione: (2024)
Vision-Language Models Create Cross-Modal Task Representations
di: Luo, Grace, et al.
Pubblicazione: (2024)
di: Luo, Grace, et al.
Pubblicazione: (2024)
Shared and Private Information Learning in Multimodal Sentiment Analysis with Deep Modal Alignment and Self-supervised Multi-Task Learning
di: Lai, Songning, et al.
Pubblicazione: (2023)
di: Lai, Songning, et al.
Pubblicazione: (2023)
$\mathcal{V}isi\mathcal{P}runer$: Decoding Discontinuous Cross-Modal Dynamics for Efficient Multimodal LLMs
di: Fan, Yingqi, et al.
Pubblicazione: (2025)
di: Fan, Yingqi, et al.
Pubblicazione: (2025)
CMAP: Cross-Modal Adaptive Prompting for Multi-Domain Task-Incremental Learning
di: Mandalika, Sriram
Pubblicazione: (2026)
di: Mandalika, Sriram
Pubblicazione: (2026)
Deciphering Cross-Modal Alignment in Large Vision-Language Models with Modality Integration Rate
di: Huang, Qidong, et al.
Pubblicazione: (2024)
di: Huang, Qidong, et al.
Pubblicazione: (2024)
MoExtend: Tuning New Experts for Modality and Task Extension
di: Zhong, Shanshan, et al.
Pubblicazione: (2024)
di: Zhong, Shanshan, et al.
Pubblicazione: (2024)
CoLA: A Choice Leakage Attack Framework to Expose Privacy Risks in Subset Training
di: Li, Qi, et al.
Pubblicazione: (2026)
di: Li, Qi, et al.
Pubblicazione: (2026)
Analyzing Reasoning Consistency in Large Multimodal Models under Cross-Modal Conflicts
di: Zhu, Zhihao, et al.
Pubblicazione: (2026)
di: Zhu, Zhihao, et al.
Pubblicazione: (2026)
CoTasks: Chain-of-Thought based Video Instruction Tuning Tasks
di: Wang, Yanan, et al.
Pubblicazione: (2025)
di: Wang, Yanan, et al.
Pubblicazione: (2025)
Agent-X: Evaluating Deep Multimodal Reasoning in Vision-Centric Agentic Tasks
di: Ashraf, Tajamul, et al.
Pubblicazione: (2025)
di: Ashraf, Tajamul, et al.
Pubblicazione: (2025)
Connect, Collapse, Corrupt: Learning Cross-Modal Tasks with Uni-Modal Data
di: Zhang, Yuhui, et al.
Pubblicazione: (2024)
di: Zhang, Yuhui, et al.
Pubblicazione: (2024)
Multimodal LLM Enhanced Cross-lingual Cross-modal Retrieval
di: Wang, Yabing, et al.
Pubblicazione: (2024)
di: Wang, Yabing, et al.
Pubblicazione: (2024)
DoRA: Weight-Decomposed Low-Rank Adaptation
di: Liu, Shih-Yang, et al.
Pubblicazione: (2024)
di: Liu, Shih-Yang, et al.
Pubblicazione: (2024)
TASO: Task-Aligned Sparse Optimization for Parameter-Efficient Model Adaptation
di: Miao, Daiye, et al.
Pubblicazione: (2025)
di: Miao, Daiye, et al.
Pubblicazione: (2025)
Multimodal Prompt Learning with Missing Modalities for Sentiment Analysis and Emotion Recognition
di: Guo, Zirun, et al.
Pubblicazione: (2024)
di: Guo, Zirun, et al.
Pubblicazione: (2024)
One Model for ALL: Low-Level Task Interaction Is a Key to Task-Agnostic Image Fusion
di: Cheng, Chunyang, et al.
Pubblicazione: (2025)
di: Cheng, Chunyang, et al.
Pubblicazione: (2025)
Cross-Modal Rationale Transfer for Explainable Humanitarian Classification on Social Media
di: Nguyen, Thi Huyen, et al.
Pubblicazione: (2026)
di: Nguyen, Thi Huyen, et al.
Pubblicazione: (2026)
Cross-Modal Obfuscation for Jailbreak Attacks on Large Vision-Language Models
di: Jiang, Lei, et al.
Pubblicazione: (2025)
di: Jiang, Lei, et al.
Pubblicazione: (2025)
Cross-Modal Retrieval for Motion and Text via DropTriple Loss
di: Yan, Sheng, et al.
Pubblicazione: (2023)
di: Yan, Sheng, et al.
Pubblicazione: (2023)
Unraveling Cross-Modality Knowledge Conflicts in Large Vision-Language Models
di: Zhu, Tinghui, et al.
Pubblicazione: (2024)
di: Zhu, Tinghui, et al.
Pubblicazione: (2024)
MANTA: Cross-Modal Semantic Alignment and Information-Theoretic Optimization for Long-form Multimodal Understanding
di: Zhong, Ziqi, et al.
Pubblicazione: (2025)
di: Zhong, Ziqi, et al.
Pubblicazione: (2025)
Cross-Modal Projection in Multimodal LLMs Doesn't Really Project Visual Attributes to Textual Space
di: Verma, Gaurav, et al.
Pubblicazione: (2024)
di: Verma, Gaurav, et al.
Pubblicazione: (2024)
Efficient Stitchable Task Adaptation
di: He, Haoyu, et al.
Pubblicazione: (2023)
di: He, Haoyu, et al.
Pubblicazione: (2023)
Summarization of Multimodal Presentations with Vision-Language Models: Study of the Effect of Modalities and Structure
di: Gigant, Théo, et al.
Pubblicazione: (2025)
di: Gigant, Théo, et al.
Pubblicazione: (2025)
Co-AttenDWG: Co-Attentive Dimension-Wise Gating and Expert Fusion for Multi-Modal Offensive Content Detection
di: Hossain, Md. Mithun, et al.
Pubblicazione: (2025)
di: Hossain, Md. Mithun, et al.
Pubblicazione: (2025)
Exploring Typographic Visual Prompts Injection Threats in Cross-Modality Generation Models
di: Cheng, Hao, et al.
Pubblicazione: (2025)
di: Cheng, Hao, et al.
Pubblicazione: (2025)
Through the Lens of Character: Resolving Modality-Role Interference in Multimodal Role-Playing Agent
di: Tang, Yihong, et al.
Pubblicazione: (2026)
di: Tang, Yihong, et al.
Pubblicazione: (2026)
Understanding Multimodal Procedural Knowledge by Sequencing Multimodal Instructional Manuals
di: Wu, Te-Lin, et al.
Pubblicazione: (2021)
di: Wu, Te-Lin, et al.
Pubblicazione: (2021)
Documenti analoghi
-
PAL: Probing Audio Encoders via LLMs -- Audio Information Transfer into LLMs
di: Alex, Tony, et al.
Pubblicazione: (2025) -
Omni-SMoLA: Boosting Generalist Multimodal Models with Soft Mixture of Low-rank Experts
di: Wu, Jialin, et al.
Pubblicazione: (2023) -
Expressive and Generalizable Low-rank Adaptation for Large Models via Slow Cascaded Learning
di: Li, Siwei, et al.
Pubblicazione: (2024) -
Domain Adaptation Without the Compute Burden for Efficient Whole Slide Image Analysis
di: Marikkar, Umar, et al.
Pubblicazione: (2026) -
CoLA: Collaborative Low-Rank Adaptation
di: Zhou, Yiyun, et al.
Pubblicazione: (2025)