Task Matrices: Linear Maps for Cross-Model Finetuning Transfer
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Brien, Darrin O', Gajulapalli, Dhikshith, Xia, Eric |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
ETHER: Efficient Finetuning of Large-Scale Models with Hyperplane Reflections
von: Bini, Massimo, et al.
Veröffentlicht: (2024)
von: Bini, Massimo, et al.
Veröffentlicht: (2024)
Vision-Language Models Create Cross-Modal Task Representations
von: Luo, Grace, et al.
Veröffentlicht: (2024)
von: Luo, Grace, et al.
Veröffentlicht: (2024)
Orthogonal Finetuning Made Scalable
von: Qiu, Zeju, et al.
Veröffentlicht: (2025)
von: Qiu, Zeju, et al.
Veröffentlicht: (2025)
Exposing and Addressing Cross-Task Inconsistency in Unified Vision-Language Models
von: Maharana, Adyasha, et al.
Veröffentlicht: (2023)
von: Maharana, Adyasha, et al.
Veröffentlicht: (2023)
Linear Alignment of Vision-language Models for Image Captioning
von: Paischer, Fabian, et al.
Veröffentlicht: (2023)
von: Paischer, Fabian, et al.
Veröffentlicht: (2023)
Parameter-Efficient Orthogonal Finetuning via Butterfly Factorization
von: Liu, Weiyang, et al.
Veröffentlicht: (2023)
von: Liu, Weiyang, et al.
Veröffentlicht: (2023)
CROME: Cross-Modal Adapters for Efficient Multimodal LLM
von: Ebrahimi, Sayna, et al.
Veröffentlicht: (2024)
von: Ebrahimi, Sayna, et al.
Veröffentlicht: (2024)
BYOM: Building Your Own Multi-Task Model For Free
von: Jiang, Weisen, et al.
Veröffentlicht: (2023)
von: Jiang, Weisen, et al.
Veröffentlicht: (2023)
Multi-Task Model Merging via Adaptive Weight Disentanglement
von: Xiong, Feng, et al.
Veröffentlicht: (2024)
von: Xiong, Feng, et al.
Veröffentlicht: (2024)
MoLEx: Mixture of Layer Experts for Finetuning with Sparse Upcycling
von: Teo, Rachel S. Y., et al.
Veröffentlicht: (2025)
von: Teo, Rachel S. Y., et al.
Veröffentlicht: (2025)
RadZero: Similarity-Based Cross-Attention for Explainable Vision-Language Alignment in Chest X-ray with Zero-Shot Multi-Task Capability
von: Park, Jonggwon, et al.
Veröffentlicht: (2025)
von: Park, Jonggwon, et al.
Veröffentlicht: (2025)
Localize-and-Stitch: Efficient Model Merging via Sparse Task Arithmetic
von: He, Yifei, et al.
Veröffentlicht: (2024)
von: He, Yifei, et al.
Veröffentlicht: (2024)
VLSM-Adapter: Finetuning Vision-Language Segmentation Efficiently with Lightweight Blocks
von: Dhakal, Manish, et al.
Veröffentlicht: (2024)
von: Dhakal, Manish, et al.
Veröffentlicht: (2024)
LLM-CXR: Instruction-Finetuned LLM for CXR Image Understanding and Generation
von: Lee, Suhyeon, et al.
Veröffentlicht: (2023)
von: Lee, Suhyeon, et al.
Veröffentlicht: (2023)
MoTe: Learning Motion-Text Diffusion Model for Multiple Generation Tasks
von: Wu, Yiming, et al.
Veröffentlicht: (2024)
von: Wu, Yiming, et al.
Veröffentlicht: (2024)
X-VILA: Cross-Modality Alignment for Large Language Model
von: Ye, Hanrong, et al.
Veröffentlicht: (2024)
von: Ye, Hanrong, et al.
Veröffentlicht: (2024)
How LoRA Remembers? A Parametric Memory Law for LLM Finetuning
von: Xu, Ziwen, et al.
Veröffentlicht: (2026)
von: Xu, Ziwen, et al.
Veröffentlicht: (2026)
Efficient Stitchable Task Adaptation
von: He, Haoyu, et al.
Veröffentlicht: (2023)
von: He, Haoyu, et al.
Veröffentlicht: (2023)
How Vision-Language Tasks Benefit from Large Pre-trained Models: A Survey
von: Qi, Yayun, et al.
Veröffentlicht: (2024)
von: Qi, Yayun, et al.
Veröffentlicht: (2024)
Mind the Gap Between Prototypes and Images in Cross-domain Finetuning
von: Tian, Hongduan, et al.
Veröffentlicht: (2024)
von: Tian, Hongduan, et al.
Veröffentlicht: (2024)
Revisiting Mixout: An Overlooked Path to Robust Finetuning
von: Aminbeidokhti, Masih, et al.
Veröffentlicht: (2025)
von: Aminbeidokhti, Masih, et al.
Veröffentlicht: (2025)
Connect, Collapse, Corrupt: Learning Cross-Modal Tasks with Uni-Modal Data
von: Zhang, Yuhui, et al.
Veröffentlicht: (2024)
von: Zhang, Yuhui, et al.
Veröffentlicht: (2024)
Transformer-VQ: Linear-Time Transformers via Vector Quantization
von: Lingle, Lucas D.
Veröffentlicht: (2023)
von: Lingle, Lucas D.
Veröffentlicht: (2023)
Transferability-Guided Cross-Domain Cross-Task Transfer Learning
von: Tan, Yang, et al.
Veröffentlicht: (2022)
von: Tan, Yang, et al.
Veröffentlicht: (2022)
Neural Style Transfer for Synthesising a Dataset of Ancient Egyptian Hieroglyphs
von: Creed, Lewis Matheson
Veröffentlicht: (2025)
von: Creed, Lewis Matheson
Veröffentlicht: (2025)
MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation
von: Joshi, Siddharth, et al.
Veröffentlicht: (2025)
von: Joshi, Siddharth, et al.
Veröffentlicht: (2025)
Task-Specific Directions: Definition, Exploration, and Utilization in Parameter Efficient Fine-Tuning
von: Si, Chongjie, et al.
Veröffentlicht: (2024)
von: Si, Chongjie, et al.
Veröffentlicht: (2024)
Language-Pretraining-Induced Bias: A Strong Foundation for General Vision Tasks
von: Luo, Yaxin, et al.
Veröffentlicht: (2026)
von: Luo, Yaxin, et al.
Veröffentlicht: (2026)
Monkey Jump : MoE-Style PEFT for Efficient Multi-Task Learning
von: Prottasha, Nusrat Jahan, et al.
Veröffentlicht: (2026)
von: Prottasha, Nusrat Jahan, et al.
Veröffentlicht: (2026)
VisualWebArena: Evaluating Multimodal Agents on Realistic Visual Web Tasks
von: Koh, Jing Yu, et al.
Veröffentlicht: (2024)
von: Koh, Jing Yu, et al.
Veröffentlicht: (2024)
GenSim: Generating Robotic Simulation Tasks via Large Language Models
von: Wang, Lirui, et al.
Veröffentlicht: (2023)
von: Wang, Lirui, et al.
Veröffentlicht: (2023)
MapIQ: Evaluating Multimodal Large Language Models for Map Question Answering
von: Srivastava, Varun, et al.
Veröffentlicht: (2025)
von: Srivastava, Varun, et al.
Veröffentlicht: (2025)
VLM Judges Can Rank but Cannot Score: Task-Dependent Uncertainty in Multimodal Evaluation
von: Kumar, Divake, et al.
Veröffentlicht: (2026)
von: Kumar, Divake, et al.
Veröffentlicht: (2026)
Crossing Language Borders: A Pipeline for Indonesian Manhwa Translation
von: Narasimhan, Nithyasri, et al.
Veröffentlicht: (2025)
von: Narasimhan, Nithyasri, et al.
Veröffentlicht: (2025)
Cross-modal Causal Relation Alignment for Video Question Grounding
von: Chen, Weixing, et al.
Veröffentlicht: (2025)
von: Chen, Weixing, et al.
Veröffentlicht: (2025)
Text-to-Image Cross-Modal Generation: A Systematic Review
von: Żelaszczyk, Maciej, et al.
Veröffentlicht: (2024)
von: Żelaszczyk, Maciej, et al.
Veröffentlicht: (2024)
Understanding Space Is Rocket Science -- Only Top Reasoning Models Can Solve Spatial Understanding Tasks
von: Hoehing, Nils, et al.
Veröffentlicht: (2025)
von: Hoehing, Nils, et al.
Veröffentlicht: (2025)
The Double Dilemma in Multi-Task Radiology Report Generation: A Gradient Dynamics Analysis and Solution
von: Zhang, Erjian, et al.
Veröffentlicht: (2026)
von: Zhang, Erjian, et al.
Veröffentlicht: (2026)
MMed-RAG: Versatile Multimodal RAG System for Medical Vision Language Models
von: Xia, Peng, et al.
Veröffentlicht: (2024)
von: Xia, Peng, et al.
Veröffentlicht: (2024)
Person-Centric Annotations of LAION-400M: Auditing Bias and Its Transfer to Models
von: Girrbach, Leander, et al.
Veröffentlicht: (2025)
von: Girrbach, Leander, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
ETHER: Efficient Finetuning of Large-Scale Models with Hyperplane Reflections
von: Bini, Massimo, et al.
Veröffentlicht: (2024) -
Vision-Language Models Create Cross-Modal Task Representations
von: Luo, Grace, et al.
Veröffentlicht: (2024) -
Orthogonal Finetuning Made Scalable
von: Qiu, Zeju, et al.
Veröffentlicht: (2025) -
Exposing and Addressing Cross-Task Inconsistency in Unified Vision-Language Models
von: Maharana, Adyasha, et al.
Veröffentlicht: (2023) -
Linear Alignment of Vision-language Models for Image Captioning
von: Paischer, Fabian, et al.
Veröffentlicht: (2023)