ETHER: Efficient Finetuning of Large-Scale Models with Hyperplane Reflections
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Bini, Massimo, Roth, Karsten, Akata, Zeynep, Khoreva, Anna |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
DeLoRA: Decoupling Angles and Strength in Low-rank Adaptation
von: Bini, Massimo, et al.
Veröffentlicht: (2025)
von: Bini, Massimo, et al.
Veröffentlicht: (2025)
MemLoRA: Distilling Expert Adapters for On-Device Memory Systems
von: Bini, Massimo, et al.
Veröffentlicht: (2025)
von: Bini, Massimo, et al.
Veröffentlicht: (2025)
Context-Aware Multimodal Pretraining
von: Roth, Karsten, et al.
Veröffentlicht: (2024)
von: Roth, Karsten, et al.
Veröffentlicht: (2024)
Reflecting on the State of Rehearsal-free Continual Learning with Pretrained Models
von: Thede, Lukas, et al.
Veröffentlicht: (2024)
von: Thede, Lukas, et al.
Veröffentlicht: (2024)
How to Merge Your Multimodal Models Over Time?
von: Dziadzio, Sebastian, et al.
Veröffentlicht: (2024)
von: Dziadzio, Sebastian, et al.
Veröffentlicht: (2024)
Improving Intervention Efficacy via Concept Realignment in Concept Bottleneck Models
von: Singhi, Nishad, et al.
Veröffentlicht: (2024)
von: Singhi, Nishad, et al.
Veröffentlicht: (2024)
A Practitioner's Guide to Continual Multimodal Pretraining
von: Roth, Karsten, et al.
Veröffentlicht: (2024)
von: Roth, Karsten, et al.
Veröffentlicht: (2024)
Fantastic Gains and Where to Find Them: On the Existence and Prospect of General Knowledge Transfer between Any Pretrained Model
von: Roth, Karsten, et al.
Veröffentlicht: (2023)
von: Roth, Karsten, et al.
Veröffentlicht: (2023)
Sparse Autoencoders are Topic Models
von: Girrbach, Leander, et al.
Veröffentlicht: (2025)
von: Girrbach, Leander, et al.
Veröffentlicht: (2025)
Disentangled Representation Learning with the Gromov-Monge Gap
von: Uscidda, Théo, et al.
Veröffentlicht: (2024)
von: Uscidda, Théo, et al.
Veröffentlicht: (2024)
Person-Centric Annotations of LAION-400M: Auditing Bias and Its Transfer to Models
von: Girrbach, Leander, et al.
Veröffentlicht: (2025)
von: Girrbach, Leander, et al.
Veröffentlicht: (2025)
Divide & Bind Your Attention for Improved Generative Semantic Nursing
von: Li, Yumeng, et al.
Veröffentlicht: (2023)
von: Li, Yumeng, et al.
Veröffentlicht: (2023)
COSMOS: Cross-Modality Self-Distillation for Vision Language Pre-training
von: Kim, Sanghwan, et al.
Veröffentlicht: (2024)
von: Kim, Sanghwan, et al.
Veröffentlicht: (2024)
Task Matrices: Linear Maps for Cross-Model Finetuning Transfer
von: Brien, Darrin O', et al.
Veröffentlicht: (2025)
von: Brien, Darrin O', et al.
Veröffentlicht: (2025)
Vision-by-Language for Training-Free Compositional Image Retrieval
von: Karthik, Shyamgopal, et al.
Veröffentlicht: (2023)
von: Karthik, Shyamgopal, et al.
Veröffentlicht: (2023)
Investigating Structural Pruning and Recovery Techniques for Compressing Multimodal Large Language Models: An Empirical Study
von: Huang, Yiran, et al.
Veröffentlicht: (2025)
von: Huang, Yiran, et al.
Veröffentlicht: (2025)
Noise Hypernetworks: Amortizing Test-Time Compute in Diffusion Models
von: Eyring, Luca, et al.
Veröffentlicht: (2025)
von: Eyring, Luca, et al.
Veröffentlicht: (2025)
Parameter-Efficient Orthogonal Finetuning via Butterfly Factorization
von: Liu, Weiyang, et al.
Veröffentlicht: (2023)
von: Liu, Weiyang, et al.
Veröffentlicht: (2023)
ReNO: Enhancing One-step Text-to-Image Models through Reward-based Noise Optimization
von: Eyring, Luca, et al.
Veröffentlicht: (2024)
von: Eyring, Luca, et al.
Veröffentlicht: (2024)
Shotluck Holmes: A Family of Efficient Small-Scale Large Language Vision Models For Video Captioning and Summarization
von: Luo, Richard, et al.
Veröffentlicht: (2024)
von: Luo, Richard, et al.
Veröffentlicht: (2024)
From Drop-off to Recovery: A Mechanistic Analysis of Segmentation in MLLMs
von: Wu, Boyong, et al.
Veröffentlicht: (2026)
von: Wu, Boyong, et al.
Veröffentlicht: (2026)
The Manifold Hypothesis for Gradient-Based Explanations
von: Bordt, Sebastian, et al.
Veröffentlicht: (2022)
von: Bordt, Sebastian, et al.
Veröffentlicht: (2022)
VLSM-Adapter: Finetuning Vision-Language Segmentation Efficiently with Lightweight Blocks
von: Dhakal, Manish, et al.
Veröffentlicht: (2024)
von: Dhakal, Manish, et al.
Veröffentlicht: (2024)
Adversarial Supervision Makes Layout-to-Image Diffusion Models Thrive
von: Li, Yumeng, et al.
Veröffentlicht: (2024)
von: Li, Yumeng, et al.
Veröffentlicht: (2024)
Orthogonal Finetuning Made Scalable
von: Qiu, Zeju, et al.
Veröffentlicht: (2025)
von: Qiu, Zeju, et al.
Veröffentlicht: (2025)
DataDream: Few-shot Guided Dataset Generation
von: Kim, Jae Myung, et al.
Veröffentlicht: (2024)
von: Kim, Jae Myung, et al.
Veröffentlicht: (2024)
Efficient Scaling of Diffusion Transformers for Text-to-Image Generation
von: Li, Hao, et al.
Veröffentlicht: (2024)
von: Li, Hao, et al.
Veröffentlicht: (2024)
Towards Efficient Large Vision-Language Models: A Comprehensive Survey on Inference Strategies
von: Pathak, Surendra, et al.
Veröffentlicht: (2026)
von: Pathak, Surendra, et al.
Veröffentlicht: (2026)
SUB: Benchmarking CBM Generalization via Synthetic Attribute Substitutions
von: Bader, Jessica, et al.
Veröffentlicht: (2025)
von: Bader, Jessica, et al.
Veröffentlicht: (2025)
Sparse Autoencoders Learn Monosemantic Features in Vision-Language Models
von: Pach, Mateusz, et al.
Veröffentlicht: (2025)
von: Pach, Mateusz, et al.
Veröffentlicht: (2025)
Post-hoc Probabilistic Vision-Language Models
von: Baumann, Anton, et al.
Veröffentlicht: (2024)
von: Baumann, Anton, et al.
Veröffentlicht: (2024)
A Large Scale Analysis of Gender Biases in Text-to-Image Generative Models
von: Girrbach, Leander, et al.
Veröffentlicht: (2025)
von: Girrbach, Leander, et al.
Veröffentlicht: (2025)
Domain-Aware Fine-Tuning of Foundation Models
von: Kaplan, Ugur Ali, et al.
Veröffentlicht: (2024)
von: Kaplan, Ugur Ali, et al.
Veröffentlicht: (2024)
MoLEx: Mixture of Layer Experts for Finetuning with Sparse Upcycling
von: Teo, Rachel S. Y., et al.
Veröffentlicht: (2025)
von: Teo, Rachel S. Y., et al.
Veröffentlicht: (2025)
WikiBigEdit: Understanding the Limits of Lifelong Knowledge Editing in LLMs
von: Thede, Lukas, et al.
Veröffentlicht: (2025)
von: Thede, Lukas, et al.
Veröffentlicht: (2025)
Dissecting Multimodal In-Context Learning: Modality Asymmetries and Circuit Dynamics in modern Transformers
von: Huang, Yiran, et al.
Veröffentlicht: (2026)
von: Huang, Yiran, et al.
Veröffentlicht: (2026)
LLM-CXR: Instruction-Finetuned LLM for CXR Image Understanding and Generation
von: Lee, Suhyeon, et al.
Veröffentlicht: (2023)
von: Lee, Suhyeon, et al.
Veröffentlicht: (2023)
Pico-Banana-400K: A Large-Scale Dataset for Text-Guided Image Editing
von: Qian, Yusu, et al.
Veröffentlicht: (2025)
von: Qian, Yusu, et al.
Veröffentlicht: (2025)
How LoRA Remembers? A Parametric Memory Law for LLM Finetuning
von: Xu, Ziwen, et al.
Veröffentlicht: (2026)
von: Xu, Ziwen, et al.
Veröffentlicht: (2026)
BanglishRev: A Large-Scale Bangla-English and Code-mixed Dataset of Product Reviews in E-Commerce
von: Shamael, Mohammad Nazmush, et al.
Veröffentlicht: (2024)
von: Shamael, Mohammad Nazmush, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
DeLoRA: Decoupling Angles and Strength in Low-rank Adaptation
von: Bini, Massimo, et al.
Veröffentlicht: (2025) -
MemLoRA: Distilling Expert Adapters for On-Device Memory Systems
von: Bini, Massimo, et al.
Veröffentlicht: (2025) -
Context-Aware Multimodal Pretraining
von: Roth, Karsten, et al.
Veröffentlicht: (2024) -
Reflecting on the State of Rehearsal-free Continual Learning with Pretrained Models
von: Thede, Lukas, et al.
Veröffentlicht: (2024) -
How to Merge Your Multimodal Models Over Time?
von: Dziadzio, Sebastian, et al.
Veröffentlicht: (2024)