Introducing Routing Functions to Vision-Language Parameter-Efficient Fine-Tuning with Low-Rank Bottlenecks
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Qu, Tingyu, Tuytelaars, Tinne, Moens, Marie-Francine |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Visually-Aware Context Modeling for News Image Captioning
von: Qu, Tingyu, et al.
Veröffentlicht: (2023)
von: Qu, Tingyu, et al.
Veröffentlicht: (2023)
TS-LLaVA: Constructing Visual Tokens through Thumbnail-and-Sampling for Training-Free Video Large Language Models
von: Qu, Tingyu, et al.
Veröffentlicht: (2024)
von: Qu, Tingyu, et al.
Veröffentlicht: (2024)
Towards More Accurate Personalized Image Generation: Addressing Overfitting and Evaluation Bias
von: Li, Mingxiao, et al.
Veröffentlicht: (2025)
von: Li, Mingxiao, et al.
Veröffentlicht: (2025)
DM-Align: Leveraging the Power of Natural Language Instructions to Make Changes to Images
von: Trusca, Maria Mihaela, et al.
Veröffentlicht: (2024)
von: Trusca, Maria Mihaela, et al.
Veröffentlicht: (2024)
Animate Your Motion: Turning Still Images into Dynamic Videos
von: Li, Mingxiao, et al.
Veröffentlicht: (2024)
von: Li, Mingxiao, et al.
Veröffentlicht: (2024)
OASIS: Online Sample Selection for Continual Visual Instruction Tuning
von: Lee, Minjae, et al.
Veröffentlicht: (2025)
von: Lee, Minjae, et al.
Veröffentlicht: (2025)
Implicit Gaussian Splatting with Efficient Multi-Level Tri-Plane Representation
von: Wu, Minye, et al.
Veröffentlicht: (2024)
von: Wu, Minye, et al.
Veröffentlicht: (2024)
Object-Attribute Binding in Text-to-Image Generation: Evaluation and Control
von: Trusca, Maria Mihaela, et al.
Veröffentlicht: (2024)
von: Trusca, Maria Mihaela, et al.
Veröffentlicht: (2024)
Unsupervised Parameter Efficient Source-free Post-pretraining
von: Jha, Abhishek, et al.
Veröffentlicht: (2025)
von: Jha, Abhishek, et al.
Veröffentlicht: (2025)
Alleviating Exposure Bias in Diffusion Models through Sampling with Shifted Time Steps
von: Li, Mingxiao, et al.
Veröffentlicht: (2023)
von: Li, Mingxiao, et al.
Veröffentlicht: (2023)
Analysis of Spatial augmentation in Self-supervised models in the purview of training and test distributions
von: Jha, Abhishek, et al.
Veröffentlicht: (2024)
von: Jha, Abhishek, et al.
Veröffentlicht: (2024)
Learning to Route for Dynamic Adapter Composition in Continual Learning with Language Models
von: Araujo, Vladimir, et al.
Veröffentlicht: (2024)
von: Araujo, Vladimir, et al.
Veröffentlicht: (2024)
Navigating the Nuances: A Fine-grained Evaluation of Vision-Language Navigation
von: Wang, Zehao, et al.
Veröffentlicht: (2024)
von: Wang, Zehao, et al.
Veröffentlicht: (2024)
Remembering by Reconstructing: Domain Incremental Learning With Test-Time Training on Video Streams
von: Swinnen, Jonathan, et al.
Veröffentlicht: (2026)
von: Swinnen, Jonathan, et al.
Veröffentlicht: (2026)
Eff-GRot: Efficient and Generalizable Rotation Estimation with Transformers
von: Mathioulakis, Fanis, et al.
Veröffentlicht: (2025)
von: Mathioulakis, Fanis, et al.
Veröffentlicht: (2025)
Charm: The Missing Piece in ViT fine-tuning for Image Aesthetic Assessment
von: Behrad, Fatemeh, et al.
Veröffentlicht: (2025)
von: Behrad, Fatemeh, et al.
Veröffentlicht: (2025)
PEO: Training-Free Aesthetic Quality Enhancement in Pre-Trained Text-to-Image Diffusion Models with Prompt Embedding Optimization
von: Margaryan, Hovhannes, et al.
Veröffentlicht: (2025)
von: Margaryan, Hovhannes, et al.
Veröffentlicht: (2025)
RGS-DR: Deferred Reflections and Residual Shading in 2D Gaussian Splatting
von: Kouros, Georgios, et al.
Veröffentlicht: (2025)
von: Kouros, Georgios, et al.
Veröffentlicht: (2025)
Object-Centric Pretraining via Target Encoder Bootstrapping
von: Đukić, Nikola, et al.
Veröffentlicht: (2025)
von: Đukić, Nikola, et al.
Veröffentlicht: (2025)
Unveiling the Ambiguity in Neural Inverse Rendering: A Parameter Compensation Analysis
von: Kouros, Georgios, et al.
Veröffentlicht: (2024)
von: Kouros, Georgios, et al.
Veröffentlicht: (2024)
Spec-Gloss Surfels and Normal-Diffuse Priors for Relightable Glossy Objects
von: Kouros, Georgios, et al.
Veröffentlicht: (2025)
von: Kouros, Georgios, et al.
Veröffentlicht: (2025)
Consistent Story Generation: Unlocking the Potential of Zigzag Sampling
von: Li, Mingxiao, et al.
Veröffentlicht: (2025)
von: Li, Mingxiao, et al.
Veröffentlicht: (2025)
Same accuracy, twice as fast: continuous training surpasses retraining from scratch
von: Verwimp, Eli, et al.
Veröffentlicht: (2025)
von: Verwimp, Eli, et al.
Veröffentlicht: (2025)
When normalization hallucinates: unseen risks in AI-powered whole slide image processing
von: Moens, Karel, et al.
Veröffentlicht: (2025)
von: Moens, Karel, et al.
Veröffentlicht: (2025)
LARGO: Low-Rank Regulated Gradient Projection for Robust Parameter Efficient Fine-Tuning
von: Zhang, Haotian, et al.
Veröffentlicht: (2025)
von: Zhang, Haotian, et al.
Veröffentlicht: (2025)
LoRAPrune: Structured Pruning Meets Low-Rank Parameter-Efficient Fine-Tuning
von: Zhang, Mingyang, et al.
Veröffentlicht: (2023)
von: Zhang, Mingyang, et al.
Veröffentlicht: (2023)
Rethinking Bottlenecks in Safety Fine-Tuning of Vision Language Models
von: Ding, Yi, et al.
Veröffentlicht: (2025)
von: Ding, Yi, et al.
Veröffentlicht: (2025)
Action-based image editing guided by human instructions
von: Trusca, Maria Mihaela, et al.
Veröffentlicht: (2024)
von: Trusca, Maria Mihaela, et al.
Veröffentlicht: (2024)
LoLDU: Low-Rank Adaptation via Lower-Diag-Upper Decomposition for Parameter-Efficient Fine-Tuning
von: Shi, Yiming, et al.
Veröffentlicht: (2024)
von: Shi, Yiming, et al.
Veröffentlicht: (2024)
Forgetting of task-specific knowledge in model merging-based continual learning
von: Hess, Timm, et al.
Veröffentlicht: (2025)
von: Hess, Timm, et al.
Veröffentlicht: (2025)
NeuroCine: Decoding Vivid Video Sequences from Human Brain Activties
von: Sun, Jingyuan, et al.
Veröffentlicht: (2024)
von: Sun, Jingyuan, et al.
Veröffentlicht: (2024)
Low-Rank Rescaled Vision Transformer Fine-Tuning: A Residual Design Approach
von: Dong, Wei, et al.
Veröffentlicht: (2024)
von: Dong, Wei, et al.
Veröffentlicht: (2024)
Is this chart lying to me? Automating the detection of misleading visualizations
von: Tonglet, Jonathan, et al.
Veröffentlicht: (2025)
von: Tonglet, Jonathan, et al.
Veröffentlicht: (2025)
SALT: Parameter-Efficient Fine-Tuning via Singular Value Adaptation with Low-Rank Transformation
von: Elsayed, Abdelrahman, et al.
Veröffentlicht: (2025)
von: Elsayed, Abdelrahman, et al.
Veröffentlicht: (2025)
Low-rank Attention Side-Tuning for Parameter-Efficient Fine-Tuning
von: Tang, Ningyuan, et al.
Veröffentlicht: (2024)
von: Tang, Ningyuan, et al.
Veröffentlicht: (2024)
PE-CLIP: A Parameter-Efficient Fine-Tuning of Vision Language Models for Dynamic Facial Expression Recognition
von: Saadi, Ibtissam, et al.
Veröffentlicht: (2025)
von: Saadi, Ibtissam, et al.
Veröffentlicht: (2025)
Adapter-X: A Novel General Parameter-Efficient Fine-Tuning Framework for Vision
von: Li, Minglei, et al.
Veröffentlicht: (2024)
von: Li, Minglei, et al.
Veröffentlicht: (2024)
Semantic Hierarchical Prompt Tuning for Parameter-Efficient Fine-Tuning
von: Zhu, Haowei, et al.
Veröffentlicht: (2024)
von: Zhu, Haowei, et al.
Veröffentlicht: (2024)
PEFT A2Z: Parameter-Efficient Fine-Tuning Survey for Large Language and Vision Models
von: Prottasha, Nusrat Jahan, et al.
Veröffentlicht: (2025)
von: Prottasha, Nusrat Jahan, et al.
Veröffentlicht: (2025)
BG-Triangle: Bézier Gaussian Triangle for 3D Vectorization and Rendering
von: Wu, Minye, et al.
Veröffentlicht: (2025)
von: Wu, Minye, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Visually-Aware Context Modeling for News Image Captioning
von: Qu, Tingyu, et al.
Veröffentlicht: (2023) -
TS-LLaVA: Constructing Visual Tokens through Thumbnail-and-Sampling for Training-Free Video Large Language Models
von: Qu, Tingyu, et al.
Veröffentlicht: (2024) -
Towards More Accurate Personalized Image Generation: Addressing Overfitting and Evaluation Bias
von: Li, Mingxiao, et al.
Veröffentlicht: (2025) -
DM-Align: Leveraging the Power of Natural Language Instructions to Make Changes to Images
von: Trusca, Maria Mihaela, et al.
Veröffentlicht: (2024) -
Animate Your Motion: Turning Still Images into Dynamic Videos
von: Li, Mingxiao, et al.
Veröffentlicht: (2024)