StrLoRA: Towards Streaming Continual Visual Instruction Tuning for MLLMs
Fuente:
arXiv
Saved in:
| Main Authors: | Che, Chang, Wang, Ziqi, Ma, Hui, Wang, Cheems, Shi, Zenglin |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
LoRA in LoRA: Towards Parameter-Efficient Architecture Expansion for Continual Visual Instruction Tuning
by: Che, Chang, et al.
Published: (2025)
by: Che, Chang, et al.
Published: (2025)
SMoLoRA: Exploring and Defying Dual Catastrophic Forgetting in Continual Visual Instruction Tuning
by: Wang, Ziqi, et al.
Published: (2024)
by: Wang, Ziqi, et al.
Published: (2024)
Harmonious Parameter Adaptation in Continual Visual Instruction Tuning for Safety-Aligned MLLMs
by: Wang, Ziqi, et al.
Published: (2025)
by: Wang, Ziqi, et al.
Published: (2025)
Selective LoRA for Visual Tokens and Attention Heads
by: Luo, Tiange, et al.
Published: (2025)
by: Luo, Tiange, et al.
Published: (2025)
MCITlib: Multimodal Continual Instruction Tuning Library and Benchmark
by: Guo, Haiyang, et al.
Published: (2025)
by: Guo, Haiyang, et al.
Published: (2025)
EMLoC: Emulator-based Memory-efficient Fine-tuning with LoRA Correction
by: Lin, Hsi-Che, et al.
Published: (2025)
by: Lin, Hsi-Che, et al.
Published: (2025)
LoRA Subtraction for Drift-Resistant Space in Exemplar-Free Continual Learning
by: Liu, Xuan, et al.
Published: (2025)
by: Liu, Xuan, et al.
Published: (2025)
Continual Diffusion: Continual Customization of Text-to-Image Diffusion with C-LoRA
by: Smith, James Seale, et al.
Published: (2023)
by: Smith, James Seale, et al.
Published: (2023)
LoRA-Based Continual Learning with Constraints on Critical Parameter Changes
by: Ling, Shimou, et al.
Published: (2025)
by: Ling, Shimou, et al.
Published: (2025)
StreamingBench: Assessing the Gap for MLLMs to Achieve Streaming Video Understanding
by: Lin, Junming, et al.
Published: (2024)
by: Lin, Junming, et al.
Published: (2024)
Dynamic Mixture of Curriculum LoRA Experts for Continual Multimodal Instruction Tuning
by: Ge, Chendi, et al.
Published: (2025)
by: Ge, Chendi, et al.
Published: (2025)
Visual-Oriented Fine-Grained Knowledge Editing for MultiModal Large Language Models
by: Zeng, Zhen, et al.
Published: (2024)
by: Zeng, Zhen, et al.
Published: (2024)
ElaLoRA: Elastic & Learnable Low-Rank Adaptation for Efficient Model Fine-Tuning
by: Chang, Huandong, et al.
Published: (2025)
by: Chang, Huandong, et al.
Published: (2025)
NAS-LoRA: Empowering Parameter-Efficient Fine-Tuning for Visual Foundation Models with Searchable Adaptation
by: Chen, Renqi, et al.
Published: (2025)
by: Chen, Renqi, et al.
Published: (2025)
Visual Prompt Tuning in Null Space for Continual Learning
by: Lu, Yue, et al.
Published: (2024)
by: Lu, Yue, et al.
Published: (2024)
Empower Vision Applications with LoRA LMM
by: Mi, Liang, et al.
Published: (2024)
by: Mi, Liang, et al.
Published: (2024)
TimeSuite: Improving MLLMs for Long Video Understanding via Grounded Tuning
by: Zeng, Xiangyu, et al.
Published: (2024)
by: Zeng, Xiangyu, et al.
Published: (2024)
When Looking Is Not Enough: Visual Attention Structure Reveals Hallucination in MLLMs
by: Cao, Fanpu, et al.
Published: (2026)
by: Cao, Fanpu, et al.
Published: (2026)
In-Video Instructions: Visual Signals as Generative Control
by: Fang, Gongfan, et al.
Published: (2025)
by: Fang, Gongfan, et al.
Published: (2025)
Block-wise LoRA: Revisiting Fine-grained LoRA for Effective Personalization and Stylization in Text-to-Image Generation
by: Li, Likun, et al.
Published: (2024)
by: Li, Likun, et al.
Published: (2024)
Shared LoRA Subspaces for almost Strict Continual Learning
by: Kaushik, Prakhar, et al.
Published: (2026)
by: Kaushik, Prakhar, et al.
Published: (2026)
Reconstructive Visual Instruction Tuning
by: Wang, Haochen, et al.
Published: (2024)
by: Wang, Haochen, et al.
Published: (2024)
Stabilizing Unsupervised Self-Evolution of MLLMs via Continuous Softened Retracing reSampling
by: Yu, Yunyao, et al.
Published: (2026)
by: Yu, Yunyao, et al.
Published: (2026)
CLIPDrag: Combining Text-based and Drag-based Instructions for Image Editing
by: Jiang, Ziqi, et al.
Published: (2024)
by: Jiang, Ziqi, et al.
Published: (2024)
FREE-Switch: Frequency-based Dynamic LoRA Switch for Style Transfer
by: Zheng, Shenghe, et al.
Published: (2026)
by: Zheng, Shenghe, et al.
Published: (2026)
VideoMind: A Chain-of-LoRA Agent for Temporal-Grounded Video Reasoning
by: Liu, Ye, et al.
Published: (2025)
by: Liu, Ye, et al.
Published: (2025)
RecycleLoRA: Rank-Revealing QR-Based Dual-LoRA Subspace Adaptation for Domain Generalized Semantic Segmentation
by: Cho, Chanseul, et al.
Published: (2026)
by: Cho, Chanseul, et al.
Published: (2026)
In-Context Meta LoRA Generation
by: Shao, Yihua, et al.
Published: (2025)
by: Shao, Yihua, et al.
Published: (2025)
DiffLoRA: Generating Personalized Low-Rank Adaptation Weights with Diffusion
by: Wu, Yujia, et al.
Published: (2024)
by: Wu, Yujia, et al.
Published: (2024)
Lifting the Veil on Visual Information Flow in MLLMs: Unlocking Pathways to Faster Inference
by: Yin, Hao, et al.
Published: (2025)
by: Yin, Hao, et al.
Published: (2025)
CDM-QTA: Quantized Training Acceleration for Efficient LoRA Fine-Tuning of Diffusion Model
by: Lu, Jinming, et al.
Published: (2025)
by: Lu, Jinming, et al.
Published: (2025)
PanoLora: Bridging Perspective and Panoramic Video Generation with LoRA Adaptation
by: Dong, Zeyu, et al.
Published: (2025)
by: Dong, Zeyu, et al.
Published: (2025)
EmoVIT: Revolutionizing Emotion Insights with Visual Instruction Tuning
by: Xie, Hongxia, et al.
Published: (2024)
by: Xie, Hongxia, et al.
Published: (2024)
Towards Unified Surgical Scene Understanding:Bridging Reasoning and Grounding via MLLMs
by: Huang, Jincai, et al.
Published: (2026)
by: Huang, Jincai, et al.
Published: (2026)
Customize Segment Anything Model for Multi-Modal Semantic Segmentation with Mixture of LoRA Experts
by: Zhu, Chenyang, et al.
Published: (2024)
by: Zhu, Chenyang, et al.
Published: (2024)
Evolutionary Negative Module Pruning for Better LoRA Merging
by: Cao, Anda, et al.
Published: (2026)
by: Cao, Anda, et al.
Published: (2026)
MapReduce LoRA: Advancing the Pareto Front in Multi-Preference Optimization for Generative Models
by: Chen, Chieh-Yun, et al.
Published: (2025)
by: Chen, Chieh-Yun, et al.
Published: (2025)
Grounding DINO-US-SAM: Text-Prompted Multi-Organ Segmentation in Ultrasound with LoRA-Tuned Vision-Language Models
by: Rasaee, Hamza, et al.
Published: (2025)
by: Rasaee, Hamza, et al.
Published: (2025)
Human Motion Instruction Tuning
by: Li, Lei, et al.
Published: (2024)
by: Li, Lei, et al.
Published: (2024)
VisNec: Measuring and Leveraging Visual Necessity for Multimodal Instruction Tuning
by: Dong, Mingkang, et al.
Published: (2026)
by: Dong, Mingkang, et al.
Published: (2026)
Similar Items
-
LoRA in LoRA: Towards Parameter-Efficient Architecture Expansion for Continual Visual Instruction Tuning
by: Che, Chang, et al.
Published: (2025) -
SMoLoRA: Exploring and Defying Dual Catastrophic Forgetting in Continual Visual Instruction Tuning
by: Wang, Ziqi, et al.
Published: (2024) -
Harmonious Parameter Adaptation in Continual Visual Instruction Tuning for Safety-Aligned MLLMs
by: Wang, Ziqi, et al.
Published: (2025) -
Selective LoRA for Visual Tokens and Attention Heads
by: Luo, Tiange, et al.
Published: (2025) -
MCITlib: Multimodal Continual Instruction Tuning Library and Benchmark
by: Guo, Haiyang, et al.
Published: (2025)