CL-VISTA: Benchmarking Continual Learning in Video Large Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Guo, Haiyang, Shi, Yichen, Zhu, Fei, Liu, Wenzhuo, Zhao, Hongbo, Zeng, Fanhu, Ma, Shijie, Wang, Da-Han, Zhang, Xu-Yao |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
MCITlib: Multimodal Continual Instruction Tuning Library and Benchmark
von: Guo, Haiyang, et al.
Veröffentlicht: (2025)
von: Guo, Haiyang, et al.
Veröffentlicht: (2025)
Continual Learning for Generative AI: From LLMs to MLLMs and Beyond
von: Guo, Haiyang, et al.
Veröffentlicht: (2025)
von: Guo, Haiyang, et al.
Veröffentlicht: (2025)
Federated Continual Instruction Tuning
von: Guo, Haiyang, et al.
Veröffentlicht: (2025)
von: Guo, Haiyang, et al.
Veröffentlicht: (2025)
MLLM-CL: Continual Learning for Multimodal Large Language Models
von: Zhao, Hongbo, et al.
Veröffentlicht: (2025)
von: Zhao, Hongbo, et al.
Veröffentlicht: (2025)
DESIRE: Dynamic Knowledge Consolidation for Rehearsal-Free Continual Learning
von: Guo, Haiyang, et al.
Veröffentlicht: (2024)
von: Guo, Haiyang, et al.
Veröffentlicht: (2024)
HiDe-LLaVA: Hierarchical Decoupling for Continual Instruction Tuning of Multimodal Large Language Model
von: Guo, Haiyang, et al.
Veröffentlicht: (2025)
von: Guo, Haiyang, et al.
Veröffentlicht: (2025)
ModalPrompt: Towards Efficient Multimodal Continual Instruction Tuning with Dual-Modality Guided Prompt
von: Zeng, Fanhu, et al.
Veröffentlicht: (2024)
von: Zeng, Fanhu, et al.
Veröffentlicht: (2024)
VTCBench: Can Vision-Language Models Understand Long Context with Vision-Text Compression?
von: Zhao, Hongbo, et al.
Veröffentlicht: (2025)
von: Zhao, Hongbo, et al.
Veröffentlicht: (2025)
RobustMerge: Parameter-Efficient Model Merging for MLLMs with Direction Robustness
von: Zeng, Fanhu, et al.
Veröffentlicht: (2025)
von: Zeng, Fanhu, et al.
Veröffentlicht: (2025)
Towards Efficient and General-Purpose Few-Shot Misclassification Detection for Vision-Language Models
von: Zeng, Fanhu, et al.
Veröffentlicht: (2025)
von: Zeng, Fanhu, et al.
Veröffentlicht: (2025)
Happy: A Debiased Learning Framework for Continual Generalized Category Discovery
von: Ma, Shijie, et al.
Veröffentlicht: (2024)
von: Ma, Shijie, et al.
Veröffentlicht: (2024)
PILoRA: Prototype Guided Incremental LoRA for Federated Class-Incremental Learning
von: Guo, Haiyang, et al.
Veröffentlicht: (2024)
von: Guo, Haiyang, et al.
Veröffentlicht: (2024)
LLaVA-c: Continual Improved Visual Instruction Tuning
von: Liu, Wenzhuo, et al.
Veröffentlicht: (2025)
von: Liu, Wenzhuo, et al.
Veröffentlicht: (2025)
Local-Prompt: Extensible Local Prompts for Few-Shot Out-of-Distribution Detection
von: Zeng, Fanhu, et al.
Veröffentlicht: (2024)
von: Zeng, Fanhu, et al.
Veröffentlicht: (2024)
MSPE: Multi-Scale Patch Embedding Prompts Vision Transformers to Any Resolution
von: Liu, Wenzhuo, et al.
Veröffentlicht: (2024)
von: Liu, Wenzhuo, et al.
Veröffentlicht: (2024)
Branch-Tuning: Balancing Stability and Plasticity for Continual Self-Supervised Learning
von: Liu, Wenzhuo, et al.
Veröffentlicht: (2024)
von: Liu, Wenzhuo, et al.
Veröffentlicht: (2024)
Global Convergence of Continual Learning on Non-IID Data
von: Zhu, Fei, et al.
Veröffentlicht: (2025)
von: Zhu, Fei, et al.
Veröffentlicht: (2025)
PrimeKG-CL: A Continual Graph Learning Benchmark on Evolving Biomedical Knowledge Graphs
von: Radwan, Yousef A., et al.
Veröffentlicht: (2026)
von: Radwan, Yousef A., et al.
Veröffentlicht: (2026)
ProtoGCD: Unified and Unbiased Prototype Learning for Generalized Category Discovery
von: Ma, Shijie, et al.
Veröffentlicht: (2025)
von: Ma, Shijie, et al.
Veröffentlicht: (2025)
Diffusion-based Perceptual Neural Video Compression with Temporal Diffusion Information Reuse
von: Ma, Wenzhuo, et al.
Veröffentlicht: (2025)
von: Ma, Wenzhuo, et al.
Veröffentlicht: (2025)
Fine-Grained Post-Training Quantization for Large Vision Language Models with Quantization-Aware Integrated Gradients
von: Xiang, Ziwei, et al.
Veröffentlicht: (2026)
von: Xiang, Ziwei, et al.
Veröffentlicht: (2026)
CL-bench: A Benchmark for Context Learning
von: Dou, Shihan, et al.
Veröffentlicht: (2026)
von: Dou, Shihan, et al.
Veröffentlicht: (2026)
EmoBench-M: Benchmarking Emotional Intelligence for Multimodal Large Language Models
von: Hu, He, et al.
Veröffentlicht: (2025)
von: Hu, He, et al.
Veröffentlicht: (2025)
Reinforcement Fine-Tuning Naturally Mitigates Forgetting in Continual Post-Training
von: Lai, Song, et al.
Veröffentlicht: (2025)
von: Lai, Song, et al.
Veröffentlicht: (2025)
DiffVC-OSD: One-Step Diffusion-based Perceptual Neural Video Compression Framework
von: Ma, Wenzhuo, et al.
Veröffentlicht: (2025)
von: Ma, Wenzhuo, et al.
Veröffentlicht: (2025)
InsCL: A Data-efficient Continual Learning Paradigm for Fine-tuning Large Language Models with Instructions
von: Wang, Yifan, et al.
Veröffentlicht: (2024)
von: Wang, Yifan, et al.
Veröffentlicht: (2024)
VISTA: Video Interaction Spatio-Temporal Analysis Benchmark
von: Aparcedo, Alejandro, et al.
Veröffentlicht: (2026)
von: Aparcedo, Alejandro, et al.
Veröffentlicht: (2026)
DiffVC-RT: Towards Practical Real-Time Diffusion-based Perceptual Neural Video Compression
von: Ma, Wenzhuo, et al.
Veröffentlicht: (2026)
von: Ma, Wenzhuo, et al.
Veröffentlicht: (2026)
Towards Non-Exemplar Semi-Supervised Class-Incremental Learning
von: Liu, Wenzhuo, et al.
Veröffentlicht: (2024)
von: Liu, Wenzhuo, et al.
Veröffentlicht: (2024)
VideoCuRL: Video Curriculum Reinforcement Learning with Orthogonal Difficulty Decomposition
von: Jin, Hongbo, et al.
Veröffentlicht: (2025)
von: Jin, Hongbo, et al.
Veröffentlicht: (2025)
Towards Trustworthy Dataset Distillation
von: Ma, Shijie, et al.
Veröffentlicht: (2023)
von: Ma, Shijie, et al.
Veröffentlicht: (2023)
VISTA: An End-to-End Benchmark for Visual Spec-to-Web-App Coding Agents
von: Guo, JunJia, et al.
Veröffentlicht: (2026)
von: Guo, JunJia, et al.
Veröffentlicht: (2026)
EIFBENCH: Extremely Complex Instruction Following Benchmark for Large Language Models
von: Zou, Tao, et al.
Veröffentlicht: (2025)
von: Zou, Tao, et al.
Veröffentlicht: (2025)
POLYCHARTQA: Benchmarking Large Vision-Language Models with Multilingual Chart Question Answering
von: Xu, Yichen, et al.
Veröffentlicht: (2025)
von: Xu, Yichen, et al.
Veröffentlicht: (2025)
Enhancing Long Video Question Answering with Scene-Localized Frame Grouping
von: Yang, Xuyi, et al.
Veröffentlicht: (2025)
von: Yang, Xuyi, et al.
Veröffentlicht: (2025)
LightCL: Compact Continual Learning with Low Memory Footprint For Edge Device
von: Wang, Zeqing, et al.
Veröffentlicht: (2024)
von: Wang, Zeqing, et al.
Veröffentlicht: (2024)
VISTA: Mitigating Semantic Inertia in Video-LLMs via Training-Free Dynamic Chain-of-Thought Routing
von: Jin, Hongbo, et al.
Veröffentlicht: (2025)
von: Jin, Hongbo, et al.
Veröffentlicht: (2025)
AdaCL:Adaptive Continual Learning
von: Yildirim, Elif Ceren Gok, et al.
Veröffentlicht: (2023)
von: Yildirim, Elif Ceren Gok, et al.
Veröffentlicht: (2023)
Reinforcement Learning with Continuous Actions Under Unmeasured Confounding
von: Li, Yuhan, et al.
Veröffentlicht: (2025)
von: Li, Yuhan, et al.
Veröffentlicht: (2025)
ChineseVideoBench: Benchmarking Multi-modal Large Models for Chinese Video Question Answering
von: Nie, Yuxiang, et al.
Veröffentlicht: (2025)
von: Nie, Yuxiang, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
MCITlib: Multimodal Continual Instruction Tuning Library and Benchmark
von: Guo, Haiyang, et al.
Veröffentlicht: (2025) -
Continual Learning for Generative AI: From LLMs to MLLMs and Beyond
von: Guo, Haiyang, et al.
Veröffentlicht: (2025) -
Federated Continual Instruction Tuning
von: Guo, Haiyang, et al.
Veröffentlicht: (2025) -
MLLM-CL: Continual Learning for Multimodal Large Language Models
von: Zhao, Hongbo, et al.
Veröffentlicht: (2025) -
DESIRE: Dynamic Knowledge Consolidation for Rehearsal-Free Continual Learning
von: Guo, Haiyang, et al.
Veröffentlicht: (2024)