OASIS: Online Sample Selection for Continual Visual Instruction Tuning
Fuente:
arXiv
Salvato in:
| Autori principali: | Lee, Minjae, Seo, Minhyuk, Qu, Tingyu, Tuytelaars, Tinne, Choi, Jonghyun |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Introducing Routing Functions to Vision-Language Parameter-Efficient Fine-Tuning with Low-Rank Bottlenecks
di: Qu, Tingyu, et al.
Pubblicazione: (2024)
di: Qu, Tingyu, et al.
Pubblicazione: (2024)
Visually-Aware Context Modeling for News Image Captioning
di: Qu, Tingyu, et al.
Pubblicazione: (2023)
di: Qu, Tingyu, et al.
Pubblicazione: (2023)
Budgeted Online Continual Learning by Adaptive Layer Freezing and Frequency-based Sampling
di: Seo, Minhyuk, et al.
Pubblicazione: (2024)
di: Seo, Minhyuk, et al.
Pubblicazione: (2024)
TS-LLaVA: Constructing Visual Tokens through Thumbnail-and-Sampling for Training-Free Video Large Language Models
di: Qu, Tingyu, et al.
Pubblicazione: (2024)
di: Qu, Tingyu, et al.
Pubblicazione: (2024)
Attention Misses Visual Risk: Risk-Adaptive Steering for Multimodal Safety Alignment
di: Park, Jonghyun, et al.
Pubblicazione: (2025)
di: Park, Jonghyun, et al.
Pubblicazione: (2025)
Towards More Accurate Personalized Image Generation: Addressing Overfitting and Evaluation Bias
di: Li, Mingxiao, et al.
Pubblicazione: (2025)
di: Li, Mingxiao, et al.
Pubblicazione: (2025)
Learning Equi-angular Representations for Online Continual Learning
di: Seo, Minhyuk, et al.
Pubblicazione: (2024)
di: Seo, Minhyuk, et al.
Pubblicazione: (2024)
GenOL: Generating Diverse Examples for Name-only Online Learning
di: Seo, Minhyuk, et al.
Pubblicazione: (2024)
di: Seo, Minhyuk, et al.
Pubblicazione: (2024)
Implicit Gaussian Splatting with Efficient Multi-Level Tri-Plane Representation
di: Wu, Minye, et al.
Pubblicazione: (2024)
di: Wu, Minye, et al.
Pubblicazione: (2024)
Analysis of Spatial augmentation in Self-supervised models in the purview of training and test distributions
di: Jha, Abhishek, et al.
Pubblicazione: (2024)
di: Jha, Abhishek, et al.
Pubblicazione: (2024)
DM-Align: Leveraging the Power of Natural Language Instructions to Make Changes to Images
di: Trusca, Maria Mihaela, et al.
Pubblicazione: (2024)
di: Trusca, Maria Mihaela, et al.
Pubblicazione: (2024)
Remembering by Reconstructing: Domain Incremental Learning With Test-Time Training on Video Streams
di: Swinnen, Jonathan, et al.
Pubblicazione: (2026)
di: Swinnen, Jonathan, et al.
Pubblicazione: (2026)
Charm: The Missing Piece in ViT fine-tuning for Image Aesthetic Assessment
di: Behrad, Fatemeh, et al.
Pubblicazione: (2025)
di: Behrad, Fatemeh, et al.
Pubblicazione: (2025)
PEO: Training-Free Aesthetic Quality Enhancement in Pre-Trained Text-to-Image Diffusion Models with Prompt Embedding Optimization
di: Margaryan, Hovhannes, et al.
Pubblicazione: (2025)
di: Margaryan, Hovhannes, et al.
Pubblicazione: (2025)
RGS-DR: Deferred Reflections and Residual Shading in 2D Gaussian Splatting
di: Kouros, Georgios, et al.
Pubblicazione: (2025)
di: Kouros, Georgios, et al.
Pubblicazione: (2025)
Object-Centric Pretraining via Target Encoder Bootstrapping
di: Đukić, Nikola, et al.
Pubblicazione: (2025)
di: Đukić, Nikola, et al.
Pubblicazione: (2025)
Spec-Gloss Surfels and Normal-Diffuse Priors for Relightable Glossy Objects
di: Kouros, Georgios, et al.
Pubblicazione: (2025)
di: Kouros, Georgios, et al.
Pubblicazione: (2025)
Multi-Level Knowledge Distillation and Dynamic Self-Supervised Learning for Continual Learning
di: Kim, Taeheon, et al.
Pubblicazione: (2025)
di: Kim, Taeheon, et al.
Pubblicazione: (2025)
Same accuracy, twice as fast: continuous training surpasses retraining from scratch
di: Verwimp, Eli, et al.
Pubblicazione: (2025)
di: Verwimp, Eli, et al.
Pubblicazione: (2025)
Eff-GRot: Efficient and Generalizable Rotation Estimation with Transformers
di: Mathioulakis, Fanis, et al.
Pubblicazione: (2025)
di: Mathioulakis, Fanis, et al.
Pubblicazione: (2025)
TTA-DAME: Test-Time Adaptation with Domain Augmentation and Model Ensemble for Dynamic Driving Conditions
di: Jeon, Dongjae, et al.
Pubblicazione: (2025)
di: Jeon, Dongjae, et al.
Pubblicazione: (2025)
Toward Ambulatory Vision: Learning Visually-Grounded Active View Selection
di: Koo, Juil, et al.
Pubblicazione: (2025)
di: Koo, Juil, et al.
Pubblicazione: (2025)
Co-LoRA: Collaborative Model Personalization on Heterogeneous Multi-Modal Clients
di: Seo, Minhyuk, et al.
Pubblicazione: (2025)
di: Seo, Minhyuk, et al.
Pubblicazione: (2025)
Unsupervised Parameter Efficient Source-free Post-pretraining
di: Jha, Abhishek, et al.
Pubblicazione: (2025)
di: Jha, Abhishek, et al.
Pubblicazione: (2025)
DAVE: Diagnostic benchmark for Audio Visual Evaluation
di: Radevski, Gorjan, et al.
Pubblicazione: (2025)
di: Radevski, Gorjan, et al.
Pubblicazione: (2025)
Forgetting of task-specific knowledge in model merging-based continual learning
di: Hess, Timm, et al.
Pubblicazione: (2025)
di: Hess, Timm, et al.
Pubblicazione: (2025)
Animate Your Motion: Turning Still Images into Dynamic Videos
di: Li, Mingxiao, et al.
Pubblicazione: (2024)
di: Li, Mingxiao, et al.
Pubblicazione: (2024)
Knowledge Accumulation in Continually Learned Representations and the Issue of Feature Forgetting
di: Hess, Timm, et al.
Pubblicazione: (2023)
di: Hess, Timm, et al.
Pubblicazione: (2023)
Diversity-Driven View Subset Selection for Indoor Novel View Synthesis
di: Wang, Zehao, et al.
Pubblicazione: (2024)
di: Wang, Zehao, et al.
Pubblicazione: (2024)
Two Complementary Perspectives to Continual Learning: Ask Not Only What to Optimize, But Also How
di: Hess, Timm, et al.
Pubblicazione: (2023)
di: Hess, Timm, et al.
Pubblicazione: (2023)
Is this chart lying to me? Automating the detection of misleading visualizations
di: Tonglet, Jonathan, et al.
Pubblicazione: (2025)
di: Tonglet, Jonathan, et al.
Pubblicazione: (2025)
Selectively Dilated Convolution for Accuracy-Preserving Sparse Pillar-based Embedded 3D Object Detection
di: Park, Seongmin, et al.
Pubblicazione: (2024)
di: Park, Seongmin, et al.
Pubblicazione: (2024)
Unveiling the Ambiguity in Neural Inverse Rendering: A Parameter Compensation Analysis
di: Kouros, Georgios, et al.
Pubblicazione: (2024)
di: Kouros, Georgios, et al.
Pubblicazione: (2024)
Continual Learning of Diffusion Models with Generative Distillation
di: Masip, Sergi, et al.
Pubblicazione: (2023)
di: Masip, Sergi, et al.
Pubblicazione: (2023)
BG-Triangle: Bézier Gaussian Triangle for 3D Vectorization and Rendering
di: Wu, Minye, et al.
Pubblicazione: (2025)
di: Wu, Minye, et al.
Pubblicazione: (2025)
On the Role of Individual Differences in Current Approaches to Computational Image Aesthetics
di: Chen, Li-Wei, et al.
Pubblicazione: (2025)
di: Chen, Li-Wei, et al.
Pubblicazione: (2025)
Tuning Large Multimodal Models for Videos using Reinforcement Learning from AI Feedback
di: Ahn, Daechul, et al.
Pubblicazione: (2024)
di: Ahn, Daechul, et al.
Pubblicazione: (2024)
LLaVA-c: Continual Improved Visual Instruction Tuning
di: Liu, Wenzhuo, et al.
Pubblicazione: (2025)
di: Liu, Wenzhuo, et al.
Pubblicazione: (2025)
The Common Stability Mechanism behind most Self-Supervised Learning Approaches
di: Jha, Abhishek, et al.
Pubblicazione: (2024)
di: Jha, Abhishek, et al.
Pubblicazione: (2024)
A Simple Framework for Open-Vocabulary Zero-Shot Segmentation
di: Stegmüller, Thomas, et al.
Pubblicazione: (2024)
di: Stegmüller, Thomas, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Introducing Routing Functions to Vision-Language Parameter-Efficient Fine-Tuning with Low-Rank Bottlenecks
di: Qu, Tingyu, et al.
Pubblicazione: (2024) -
Visually-Aware Context Modeling for News Image Captioning
di: Qu, Tingyu, et al.
Pubblicazione: (2023) -
Budgeted Online Continual Learning by Adaptive Layer Freezing and Frequency-based Sampling
di: Seo, Minhyuk, et al.
Pubblicazione: (2024) -
TS-LLaVA: Constructing Visual Tokens through Thumbnail-and-Sampling for Training-Free Video Large Language Models
di: Qu, Tingyu, et al.
Pubblicazione: (2024) -
Attention Misses Visual Risk: Risk-Adaptive Steering for Multimodal Safety Alignment
di: Park, Jonghyun, et al.
Pubblicazione: (2025)