Enhancing Multimodal In-Context Learning for Image Classification through Coreset Optimization
Fuente:
arXiv
Saved in:
| Main Authors: | Chen, Huiyi, Peng, Jiawei, Tang, Kaihua, Geng, Xin, Yang, Xu |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Geometry-Aware Uncertainty Coresets for Robust Visual In-Context Learning in Histopathology
by: Erick, Franciskus Xaverius, et al.
Published: (2026)
by: Erick, Franciskus Xaverius, et al.
Published: (2026)
Natural Language Supervision for Low-light Image Enhancement
by: Tang, Jiahui, et al.
Published: (2025)
by: Tang, Jiahui, et al.
Published: (2025)
Enhancing Multimodal In-Context Learning via Inductive-Deductive Reasoning
by: Wang, Haoyu, et al.
Published: (2026)
by: Wang, Haoyu, et al.
Published: (2026)
CADIC: Continual Anomaly Detection Based on Incremental Coreset
by: Yang, Gen, et al.
Published: (2025)
by: Yang, Gen, et al.
Published: (2025)
Enhanced Convolutional Neural Networks for Improved Image Classification
by: Yang, Xiaoran, et al.
Published: (2025)
by: Yang, Xiaoran, et al.
Published: (2025)
Multimodal Medical Image Classification via Synergistic Learning Pre-training
by: Lin, Qinghua, et al.
Published: (2025)
by: Lin, Qinghua, et al.
Published: (2025)
Large Vision-Language Models as Emotion Recognizers in Context Awareness
by: Lei, Yuxuan, et al.
Published: (2024)
by: Lei, Yuxuan, et al.
Published: (2024)
COHERENCE: Benchmarking Fine-Grained Image-Text Alignment in Interleaved Multimodal Contexts
by: Wang, Bingli, et al.
Published: (2026)
by: Wang, Bingli, et al.
Published: (2026)
True Multimodal In-Context Learning Needs Attention to the Visual Context
by: Chen, Shuo, et al.
Published: (2025)
by: Chen, Shuo, et al.
Published: (2025)
Diffusion-Augmented Coreset Expansion for Scalable Dataset Distillation
by: Abbasi, Ali, et al.
Published: (2024)
by: Abbasi, Ali, et al.
Published: (2024)
TextMatch: Enhancing Image-Text Consistency Through Multimodal Optimization
by: Luo, Yucong, et al.
Published: (2024)
by: Luo, Yucong, et al.
Published: (2024)
MMCL-Bench: Multimodal Context Learning from Visual Rules, Procedures, and Evidence
by: Chen, Yifan, et al.
Published: (2026)
by: Chen, Yifan, et al.
Published: (2026)
Enhancing Whole Slide Image Classification through Supervised Contrastive Domain Adaptation
by: Carretero, Ilán, et al.
Published: (2024)
by: Carretero, Ilán, et al.
Published: (2024)
ELFS: Label-Free Coreset Selection with Proxy Training Dynamics
by: Zheng, Haizhong, et al.
Published: (2024)
by: Zheng, Haizhong, et al.
Published: (2024)
Reviving the Context: Camera Trap Species Classification as Link Prediction on Multimodal Knowledge Graphs
by: Pahuja, Vardaan, et al.
Published: (2023)
by: Pahuja, Vardaan, et al.
Published: (2023)
Discrete Diffusion Models with MLLMs for Unified Medical Multimodal Generation
by: Mao, Jiawei, et al.
Published: (2025)
by: Mao, Jiawei, et al.
Published: (2025)
Flatten: Video Action Recognition is an Image Classification task
by: Chen, Junlin, et al.
Published: (2024)
by: Chen, Junlin, et al.
Published: (2024)
SegICL: A Multimodal In-context Learning Framework for Enhanced Segmentation in Medical Imaging
by: Shen, Lingdong, et al.
Published: (2024)
by: Shen, Lingdong, et al.
Published: (2024)
Multimodal Contrastive Pretraining of CBCT and IOS for Enhanced Tooth Segmentation
by: Son, Moo Hyun, et al.
Published: (2025)
by: Son, Moo Hyun, et al.
Published: (2025)
What Makes Multimodal In-Context Learning Work?
by: Baldassini, Folco Bertini, et al.
Published: (2024)
by: Baldassini, Folco Bertini, et al.
Published: (2024)
Enhancing Few-Shot Image Classification through Learnable Multi-Scale Embedding and Attention Mechanisms
by: Askari, Fatemeh, et al.
Published: (2024)
by: Askari, Fatemeh, et al.
Published: (2024)
Case-Aware Medical Image Classification with Multimodal Knowledge Graphs and Reliability-Guided Refinement
by: Xu, Yiming, et al.
Published: (2026)
by: Xu, Yiming, et al.
Published: (2026)
Enhancing Orthopox Image Classification Using Hybrid Machine Learning and Deep Learning Models
by: Puente-Castro, Alejandro, et al.
Published: (2025)
by: Puente-Castro, Alejandro, et al.
Published: (2025)
Parameter Interpolation Adversarial Training for Robust Image Classification
by: Liu, Xin, et al.
Published: (2025)
by: Liu, Xin, et al.
Published: (2025)
Looking Back and Forth: Cross-Image Attention Calibration and Attentive Preference Learning for Multi-Image Hallucination Mitigation
by: Yang, Xiaochen, et al.
Published: (2026)
by: Yang, Xiaochen, et al.
Published: (2026)
Multimodal Deep Learning for Subtype Classification in Breast Cancer Using Histopathological Images and Gene Expression Data
by: Shandiz, Amin Honarmandi
Published: (2025)
by: Shandiz, Amin Honarmandi
Published: (2025)
Harnessing Shared Relations via Multimodal Mixup Contrastive Learning for Multimodal Classification
by: Kumar, Raja, et al.
Published: (2024)
by: Kumar, Raja, et al.
Published: (2024)
Multimodal Continual Learning with MLLMs from Multi-scenario Perspectives
by: Jiang, Kai, et al.
Published: (2025)
by: Jiang, Kai, et al.
Published: (2025)
How Does the Textual Information Affect the Retrieval of Multimodal In-Context Learning?
by: Luo, Yang, et al.
Published: (2024)
by: Luo, Yang, et al.
Published: (2024)
Enhanced Multi-Class Classification of Gastrointestinal Endoscopic Images with Interpretable Deep Learning Model
by: Kamble, Astitva, et al.
Published: (2025)
by: Kamble, Astitva, et al.
Published: (2025)
M$^2$IV: Towards Efficient and Fine-grained Multimodal In-Context Learning via Representation Engineering
by: Li, Yanshu, et al.
Published: (2025)
by: Li, Yanshu, et al.
Published: (2025)
MedSAM-Agent: Empowering Interactive Medical Image Segmentation with Multi-turn Agentic Reinforcement Learning
by: Liu, Shengyuan, et al.
Published: (2026)
by: Liu, Shengyuan, et al.
Published: (2026)
Dual-Model Distillation for Efficient Action Classification with Hybrid Edge-Cloud Solution
by: Wei, Timothy, et al.
Published: (2024)
by: Wei, Timothy, et al.
Published: (2024)
Revealing Temporal Label Noise in Multimodal Hateful Video Classification
by: Yang, Shuonan, et al.
Published: (2025)
by: Yang, Shuonan, et al.
Published: (2025)
Few-shot Writer Adaptation via Multimodal In-Context Learning
by: Simon, Tom, et al.
Published: (2026)
by: Simon, Tom, et al.
Published: (2026)
INSIGHT: Enhancing Autonomous Driving Safety through Vision-Language Models on Context-Aware Hazard Detection and Edge Case Evaluation
by: Chen, Dianwei, et al.
Published: (2025)
by: Chen, Dianwei, et al.
Published: (2025)
Enlighten-Your-Voice: When Multimodal Meets Zero-shot Low-light Image Enhancement
by: Zhang, Xiaofeng, et al.
Published: (2023)
by: Zhang, Xiaofeng, et al.
Published: (2023)
ArtAug: Enhancing Text-to-Image Generation through Synthesis-Understanding Interaction
by: Duan, Zhongjie, et al.
Published: (2024)
by: Duan, Zhongjie, et al.
Published: (2024)
Metis-HOME: Hybrid Optimized Mixture-of-Experts for Multimodal Reasoning
by: Lan, Xiaohan, et al.
Published: (2025)
by: Lan, Xiaohan, et al.
Published: (2025)
Active Generation for Image Classification
by: Huang, Tao, et al.
Published: (2024)
by: Huang, Tao, et al.
Published: (2024)
Similar Items
-
Geometry-Aware Uncertainty Coresets for Robust Visual In-Context Learning in Histopathology
by: Erick, Franciskus Xaverius, et al.
Published: (2026) -
Natural Language Supervision for Low-light Image Enhancement
by: Tang, Jiahui, et al.
Published: (2025) -
Enhancing Multimodal In-Context Learning via Inductive-Deductive Reasoning
by: Wang, Haoyu, et al.
Published: (2026) -
CADIC: Continual Anomaly Detection Based on Incremental Coreset
by: Yang, Gen, et al.
Published: (2025) -
Enhanced Convolutional Neural Networks for Improved Image Classification
by: Yang, Xiaoran, et al.
Published: (2025)