CM1 -- A Dataset for Evaluating Few-Shot Information Extraction with Large Vision Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wolf, Fabian, Tüselmann, Oliver, Matei, Arthur, Hennies, Lukas, Rass, Christoph, Fink, Gernot A. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Self-Supervised Vision Transformers for Writer Retrieval
von: Raven, Tim, et al.
Veröffentlicht: (2024)
von: Raven, Tim, et al.
Veröffentlicht: (2024)
Exploring Architectures for CNN-Based Word Spotting
von: Rusakov, Eugen, et al.
Veröffentlicht: (2018)
von: Rusakov, Eugen, et al.
Veröffentlicht: (2018)
Boosting Few-Shot Detection with Large Language Models and Layout-to-Image Synthesis
von: Abdullah, Ahmed, et al.
Veröffentlicht: (2024)
von: Abdullah, Ahmed, et al.
Veröffentlicht: (2024)
A Closer Look at the Few-Shot Adaptation of Large Vision-Language Models
von: Silva-Rodríguez, Julio, et al.
Veröffentlicht: (2023)
von: Silva-Rodríguez, Julio, et al.
Veröffentlicht: (2023)
Generative Compositor for Few-Shot Visual Information Extraction
von: Yang, Zhibo, et al.
Veröffentlicht: (2025)
von: Yang, Zhibo, et al.
Veröffentlicht: (2025)
Semi-Supervised Few-Shot Adaptation of Vision-Language Models
von: Silva-Rodríguez, Julio, et al.
Veröffentlicht: (2026)
von: Silva-Rodríguez, Julio, et al.
Veröffentlicht: (2026)
Low-Rank Few-Shot Adaptation of Vision-Language Models
von: Zanella, Maxime, et al.
Veröffentlicht: (2024)
von: Zanella, Maxime, et al.
Veröffentlicht: (2024)
Revisiting Few-Shot Object Detection with Vision-Language Models
von: Madan, Anish, et al.
Veröffentlicht: (2023)
von: Madan, Anish, et al.
Veröffentlicht: (2023)
Calibrated Cache Model for Few-Shot Vision-Language Model Adaptation
von: Ding, Kun, et al.
Veröffentlicht: (2024)
von: Ding, Kun, et al.
Veröffentlicht: (2024)
Few-Shot Adaptation Benchmark for Remote Sensing Vision-Language Models
von: Khoury, Karim El, et al.
Veröffentlicht: (2025)
von: Khoury, Karim El, et al.
Veröffentlicht: (2025)
Auxiliary Descriptive Knowledge for Few-Shot Adaptation of Vision-Language Model
von: Lee, SuBeen, et al.
Veröffentlicht: (2025)
von: Lee, SuBeen, et al.
Veröffentlicht: (2025)
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features
von: Mitra, Chancharik, et al.
Veröffentlicht: (2024)
von: Mitra, Chancharik, et al.
Veröffentlicht: (2024)
Efficient Few-Shot Continual Learning in Vision-Language Models
von: Panos, Aristeidis, et al.
Veröffentlicht: (2025)
von: Panos, Aristeidis, et al.
Veröffentlicht: (2025)
Benchmarking Vision-Language and Multimodal Large Language Models in Zero-shot and Few-shot Scenarios: A study on Christian Iconography
von: Spinaci, Gianmarco, et al.
Veröffentlicht: (2025)
von: Spinaci, Gianmarco, et al.
Veröffentlicht: (2025)
FSDAM: Few-Shot Driving Attention Modeling via Vision-Language Coupling
von: Hamid, Kaiser, et al.
Veröffentlicht: (2025)
von: Hamid, Kaiser, et al.
Veröffentlicht: (2025)
Few-Shot Image Quality Assessment via Adaptation of Vision-Language Models
von: Li, Xudong, et al.
Veröffentlicht: (2024)
von: Li, Xudong, et al.
Veröffentlicht: (2024)
Vision-Language In-Context Learning Driven Few-Shot Visual Inspection Model
von: Ueno, Shiryu, et al.
Veröffentlicht: (2025)
von: Ueno, Shiryu, et al.
Veröffentlicht: (2025)
Noise-Tolerant Few-Shot Unsupervised Adapter for Vision-Language Models
von: Ali, Eman, et al.
Veröffentlicht: (2023)
von: Ali, Eman, et al.
Veröffentlicht: (2023)
ContextVLM: Zero-Shot and Few-Shot Context Understanding for Autonomous Driving using Vision Language Models
von: Sural, Shounak, et al.
Veröffentlicht: (2024)
von: Sural, Shounak, et al.
Veröffentlicht: (2024)
Cluster-Aware Prompt Ensemble Learning for Few-Shot Vision-Language Model Adaptation
von: Chen, Zhi, et al.
Veröffentlicht: (2025)
von: Chen, Zhi, et al.
Veröffentlicht: (2025)
Complementary Subspace Low-Rank Adaptation of Vision-Language Models for Few-Shot Classification
von: Wang, Zhongqi, et al.
Veröffentlicht: (2025)
von: Wang, Zhongqi, et al.
Veröffentlicht: (2025)
Pre-trained Vision and Language Transformers Are Few-Shot Incremental Learners
von: Park, Keon-Hee, et al.
Veröffentlicht: (2024)
von: Park, Keon-Hee, et al.
Veröffentlicht: (2024)
LLaFS: When Large Language Models Meet Few-Shot Segmentation
von: Zhu, Lanyun, et al.
Veröffentlicht: (2023)
von: Zhu, Lanyun, et al.
Veröffentlicht: (2023)
Rethinking Few-Shot Adaptation of Vision-Language Models in Two Stages
von: Farina, Matteo, et al.
Veröffentlicht: (2025)
von: Farina, Matteo, et al.
Veröffentlicht: (2025)
Preserve and Sculpt: Manifold-Aligned Fine-tuning of Vision-Language Models for Few-Shot Learning
von: Chen, Dexia, et al.
Veröffentlicht: (2025)
von: Chen, Dexia, et al.
Veröffentlicht: (2025)
Few-Shot Learning from Gigapixel Images via Hierarchical Vision-Language Alignment and Modeling
von: Wong, Bryan, et al.
Veröffentlicht: (2025)
von: Wong, Bryan, et al.
Veröffentlicht: (2025)
Few-Shot Image Classification and Segmentation as Visual Question Answering Using Vision-Language Models
von: Meng, Tian, et al.
Veröffentlicht: (2024)
von: Meng, Tian, et al.
Veröffentlicht: (2024)
DDFAV: Remote Sensing Large Vision Language Models Dataset and Evaluation Benchmark
von: Li, Haodong, et al.
Veröffentlicht: (2024)
von: Li, Haodong, et al.
Veröffentlicht: (2024)
Enhancing Vision-Language Few-Shot Adaptation with Negative Learning
von: Zhang, Ce, et al.
Veröffentlicht: (2024)
von: Zhang, Ce, et al.
Veröffentlicht: (2024)
Language-Aware Information Maximization for Transductive Few-Shot CLIP
von: Baklouti, Ghassen, et al.
Veröffentlicht: (2025)
von: Baklouti, Ghassen, et al.
Veröffentlicht: (2025)
Proto-CLIP: Vision-Language Prototypical Network for Few-Shot Learning
von: P, Jishnu Jaykumar, et al.
Veröffentlicht: (2023)
von: P, Jishnu Jaykumar, et al.
Veröffentlicht: (2023)
Few-Shot Vision-Language Reasoning for Satellite Imagery via Verifiable Rewards
von: Koksal, Aybora, et al.
Veröffentlicht: (2025)
von: Koksal, Aybora, et al.
Veröffentlicht: (2025)
Towards Fine-Grained Vision-Language Alignment for Few-Shot Anomaly Detection
von: Fan, Yuanting, et al.
Veröffentlicht: (2025)
von: Fan, Yuanting, et al.
Veröffentlicht: (2025)
Efficient Few-Shot Learning in Remote Sensing: Fusing Vision and Vision-Language Models
von: Chua, Jia Yun, et al.
Veröffentlicht: (2025)
von: Chua, Jia Yun, et al.
Veröffentlicht: (2025)
Few-Shot Adversarial Prompt Learning on Vision-Language Models
von: Zhou, Yiwei, et al.
Veröffentlicht: (2024)
von: Zhou, Yiwei, et al.
Veröffentlicht: (2024)
ProKeR: A Kernel Perspective on Few-Shot Adaptation of Large Vision-Language Models
von: Bendou, Yassir, et al.
Veröffentlicht: (2025)
von: Bendou, Yassir, et al.
Veröffentlicht: (2025)
Making Large Vision Language Models to be Good Few-shot Learners
von: Liu, Fan, et al.
Veröffentlicht: (2024)
von: Liu, Fan, et al.
Veröffentlicht: (2024)
Mechanistic Finetuning of Vision-Language-Action Models via Few-Shot Demonstrations
von: Mitra, Chancharik, et al.
Veröffentlicht: (2025)
von: Mitra, Chancharik, et al.
Veröffentlicht: (2025)
Few-Shot Relation Extraction with Hybrid Visual Evidence
von: Gong, Jiaying, et al.
Veröffentlicht: (2024)
von: Gong, Jiaying, et al.
Veröffentlicht: (2024)
Towards Efficient and General-Purpose Few-Shot Misclassification Detection for Vision-Language Models
von: Zeng, Fanhu, et al.
Veröffentlicht: (2025)
von: Zeng, Fanhu, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Self-Supervised Vision Transformers for Writer Retrieval
von: Raven, Tim, et al.
Veröffentlicht: (2024) -
Exploring Architectures for CNN-Based Word Spotting
von: Rusakov, Eugen, et al.
Veröffentlicht: (2018) -
Boosting Few-Shot Detection with Large Language Models and Layout-to-Image Synthesis
von: Abdullah, Ahmed, et al.
Veröffentlicht: (2024) -
A Closer Look at the Few-Shot Adaptation of Large Vision-Language Models
von: Silva-Rodríguez, Julio, et al.
Veröffentlicht: (2023) -
Generative Compositor for Few-Shot Visual Information Extraction
von: Yang, Zhibo, et al.
Veröffentlicht: (2025)