Few-Shot Relation Extraction with Hybrid Visual Evidence
Fuente:
arXiv
Salvato in:
| Autori principali: | Gong, Jiaying, Eldardiry, Hoda |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
VideoAVE: A Multi-Attribute Video-to-Text Attribute Value Extraction Dataset and Benchmark Models
di: Cheng, Ming, et al.
Pubblicazione: (2025)
di: Cheng, Ming, et al.
Pubblicazione: (2025)
Visual Zero-Shot E-Commerce Product Attribute Value Extraction
di: Gong, Jiaying, et al.
Pubblicazione: (2025)
di: Gong, Jiaying, et al.
Pubblicazione: (2025)
Prompt-based Zero-shot Relation Extraction with Semantic Knowledge Augmentation
di: Gong, Jiaying, et al.
Pubblicazione: (2021)
di: Gong, Jiaying, et al.
Pubblicazione: (2021)
Investigating Prompting Techniques for Zero- and Few-Shot Visual Question Answering
di: Awal, Rabiul, et al.
Pubblicazione: (2023)
di: Awal, Rabiul, et al.
Pubblicazione: (2023)
Generative Compositor for Few-Shot Visual Information Extraction
di: Yang, Zhibo, et al.
Pubblicazione: (2025)
di: Yang, Zhibo, et al.
Pubblicazione: (2025)
Enhancing Vision-Language Few-Shot Adaptation with Negative Learning
di: Zhang, Ce, et al.
Pubblicazione: (2024)
di: Zhang, Ce, et al.
Pubblicazione: (2024)
Multi-Level Correlation Network For Few-Shot Image Classification
di: Dang, Yunkai, et al.
Pubblicazione: (2024)
di: Dang, Yunkai, et al.
Pubblicazione: (2024)
Few-Shot VQA with Frozen LLMs: A Tale of Two Approaches
di: Sterner, Igor, et al.
Pubblicazione: (2024)
di: Sterner, Igor, et al.
Pubblicazione: (2024)
Evaluating Linguistic Capabilities of Multimodal LLMs in the Lens of Few-Shot Learning
di: Dogan, Mustafa, et al.
Pubblicazione: (2024)
di: Dogan, Mustafa, et al.
Pubblicazione: (2024)
Sci-LoRA: Mixture of Scientific LoRAs for Cross-Domain Lay Paraphrasing
di: Cheng, Ming, et al.
Pubblicazione: (2025)
di: Cheng, Ming, et al.
Pubblicazione: (2025)
Efficient Few-Shot Medical Image Analysis via Hierarchical Contrastive Vision-Language Learning
di: Fuller, Harrison, et al.
Pubblicazione: (2025)
di: Fuller, Harrison, et al.
Pubblicazione: (2025)
See, Explain, and Intervene: A Few-Shot Multimodal Agent Framework for Hateful Meme Moderation
di: Rizwan, Naquee, et al.
Pubblicazione: (2026)
di: Rizwan, Naquee, et al.
Pubblicazione: (2026)
Meta-Adaptive Prompt Distillation for Few-Shot Visual Question Answering
di: Gupta, Akash, et al.
Pubblicazione: (2025)
di: Gupta, Akash, et al.
Pubblicazione: (2025)
Learning to Adapt Category Consistent Meta-Feature of CLIP for Few-Shot Classification
di: Shi, Jiaying, et al.
Pubblicazione: (2024)
di: Shi, Jiaying, et al.
Pubblicazione: (2024)
Joint Extraction Matters: Prompt-Based Visual Question Answering for Multi-Field Document Information Extraction
di: Loem, Mengsay, et al.
Pubblicazione: (2025)
di: Loem, Mengsay, et al.
Pubblicazione: (2025)
Verbalized Representation Learning for Interpretable Few-Shot Generalization
di: Yang, Cheng-Fu, et al.
Pubblicazione: (2024)
di: Yang, Cheng-Fu, et al.
Pubblicazione: (2024)
Mechanistic Finetuning of Vision-Language-Action Models via Few-Shot Demonstrations
di: Mitra, Chancharik, et al.
Pubblicazione: (2025)
di: Mitra, Chancharik, et al.
Pubblicazione: (2025)
Hybrid Mamba for Few-Shot Segmentation
di: Xu, Qianxiong, et al.
Pubblicazione: (2024)
di: Xu, Qianxiong, et al.
Pubblicazione: (2024)
Recurrent Visual Feature Extraction and Stereo Attentions for CT Report Generation
di: Tian, Yuanhe, et al.
Pubblicazione: (2025)
di: Tian, Yuanhe, et al.
Pubblicazione: (2025)
Pose2Gest: A Few-Shot Model-Free Approach Applied In South Indian Classical Dance Gesture Recognition
di: Raju, Kavitha, et al.
Pubblicazione: (2024)
di: Raju, Kavitha, et al.
Pubblicazione: (2024)
UNIKIE-BENCH: Benchmarking Large Multimodal Models for Key Information Extraction in Visual Documents
di: Ji, Yifan, et al.
Pubblicazione: (2026)
di: Ji, Yifan, et al.
Pubblicazione: (2026)
InstructDoc: A Dataset for Zero-Shot Generalization of Visual Document Understanding with Instructions
di: Tanaka, Ryota, et al.
Pubblicazione: (2024)
di: Tanaka, Ryota, et al.
Pubblicazione: (2024)
Enhancing Few-Shot Vision-Language Classification with Large Multimodal Model Features
di: Mitra, Chancharik, et al.
Pubblicazione: (2024)
di: Mitra, Chancharik, et al.
Pubblicazione: (2024)
RMPL: Relation-aware Multi-task Progressive Learning with Stage-wise Training for Multimedia Event Extraction
di: Jin, Yongkang, et al.
Pubblicazione: (2026)
di: Jin, Yongkang, et al.
Pubblicazione: (2026)
EviMem: Evidence-Gap-Driven Iterative Retrieval for Long-Term Conversational Memory
di: Li, Yuyang, et al.
Pubblicazione: (2026)
di: Li, Yuyang, et al.
Pubblicazione: (2026)
VERA: Identifying and Leveraging Visual Evidence Retrieval Heads in Long-Context Understanding
di: Pei, Rongcan, et al.
Pubblicazione: (2026)
di: Pei, Rongcan, et al.
Pubblicazione: (2026)
Few-Shot Image Generation by Conditional Relaxing Diffusion Inversion
di: Cao, Yu, et al.
Pubblicazione: (2024)
di: Cao, Yu, et al.
Pubblicazione: (2024)
Few Shot Class Incremental Learning using Vision-Language models
di: Kumar, Anurag, et al.
Pubblicazione: (2024)
di: Kumar, Anurag, et al.
Pubblicazione: (2024)
The Devil is in the Few Shots: Iterative Visual Knowledge Completion for Few-shot Learning
di: Li, Yaohui, et al.
Pubblicazione: (2024)
di: Li, Yaohui, et al.
Pubblicazione: (2024)
PMCE: Probabilistic Multi-Granularity Semantics with Caption-Guided Enhancement for Few-Shot Learning
di: Wu, Jiaying, et al.
Pubblicazione: (2026)
di: Wu, Jiaying, et al.
Pubblicazione: (2026)
Grounding Language Models for Visual Entity Recognition
di: Xiao, Zilin, et al.
Pubblicazione: (2024)
di: Xiao, Zilin, et al.
Pubblicazione: (2024)
You May Speak Freely: Improving the Fine-Grained Visual Recognition Capabilities of Multimodal Large Language Models with Answer Extraction
di: Lawrence, Logan, et al.
Pubblicazione: (2025)
di: Lawrence, Logan, et al.
Pubblicazione: (2025)
VisuCraft: Enhancing Large Vision-Language Models for Complex Visual-Guided Creative Content Generation via Structured Information Extraction
di: Jiang, Rongxin, et al.
Pubblicazione: (2025)
di: Jiang, Rongxin, et al.
Pubblicazione: (2025)
Mitigating the Modality Gap: Few-Shot Out-of-Distribution Detection with Multi-modal Prototypes and Image Bias Estimation
di: Wang, Yimu, et al.
Pubblicazione: (2025)
di: Wang, Yimu, et al.
Pubblicazione: (2025)
VisRAG 2.0: Evidence-Guided Multi-Image Reasoning in Visual Retrieval-Augmented Generation
di: Sun, Yubo, et al.
Pubblicazione: (2025)
di: Sun, Yubo, et al.
Pubblicazione: (2025)
Few-Shot Adversarial Prompt Learning on Vision-Language Models
di: Zhou, Yiwei, et al.
Pubblicazione: (2024)
di: Zhou, Yiwei, et al.
Pubblicazione: (2024)
Controllable Relation Disentanglement for Few-Shot Class-Incremental Learning
di: Zhou, Yuan, et al.
Pubblicazione: (2024)
di: Zhou, Yuan, et al.
Pubblicazione: (2024)
Modularized Networks for Few-shot Hateful Meme Detection
di: Cao, Rui, et al.
Pubblicazione: (2024)
di: Cao, Rui, et al.
Pubblicazione: (2024)
Reasoning Resides in Layers: Restoring Temporal Reasoning in Video-Language Models with Layer-Selective Merging
di: Fu, Zihang, et al.
Pubblicazione: (2026)
di: Fu, Zihang, et al.
Pubblicazione: (2026)
ARPA: A Novel Hybrid Model for Advancing Visual Word Disambiguation Using Large Language Models and Transformers
di: Papastavrou, Aristi, et al.
Pubblicazione: (2024)
di: Papastavrou, Aristi, et al.
Pubblicazione: (2024)
Documenti analoghi
-
VideoAVE: A Multi-Attribute Video-to-Text Attribute Value Extraction Dataset and Benchmark Models
di: Cheng, Ming, et al.
Pubblicazione: (2025) -
Visual Zero-Shot E-Commerce Product Attribute Value Extraction
di: Gong, Jiaying, et al.
Pubblicazione: (2025) -
Prompt-based Zero-shot Relation Extraction with Semantic Knowledge Augmentation
di: Gong, Jiaying, et al.
Pubblicazione: (2021) -
Investigating Prompting Techniques for Zero- and Few-Shot Visual Question Answering
di: Awal, Rabiul, et al.
Pubblicazione: (2023) -
Generative Compositor for Few-Shot Visual Information Extraction
di: Yang, Zhibo, et al.
Pubblicazione: (2025)