Generative Compositor for Few-Shot Visual Information Extraction
Fuente:
arXiv
Saved in:
| Main Authors: | Yang, Zhibo, Hua, Wei, Song, Sibo, Yao, Cong, Zhu, Yingying, Cheng, Wenqing, Bai, Xiang |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
OmniParser: A Unified Framework for Text Spotting, Key Information Extraction and Table Recognition
by: Wan, Jianqiang, et al.
Published: (2024)
by: Wan, Jianqiang, et al.
Published: (2024)
OmniParser V2: Structured-Points-of-Thought for Unified Visual Text Parsing and Its Generality to Multimodal Large Language Models
by: Yu, Wenwen, et al.
Published: (2025)
by: Yu, Wenwen, et al.
Published: (2025)
HIP: Hierarchical Point Modeling and Pre-training for Visual Information Extraction
by: Long, Rujiao, et al.
Published: (2024)
by: Long, Rujiao, et al.
Published: (2024)
GenCompositor: Generative Video Compositing with Diffusion Transformer
by: Yang, Shuzhou, et al.
Published: (2025)
by: Yang, Shuzhou, et al.
Published: (2025)
Relation-Rich Visual Document Generator for Visual Information Extraction
by: Jiang, Zi-Han, et al.
Published: (2025)
by: Jiang, Zi-Han, et al.
Published: (2025)
Few-Shot Relation Extraction with Hybrid Visual Evidence
by: Gong, Jiaying, et al.
Published: (2024)
by: Gong, Jiaying, et al.
Published: (2024)
VL-Reader: Vision and Language Reconstructor is an Effective Scene Text Recognizer
by: Zhong, Humen, et al.
Published: (2024)
by: Zhong, Humen, et al.
Published: (2024)
Sequential Visual and Semantic Consistency for Semi-supervised Text Recognition
by: Yang, Mingkun, et al.
Published: (2024)
by: Yang, Mingkun, et al.
Published: (2024)
Query-guided Prototype Evolution Network for Few-Shot Segmentation
by: Cong, Runmin, et al.
Published: (2024)
by: Cong, Runmin, et al.
Published: (2024)
Visual Text Generation in the Wild
by: Zhu, Yuanzhi, et al.
Published: (2024)
by: Zhu, Yuanzhi, et al.
Published: (2024)
Local-Prompt: Extensible Local Prompts for Few-Shot Out-of-Distribution Detection
by: Zeng, Fanhu, et al.
Published: (2024)
by: Zeng, Fanhu, et al.
Published: (2024)
Divide-and-Conquer Decoupled Network for Cross-Domain Few-Shot Segmentation
by: Cong, Runmin, et al.
Published: (2025)
by: Cong, Runmin, et al.
Published: (2025)
Towards Generalized Few-Shot Open-Set Object Detection
by: Su, Binyi, et al.
Published: (2022)
by: Su, Binyi, et al.
Published: (2022)
Towards Efficient and General-Purpose Few-Shot Misclassification Detection for Vision-Language Models
by: Zeng, Fanhu, et al.
Published: (2025)
by: Zeng, Fanhu, et al.
Published: (2025)
Generalized Semantic Contrastive Learning via Embedding Side Information for Few-Shot Object Detection
by: Chen, Ruoyu, et al.
Published: (2025)
by: Chen, Ruoyu, et al.
Published: (2025)
FiLo++: Zero-/Few-Shot Anomaly Detection by Fused Fine-Grained Descriptions and Deformable Localization
by: Gu, Zhaopeng, et al.
Published: (2025)
by: Gu, Zhaopeng, et al.
Published: (2025)
1S-DAug: One-Shot Data Augmentation for Robust Few-Shot Generalization
by: Bai, Yunwei, et al.
Published: (2026)
by: Bai, Yunwei, et al.
Published: (2026)
Bayesian Evidential Learning for Few-Shot Classification
by: Linghu, Xiongkun, et al.
Published: (2022)
by: Linghu, Xiongkun, et al.
Published: (2022)
Hybrid Mamba for Few-Shot Segmentation
by: Xu, Qianxiong, et al.
Published: (2024)
by: Xu, Qianxiong, et al.
Published: (2024)
Rethinking Prior Information Generation with CLIP for Few-Shot Segmentation
by: Wang, Jin, et al.
Published: (2024)
by: Wang, Jin, et al.
Published: (2024)
CM1 -- A Dataset for Evaluating Few-Shot Information Extraction with Large Vision Language Models
by: Wolf, Fabian, et al.
Published: (2025)
by: Wolf, Fabian, et al.
Published: (2025)
Beyond Visual Cues: Leveraging General Semantics as Support for Few-Shot Segmentation
by: Wang, Jin, et al.
Published: (2025)
by: Wang, Jin, et al.
Published: (2025)
Dual Distillation for Few-Shot Anomaly Detection
by: Dong, Le, et al.
Published: (2026)
by: Dong, Le, et al.
Published: (2026)
The Devil is in the Few Shots: Iterative Visual Knowledge Completion for Few-shot Learning
by: Li, Yaohui, et al.
Published: (2024)
by: Li, Yaohui, et al.
Published: (2024)
UniVAD: A Training-free Unified Model for Few-shot Visual Anomaly Detection
by: Gu, Zhaopeng, et al.
Published: (2024)
by: Gu, Zhaopeng, et al.
Published: (2024)
Boosting Consistency in Story Visualization with Rich-Contextual Conditional Diffusion Models
by: Shen, Fei, et al.
Published: (2024)
by: Shen, Fei, et al.
Published: (2024)
Verbalized Representation Learning for Interpretable Few-Shot Generalization
by: Yang, Cheng-Fu, et al.
Published: (2024)
by: Yang, Cheng-Fu, et al.
Published: (2024)
Generalization-Enhanced Few-Shot Object Detection in Remote Sensing
by: Lin, Hui, et al.
Published: (2025)
by: Lin, Hui, et al.
Published: (2025)
Few-Part-Shot Font Generation
by: Akiba, Masaki, et al.
Published: (2025)
by: Akiba, Masaki, et al.
Published: (2025)
Exploring Few-Shot Defect Segmentation in General Industrial Scenarios with Metric Learning and Vision Foundation Models
by: Liu, Tongkun, et al.
Published: (2025)
by: Liu, Tongkun, et al.
Published: (2025)
Unlocking the Power of SAM 2 for Few-Shot Segmentation
by: Xu, Qianxiong, et al.
Published: (2025)
by: Xu, Qianxiong, et al.
Published: (2025)
A Turn Toward Better Alignment: Few-Shot Generative Adaptation with Equivariant Feature Rotation
by: Xu, Chenghao, et al.
Published: (2025)
by: Xu, Chenghao, et al.
Published: (2025)
Class-Aware Mask-Guided Feature Refinement for Scene Text Recognition
by: Yang, Mingkun, et al.
Published: (2024)
by: Yang, Mingkun, et al.
Published: (2024)
DocThinker: Explainable Multimodal Large Language Models with Rule-based Reinforcement Learning for Document Understanding
by: Yu, Wenwen, et al.
Published: (2025)
by: Yu, Wenwen, et al.
Published: (2025)
Adaptive Decision Boundary for Few-Shot Class-Incremental Learning
by: Li, Linhao, et al.
Published: (2025)
by: Li, Linhao, et al.
Published: (2025)
MPA: Multimodal Prototype Augmentation for Few-Shot Learning
by: Wu, Liwen, et al.
Published: (2026)
by: Wu, Liwen, et al.
Published: (2026)
Object-level Correlation for Few-Shot Segmentation
by: Wen, Chunlin, et al.
Published: (2025)
by: Wen, Chunlin, et al.
Published: (2025)
Generalized Few-Shot Semantic Segmentation in Remote Sensing: Challenge and Benchmark
by: Broni-Bediako, Clifford, et al.
Published: (2024)
by: Broni-Bediako, Clifford, et al.
Published: (2024)
Attend and Enrich: Enhanced Visual Prompt for Zero-Shot Learning
by: Liu, Man, et al.
Published: (2024)
by: Liu, Man, et al.
Published: (2024)
A Feature Generator for Few-Shot Learning
by: Kanagalingam, Heethanjan, et al.
Published: (2024)
by: Kanagalingam, Heethanjan, et al.
Published: (2024)
Similar Items
-
OmniParser: A Unified Framework for Text Spotting, Key Information Extraction and Table Recognition
by: Wan, Jianqiang, et al.
Published: (2024) -
OmniParser V2: Structured-Points-of-Thought for Unified Visual Text Parsing and Its Generality to Multimodal Large Language Models
by: Yu, Wenwen, et al.
Published: (2025) -
HIP: Hierarchical Point Modeling and Pre-training for Visual Information Extraction
by: Long, Rujiao, et al.
Published: (2024) -
GenCompositor: Generative Video Compositing with Diffusion Transformer
by: Yang, Shuzhou, et al.
Published: (2025) -
Relation-Rich Visual Document Generator for Visual Information Extraction
by: Jiang, Zi-Han, et al.
Published: (2025)