Boosting Few-Shot Detection with Large Language Models and Layout-to-Image Synthesis
Fuente:
arXiv
Saved in:
| Main Authors: | Abdullah, Ahmed, Ebert, Nikolas, Wasenmüller, Oliver |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
TAP into the Patch Tokens: Leveraging Vision Foundation Model Features for AI-Generated Image Detection
by: Abdullah, Ahmed, et al.
Published: (2026)
by: Abdullah, Ahmed, et al.
Published: (2026)
GenFormer -- Generated Images are All You Need to Improve Robustness of Transformers on Small Datasets
by: Oehri, Sven, et al.
Published: (2024)
by: Oehri, Sven, et al.
Published: (2024)
SSFT: A Lightweight Spectral-Spatial Fusion Transformer for Generic Hyperspectral Classification
by: Musiat, Alexander, et al.
Published: (2026)
by: Musiat, Alexander, et al.
Published: (2026)
PointTransformerX: Portable and Efficient 3D Point Cloud Processing without Sparse Algorithms
by: Reichardt, Laurenz, et al.
Published: (2026)
by: Reichardt, Laurenz, et al.
Published: (2026)
Classifier Ensemble for Efficient Uncertainty Calibration of Deep Neural Networks for Image Classification
by: Schulze, Michael, et al.
Published: (2025)
by: Schulze, Michael, et al.
Published: (2025)
IonMorphNet: Generalizable Learning of Ion Image Morphologies for Peak Picking in Mass Spectrometry Imaging
by: Weigand, Philipp, et al.
Published: (2026)
by: Weigand, Philipp, et al.
Published: (2026)
D-PLS: Decoupled Semantic Segmentation for 4D-Panoptic-LiDAR-Segmentation
by: Steinhauser, Maik, et al.
Published: (2025)
by: Steinhauser, Maik, et al.
Published: (2025)
Spatial self-supervised Peak Learning and correlation-based Evaluation of peak picking in Mass Spectrometry Imaging
by: Weigand, Philipp, et al.
Published: (2026)
by: Weigand, Philipp, et al.
Published: (2026)
Generative Texture Diversification of 3D Pedestrians for Robust Autonomous Driving Perception
by: Bhowmick, Arka, et al.
Published: (2026)
by: Bhowmick, Arka, et al.
Published: (2026)
RadarPillars: Efficient Object Detection from 4D Radar Point Clouds
by: Musiat, Alexander, et al.
Published: (2024)
by: Musiat, Alexander, et al.
Published: (2024)
Conditional Distribution Modelling for Few-Shot Image Synthesis with Diffusion Models
by: Gupta, Parul, et al.
Published: (2024)
by: Gupta, Parul, et al.
Published: (2024)
Image to Pseudo-Episode: Boosting Few-Shot Segmentation by Unlabeled Data
by: Zhang, Jie, et al.
Published: (2024)
by: Zhang, Jie, et al.
Published: (2024)
Revisiting Few-Shot Object Detection with Vision-Language Models
by: Madan, Anish, et al.
Published: (2023)
by: Madan, Anish, et al.
Published: (2023)
Zero-Shot Prompting and Few-Shot Fine-Tuning: Revisiting Document Image Classification Using Large Language Models
by: Scius-Bertrand, Anna, et al.
Published: (2024)
by: Scius-Bertrand, Anna, et al.
Published: (2024)
CM1 -- A Dataset for Evaluating Few-Shot Information Extraction with Large Vision Language Models
by: Wolf, Fabian, et al.
Published: (2025)
by: Wolf, Fabian, et al.
Published: (2025)
ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models
by: Gong, Bingchen, et al.
Published: (2024)
by: Gong, Bingchen, et al.
Published: (2024)
Boosting Few-Shot Semantic Segmentation Via Segment Anything Model
by: Feng, Chen-Bin, et al.
Published: (2024)
by: Feng, Chen-Bin, et al.
Published: (2024)
Layered Diffusion Model for One-Shot High Resolution Text-to-Image Synthesis
by: Khwaja, Emaad, et al.
Published: (2024)
by: Khwaja, Emaad, et al.
Published: (2024)
OmniDFA: A Unified Framework for Open Set Synthesis Image Detection and Few-Shot Attribution
by: Wu, Shiyu, et al.
Published: (2025)
by: Wu, Shiyu, et al.
Published: (2025)
Few-Shot Medical Image Segmentation with Large Kernel Attention
by: Wu, Xiaoxiao, et al.
Published: (2024)
by: Wu, Xiaoxiao, et al.
Published: (2024)
Text3DAug -- Prompted Instance Augmentation for LiDAR Perception
by: Reichardt, Laurenz, et al.
Published: (2024)
by: Reichardt, Laurenz, et al.
Published: (2024)
Few-Shot Image Quality Assessment via Adaptation of Vision-Language Models
by: Li, Xudong, et al.
Published: (2024)
by: Li, Xudong, et al.
Published: (2024)
A Closer Look at the Few-Shot Adaptation of Large Vision-Language Models
by: Silva-Rodríguez, Julio, et al.
Published: (2023)
by: Silva-Rodríguez, Julio, et al.
Published: (2023)
LLaFS: When Large Language Models Meet Few-Shot Segmentation
by: Zhu, Lanyun, et al.
Published: (2023)
by: Zhu, Lanyun, et al.
Published: (2023)
Boosting Few-Shot Learning via Attentive Feature Regularization
by: Zhu, Xingyu, et al.
Published: (2024)
by: Zhu, Xingyu, et al.
Published: (2024)
Layout Agnostic Scene Text Image Synthesis with Diffusion Models
by: Zhangli, Qilong, et al.
Published: (2024)
by: Zhangli, Qilong, et al.
Published: (2024)
Layout Anything: One Transformer for Universal Room Layout Estimation
by: Mia, Md Sohag, et al.
Published: (2025)
by: Mia, Md Sohag, et al.
Published: (2025)
LayoutCoT: Unleashing the Deep Reasoning Potential of Large Language Models for Layout Generation
by: Shi, Hengyu, et al.
Published: (2025)
by: Shi, Hengyu, et al.
Published: (2025)
Few-Shot Learner Generalizes Across AI-Generated Image Detection
by: Wu, Shiyu, et al.
Published: (2025)
by: Wu, Shiyu, et al.
Published: (2025)
Boosting Few-Shot Open-Set Object Detection via Prompt Learning and Robust Decision Boundary
by: Wu, Zhaowei, et al.
Published: (2024)
by: Wu, Zhaowei, et al.
Published: (2024)
Re-Scoring Using Image-Language Similarity for Few-Shot Object Detection
by: Jung, Min Jae, et al.
Published: (2023)
by: Jung, Min Jae, et al.
Published: (2023)
LayoutLLM: Layout Instruction Tuning with Large Language Models for Document Understanding
by: Luo, Chuwei, et al.
Published: (2024)
by: Luo, Chuwei, et al.
Published: (2024)
AirShot: Efficient Few-Shot Detection for Autonomous Exploration
by: Wang, Zihan, et al.
Published: (2024)
by: Wang, Zihan, et al.
Published: (2024)
Few-Shot Image Classification and Segmentation as Visual Question Answering Using Vision-Language Models
by: Meng, Tian, et al.
Published: (2024)
by: Meng, Tian, et al.
Published: (2024)
Few-Shot Learning from Gigapixel Images via Hierarchical Vision-Language Alignment and Modeling
by: Wong, Bryan, et al.
Published: (2025)
by: Wong, Bryan, et al.
Published: (2025)
InfRS: Incremental Few-Shot Object Detection in Remote Sensing Images
by: Li, Wuzhou, et al.
Published: (2024)
by: Li, Wuzhou, et al.
Published: (2024)
Low-Rank Few-Shot Adaptation of Vision-Language Models
by: Zanella, Maxime, et al.
Published: (2024)
by: Zanella, Maxime, et al.
Published: (2024)
Semi-Supervised Few-Shot Adaptation of Vision-Language Models
by: Silva-Rodríguez, Julio, et al.
Published: (2026)
by: Silva-Rodríguez, Julio, et al.
Published: (2026)
Calibrated Cache Model for Few-Shot Vision-Language Model Adaptation
by: Ding, Kun, et al.
Published: (2024)
by: Ding, Kun, et al.
Published: (2024)
Dual Distillation for Few-Shot Anomaly Detection
by: Dong, Le, et al.
Published: (2026)
by: Dong, Le, et al.
Published: (2026)
Similar Items
-
TAP into the Patch Tokens: Leveraging Vision Foundation Model Features for AI-Generated Image Detection
by: Abdullah, Ahmed, et al.
Published: (2026) -
GenFormer -- Generated Images are All You Need to Improve Robustness of Transformers on Small Datasets
by: Oehri, Sven, et al.
Published: (2024) -
SSFT: A Lightweight Spectral-Spatial Fusion Transformer for Generic Hyperspectral Classification
by: Musiat, Alexander, et al.
Published: (2026) -
PointTransformerX: Portable and Efficient 3D Point Cloud Processing without Sparse Algorithms
by: Reichardt, Laurenz, et al.
Published: (2026) -
Classifier Ensemble for Efficient Uncertainty Calibration of Deep Neural Networks for Image Classification
by: Schulze, Michael, et al.
Published: (2025)