A Visual RAG Pipeline for Few-Shot Fine-Grained Product Classification
Fuente:
arXiv
Salvato in:
| Autori principali: | Lamm, Bianca, Keuper, Janis |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Can Visual Language Models Replace OCR-Based Visual Question Answering Pipelines in Production? A Case Study in Retail
di: Lamm, Bianca, et al.
Pubblicazione: (2024)
di: Lamm, Bianca, et al.
Pubblicazione: (2024)
Retail-786k: a Large-Scale Dataset for Visual Entity Matching
di: Lamm, Bianca, et al.
Pubblicazione: (2023)
di: Lamm, Bianca, et al.
Pubblicazione: (2023)
Deepfakes: we need to re-think the concept of "real" images
di: Keuper, Janis, et al.
Pubblicazione: (2025)
di: Keuper, Janis, et al.
Pubblicazione: (2025)
As large as it gets: Learning infinitely large Filters via Neural Implicit Functions in the Fourier Domain
di: Grabinski, Julia, et al.
Pubblicazione: (2023)
di: Grabinski, Julia, et al.
Pubblicazione: (2023)
A New Kind of Network? Review and Reference Implementation of Neural Cellular Automata
di: Spitznagel, Martin, et al.
Pubblicazione: (2026)
di: Spitznagel, Martin, et al.
Pubblicazione: (2026)
Unfolding Local Growth Rate Estimates for (Almost) Perfect Adversarial Detection
di: Lorenz, Peter, et al.
Pubblicazione: (2022)
di: Lorenz, Peter, et al.
Pubblicazione: (2022)
Frequency-Enhanced Dual-Subspace Networks for Few-Shot Fine-Grained Image Classification
di: Wang, Meijia, et al.
Pubblicazione: (2026)
di: Wang, Meijia, et al.
Pubblicazione: (2026)
Hybrid Feature Collaborative Reconstruction Network for Few-Shot Fine-Grained Image Classification
di: Qiu, Shulei, et al.
Pubblicazione: (2024)
di: Qiu, Shulei, et al.
Pubblicazione: (2024)
UltraAD: Fine-Grained Ultrasound Anomaly Classification via Few-Shot CLIP Adaptation
di: Zhou, Yue, et al.
Pubblicazione: (2025)
di: Zhou, Yue, et al.
Pubblicazione: (2025)
Hierarchical Mask-Enhanced Dual Reconstruction Network for Few-Shot Fine-Grained Image Classification
di: Luo, Ning, et al.
Pubblicazione: (2025)
di: Luo, Ning, et al.
Pubblicazione: (2025)
CausalFSFG: Rethinking Few-Shot Fine-Grained Visual Categorization from Causal Perspective
di: Yang, Zhiwen, et al.
Pubblicazione: (2025)
di: Yang, Zhiwen, et al.
Pubblicazione: (2025)
Detail Reinforcement Diffusion Model: Augmentation Fine-Grained Visual Categorization in Few-Shot Conditions
di: Wu, Tianxu, et al.
Pubblicazione: (2023)
di: Wu, Tianxu, et al.
Pubblicazione: (2023)
Fine-Grained Prototypes Distillation for Few-Shot Object Detection
di: Wang, Zichen, et al.
Pubblicazione: (2024)
di: Wang, Zichen, et al.
Pubblicazione: (2024)
PhysicsGen: Can Generative Models Learn from Images to Predict Complex Physical Relations?
di: Spitznagel, Martin, et al.
Pubblicazione: (2025)
di: Spitznagel, Martin, et al.
Pubblicazione: (2025)
Assessing Foundation Models for Mold Colony Detection with Limited Training Data
di: Pichler, Henrik, et al.
Pubblicazione: (2025)
di: Pichler, Henrik, et al.
Pubblicazione: (2025)
Is RobustBench/AutoAttack a suitable Benchmark for Adversarial Robustness?
di: Lorenz, Peter, et al.
Pubblicazione: (2021)
di: Lorenz, Peter, et al.
Pubblicazione: (2021)
Transfer Learning and Mixup for Fine-Grained Few-Shot Fungi Classification
di: Tam, Jason Kahei, et al.
Pubblicazione: (2025)
di: Tam, Jason Kahei, et al.
Pubblicazione: (2025)
Benchmarking Document Parsers on Mathematical Formula Extraction from PDFs
di: Horn, Pius, et al.
Pubblicazione: (2025)
di: Horn, Pius, et al.
Pubblicazione: (2025)
Can Biases in ImageNet Models Explain Generalization?
di: Gavrikov, Paul, et al.
Pubblicazione: (2024)
di: Gavrikov, Paul, et al.
Pubblicazione: (2024)
Urban Sound Propagation: a Benchmark for 1-Step Generative Modeling of Complex Physical Systems
di: Spitznagel, Martin, et al.
Pubblicazione: (2024)
di: Spitznagel, Martin, et al.
Pubblicazione: (2024)
Beyond String Matching: Semantic Evaluation of PDF Table Extraction
di: Horn, Pius, et al.
Pubblicazione: (2026)
di: Horn, Pius, et al.
Pubblicazione: (2026)
CoFi: A Fast Coarse-to-Fine Few-Shot Pipeline for Glomerular Basement Membrane Segmentation
di: Fang, Hongjin, et al.
Pubblicazione: (2025)
di: Fang, Hongjin, et al.
Pubblicazione: (2025)
Fix your downsampling ASAP! Be natively more robust via Aliasing and Spectral Artifact free Pooling
di: Grabinski, Julia, et al.
Pubblicazione: (2023)
di: Grabinski, Julia, et al.
Pubblicazione: (2023)
Towards Fine-Grained Vision-Language Alignment for Few-Shot Anomaly Detection
di: Fan, Yuanting, et al.
Pubblicazione: (2025)
di: Fan, Yuanting, et al.
Pubblicazione: (2025)
Few-Shot Learning Pipeline for Monkeypox Skin Disease Classification Using CNN Feature Extractors
di: Rashid, Md. Safirur, et al.
Pubblicazione: (2026)
di: Rashid, Md. Safirur, et al.
Pubblicazione: (2026)
Detecting AutoAttack Perturbations in the Frequency Domain
di: Lorenz, Peter, et al.
Pubblicazione: (2021)
di: Lorenz, Peter, et al.
Pubblicazione: (2021)
How Do Training Methods Influence the Utilization of Vision Models?
di: Gavrikov, Paul, et al.
Pubblicazione: (2024)
di: Gavrikov, Paul, et al.
Pubblicazione: (2024)
A streamlined Approach to Multimodal Few-Shot Class Incremental Learning for Fine-Grained Datasets
di: Doan, Thang, et al.
Pubblicazione: (2024)
di: Doan, Thang, et al.
Pubblicazione: (2024)
Shallow Deep Learning Can Still Excel in Fine-Grained Few-Shot Learning
di: Qi, Chaofei, et al.
Pubblicazione: (2025)
di: Qi, Chaofei, et al.
Pubblicazione: (2025)
Real-time Prediction of Urban Sound Propagation with Conditioned Normalizing Flows
di: Eckerle, Achim, et al.
Pubblicazione: (2025)
di: Eckerle, Achim, et al.
Pubblicazione: (2025)
Adversarial Examples are Misaligned in Diffusion Model Manifolds
di: Lorenz, Peter, et al.
Pubblicazione: (2024)
di: Lorenz, Peter, et al.
Pubblicazione: (2024)
Top-GAP: Integrating Size Priors in CNNs for more Interpretability, Robustness, and Bias Mitigation
di: Nieradzik, Lars, et al.
Pubblicazione: (2024)
di: Nieradzik, Lars, et al.
Pubblicazione: (2024)
Ambiguous Annotations: When is a Pedestrian not a Pedestrian?
di: Schwirten, Luisa, et al.
Pubblicazione: (2024)
di: Schwirten, Luisa, et al.
Pubblicazione: (2024)
Exploration of Class Center for Fine-Grained Visual Classification
di: Yao, Hang, et al.
Pubblicazione: (2024)
di: Yao, Hang, et al.
Pubblicazione: (2024)
Object-Centric Cropping for Visual Few-Shot Classification
di: Abdali, Aymane, et al.
Pubblicazione: (2025)
di: Abdali, Aymane, et al.
Pubblicazione: (2025)
Beware of Aliases -- Signal Preservation is Crucial for Robust Image Restoration
di: Agnihotri, Shashank, et al.
Pubblicazione: (2024)
di: Agnihotri, Shashank, et al.
Pubblicazione: (2024)
Zero-Shot Prompting and Few-Shot Fine-Tuning: Revisiting Document Image Classification Using Large Language Models
di: Scius-Bertrand, Anna, et al.
Pubblicazione: (2024)
di: Scius-Bertrand, Anna, et al.
Pubblicazione: (2024)
Foundation Models For Seismic Data Processing: An Extensive Review
di: Fuchs, Fabian, et al.
Pubblicazione: (2025)
di: Fuchs, Fabian, et al.
Pubblicazione: (2025)
Reliable Evaluation of Attribution Maps in CNNs: A Perturbation-Based Approach
di: Nieradzik, Lars, et al.
Pubblicazione: (2024)
di: Nieradzik, Lars, et al.
Pubblicazione: (2024)
FiLo++: Zero-/Few-Shot Anomaly Detection by Fused Fine-Grained Descriptions and Deformable Localization
di: Gu, Zhaopeng, et al.
Pubblicazione: (2025)
di: Gu, Zhaopeng, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Can Visual Language Models Replace OCR-Based Visual Question Answering Pipelines in Production? A Case Study in Retail
di: Lamm, Bianca, et al.
Pubblicazione: (2024) -
Retail-786k: a Large-Scale Dataset for Visual Entity Matching
di: Lamm, Bianca, et al.
Pubblicazione: (2023) -
Deepfakes: we need to re-think the concept of "real" images
di: Keuper, Janis, et al.
Pubblicazione: (2025) -
As large as it gets: Learning infinitely large Filters via Neural Implicit Functions in the Fourier Domain
di: Grabinski, Julia, et al.
Pubblicazione: (2023) -
A New Kind of Network? Review and Reference Implementation of Neural Cellular Automata
di: Spitznagel, Martin, et al.
Pubblicazione: (2026)