Retail-786k: a Large-Scale Dataset for Visual Entity Matching
Fuente:
arXiv
Saved in:
| Main Authors: | Lamm, Bianca, Keuper, Janis |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Can Visual Language Models Replace OCR-Based Visual Question Answering Pipelines in Production? A Case Study in Retail
by: Lamm, Bianca, et al.
Published: (2024)
by: Lamm, Bianca, et al.
Published: (2024)
A Visual RAG Pipeline for Few-Shot Fine-Grained Product Classification
by: Lamm, Bianca, et al.
Published: (2025)
by: Lamm, Bianca, et al.
Published: (2025)
Deepfakes: we need to re-think the concept of "real" images
by: Keuper, Janis, et al.
Published: (2025)
by: Keuper, Janis, et al.
Published: (2025)
Beyond String Matching: Semantic Evaluation of PDF Table Extraction
by: Horn, Pius, et al.
Published: (2026)
by: Horn, Pius, et al.
Published: (2026)
As large as it gets: Learning infinitely large Filters via Neural Implicit Functions in the Fourier Domain
by: Grabinski, Julia, et al.
Published: (2023)
by: Grabinski, Julia, et al.
Published: (2023)
A New Kind of Network? Review and Reference Implementation of Neural Cellular Automata
by: Spitznagel, Martin, et al.
Published: (2026)
by: Spitznagel, Martin, et al.
Published: (2026)
Unfolding Local Growth Rate Estimates for (Almost) Perfect Adversarial Detection
by: Lorenz, Peter, et al.
Published: (2022)
by: Lorenz, Peter, et al.
Published: (2022)
Is RobustBench/AutoAttack a suitable Benchmark for Adversarial Robustness?
by: Lorenz, Peter, et al.
Published: (2021)
by: Lorenz, Peter, et al.
Published: (2021)
PhysicsGen: Can Generative Models Learn from Images to Predict Complex Physical Relations?
by: Spitznagel, Martin, et al.
Published: (2025)
by: Spitznagel, Martin, et al.
Published: (2025)
Assessing Foundation Models for Mold Colony Detection with Limited Training Data
by: Pichler, Henrik, et al.
Published: (2025)
by: Pichler, Henrik, et al.
Published: (2025)
Urban Sound Propagation: a Benchmark for 1-Step Generative Modeling of Complex Physical Systems
by: Spitznagel, Martin, et al.
Published: (2024)
by: Spitznagel, Martin, et al.
Published: (2024)
Benchmarking Document Parsers on Mathematical Formula Extraction from PDFs
by: Horn, Pius, et al.
Published: (2025)
by: Horn, Pius, et al.
Published: (2025)
Can Biases in ImageNet Models Explain Generalization?
by: Gavrikov, Paul, et al.
Published: (2024)
by: Gavrikov, Paul, et al.
Published: (2024)
Ambiguous Annotations: When is a Pedestrian not a Pedestrian?
by: Schwirten, Luisa, et al.
Published: (2024)
by: Schwirten, Luisa, et al.
Published: (2024)
Fix your downsampling ASAP! Be natively more robust via Aliasing and Spectral Artifact free Pooling
by: Grabinski, Julia, et al.
Published: (2023)
by: Grabinski, Julia, et al.
Published: (2023)
Detecting AutoAttack Perturbations in the Frequency Domain
by: Lorenz, Peter, et al.
Published: (2021)
by: Lorenz, Peter, et al.
Published: (2021)
How Do Training Methods Influence the Utilization of Vision Models?
by: Gavrikov, Paul, et al.
Published: (2024)
by: Gavrikov, Paul, et al.
Published: (2024)
Adversarial Examples are Misaligned in Diffusion Model Manifolds
by: Lorenz, Peter, et al.
Published: (2024)
by: Lorenz, Peter, et al.
Published: (2024)
Top-GAP: Integrating Size Priors in CNNs for more Interpretability, Robustness, and Bias Mitigation
by: Nieradzik, Lars, et al.
Published: (2024)
by: Nieradzik, Lars, et al.
Published: (2024)
Real-time Prediction of Urban Sound Propagation with Conditioned Normalizing Flows
by: Eckerle, Achim, et al.
Published: (2025)
by: Eckerle, Achim, et al.
Published: (2025)
Beware of Aliases -- Signal Preservation is Crucial for Robust Image Restoration
by: Agnihotri, Shashank, et al.
Published: (2024)
by: Agnihotri, Shashank, et al.
Published: (2024)
Fake or JPEG? Revealing Common Biases in Generated Image Detection Datasets
by: Grommelt, Patrick, et al.
Published: (2024)
by: Grommelt, Patrick, et al.
Published: (2024)
Foundation Models For Seismic Data Processing: An Extensive Review
by: Fuchs, Fabian, et al.
Published: (2025)
by: Fuchs, Fabian, et al.
Published: (2025)
Reliable Evaluation of Attribution Maps in CNNs: A Perturbation-Based Approach
by: Nieradzik, Lars, et al.
Published: (2024)
by: Nieradzik, Lars, et al.
Published: (2024)
In-Context Learning for Seismic Data Processing
by: Fuchs, Fabian, et al.
Published: (2025)
by: Fuchs, Fabian, et al.
Published: (2025)
Teddy: Efficient Large-Scale Dataset Distillation via Taylor-Approximated Matching
by: Yu, Ruonan, et al.
Published: (2024)
by: Yu, Ruonan, et al.
Published: (2024)
A Generative Approach for Wikipedia-Scale Visual Entity Recognition
by: Caron, Mathilde, et al.
Published: (2024)
by: Caron, Mathilde, et al.
Published: (2024)
Entity6K: A Large Open-Domain Evaluation Dataset for Real-World Entity Recognition
by: Qiu, Jielin, et al.
Published: (2024)
by: Qiu, Jielin, et al.
Published: (2024)
I Spy With My Little Eye: A Minimum Cost Multicut Investigation of Dataset Frames
by: Prasse, Katharina, et al.
Published: (2024)
by: Prasse, Katharina, et al.
Published: (2024)
Web-Scale Visual Entity Recognition: An LLM-Driven Data Approach
by: Caron, Mathilde, et al.
Published: (2024)
by: Caron, Mathilde, et al.
Published: (2024)
Understanding Bias in Large-Scale Visual Datasets
by: Zeng, Boya, et al.
Published: (2024)
by: Zeng, Boya, et al.
Published: (2024)
DCBM: Data-Efficient Visual Concept Bottleneck Models
by: Prasse, Katharina, et al.
Published: (2024)
by: Prasse, Katharina, et al.
Published: (2024)
EntityCLIP: Entity-Centric Image-Text Matching via Multimodal Attentive Contrastive Learning
by: Wang, Yaxiong, et al.
Published: (2024)
by: Wang, Yaxiong, et al.
Published: (2024)
Compositional Image-Text Matching and Retrieval by Grounding Entities
by: Vongala, Madhukar Reddy, et al.
Published: (2025)
by: Vongala, Madhukar Reddy, et al.
Published: (2025)
TABLET: A Large-Scale Dataset for Robust Visual Table Understanding
by: Alonso, Iñigo, et al.
Published: (2025)
by: Alonso, Iñigo, et al.
Published: (2025)
Car-1000: A New Large Scale Fine-Grained Visual Categorization Dataset
by: Hu, Yutao, et al.
Published: (2025)
by: Hu, Yutao, et al.
Published: (2025)
Contrastive Learning-Enhanced Trajectory Matching for Small-Scale Dataset Distillation
by: Li, Wenmin, et al.
Published: (2025)
by: Li, Wenmin, et al.
Published: (2025)
Challenging the Black Box: A Comprehensive Evaluation of Attribution Maps of CNN Applications in Agriculture and Forestry
by: Nieradzik, Lars, et al.
Published: (2024)
by: Nieradzik, Lars, et al.
Published: (2024)
PM25Vision: A Large-Scale Benchmark Dataset for Visual Estimation of Air Quality
by: Han, Yang
Published: (2025)
by: Han, Yang
Published: (2025)
UWStereo: A Large Synthetic Dataset for Underwater Stereo Matching
by: Lv, Qingxuan, et al.
Published: (2024)
by: Lv, Qingxuan, et al.
Published: (2024)
Similar Items
-
Can Visual Language Models Replace OCR-Based Visual Question Answering Pipelines in Production? A Case Study in Retail
by: Lamm, Bianca, et al.
Published: (2024) -
A Visual RAG Pipeline for Few-Shot Fine-Grained Product Classification
by: Lamm, Bianca, et al.
Published: (2025) -
Deepfakes: we need to re-think the concept of "real" images
by: Keuper, Janis, et al.
Published: (2025) -
Beyond String Matching: Semantic Evaluation of PDF Table Extraction
by: Horn, Pius, et al.
Published: (2026) -
As large as it gets: Learning infinitely large Filters via Neural Implicit Functions in the Fourier Domain
by: Grabinski, Julia, et al.
Published: (2023)