The devil is in the fine-grained details: Evaluating open-vocabulary object detectors for fine-grained understanding
Fuente:
arXiv
Saved in:
| Main Authors: | Bianchi, Lorenzo, Carrara, Fabio, Messina, Nicola, Gennaro, Claudio, Falchi, Fabrizio |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Is CLIP the main roadblock for fine-grained open-world perception?
by: Bianchi, Lorenzo, et al.
Published: (2024)
by: Bianchi, Lorenzo, et al.
Published: (2024)
Talking to DINO: Bridging Self-Supervised Vision Backbones with Language for Open-Vocabulary Segmentation
by: Barsellotti, Luca, et al.
Published: (2024)
by: Barsellotti, Luca, et al.
Published: (2024)
Learning Egocentric In-Hand Object Segmentation through Weak Supervision from Human Narrations
by: Messina, Nicola, et al.
Published: (2025)
by: Messina, Nicola, et al.
Published: (2025)
One Patch to Caption Them All: A Unified Zero-Shot Captioning Framework
by: Bianchi, Lorenzo, et al.
Published: (2025)
by: Bianchi, Lorenzo, et al.
Published: (2025)
Improving fine-grained understanding in image-text pre-training
by: Bica, Ioana, et al.
Published: (2024)
by: Bica, Ioana, et al.
Published: (2024)
Adversarial Magnification to Deceive Deepfake Detection through Super Resolution
by: Coccomini, Davide Alessandro, et al.
Published: (2024)
by: Coccomini, Davide Alessandro, et al.
Published: (2024)
Deepfake Detection without Deepfakes: Generalization via Synthetic Frequency Patterns Injection
by: Coccomini, Davide Alessandro, et al.
Published: (2024)
by: Coccomini, Davide Alessandro, et al.
Published: (2024)
Eliminating the Language Bias for Visual Question Answering with fine-grained Causal Intervention
by: Liu, Ying, et al.
Published: (2024)
by: Liu, Ying, et al.
Published: (2024)
Performance of computer vision algorithms for fine-grained classification using crowdsourced insect images
by: Pucci, Rita, et al.
Published: (2024)
by: Pucci, Rita, et al.
Published: (2024)
ViSketch-GPT: Collaborative Multi-Scale Feature Extraction for Sketch Recognition and Generation
by: Federico, Giulio, et al.
Published: (2025)
by: Federico, Giulio, et al.
Published: (2025)
Demographic-aware fine-grained visual recognition of pediatric wrist pathologies
by: Ahmed, Ammar, et al.
Published: (2025)
by: Ahmed, Ammar, et al.
Published: (2025)
GCAM: Gaussian and causal-attention model of food fine-grained recognition
by: Zhuang, Guohang, et al.
Published: (2024)
by: Zhuang, Guohang, et al.
Published: (2024)
FiCo-ITR: bridging fine-grained and coarse-grained image-text retrieval for comparative performance analysis
by: Williams-Lekuona, Mikel, et al.
Published: (2024)
by: Williams-Lekuona, Mikel, et al.
Published: (2024)
SEPS: Semantic-enhanced Patch Slimming Framework for fine-grained cross-modal alignment
by: Mao, Xinyu, et al.
Published: (2025)
by: Mao, Xinyu, et al.
Published: (2025)
CountingDINO: A Training-free Pipeline for Class-Agnostic Counting using Unsupervised Backbones
by: Pacini, Giacomo, et al.
Published: (2025)
by: Pacini, Giacomo, et al.
Published: (2025)
Adaptive receptive field-based spatial-frequency feature reconstruction network for few-shot fine-grained image classification
by: Zhang, Linyue, et al.
Published: (2026)
by: Zhang, Linyue, et al.
Published: (2026)
Large Language Models estimate fine-grained human color-concept associations
by: Mukherjee, Kushin, et al.
Published: (2024)
by: Mukherjee, Kushin, et al.
Published: (2024)
CA3D: Convolutional-Attentional 3D Nets for Efficient Video Activity Recognition on the Edge
by: Lagani, Gabriele, et al.
Published: (2025)
by: Lagani, Gabriele, et al.
Published: (2025)
Specificity-aware reinforcement learning for fine-grained open-world classification
by: Angheben, Samuele, et al.
Published: (2026)
by: Angheben, Samuele, et al.
Published: (2026)
Color histogram equalization and fine-tuning to improve expression recognition of (partially occluded) faces on sign language datasets
by: Nunnari, Fabrizio, et al.
Published: (2025)
by: Nunnari, Fabrizio, et al.
Published: (2025)
Joint-Dataset Learning and Cross-Consistent Regularization for Text-to-Motion Retrieval
by: Messina, Nicola, et al.
Published: (2024)
by: Messina, Nicola, et al.
Published: (2024)
Comparing fine-grained and coarse-grained object detection for ecology
by: Tam, Jess, et al.
Published: (2024)
by: Tam, Jess, et al.
Published: (2024)
Does it Really Count? Assessing Semantic Grounding in Text-Guided Class-Agnostic Counting
by: Pacini, Giacomo, et al.
Published: (2026)
by: Pacini, Giacomo, et al.
Published: (2026)
FingER: Content Aware Fine-grained Evaluation with Reasoning for AI-Generated Videos
by: Chen, Rui, et al.
Published: (2025)
by: Chen, Rui, et al.
Published: (2025)
Maybe you are looking for CroQS: Cross-modal Query Suggestion for Text-to-Image Retrieval
by: Pacini, Giacomo, et al.
Published: (2024)
by: Pacini, Giacomo, et al.
Published: (2024)
Enhancing targeted transferability via feature space fine-tuning
by: Zeng, Hui, et al.
Published: (2024)
by: Zeng, Hui, et al.
Published: (2024)
Reviving ConvNeXt for Efficient Convolutional Diffusion Models
by: Kwon, Taesung, et al.
Published: (2026)
by: Kwon, Taesung, et al.
Published: (2026)
FORGE: Fine-grained Multimodal Evaluation for Manufacturing Scenarios
by: Jian, Xiangru, et al.
Published: (2026)
by: Jian, Xiangru, et al.
Published: (2026)
Gamified crowd-sourcing of high-quality data for visual fine-tuning
by: Yadav, Shashank, et al.
Published: (2024)
by: Yadav, Shashank, et al.
Published: (2024)
How Texts Help? A Fine-grained Evaluation to Reveal the Role of Language in Vision-Language Tracking
by: Li, Xuchen, et al.
Published: (2024)
by: Li, Xuchen, et al.
Published: (2024)
Pairwise Matching of Intermediate Representations for Fine-grained Explainability
by: Shrack, Lauren, et al.
Published: (2025)
by: Shrack, Lauren, et al.
Published: (2025)
Mind the Prompt: A Novel Benchmark for Prompt-based Class-Agnostic Counting
by: Ciampi, Luca, et al.
Published: (2024)
by: Ciampi, Luca, et al.
Published: (2024)
Fine-grained Background Representation for Weakly Supervised Semantic Segmentation
by: Yin, Xu, et al.
Published: (2024)
by: Yin, Xu, et al.
Published: (2024)
FLAIR: VLM with Fine-grained Language-informed Image Representations
by: Xiao, Rui, et al.
Published: (2024)
by: Xiao, Rui, et al.
Published: (2024)
FINER: MLLMs Hallucinate under Fine-grained Negative Queries
by: Xiao, Rui, et al.
Published: (2026)
by: Xiao, Rui, et al.
Published: (2026)
Infusing fine-grained visual knowledge to Vision-Language Models
by: Ypsilantis, Nikolaos-Antonios, et al.
Published: (2025)
by: Ypsilantis, Nikolaos-Antonios, et al.
Published: (2025)
Uncertainty modeling for fine-tuned implicit functions
by: Susmelj, Anna, et al.
Published: (2024)
by: Susmelj, Anna, et al.
Published: (2024)
Navigating the Nuances: A Fine-grained Evaluation of Vision-Language Navigation
by: Wang, Zehao, et al.
Published: (2024)
by: Wang, Zehao, et al.
Published: (2024)
MedFILIP: Medical Fine-grained Language-Image Pre-training
by: Liang, Xinjie, et al.
Published: (2025)
by: Liang, Xinjie, et al.
Published: (2025)
Respect the model: Fine-grained and Robust Explanation with Sharing Ratio Decomposition
by: Han, Sangyu, et al.
Published: (2024)
by: Han, Sangyu, et al.
Published: (2024)
Similar Items
-
Is CLIP the main roadblock for fine-grained open-world perception?
by: Bianchi, Lorenzo, et al.
Published: (2024) -
Talking to DINO: Bridging Self-Supervised Vision Backbones with Language for Open-Vocabulary Segmentation
by: Barsellotti, Luca, et al.
Published: (2024) -
Learning Egocentric In-Hand Object Segmentation through Weak Supervision from Human Narrations
by: Messina, Nicola, et al.
Published: (2025) -
One Patch to Caption Them All: A Unified Zero-Shot Captioning Framework
by: Bianchi, Lorenzo, et al.
Published: (2025) -
Improving fine-grained understanding in image-text pre-training
by: Bica, Ioana, et al.
Published: (2024)