A Survey on Class-Agnostic Counting: Advancements from Reference-Based to Open-World Text-Guided Approaches
Fuente:
arXiv
Saved in:
| Main Authors: | Ciampi, Luca, Azmoudeh, Ali, Akbaba, Elif Ecem, Sarıtaş, Erdi, Yazıcı, Ziya Ata, Ekenel, Hazım Kemal, Amato, Giuseppe, Falchi, Fabrizio |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
On Applicability of Synthetic Datasets for Facial Expression Recognition
by: Azmoudeh, Ali, et al.
Published: (2026)
by: Azmoudeh, Ali, et al.
Published: (2026)
Analyzing the Effect of Combined Degradations on Face Recognition
by: Sarıtaş, Erdi, et al.
Published: (2024)
by: Sarıtaş, Erdi, et al.
Published: (2024)
Analyzing the Feature Extractor Networks for Face Image Synthesis
by: Sarıtaş, Erdi, et al.
Published: (2024)
by: Sarıtaş, Erdi, et al.
Published: (2024)
Attention-Enhanced Hybrid Feature Aggregation Network for 3D Brain Tumor Segmentation
by: Yazıcı, Ziya Ata, et al.
Published: (2024)
by: Yazıcı, Ziya Ata, et al.
Published: (2024)
GLIMS: Attention-Guided Lightweight Multi-Scale Hybrid Network for Volumetric Semantic Segmentation
by: Yazıcı, Ziya Ata, et al.
Published: (2024)
by: Yazıcı, Ziya Ata, et al.
Published: (2024)
In-Bed Pose Estimation: A Review
by: Yazıcı, Ziya Ata, et al.
Published: (2024)
by: Yazıcı, Ziya Ata, et al.
Published: (2024)
Does it Really Count? Assessing Semantic Grounding in Text-Guided Class-Agnostic Counting
by: Pacini, Giacomo, et al.
Published: (2026)
by: Pacini, Giacomo, et al.
Published: (2026)
Impact of Face Alignment on Face Image Quality
by: Onaran, Eren, et al.
Published: (2024)
by: Onaran, Eren, et al.
Published: (2024)
CountingDINO: A Training-free Pipeline for Class-Agnostic Counting using Unsupervised Backbones
by: Pacini, Giacomo, et al.
Published: (2025)
by: Pacini, Giacomo, et al.
Published: (2025)
Mind the Prompt: A Novel Benchmark for Prompt-based Class-Agnostic Counting
by: Ciampi, Luca, et al.
Published: (2024)
by: Ciampi, Luca, et al.
Published: (2024)
Employing Vision-Language Models for Face Image Quality Assessment
by: Sarıtaş, Erdi, et al.
Published: (2026)
by: Sarıtaş, Erdi, et al.
Published: (2026)
Semi-Supervised Biomedical Image Segmentation via Diffusion Models and Teacher-Student Co-Training
by: Ciampi, Luca, et al.
Published: (2025)
by: Ciampi, Luca, et al.
Published: (2025)
Biologically-inspired Semi-supervised Semantic Segmentation for Biomedical Imaging
by: Ciampi, Luca, et al.
Published: (2024)
by: Ciampi, Luca, et al.
Published: (2024)
Facial Attribute Based Text Guided Face Anonymization
by: Muştu, Mustafa İzzet, et al.
Published: (2025)
by: Muştu, Mustafa İzzet, et al.
Published: (2025)
Yolo-Key-6D: Single Stage Monocular 6D Pose Estimation with Keypoint Enhancements
by: Çetiner, Kemal Alperen, et al.
Published: (2026)
by: Çetiner, Kemal Alperen, et al.
Published: (2026)
Impact of Surface Reflections in Maritime Obstacle Detection
by: Yalçın, Samed, et al.
Published: (2024)
by: Yalçın, Samed, et al.
Published: (2024)
Improved MambdaBDA Framework for Robust Building Damage Assessment Across Disaster Domains
by: Gençoğlu, Alp Eren, et al.
Published: (2026)
by: Gençoğlu, Alp Eren, et al.
Published: (2026)
Assessing the Use of Face Swapping Methods as Face Anonymizers in Videos
by: Muştu, Mustafa İzzet, et al.
Published: (2025)
by: Muştu, Mustafa İzzet, et al.
Published: (2025)
A Multimodal Depth-Aware Method For Embodied Reference Understanding
by: Eyiokur, Fevziye Irem, et al.
Published: (2025)
by: Eyiokur, Fevziye Irem, et al.
Published: (2025)
CAPE: A CLIP-Aware Pointing Ensemble of Complementary Heatmap Cues for Embodied Reference Understanding
by: Eyiokur, Fevziye Irem, et al.
Published: (2025)
by: Eyiokur, Fevziye Irem, et al.
Published: (2025)
Bias-Aware Face Mask Detection Dataset
by: Kantarcı, Alperen, et al.
Published: (2022)
by: Kantarcı, Alperen, et al.
Published: (2022)
Assessing Identity Leakage in Talking Face Generation: Metrics and Evaluation Framework
by: Yaman, Dogucan, et al.
Published: (2025)
by: Yaman, Dogucan, et al.
Published: (2025)
CA3D: Convolutional-Attentional 3D Nets for Efficient Video Activity Recognition on the Edge
by: Lagani, Gabriele, et al.
Published: (2025)
by: Lagani, Gabriele, et al.
Published: (2025)
Maybe you are looking for CroQS: Cross-modal Query Suggestion for Text-to-Image Retrieval
by: Pacini, Giacomo, et al.
Published: (2024)
by: Pacini, Giacomo, et al.
Published: (2024)
Audio-driven Talking Face Generation with Stabilized Synchronization Loss
by: Yaman, Dogucan, et al.
Published: (2023)
by: Yaman, Dogucan, et al.
Published: (2023)
Mask-Free Audio-driven Talking Face Generation for Enhanced Visual Quality and Identity Preservation
by: Yaman, Dogucan, et al.
Published: (2025)
by: Yaman, Dogucan, et al.
Published: (2025)
Adversarial Magnification to Deceive Deepfake Detection through Super Resolution
by: Coccomini, Davide Alessandro, et al.
Published: (2024)
by: Coccomini, Davide Alessandro, et al.
Published: (2024)
Exploring Strengths and Weaknesses of Super-Resolution Attack in Deepfake Detection
by: Coccomini, Davide Alessandro, et al.
Published: (2024)
by: Coccomini, Davide Alessandro, et al.
Published: (2024)
Audio-Visual Speech Representation Expert for Enhanced Talking Face Video Generation and Evaluation
by: Yaman, Dogucan, et al.
Published: (2024)
by: Yaman, Dogucan, et al.
Published: (2024)
Deepfake Detection without Deepfakes: Generalization via Synthetic Frequency Patterns Injection
by: Coccomini, Davide Alessandro, et al.
Published: (2024)
by: Coccomini, Davide Alessandro, et al.
Published: (2024)
Learning Egocentric In-Hand Object Segmentation through Weak Supervision from Human Narrations
by: Messina, Nicola, et al.
Published: (2025)
by: Messina, Nicola, et al.
Published: (2025)
Comparison of Different Deep Neural Network Models in the Cultural Heritage Domain
by: Boyadzhiev, Teodor, et al.
Published: (2025)
by: Boyadzhiev, Teodor, et al.
Published: (2025)
One Patch to Caption Them All: A Unified Zero-Shot Captioning Framework
by: Bianchi, Lorenzo, et al.
Published: (2025)
by: Bianchi, Lorenzo, et al.
Published: (2025)
FocalCount: Towards Class-Count Imbalance in Class-Agnostic Counting
by: Zhu, Huilin, et al.
Published: (2025)
by: Zhu, Huilin, et al.
Published: (2025)
Neuro-Inspired Visual Pattern Recognition via Biological Reservoir Computing
by: Ciampi, Luca, et al.
Published: (2026)
by: Ciampi, Luca, et al.
Published: (2026)
Health Perceptions and Risk of Metabolic Syndrome and Diabetes in Psychiatric Patients
by: Şenay Öztürk, et al.
Published: (2024)
by: Şenay Öztürk, et al.
Published: (2024)
Joint-Dataset Learning and Cross-Consistent Regularization for Text-to-Motion Retrieval
by: Messina, Nicola, et al.
Published: (2024)
by: Messina, Nicola, et al.
Published: (2024)
Enhanced Phishing Website Detection Through Effective Feature Selection With Time‐Varying Mirrored S‐Shaped Transfer Function
by: Fatih Kılıç, et al.
Published: (2025)
by: Fatih Kılıç, et al.
Published: (2025)
From Neural Activity to Computation: Biological Reservoirs for Pattern Recognition in Digit Classification
by: Iannello, Ludovico, et al.
Published: (2025)
by: Iannello, Ludovico, et al.
Published: (2025)
Astrophysical Parameters of 5056 Open Star Clusters from Bayesian Nested Sampling with PARSEC Isochrones
by: Plevne, Olcay, et al.
Published: (2026)
by: Plevne, Olcay, et al.
Published: (2026)
Similar Items
-
On Applicability of Synthetic Datasets for Facial Expression Recognition
by: Azmoudeh, Ali, et al.
Published: (2026) -
Analyzing the Effect of Combined Degradations on Face Recognition
by: Sarıtaş, Erdi, et al.
Published: (2024) -
Analyzing the Feature Extractor Networks for Face Image Synthesis
by: Sarıtaş, Erdi, et al.
Published: (2024) -
Attention-Enhanced Hybrid Feature Aggregation Network for 3D Brain Tumor Segmentation
by: Yazıcı, Ziya Ata, et al.
Published: (2024) -
GLIMS: Attention-Guided Lightweight Multi-Scale Hybrid Network for Volumetric Semantic Segmentation
by: Yazıcı, Ziya Ata, et al.
Published: (2024)