Estimating the distribution of numerosity and non-numerical visual magnitudes in natural scenes using computer vision
Fuente:
arXiv
Saved in:
| Main Authors: | Hou, Kuinan, Zorzi, Marco, Testolin, Alberto |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Visual Enumeration Remains Challenging for Multimodal Generative AI
by: Testolin, Alberto, et al.
Published: (2024)
by: Testolin, Alberto, et al.
Published: (2024)
Assessing the Visual Enumeration Abilities of Specialized Counting Architectures and Vision-Language Models
by: Hou, Kuinan, et al.
Published: (2025)
by: Hou, Kuinan, et al.
Published: (2025)
Sequential Enumeration in Large Language Models
by: Hou, Kuinan, et al.
Published: (2025)
by: Hou, Kuinan, et al.
Published: (2025)
A transition towards virtual representations of visual scenes
by: Pereira, Américo, et al.
Published: (2024)
by: Pereira, Américo, et al.
Published: (2024)
Indoor scene recognition from images under visual corruptions
by: Costa, Willams de Lima, et al.
Published: (2024)
by: Costa, Willams de Lima, et al.
Published: (2024)
For a semiotic AI: Bridging computer vision and visual semiotics for computational observation of large scale facial image archives
by: Morra, Lia, et al.
Published: (2024)
by: Morra, Lia, et al.
Published: (2024)
Do computer vision foundation models learn the low-level characteristics of the human visual system?
by: Cai, Yancheng, et al.
Published: (2025)
by: Cai, Yancheng, et al.
Published: (2025)
Motion-guided small MAV detection in complex and non-planar scenes
by: Guo, Hanqing, et al.
Published: (2024)
by: Guo, Hanqing, et al.
Published: (2024)
Features extraction for image identification using computer vision
by: Niyonkuru, Venant, et al.
Published: (2025)
by: Niyonkuru, Venant, et al.
Published: (2025)
3D scene generation from scene graphs and self-attention
by: Bonazzi, Pietro, et al.
Published: (2024)
by: Bonazzi, Pietro, et al.
Published: (2024)
RobustSurg: Tackling domain generalisation for out-of-distribution surgical scene segmentation
by: Ali, Mansoor, et al.
Published: (2025)
by: Ali, Mansoor, et al.
Published: (2025)
The effects of using created synthetic images in computer vision training
by: Smutny, John W.
Published: (2025)
by: Smutny, John W.
Published: (2025)
Predicting upcoming visual features during eye movements yields scene representations aligned with human visual cortex
by: Thorat, Sushrut, et al.
Published: (2025)
by: Thorat, Sushrut, et al.
Published: (2025)
Near, far: Patch-ordering enhances vision foundation models' scene understanding
by: Pariza, Valentinos, et al.
Published: (2024)
by: Pariza, Valentinos, et al.
Published: (2024)
Automated high-frequency quantification of fish communities and biomass using computer vision
by: Ishikawa, Kota, et al.
Published: (2026)
by: Ishikawa, Kota, et al.
Published: (2026)
Learning rigid-body simulators over implicit shapes for large-scale scenes and vision
by: Rubanova, Yulia, et al.
Published: (2024)
by: Rubanova, Yulia, et al.
Published: (2024)
The possibility of making \$138,000 from shredded banknote pieces using computer vision
by: Kong, Chung To
Published: (2023)
by: Kong, Chung To
Published: (2023)
A survey of datasets for computer vision in agriculture
by: Heider, Nico, et al.
Published: (2025)
by: Heider, Nico, et al.
Published: (2025)
Feature boosting with efficient attention for scene parsing
by: Singh, Vivek, et al.
Published: (2024)
by: Singh, Vivek, et al.
Published: (2024)
Functionality understanding and segmentation in 3D scenes
by: Corsetti, Jaime, et al.
Published: (2024)
by: Corsetti, Jaime, et al.
Published: (2024)
Ultra-low-light computer vision using trained photon correlations
by: Sohoni, Mandar M., et al.
Published: (2026)
by: Sohoni, Mandar M., et al.
Published: (2026)
Hebrew letters Detection and Cuneiform tablets Classification by using the yolov8 computer vision model
by: Saeed, Elaf A., et al.
Published: (2024)
by: Saeed, Elaf A., et al.
Published: (2024)
Site-specific weed management in corn using UAS imagery analysis and computer vision techniques
by: Sapkota, Ranjan, et al.
Published: (2022)
by: Sapkota, Ranjan, et al.
Published: (2022)
Are foundation models for computer vision good conformal predictors?
by: Fillioux, Leo, et al.
Published: (2024)
by: Fillioux, Leo, et al.
Published: (2024)
Enhancing seeding efficiency using a computer vision system to monitor furrow quality in real-time
by: Rai, Sidharth, et al.
Published: (2025)
by: Rai, Sidharth, et al.
Published: (2025)
Markerless Stride Length estimation in Athletic using Pose Estimation with monocular vision
by: Skorupski, Patryk, et al.
Published: (2025)
by: Skorupski, Patryk, et al.
Published: (2025)
Understanding and evaluating computer vision models through the lens of counterfactuals
by: Shukla, Pushkar
Published: (2025)
by: Shukla, Pushkar
Published: (2025)
A Versatile Framework for Multi-scene Person Re-identification
by: Zheng, Wei-Shi, et al.
Published: (2024)
by: Zheng, Wei-Shi, et al.
Published: (2024)
Efficient scene text image super-resolution with semantic guidance
by: TomyEnrique, LeoWu, et al.
Published: (2024)
by: TomyEnrique, LeoWu, et al.
Published: (2024)
Semantic UV mapping to improve texture inpainting for indoor scenes
by: Vermandere, Jelle, et al.
Published: (2024)
by: Vermandere, Jelle, et al.
Published: (2024)
ExploreGS: a vision-based low overhead framework for 3D scene reconstruction
by: Feng, Yunji, et al.
Published: (2025)
by: Feng, Yunji, et al.
Published: (2025)
Structured prototype regularization for synthetic-to-real driving scene parsing
by: Fan, Jiahe, et al.
Published: (2026)
by: Fan, Jiahe, et al.
Published: (2026)
Consistent text-to-image generation via scene de-contextualization
by: Tang, Song, et al.
Published: (2025)
by: Tang, Song, et al.
Published: (2025)
Open-vocabulary 3D scene perception in industrial environments
by: Moenck, Keno, et al.
Published: (2026)
by: Moenck, Keno, et al.
Published: (2026)
Teaching in adverse scenes: a statistically feedback-driven threshold and mask adjustment teacher-student framework for object detection in UAV images under adverse scenes
by: Chen, Hongyu, et al.
Published: (2025)
by: Chen, Hongyu, et al.
Published: (2025)
Implicit 3D scene reconstruction using deep learning towards efficient collision understanding in autonomous driving
by: Ramanayake, Akarshani, et al.
Published: (2025)
by: Ramanayake, Akarshani, et al.
Published: (2025)
VWise: A novel benchmark for evaluating scene classification for vehicular applications
by: Azevedo, Pedro, et al.
Published: (2024)
by: Azevedo, Pedro, et al.
Published: (2024)
Visual enhancement and 3D representation for underwater scenes: a review
by: Huang, Guoxi, et al.
Published: (2025)
by: Huang, Guoxi, et al.
Published: (2025)
Technical note: ShinyAnimalCV: open-source cloud-based web application for object detection, segmentation, and three-dimensional visualization of animals using computer vision
by: Wang, Jin, et al.
Published: (2023)
by: Wang, Jin, et al.
Published: (2023)
Object Depth and Size Estimation using Stereo-vision and Integration with SLAM
by: Hamad, Layth, et al.
Published: (2024)
by: Hamad, Layth, et al.
Published: (2024)
Similar Items
-
Visual Enumeration Remains Challenging for Multimodal Generative AI
by: Testolin, Alberto, et al.
Published: (2024) -
Assessing the Visual Enumeration Abilities of Specialized Counting Architectures and Vision-Language Models
by: Hou, Kuinan, et al.
Published: (2025) -
Sequential Enumeration in Large Language Models
by: Hou, Kuinan, et al.
Published: (2025) -
A transition towards virtual representations of visual scenes
by: Pereira, Américo, et al.
Published: (2024) -
Indoor scene recognition from images under visual corruptions
by: Costa, Willams de Lima, et al.
Published: (2024)