What could go wrong? Discovering and describing failure modes in computer vision
Fuente:
arXiv
Salvato in:
| Autori principali: | Csurka, Gabriela, Hayes, Tyler L., Larlus, Diane, Volpi, Riccardo |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
SHiNe: Semantic Hierarchy Nexus for Open-vocabulary Object Detection
di: Liu, Mingxuan, et al.
Pubblicazione: (2024)
di: Liu, Mingxuan, et al.
Pubblicazione: (2024)
Test-time Vocabulary Adaptation for Language-driven Object Detection
di: Liu, Mingxuan, et al.
Pubblicazione: (2025)
di: Liu, Mingxuan, et al.
Pubblicazione: (2025)
PANDAS: Prototype-based Novel Class Discovery and Detection
di: Hayes, Tyler L., et al.
Pubblicazione: (2024)
di: Hayes, Tyler L., et al.
Pubblicazione: (2024)
Weatherproofing Retrieval for Localization with Generative AI and Geometric Consistency
di: Kalantidis, Yannis, et al.
Pubblicazione: (2024)
di: Kalantidis, Yannis, et al.
Pubblicazione: (2024)
Task Alignment: A simple and effective proxy for model merging in computer vision
di: de Jorge, Pau, et al.
Pubblicazione: (2026)
di: de Jorge, Pau, et al.
Pubblicazione: (2026)
Gaussian Splatting Feature Fields for Privacy-Preserving Visual Localization
di: Pietrantoni, Maxime, et al.
Pubblicazione: (2025)
di: Pietrantoni, Maxime, et al.
Pubblicazione: (2025)
On Good Practices for Task-Specific Distillation of Large Pretrained Visual Models
di: Marrie, Juliette, et al.
Pubblicazione: (2024)
di: Marrie, Juliette, et al.
Pubblicazione: (2024)
Layered Motion Fusion: Lifting Motion Segmentation to 3D in Egocentric Videos
di: Tschernezki, Vadim, et al.
Pubblicazione: (2025)
di: Tschernezki, Vadim, et al.
Pubblicazione: (2025)
Self-supervised Learning of Neural Implicit Feature Fields for Camera Pose Refinement
di: Pietrantoni, Maxime, et al.
Pubblicazione: (2024)
di: Pietrantoni, Maxime, et al.
Pubblicazione: (2024)
Can we make NeRF-based visual localization privacy-preserving?
di: Pietrantoni, Maxime, et al.
Pubblicazione: (2025)
di: Pietrantoni, Maxime, et al.
Pubblicazione: (2025)
LUDVIG: Learning-Free Uplifting of 2D Visual Features to Gaussian Splatting Scenes
di: Marrie, Juliette, et al.
Pubblicazione: (2024)
di: Marrie, Juliette, et al.
Pubblicazione: (2024)
Mesh4D: 4D Mesh Reconstruction and Tracking from Monocular Video
di: Jiang, Zeren, et al.
Pubblicazione: (2026)
di: Jiang, Zeren, et al.
Pubblicazione: (2026)
UNIC: Universal Classification Models via Multi-teacher Distillation
di: Sariyildiz, Mert Bulent, et al.
Pubblicazione: (2024)
di: Sariyildiz, Mert Bulent, et al.
Pubblicazione: (2024)
Geo4D: Leveraging Video Generators for Geometric 4D Scene Reconstruction
di: Jiang, Zeren, et al.
Pubblicazione: (2025)
di: Jiang, Zeren, et al.
Pubblicazione: (2025)
EPIC Fields: Marrying 3D Geometry and Video Understanding
di: Tschernezki, Vadim, et al.
Pubblicazione: (2023)
di: Tschernezki, Vadim, et al.
Pubblicazione: (2023)
PanSt3R: Multi-view Consistent Panoptic Segmentation
di: Zust, Lojze, et al.
Pubblicazione: (2025)
di: Zust, Lojze, et al.
Pubblicazione: (2025)
MUSt3R: Multi-view Network for Stereo 3D Reconstruction
di: Cabon, Yohann, et al.
Pubblicazione: (2025)
di: Cabon, Yohann, et al.
Pubblicazione: (2025)
DUNE: Distilling a Universal Encoder from Heterogeneous 2D and 3D Teachers
di: Sariyildiz, Mert Bulent, et al.
Pubblicazione: (2025)
di: Sariyildiz, Mert Bulent, et al.
Pubblicazione: (2025)
Placing Objects in Context via Inpainting for Out-of-distribution Segmentation
di: de Jorge, Pau, et al.
Pubblicazione: (2024)
di: de Jorge, Pau, et al.
Pubblicazione: (2024)
Syn4D: A Multiview Synthetic 4D Dataset
di: Jiang, Zeren, et al.
Pubblicazione: (2026)
di: Jiang, Zeren, et al.
Pubblicazione: (2026)
A survey of datasets for computer vision in agriculture
di: Heider, Nico, et al.
Pubblicazione: (2025)
di: Heider, Nico, et al.
Pubblicazione: (2025)
Anatomy of a failure: When, how, and why deep vision fails in scientific domains
di: Oh, Ji-Hun, et al.
Pubblicazione: (2026)
di: Oh, Ji-Hun, et al.
Pubblicazione: (2026)
Are foundation models for computer vision good conformal predictors?
di: Fillioux, Leo, et al.
Pubblicazione: (2024)
di: Fillioux, Leo, et al.
Pubblicazione: (2024)
Features extraction for image identification using computer vision
di: Niyonkuru, Venant, et al.
Pubblicazione: (2025)
di: Niyonkuru, Venant, et al.
Pubblicazione: (2025)
What You See is What You Classify: Black Box Attributions
di: Stalder, Steven, et al.
Pubblicazione: (2022)
di: Stalder, Steven, et al.
Pubblicazione: (2022)
The effects of using created synthetic images in computer vision training
di: Smutny, John W.
Pubblicazione: (2025)
di: Smutny, John W.
Pubblicazione: (2025)
Understanding and evaluating computer vision models through the lens of counterfactuals
di: Shukla, Pushkar
Pubblicazione: (2025)
di: Shukla, Pushkar
Pubblicazione: (2025)
What Do You See in Common? Learning Hierarchical Prototypes over Tree-of-Life to Discover Evolutionary Traits
di: Manogaran, Harish Babu, et al.
Pubblicazione: (2024)
di: Manogaran, Harish Babu, et al.
Pubblicazione: (2024)
Evaluating the Role of Training Data Origin for Country-Scale Cropland Mapping in Data-Scarce Regions: A Case Study of Nigeria
di: Gajardo, Joaquin, et al.
Pubblicazione: (2023)
di: Gajardo, Joaquin, et al.
Pubblicazione: (2023)
RANa: Retrieval-Augmented Navigation
di: Monaci, Gianluca, et al.
Pubblicazione: (2025)
di: Monaci, Gianluca, et al.
Pubblicazione: (2025)
For a semiotic AI: Bridging computer vision and visual semiotics for computational observation of large scale facial image archives
di: Morra, Lia, et al.
Pubblicazione: (2024)
di: Morra, Lia, et al.
Pubblicazione: (2024)
Ultra-low-light computer vision using trained photon correlations
di: Sohoni, Mandar M., et al.
Pubblicazione: (2026)
di: Sohoni, Mandar M., et al.
Pubblicazione: (2026)
Automated high-frequency quantification of fish communities and biomass using computer vision
di: Ishikawa, Kota, et al.
Pubblicazione: (2026)
di: Ishikawa, Kota, et al.
Pubblicazione: (2026)
Track Anything Annotate: Video annotation and dataset generation of computer vision models
di: Ivanov, Nikita, et al.
Pubblicazione: (2025)
di: Ivanov, Nikita, et al.
Pubblicazione: (2025)
Synthetic Image Verification in the Era of Generative AI: What Works and What Isn't There Yet
di: Tariang, Diangarti, et al.
Pubblicazione: (2024)
di: Tariang, Diangarti, et al.
Pubblicazione: (2024)
How far can we go with ImageNet for Text-to-Image generation?
di: Degeorge, L., et al.
Pubblicazione: (2025)
di: Degeorge, L., et al.
Pubblicazione: (2025)
A hierarchical semantic segmentation framework for computer vision-based bridge damage detection
di: Liu, Jingxiao, et al.
Pubblicazione: (2022)
di: Liu, Jingxiao, et al.
Pubblicazione: (2022)
The possibility of making \$138,000 from shredded banknote pieces using computer vision
di: Kong, Chung To
Pubblicazione: (2023)
di: Kong, Chung To
Pubblicazione: (2023)
Hebrew letters Detection and Cuneiform tablets Classification by using the yolov8 computer vision model
di: Saeed, Elaf A., et al.
Pubblicazione: (2024)
di: Saeed, Elaf A., et al.
Pubblicazione: (2024)
Estimating the distribution of numerosity and non-numerical visual magnitudes in natural scenes using computer vision
di: Hou, Kuinan, et al.
Pubblicazione: (2024)
di: Hou, Kuinan, et al.
Pubblicazione: (2024)
Documenti analoghi
-
SHiNe: Semantic Hierarchy Nexus for Open-vocabulary Object Detection
di: Liu, Mingxuan, et al.
Pubblicazione: (2024) -
Test-time Vocabulary Adaptation for Language-driven Object Detection
di: Liu, Mingxuan, et al.
Pubblicazione: (2025) -
PANDAS: Prototype-based Novel Class Discovery and Detection
di: Hayes, Tyler L., et al.
Pubblicazione: (2024) -
Weatherproofing Retrieval for Localization with Generative AI and Geometric Consistency
di: Kalantidis, Yannis, et al.
Pubblicazione: (2024) -
Task Alignment: A simple and effective proxy for model merging in computer vision
di: de Jorge, Pau, et al.
Pubblicazione: (2026)