Uncertainties of Latent Representations in Computer Vision
Fuente:
arXiv
Saved in:
| Main Author: | Kirchhof, Michael |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Probing the Latent World: Emergent Discrete Symbols and Physical Structure in Latent Representations
by: ming, Liu hung
Published: (2026)
by: ming, Liu hung
Published: (2026)
On the Generalization of Representation Uncertainty in Earth Observation
by: Kondylatos, Spyros, et al.
Published: (2025)
by: Kondylatos, Spyros, et al.
Published: (2025)
A Review of Latent Representation Models in Neuroimaging
by: Vázquez-García, C., et al.
Published: (2024)
by: Vázquez-García, C., et al.
Published: (2024)
VisMem: Latent Vision Memory Unlocks Potential of Vision-Language Models
by: Yu, Xinlei, et al.
Published: (2025)
by: Yu, Xinlei, et al.
Published: (2025)
MMCORE: MultiModal COnnection with Representation Aligned Latent Embeddings
by: Li, Zijie, et al.
Published: (2026)
by: Li, Zijie, et al.
Published: (2026)
i-MAE: Are Latent Representations in Masked Autoencoders Linearly Separable?
by: Zhang, Kevin, et al.
Published: (2022)
by: Zhang, Kevin, et al.
Published: (2022)
Temporal Embeddings: Scalable Self-Supervised Temporal Representation Learning from Spatiotemporal Data for Multimodal Computer Vision
by: Cao, Yi, et al.
Published: (2023)
by: Cao, Yi, et al.
Published: (2023)
Vision-Language Models Unlock Task-Centric Latent Actions
by: Nikulin, Alexander, et al.
Published: (2026)
by: Nikulin, Alexander, et al.
Published: (2026)
Unsupervised Dynamic Feature Selection for Robust Latent Spaces in Vision Tasks
by: Corcuera, Bruno, et al.
Published: (2025)
by: Corcuera, Bruno, et al.
Published: (2025)
Pretrained Visual Uncertainties
by: Kirchhof, Michael, et al.
Published: (2024)
by: Kirchhof, Michael, et al.
Published: (2024)
Disentangling Disentangled Representations: Towards Improved Latent Units via Diffusion Models
by: Jun, Youngjun, et al.
Published: (2024)
by: Jun, Youngjun, et al.
Published: (2024)
Uncertainty Quantification via Hölder Divergence for Multi-View Representation Learning
by: Zhang, Yan, et al.
Published: (2024)
by: Zhang, Yan, et al.
Published: (2024)
When Multi-Task Learning Meets Partial Supervision: A Computer Vision Review
by: Fontana, Maxime, et al.
Published: (2023)
by: Fontana, Maxime, et al.
Published: (2023)
Enhancing Vision-Language Model Reliability with Uncertainty-Guided Dropout Decoding
by: Fang, Yixiong, et al.
Published: (2024)
by: Fang, Yixiong, et al.
Published: (2024)
Gaze-Informed Vision Transformers: Predicting Driving Decisions Under Uncertainty
by: Koorathota, Sharath, et al.
Published: (2023)
by: Koorathota, Sharath, et al.
Published: (2023)
Latent Zoning Network: A Unified Principle for Generative Modeling, Representation Learning, and Classification
by: Lin, Zinan, et al.
Published: (2025)
by: Lin, Zinan, et al.
Published: (2025)
Vision-Language Models Provide Promptable Representations for Reinforcement Learning
by: Chen, William, et al.
Published: (2024)
by: Chen, William, et al.
Published: (2024)
Hypersolid: Emergent Vision Representations via Short-Range Repulsion
by: Rodríguez-Betancourt, Esteban, et al.
Published: (2026)
by: Rodríguez-Betancourt, Esteban, et al.
Published: (2026)
Steering Sparse Autoencoder Latents to Control Dynamic Head Pruning in Vision Transformers (Student Abstract)
by: Lee, Yousung, et al.
Published: (2026)
by: Lee, Yousung, et al.
Published: (2026)
Test-Time Spectrum-Aware Latent Steering for Zero-Shot Generalization in Vision-Language Models
by: Dafnis, Konstantinos M., et al.
Published: (2025)
by: Dafnis, Konstantinos M., et al.
Published: (2025)
Computer Vision Approaches for Automated Bee Counting Application
by: Bilik, Simon, et al.
Published: (2024)
by: Bilik, Simon, et al.
Published: (2024)
SILO: Solving Inverse Problems with Latent Operators
by: Raphaeli, Ron, et al.
Published: (2025)
by: Raphaeli, Ron, et al.
Published: (2025)
Time Series Representations for Classification Lie Hidden in Pretrained Vision Transformers
by: Roschmann, Simon, et al.
Published: (2025)
by: Roschmann, Simon, et al.
Published: (2025)
Linking Robustness and Generalization: A k* Distribution Analysis of Concept Clustering in Latent Space for Vision Models
by: Kotyan, Shashank, et al.
Published: (2024)
by: Kotyan, Shashank, et al.
Published: (2024)
Modelling and Simulation of Neuromorphic Datasets for Anomaly Detection in Computer Vision
by: Middleton, Mike, et al.
Published: (2026)
by: Middleton, Mike, et al.
Published: (2026)
Driver Activity Classification Using Generalizable Representations from Vision-Language Models
by: Greer, Ross, et al.
Published: (2024)
by: Greer, Ross, et al.
Published: (2024)
Uncertainty-Informed Volume Visualization using Implicit Neural Representation
by: Saklani, Shanu, et al.
Published: (2024)
by: Saklani, Shanu, et al.
Published: (2024)
SynthVision -- Harnessing Minimal Input for Maximal Output in Computer Vision Models using Synthetic Image data
by: Kularathne, Yudara, et al.
Published: (2024)
by: Kularathne, Yudara, et al.
Published: (2024)
Reducing Hallucinations in Vision-Language Models via Latent Space Steering
by: Liu, Sheng, et al.
Published: (2024)
by: Liu, Sheng, et al.
Published: (2024)
Unified Supervision For Vision-Language Modeling in 3D Computed Tomography
by: Lee, Hao-Chih, et al.
Published: (2025)
by: Lee, Hao-Chih, et al.
Published: (2025)
On Background Bias of Post-Hoc Concept Embeddings in Computer Vision DNNs
by: Schwalbe, Gesina, et al.
Published: (2025)
by: Schwalbe, Gesina, et al.
Published: (2025)
Decipher-MR: A Vision-Language Foundation Model for 3D MRI Representations
by: Yang, Zhijian, et al.
Published: (2025)
by: Yang, Zhijian, et al.
Published: (2025)
Compositional Discrete Latent Code for High Fidelity, Productive Diffusion Models
by: Lavoie, Samuel, et al.
Published: (2025)
by: Lavoie, Samuel, et al.
Published: (2025)
How Well Does GPT-4o Understand Vision? Evaluating Multimodal Foundation Models on Standard Computer Vision Tasks
by: Ramachandran, Rahul, et al.
Published: (2025)
by: Ramachandran, Rahul, et al.
Published: (2025)
FM-G-CAM: A Holistic Approach for Explainable AI in Computer Vision
by: Silva, Ravidu Suien Rammuni, et al.
Published: (2023)
by: Silva, Ravidu Suien Rammuni, et al.
Published: (2023)
Caregiver Talk Shapes Toddler Vision: A Computational Study of Dyadic Play
by: Schaumlöffel, Timothy, et al.
Published: (2023)
by: Schaumlöffel, Timothy, et al.
Published: (2023)
Back Home: A Computer Vision Solution to Seashell Identification for Ecological Restoration
by: Valverde, Alexander, et al.
Published: (2025)
by: Valverde, Alexander, et al.
Published: (2025)
Multimodal AI for Body Fat Estimation: Computer Vision and Anthropometry with DEXA Benchmarks
by: Aldajani, Rayan
Published: (2025)
by: Aldajani, Rayan
Published: (2025)
LucidPPN: Unambiguous Prototypical Parts Network for User-centric Interpretable Computer Vision
by: Pach, Mateusz, et al.
Published: (2024)
by: Pach, Mateusz, et al.
Published: (2024)
DetoxAI: a Python Toolkit for Debiasing Deep Learning Models in Computer Vision
by: Stępka, Ignacy, et al.
Published: (2025)
by: Stępka, Ignacy, et al.
Published: (2025)
Similar Items
-
Probing the Latent World: Emergent Discrete Symbols and Physical Structure in Latent Representations
by: ming, Liu hung
Published: (2026) -
On the Generalization of Representation Uncertainty in Earth Observation
by: Kondylatos, Spyros, et al.
Published: (2025) -
A Review of Latent Representation Models in Neuroimaging
by: Vázquez-García, C., et al.
Published: (2024) -
VisMem: Latent Vision Memory Unlocks Potential of Vision-Language Models
by: Yu, Xinlei, et al.
Published: (2025) -
MMCORE: MultiModal COnnection with Representation Aligned Latent Embeddings
by: Li, Zijie, et al.
Published: (2026)