Post-hoc Probabilistic Vision-Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Baumann, Anton, Li, Rui, Klasson, Marcus, Mentu, Santeri, Karthik, Shyamgopal, Akata, Zeynep, Solin, Arno, Trapp, Martin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Flatness Improves Backbone Generalisation in Few-shot Classification
von: Li, Rui, et al.
Veröffentlicht: (2024)
von: Li, Rui, et al.
Veröffentlicht: (2024)
Sparse Autoencoders Learn Monosemantic Features in Vision-Language Models
von: Pach, Mateusz, et al.
Veröffentlicht: (2025)
von: Pach, Mateusz, et al.
Veröffentlicht: (2025)
Vision-by-Language for Training-Free Compositional Image Retrieval
von: Karthik, Shyamgopal, et al.
Veröffentlicht: (2023)
von: Karthik, Shyamgopal, et al.
Veröffentlicht: (2023)
Sources of Uncertainty in 3D Scene Reconstruction
von: Klasson, Marcus, et al.
Veröffentlicht: (2024)
von: Klasson, Marcus, et al.
Veröffentlicht: (2024)
Noise Hypernetworks: Amortizing Test-Time Compute in Diffusion Models
von: Eyring, Luca, et al.
Veröffentlicht: (2025)
von: Eyring, Luca, et al.
Veröffentlicht: (2025)
DeSplat: Decomposed Gaussian Splatting for Distractor-Free Rendering
von: Wang, Yihao, et al.
Veröffentlicht: (2024)
von: Wang, Yihao, et al.
Veröffentlicht: (2024)
EgoCVR: An Egocentric Benchmark for Fine-Grained Composed Video Retrieval
von: Hummel, Thomas, et al.
Veröffentlicht: (2024)
von: Hummel, Thomas, et al.
Veröffentlicht: (2024)
ReNO: Enhancing One-step Text-to-Image Models through Reward-based Noise Optimization
von: Eyring, Luca, et al.
Veröffentlicht: (2024)
von: Eyring, Luca, et al.
Veröffentlicht: (2024)
Streamlining Prediction in Bayesian Deep Learning
von: Li, Rui, et al.
Veröffentlicht: (2024)
von: Li, Rui, et al.
Veröffentlicht: (2024)
Concept-Guided Interpretability via Neural Chunking
von: Wu, Shuchen, et al.
Veröffentlicht: (2025)
von: Wu, Shuchen, et al.
Veröffentlicht: (2025)
Scalable Ranked Preference Optimization for Text-to-Image Generation
von: Karthik, Shyamgopal, et al.
Veröffentlicht: (2024)
von: Karthik, Shyamgopal, et al.
Veröffentlicht: (2024)
Sparse Autoencoders are Topic Models
von: Girrbach, Leander, et al.
Veröffentlicht: (2025)
von: Girrbach, Leander, et al.
Veröffentlicht: (2025)
Simplifying Knowledge Transfer in Pretrained Models
von: Jain, Siddharth, et al.
Veröffentlicht: (2025)
von: Jain, Siddharth, et al.
Veröffentlicht: (2025)
Road Obstacle Video Segmentation
von: Rai, Shyam Nandan, et al.
Veröffentlicht: (2025)
von: Rai, Shyam Nandan, et al.
Veröffentlicht: (2025)
COSMOS: Cross-Modality Self-Distillation for Vision Language Pre-training
von: Kim, Sanghwan, et al.
Veröffentlicht: (2024)
von: Kim, Sanghwan, et al.
Veröffentlicht: (2024)
Compressing 3D Gaussian Splatting by Noise-Substituted Vector Quantization
von: Wang, Haishan, et al.
Veröffentlicht: (2025)
von: Wang, Haishan, et al.
Veröffentlicht: (2025)
It's Never Too Late: Noise Optimization for Collapse Recovery in Trained Diffusion Models
von: Harrington, Anne, et al.
Veröffentlicht: (2025)
von: Harrington, Anne, et al.
Veröffentlicht: (2025)
DeLoRA: Decoupling Angles and Strength in Low-rank Adaptation
von: Bini, Massimo, et al.
Veröffentlicht: (2025)
von: Bini, Massimo, et al.
Veröffentlicht: (2025)
From Drop-off to Recovery: A Mechanistic Analysis of Segmentation in MLLMs
von: Wu, Boyong, et al.
Veröffentlicht: (2026)
von: Wu, Boyong, et al.
Veröffentlicht: (2026)
Time Series Representations for Classification Lie Hidden in Pretrained Vision Transformers
von: Roschmann, Simon, et al.
Veröffentlicht: (2025)
von: Roschmann, Simon, et al.
Veröffentlicht: (2025)
ETHER: Efficient Finetuning of Large-Scale Models with Hyperplane Reflections
von: Bini, Massimo, et al.
Veröffentlicht: (2024)
von: Bini, Massimo, et al.
Veröffentlicht: (2024)
Reflecting on the State of Rehearsal-free Continual Learning with Pretrained Models
von: Thede, Lukas, et al.
Veröffentlicht: (2024)
von: Thede, Lukas, et al.
Veröffentlicht: (2024)
The Manifold Hypothesis for Gradient-Based Explanations
von: Bordt, Sebastian, et al.
Veröffentlicht: (2022)
von: Bordt, Sebastian, et al.
Veröffentlicht: (2022)
A Good CREPE needs more than just Sugar: Investigating Biases in Compositional Vision-Language Benchmarks
von: Udandarao, Vishaal, et al.
Veröffentlicht: (2025)
von: Udandarao, Vishaal, et al.
Veröffentlicht: (2025)
Solving Spatial Supersensing Without Spatial Supersensing
von: Udandarao, Vishaal, et al.
Veröffentlicht: (2025)
von: Udandarao, Vishaal, et al.
Veröffentlicht: (2025)
Improving Intervention Efficacy via Concept Realignment in Concept Bottleneck Models
von: Singhi, Nishad, et al.
Veröffentlicht: (2024)
von: Singhi, Nishad, et al.
Veröffentlicht: (2024)
Post-hoc Self-explanation of CNNs
von: Boubekki, Ahcène, et al.
Veröffentlicht: (2026)
von: Boubekki, Ahcène, et al.
Veröffentlicht: (2026)
Interpretable Generative Models through Post-hoc Concept Bottlenecks
von: Kulkarni, Akshay, et al.
Veröffentlicht: (2025)
von: Kulkarni, Akshay, et al.
Veröffentlicht: (2025)
SUB: Benchmarking CBM Generalization via Synthetic Attribute Substitutions
von: Bader, Jessica, et al.
Veröffentlicht: (2025)
von: Bader, Jessica, et al.
Veröffentlicht: (2025)
Fantastic Gains and Where to Find Them: On the Existence and Prospect of General Knowledge Transfer between Any Pretrained Model
von: Roth, Karsten, et al.
Veröffentlicht: (2023)
von: Roth, Karsten, et al.
Veröffentlicht: (2023)
DataDream: Few-shot Guided Dataset Generation
von: Kim, Jae Myung, et al.
Veröffentlicht: (2024)
von: Kim, Jae Myung, et al.
Veröffentlicht: (2024)
Smol-GS: Compact Representations for Abstract 3D Gaussian Splatting
von: Wang, Haishan, et al.
Veröffentlicht: (2025)
von: Wang, Haishan, et al.
Veröffentlicht: (2025)
Disentangled Representation Learning with the Gromov-Monge Gap
von: Uscidda, Théo, et al.
Veröffentlicht: (2024)
von: Uscidda, Théo, et al.
Veröffentlicht: (2024)
Bias Is a Subspace, Not a Coordinate: A Geometric Rethinking of Post-hoc Debiasing in Vision-Language Models
von: Zhao, Dachuan, et al.
Veröffentlicht: (2025)
von: Zhao, Dachuan, et al.
Veröffentlicht: (2025)
Post-hoc Selective Classification for Reliable Synthetic Image Detection
von: Zheng, Kaixiang, et al.
Veröffentlicht: (2026)
von: Zheng, Kaixiang, et al.
Veröffentlicht: (2026)
The Latent Color Subspace: Emergent Order in High-Dimensional Chaos
von: Pach, Mateusz, et al.
Veröffentlicht: (2026)
von: Pach, Mateusz, et al.
Veröffentlicht: (2026)
Person-Centric Annotations of LAION-400M: Auditing Bias and Its Transfer to Models
von: Girrbach, Leander, et al.
Veröffentlicht: (2025)
von: Girrbach, Leander, et al.
Veröffentlicht: (2025)
Training-free Uncertainty Guidance for Complex Visual Tasks with MLLMs
von: Kim, Sanghwan, et al.
Veröffentlicht: (2025)
von: Kim, Sanghwan, et al.
Veröffentlicht: (2025)
FLAIR: VLM with Fine-grained Language-informed Image Representations
von: Xiao, Rui, et al.
Veröffentlicht: (2024)
von: Xiao, Rui, et al.
Veröffentlicht: (2024)
Innovative Silicosis and Pneumonia Classification: Leveraging Graph Transformer Post-hoc Modeling and Ensemble Techniques
von: Bui, Bao Q., et al.
Veröffentlicht: (2024)
von: Bui, Bao Q., et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Flatness Improves Backbone Generalisation in Few-shot Classification
von: Li, Rui, et al.
Veröffentlicht: (2024) -
Sparse Autoencoders Learn Monosemantic Features in Vision-Language Models
von: Pach, Mateusz, et al.
Veröffentlicht: (2025) -
Vision-by-Language for Training-Free Compositional Image Retrieval
von: Karthik, Shyamgopal, et al.
Veröffentlicht: (2023) -
Sources of Uncertainty in 3D Scene Reconstruction
von: Klasson, Marcus, et al.
Veröffentlicht: (2024) -
Noise Hypernetworks: Amortizing Test-Time Compute in Diffusion Models
von: Eyring, Luca, et al.
Veröffentlicht: (2025)