HARMONY: Hidden Activation Representations and Model Output-Aware Uncertainty Estimation for Vision-Language Models
Fuente:
arXiv
Guardado en:
| Autores principales: | Mushtaq, Erum, Fabian, Zalan, Bakman, Yavuz Faruk, Ramakrishna, Anil, Soltanolkotabi, Mahdi, Avestimehr, Salman |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
CroMo-Mixup: Augmenting Cross-Model Representations for Continual Self-Supervised Learning
por: Mushtaq, Erum, et al.
Publicado: (2024)
por: Mushtaq, Erum, et al.
Publicado: (2024)
Emergence and Evolution of Interpretable Concepts in Diffusion Models
por: Tinaz, Berk, et al.
Publicado: (2025)
por: Tinaz, Berk, et al.
Publicado: (2025)
Uncertainty Quantification for Hallucination Detection in Large Language Models: Foundations, Methodology, and Future Directions
por: Kang, Sungmin, et al.
Publicado: (2025)
por: Kang, Sungmin, et al.
Publicado: (2025)
Do Not Design, Learn: A Trainable Scoring Function for Uncertainty Estimation in Generative LLMs
por: Yaldiz, Duygu Nur, et al.
Publicado: (2024)
por: Yaldiz, Duygu Nur, et al.
Publicado: (2024)
MARS: Meaning-Aware Response Scoring for Uncertainty Estimation in Generative LLMs
por: Bakman, Yavuz Faruk, et al.
Publicado: (2024)
por: Bakman, Yavuz Faruk, et al.
Publicado: (2024)
MediConfusion: Can you trust your AI radiologist? Probing the reliability of multimodal medical foundation models
por: Sepehri, Mohammad Shahab, et al.
Publicado: (2024)
por: Sepehri, Mohammad Shahab, et al.
Publicado: (2024)
Serpent: Scalable and Efficient Image Restoration via Multi-scale Structured State Space Models
por: Sepehri, Mohammad Shahab, et al.
Publicado: (2024)
por: Sepehri, Mohammad Shahab, et al.
Publicado: (2024)
DiracDiffusion: Denoising and Incremental Reconstruction with Assured Data-Consistency
por: Fabian, Zalan, et al.
Publicado: (2023)
por: Fabian, Zalan, et al.
Publicado: (2023)
ATHENA: Adaptive Test-Time Steering for Improving Count Fidelity in Diffusion Models
por: Sepehri, Mohammad Shahab, et al.
Publicado: (2026)
por: Sepehri, Mohammad Shahab, et al.
Publicado: (2026)
Adapt and Diffuse: Sample-adaptive Reconstruction via Latent Diffusion Models
por: Fabian, Zalan, et al.
Publicado: (2023)
por: Fabian, Zalan, et al.
Publicado: (2023)
Hyperphantasia: A Benchmark for Evaluating the Mental Visualization Capabilities of Multimodal LLMs
por: Sepehri, Mohammad Shahab, et al.
Publicado: (2025)
por: Sepehri, Mohammad Shahab, et al.
Publicado: (2025)
ConceptMix++: Leveling the Playing Field in Text-to-Image Benchmarking via Iterative Prompt Optimization
por: Gan, Haosheng, et al.
Publicado: (2025)
por: Gan, Haosheng, et al.
Publicado: (2025)
CryptoMamba: Leveraging State Space Models for Accurate Bitcoin Price Prediction
por: Sepehri, Mohammad Shahab, et al.
Publicado: (2025)
por: Sepehri, Mohammad Shahab, et al.
Publicado: (2025)
Reconsidering LLM Uncertainty Estimation Methods in the Wild
por: Bakman, Yavuz, et al.
Publicado: (2025)
por: Bakman, Yavuz, et al.
Publicado: (2025)
Reject Only Critical Tokens: Pivot-Aware Speculative Decoding
por: Ziashahabi, Amir, et al.
Publicado: (2025)
por: Ziashahabi, Amir, et al.
Publicado: (2025)
Hair-Trigger Alignment: Black-Box Evaluation Cannot Guarantee Post-Update Alignment
por: Bakman, Yavuz, et al.
Publicado: (2026)
por: Bakman, Yavuz, et al.
Publicado: (2026)
Assessing Visual Privacy Risks in Multimodal AI: A Novel Taxonomy-Grounded Evaluation of Vision-Language Models
por: Tsaprazlis, Efthymios, et al.
Publicado: (2025)
por: Tsaprazlis, Efthymios, et al.
Publicado: (2025)
Do LLMs Recognize me, When I is not me: Assessment of LLMs Understanding of Turkish Indexical Pronouns in Indexical Shift Contexts
por: Oğuz, Metehan, et al.
Publicado: (2024)
por: Oğuz, Metehan, et al.
Publicado: (2024)
Uncertainty-Aware Evaluation for Vision-Language Models
por: Kostumov, Vasily, et al.
Publicado: (2024)
por: Kostumov, Vasily, et al.
Publicado: (2024)
TruthTorchLM: A Comprehensive Library for Predicting Truthfulness in LLM Outputs
por: Yaldiz, Duygu Nur, et al.
Publicado: (2025)
por: Yaldiz, Duygu Nur, et al.
Publicado: (2025)
MosaicMRI: A Diverse Dataset and Benchmark for Raw Musculoskeletal MRI
por: Arguello, Paula, et al.
Publicado: (2026)
por: Arguello, Paula, et al.
Publicado: (2026)
Leveraging Uncertainty Estimation for Efficient LLM Routing
por: Zhang, Tuo, et al.
Publicado: (2025)
por: Zhang, Tuo, et al.
Publicado: (2025)
VL-Uncertainty: Detecting Hallucination in Large Vision-Language Model via Uncertainty Estimation
por: Zhang, Ruiyang, et al.
Publicado: (2024)
por: Zhang, Ruiyang, et al.
Publicado: (2024)
Bridging Hidden States in Vision-Language Models
por: Fein-Ashley, Benjamin, et al.
Publicado: (2025)
por: Fein-Ashley, Benjamin, et al.
Publicado: (2025)
Uncertainty as Feature Gaps: Epistemic Uncertainty Quantification of LLMs in Contextual Question-Answering
por: Bakman, Yavuz, et al.
Publicado: (2025)
por: Bakman, Yavuz, et al.
Publicado: (2025)
Intra-Class Probabilistic Embeddings for Uncertainty Estimation in Vision-Language Models
por: Lin, Zhenxiang, et al.
Publicado: (2025)
por: Lin, Zhenxiang, et al.
Publicado: (2025)
Memory-Efficient Vision Transformers: An Activation-Aware Mixed-Rank Compression Strategy
por: Azizi, Seyedarmin, et al.
Publicado: (2024)
por: Azizi, Seyedarmin, et al.
Publicado: (2024)
ColorFoil: Investigating Color Blindness in Large Vision and Language Models
por: Samin, Ahnaf Mozib, et al.
Publicado: (2024)
por: Samin, Ahnaf Mozib, et al.
Publicado: (2024)
MMRL++: Parameter-Efficient and Interaction-Aware Representation Learning for Vision-Language Models
por: Guo, Yuncheng, et al.
Publicado: (2025)
por: Guo, Yuncheng, et al.
Publicado: (2025)
GeoToken: Hierarchical Geolocalization of Images via Next Token Prediction
por: Ghasemi, Narges, et al.
Publicado: (2025)
por: Ghasemi, Narges, et al.
Publicado: (2025)
Representation Calibration and Uncertainty Guidance for Class-Incremental Learning based on Vision Language Model
por: Tan, Jiantao, et al.
Publicado: (2025)
por: Tan, Jiantao, et al.
Publicado: (2025)
GesVLA: Gesture-Aware Vision-Language-Action Model Embedded Representations
por: Guo, Wenxuan, et al.
Publicado: (2026)
por: Guo, Wenxuan, et al.
Publicado: (2026)
Uncertainty-Aware Gaussian Map for Vision-Language Navigation
por: Gao, Jianzhe, et al.
Publicado: (2026)
por: Gao, Jianzhe, et al.
Publicado: (2026)
Data Pruning via Separability, Integrity, and Model Uncertainty-Aware Importance Sampling
por: Grosz, Steven, et al.
Publicado: (2024)
por: Grosz, Steven, et al.
Publicado: (2024)
See, Plan, Rewind: Progress-Aware Vision-Language-Action Models for Robust Robotic Manipulation
por: Dai, Tingjun, et al.
Publicado: (2026)
por: Dai, Tingjun, et al.
Publicado: (2026)
LANTERN: A Machine Learning Framework for Lipid Nanoparticle Transfection Efficiency Prediction
por: Mehradfar, Asal, et al.
Publicado: (2025)
por: Mehradfar, Asal, et al.
Publicado: (2025)
Edge-Optimized Vision-Language Models for Underground Infrastructure Assessment
por: Lopez, Johny J., et al.
Publicado: (2026)
por: Lopez, Johny J., et al.
Publicado: (2026)
Uncertainty-Aware Vision-Language Segmentation for Medical Imaging
por: Das, Aryan, et al.
Publicado: (2026)
por: Das, Aryan, et al.
Publicado: (2026)
FLARE: Learning Future-Aware Latent Representations from Vision-Language Models for Autonomous Driving
por: Xie, Chengen, et al.
Publicado: (2026)
por: Xie, Chengen, et al.
Publicado: (2026)
Interpretable Zero-Shot Learning with Locally-Aligned Vision-Language Model
por: Chen, Shiming, et al.
Publicado: (2025)
por: Chen, Shiming, et al.
Publicado: (2025)
Ejemplares similares
-
CroMo-Mixup: Augmenting Cross-Model Representations for Continual Self-Supervised Learning
por: Mushtaq, Erum, et al.
Publicado: (2024) -
Emergence and Evolution of Interpretable Concepts in Diffusion Models
por: Tinaz, Berk, et al.
Publicado: (2025) -
Uncertainty Quantification for Hallucination Detection in Large Language Models: Foundations, Methodology, and Future Directions
por: Kang, Sungmin, et al.
Publicado: (2025) -
Do Not Design, Learn: A Trainable Scoring Function for Uncertainty Estimation in Generative LLMs
por: Yaldiz, Duygu Nur, et al.
Publicado: (2024) -
MARS: Meaning-Aware Response Scoring for Uncertainty Estimation in Generative LLMs
por: Bakman, Yavuz Faruk, et al.
Publicado: (2024)