The Illusion of AI Expertise Under Uncertainty: Navigating Elusive Ground Truth via a Probabilistic Paradigm
Fuente:
arXiv
Guardado en:
| Autores principales: | Elangovan, Aparna, Xu, Lei, Elyasi, Mahsa, Akdulum, Ismail, Aksakal, Mehmet, Gurun, Enes, Hur, Brian, Mansour, Saab, Ziv, Ravid Shwartz, Verspoor, Karin, Roth, Dan |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Principles from Clinical Research for NLP Model Generalization
por: Elangovan, Aparna, et al.
Publicado: (2023)
por: Elangovan, Aparna, et al.
Publicado: (2023)
Ultrasonography Combined With Strain Elastography for Differentiation of Parotid Masses
por: Enes Gurun
Publicado: (2025)
por: Enes Gurun
Publicado: (2025)
Learning to Compress: Local Rank and Information Compression in Deep Neural Networks
por: Patel, Niket, et al.
Publicado: (2024)
por: Patel, Niket, et al.
Publicado: (2024)
The Illusion of Progress: Re-evaluating Hallucination Detection in LLMs
por: Janiak, Denis, et al.
Publicado: (2025)
por: Janiak, Denis, et al.
Publicado: (2025)
Beyond correlation: The Impact of Human Uncertainty in Measuring the Effectiveness of Automatic Evaluation and LLM-as-a-Judge
por: Elangovan, Aparna, et al.
Publicado: (2024)
por: Elangovan, Aparna, et al.
Publicado: (2024)
The Added Value of Shear Wave Elastography in the Diagnosis and Grading of Carpal Tunnel Syndrome
por: Enes Gurun, et al.
Publicado: (2025)
por: Enes Gurun, et al.
Publicado: (2025)
Added Value of Ultra Micro Angiography for Renal Tumor Assessment
por: Mustafa Basaran, et al.
Publicado: (2026)
por: Mustafa Basaran, et al.
Publicado: (2026)
Video Representation Learning with Joint-Embedding Predictive Architectures
por: Drozdov, Katrina, et al.
Publicado: (2024)
por: Drozdov, Katrina, et al.
Publicado: (2024)
Interpreting Peritumoral Elastography—Technical and Pathologic Considerations
por: Mahsima Gül Gümrükçü, et al.
Publicado: (2025)
por: Mahsima Gül Gümrükçü, et al.
Publicado: (2025)
Antislop: A Comprehensive Framework for Identifying and Eliminating Repetitive Patterns in Language Models
por: Paech, Samuel, et al.
Publicado: (2025)
por: Paech, Samuel, et al.
Publicado: (2025)
When Attention Collapses: How Degenerate Layers in LLMs Enable Smaller, Stronger Models
por: Sanyal, Sunny, et al.
Publicado: (2024)
por: Sanyal, Sunny, et al.
Publicado: (2024)
AI Must Embrace Specialization via Superhuman Adaptable Intelligence
por: Goldfeder, Judah, et al.
Publicado: (2026)
por: Goldfeder, Judah, et al.
Publicado: (2026)
The Entropy Enigma: Success and Failure of Entropy Minimization
por: Press, Ori, et al.
Publicado: (2024)
por: Press, Ori, et al.
Publicado: (2024)
Does Representation Matter? Exploring Intermediate Layers in Large Language Models
por: Skean, Oscar, et al.
Publicado: (2024)
por: Skean, Oscar, et al.
Publicado: (2024)
Latent Transfer Attack: Adversarial Examples via Generative Latent Spaces
por: Shaar, Eitan, et al.
Publicado: (2026)
por: Shaar, Eitan, et al.
Publicado: (2026)
EMBRE: Entity-aware Masking for Biomedical Relation Extraction
por: Li, Mingjie, et al.
Publicado: (2024)
por: Li, Mingjie, et al.
Publicado: (2024)
From Tokens to Thoughts: How LLMs and Humans Trade Compression for Meaning
por: Shani, Chen, et al.
Publicado: (2025)
por: Shani, Chen, et al.
Publicado: (2025)
On Training in Imagination
por: Timor, Nadav, et al.
Publicado: (2026)
por: Timor, Nadav, et al.
Publicado: (2026)
Variance-Covariance Regularization Improves Representation Learning
por: Zhu, Jiachen, et al.
Publicado: (2023)
por: Zhu, Jiachen, et al.
Publicado: (2023)
Sudden Drops in the Loss: Syntax Acquisition, Phase Transitions, and Simplicity Bias in MLMs
por: Chen, Angelica, et al.
Publicado: (2023)
por: Chen, Angelica, et al.
Publicado: (2023)
Truth and the Will to Illusion
por: Filipowicz, Stanisław
Publicado: (2024)
por: Filipowicz, Stanisław
Publicado: (2024)
Exploring Human-AI Conceptual Alignment through the Prism of Chess
por: Lomasov, Semyon, et al.
Publicado: (2025)
por: Lomasov, Semyon, et al.
Publicado: (2025)
An Information-Theoretic Perspective on Variance-Invariance-Covariance Regularization
por: Shwartz-Ziv, Ravid, et al.
Publicado: (2023)
por: Shwartz-Ziv, Ravid, et al.
Publicado: (2023)
Turning Up the Heat: Min-p Sampling for Creative and Coherent LLM Outputs
por: Nguyen, Minh Nhat, et al.
Publicado: (2024)
por: Nguyen, Minh Nhat, et al.
Publicado: (2024)
You Had One Job: Per-Task Quantization Using LLMs' Hidden Representations
por: LeVi, Amit, et al.
Publicado: (2025)
por: LeVi, Amit, et al.
Publicado: (2025)
Rate-In: Information-Driven Adaptive Dropout Rates for Improved Inference-Time Uncertainty Estimation
por: Zeevi, Tal, et al.
Publicado: (2024)
por: Zeevi, Tal, et al.
Publicado: (2024)
ConSiDERS-The-Human Evaluation Framework: Rethinking Human Evaluation for Generative Large Language Models
por: Elangovan, Aparna, et al.
Publicado: (2024)
por: Elangovan, Aparna, et al.
Publicado: (2024)
Between Institutional Loneliness and Visibility: Low‐Income Families Navigating Housing Insecurity in Social Welfare Programs
por: Tamar Shwartz‐Ziv, et al.
Publicado: (2025)
por: Tamar Shwartz‐Ziv, et al.
Publicado: (2025)
Beyond the Loss Curve: Scaling Laws, Active Learning, and the Limits of Learning from Exact Posteriors
por: Khorasani, Arian, et al.
Publicado: (2026)
por: Khorasani, Arian, et al.
Publicado: (2026)
Layer Importance for Mathematical Reasoning is Forged in Pre-Training and Invariant after Post-Training
por: Nepal, Aadim, et al.
Publicado: (2025)
por: Nepal, Aadim, et al.
Publicado: (2025)
Zero‐ and few‐shot prompting of generative large language models provides weak assessment of risk of bias in clinical trials
por: Simon Šuster, et al.
Publicado: (2024)
por: Simon Šuster, et al.
Publicado: (2024)
JEPA as a Neural Tokenizer: Learning Robust Speech Representations with Density Adaptive Attention
por: Ioannides, Georgios, et al.
Publicado: (2025)
por: Ioannides, Georgios, et al.
Publicado: (2025)
Just How Flexible are Neural Networks in Practice?
por: Shwartz-Ziv, Ravid, et al.
Publicado: (2024)
por: Shwartz-Ziv, Ravid, et al.
Publicado: (2024)
UAT-LITE: Inference-Time Uncertainty-Aware Attention for Pretrained Transformers
por: Hossain, Elias, et al.
Publicado: (2026)
por: Hossain, Elias, et al.
Publicado: (2026)
A superpersuasive autonomous policy debating system
por: Roush, Allen, et al.
Publicado: (2025)
por: Roush, Allen, et al.
Publicado: (2025)
A chance to learn : knowledge and finance for education in Sub-Saharan Africa / Adriaan Verspoor
por: Verspoor, Adriaan
Publicado: (2001)
por: Verspoor, Adriaan
Publicado: (2001)
Pathways two change : improving the quality of education in developing countries / Adriaan Verspoor
por: Verspoor, Adriaan
Publicado: (1989)
por: Verspoor, Adriaan
Publicado: (1989)
At the crossroads : choices for secondary education in Sub-Saharan Africa / Adriaan Verspoor
por: Verspoor, Adriaan
por: Verspoor, Adriaan
Challenges to the planning of education / Adriaan Verspoor
por: Verspoor, Adriaan
Publicado: (1992)
por: Verspoor, Adriaan
Publicado: (1992)
Disambiguating Complexity: From CAF to CAFIC: A Commentary on “Complexity and Difficulty in Second Language Acquisition: A Theoretical and Methodological Overview”
por: Marjolijn Verspoor
Publicado: (2024)
por: Marjolijn Verspoor
Publicado: (2024)
Ejemplares similares
-
Principles from Clinical Research for NLP Model Generalization
por: Elangovan, Aparna, et al.
Publicado: (2023) -
Ultrasonography Combined With Strain Elastography for Differentiation of Parotid Masses
por: Enes Gurun
Publicado: (2025) -
Learning to Compress: Local Rank and Information Compression in Deep Neural Networks
por: Patel, Niket, et al.
Publicado: (2024) -
The Illusion of Progress: Re-evaluating Hallucination Detection in LLMs
por: Janiak, Denis, et al.
Publicado: (2025) -
Beyond correlation: The Impact of Human Uncertainty in Measuring the Effectiveness of Automatic Evaluation and LLM-as-a-Judge
por: Elangovan, Aparna, et al.
Publicado: (2024)