Entropy Alone is Insufficient for Safe Selective Prediction in LLMs
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Phillips, Edward, Gustafsson, Fredrik K., Wu, Sean, Thakur, Anshul, Clifton, David A. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
BAS: A Decision-Theoretic Approach to Evaluating Large Language Model Confidence
von: Wu, Sean, et al.
Veröffentlicht: (2026)
von: Wu, Sean, et al.
Veröffentlicht: (2026)
Geometric Uncertainty for Detecting and Correcting Hallucinations in LLMs
von: Phillips, Edward, et al.
Veröffentlicht: (2025)
von: Phillips, Edward, et al.
Veröffentlicht: (2025)
Semantic Self-Distillation for Language Model Uncertainty
von: Phillips, Edward, et al.
Veröffentlicht: (2026)
von: Phillips, Edward, et al.
Veröffentlicht: (2026)
Benchmarking Pathology Foundation Models for Breast Cancer Survival Prediction
von: Gustafsson, Fredrik K., et al.
Veröffentlicht: (2026)
von: Gustafsson, Fredrik K., et al.
Veröffentlicht: (2026)
TECP: Token-Entropy Conformal Prediction for LLMs
von: Xu, Beining, et al.
Veröffentlicht: (2025)
von: Xu, Beining, et al.
Veröffentlicht: (2025)
Evaluating Deep Regression Models for WSI-Based Gene-Expression Prediction
von: Gustafsson, Fredrik K., et al.
Veröffentlicht: (2024)
von: Gustafsson, Fredrik K., et al.
Veröffentlicht: (2024)
Improving Clinical Dataset Condensation with Mode Connectivity-based Trajectory Surrogates
von: Nganjimi, Pafue Christy, et al.
Veröffentlicht: (2025)
von: Nganjimi, Pafue Christy, et al.
Veröffentlicht: (2025)
Invisible Entropy: Towards Safe and Efficient Low-Entropy LLM Watermarking
von: Gu, Tianle, et al.
Veröffentlicht: (2025)
von: Gu, Tianle, et al.
Veröffentlicht: (2025)
Medical Reasoning in LLMs: An In-Depth Analysis of DeepSeek R1
von: Moell, Birger, et al.
Veröffentlicht: (2025)
von: Moell, Birger, et al.
Veröffentlicht: (2025)
Slow Tuning and Low-Entropy Masking for Safe Chain-of-Thought Distillation
von: Ma, Ziyang, et al.
Veröffentlicht: (2025)
von: Ma, Ziyang, et al.
Veröffentlicht: (2025)
Large Language Models in the Clinic: A Comprehensive Benchmark
von: Liu, Fenglin, et al.
Veröffentlicht: (2024)
von: Liu, Fenglin, et al.
Veröffentlicht: (2024)
Paradox of De-identification: A Critique of HIPAA Safe Harbour in the Age of LLMs
von: Jiang, Lavender Y., et al.
Veröffentlicht: (2026)
von: Jiang, Lavender Y., et al.
Veröffentlicht: (2026)
Cochain: Balancing Insufficient and Excessive Collaboration in LLM Agent Workflows
von: Zhao, Jiaxing, et al.
Veröffentlicht: (2025)
von: Zhao, Jiaxing, et al.
Veröffentlicht: (2025)
Is Sanskrit the most token-efficient language? A quantitative study using GPT, Gemini, and SentencePiece
von: Kumar, Anshul
Veröffentlicht: (2026)
von: Kumar, Anshul
Veröffentlicht: (2026)
Pre-Training Multimodal Hallucination Detectors with Corrupted Grounding Data
von: Whitehead, Spencer, et al.
Veröffentlicht: (2024)
von: Whitehead, Spencer, et al.
Veröffentlicht: (2024)
Language-Aware Distillation for Multilingual Instruction-Following Speech LLMs with ASR-Only Supervision
von: Gopal, Shreyas, et al.
Veröffentlicht: (2026)
von: Gopal, Shreyas, et al.
Veröffentlicht: (2026)
Reasoning Isn't Enough: Examining Truth-Bias and Sycophancy in LLMs
von: Barkett, Emilio, et al.
Veröffentlicht: (2025)
von: Barkett, Emilio, et al.
Veröffentlicht: (2025)
JT-Safe: Intrinsically Enhancing the Safety and Trustworthiness of LLMs
von: Feng, Junlan, et al.
Veröffentlicht: (2025)
von: Feng, Junlan, et al.
Veröffentlicht: (2025)
CalibraEval: Calibrating Prediction Distribution to Mitigate Selection Bias in LLMs-as-Judges
von: Li, Haitao, et al.
Veröffentlicht: (2024)
von: Li, Haitao, et al.
Veröffentlicht: (2024)
Reinforcement Learning for Diffusion LLMs with Entropy-Guided Step Selection and Stepwise Advantages
von: Kunde, Vishnu Teja, et al.
Veröffentlicht: (2026)
von: Kunde, Vishnu Teja, et al.
Veröffentlicht: (2026)
Chatting Up Attachment: Using LLMs to Predict Adult Bonds
von: Soares, Paulo, et al.
Veröffentlicht: (2024)
von: Soares, Paulo, et al.
Veröffentlicht: (2024)
Across the Levels of Analysis: Explaining Predictive Processing in Humans Requires More Than Machine-Estimated Probabilities
von: Nair, Sathvik, et al.
Veröffentlicht: (2026)
von: Nair, Sathvik, et al.
Veröffentlicht: (2026)
CLARity: Reasoning Consistency Alone Can Teach Reinforced Experts
von: Lin, Jiuheng, et al.
Veröffentlicht: (2025)
von: Lin, Jiuheng, et al.
Veröffentlicht: (2025)
Entropy-Based Data Selection for Language Models
von: Li, Hongming, et al.
Veröffentlicht: (2026)
von: Li, Hongming, et al.
Veröffentlicht: (2026)
Culturally-Grounded Chain-of-Thought (CG-CoT):Enhancing LLM Performance on Culturally-Specific Tasks in Low-Resource Languages
von: Thakur, Madhavendra
Veröffentlicht: (2025)
von: Thakur, Madhavendra
Veröffentlicht: (2025)
Towards Neural No-Resource Language Translation: A Comparative Evaluation of Approaches
von: Thakur, Madhavendra
Veröffentlicht: (2024)
von: Thakur, Madhavendra
Veröffentlicht: (2024)
Measurement in the Age of LLMs: An Application to Ideological Scaling
von: O'Hagan, Sean, et al.
Veröffentlicht: (2023)
von: O'Hagan, Sean, et al.
Veröffentlicht: (2023)
Bribery's Influence on Ranked Aggregation
von: Jain, Pallavi, et al.
Veröffentlicht: (2026)
von: Jain, Pallavi, et al.
Veröffentlicht: (2026)
Randomise Alone, Reach as a Team
von: Brice, Léonard, et al.
Veröffentlicht: (2026)
von: Brice, Léonard, et al.
Veröffentlicht: (2026)
SPINE: Token-Selective Test-Time Reinforcement Learning with Entropy-Band Regularization
von: Wu, Jianghao, et al.
Veröffentlicht: (2025)
von: Wu, Jianghao, et al.
Veröffentlicht: (2025)
Actor Identification in Discourse: A Challenge for LLMs?
von: Barić, Ana, et al.
Veröffentlicht: (2024)
von: Barić, Ana, et al.
Veröffentlicht: (2024)
A Dual-Layered Evaluation of Geopolitical and Cultural Bias in LLMs
von: Kim, Sean, et al.
Veröffentlicht: (2025)
von: Kim, Sean, et al.
Veröffentlicht: (2025)
GPT-4 Is Too Smart To Be Safe: Stealthy Chat with LLMs via Cipher
von: Yuan, Youliang, et al.
Veröffentlicht: (2023)
von: Yuan, Youliang, et al.
Veröffentlicht: (2023)
Being Kind Isn't Always Being Safe: Diagnosing Affective Hallucination in LLMs
von: Kim, Sewon, et al.
Veröffentlicht: (2025)
von: Kim, Sewon, et al.
Veröffentlicht: (2025)
Intrinsic Entropy of Context Length Scaling in LLMs
von: Shi, Jingzhe, et al.
Veröffentlicht: (2025)
von: Shi, Jingzhe, et al.
Veröffentlicht: (2025)
Structural Entropy Guided Agent for Detecting and Repairing Knowledge Deficiencies in LLMs
von: Wei, Yifan, et al.
Veröffentlicht: (2025)
von: Wei, Yifan, et al.
Veröffentlicht: (2025)
SafeConstellations: Mitigating Over-Refusals in LLMs Through Task-Aware Representation Steering
von: Maskey, Utsav, et al.
Veröffentlicht: (2025)
von: Maskey, Utsav, et al.
Veröffentlicht: (2025)
The Hyperfitting Phenomenon: Sharpening and Stabilizing LLMs for Open-Ended Text Generation
von: Carlsson, Fredrik, et al.
Veröffentlicht: (2024)
von: Carlsson, Fredrik, et al.
Veröffentlicht: (2024)
LADM: Long-context Training Data Selection with Attention-based Dependency Measurement for LLMs
von: Chen, Jianghao, et al.
Veröffentlicht: (2025)
von: Chen, Jianghao, et al.
Veröffentlicht: (2025)
SafeLawBench: Towards Safe Alignment of Large Language Models
von: Cao, Chuxue, et al.
Veröffentlicht: (2025)
von: Cao, Chuxue, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
BAS: A Decision-Theoretic Approach to Evaluating Large Language Model Confidence
von: Wu, Sean, et al.
Veröffentlicht: (2026) -
Geometric Uncertainty for Detecting and Correcting Hallucinations in LLMs
von: Phillips, Edward, et al.
Veröffentlicht: (2025) -
Semantic Self-Distillation for Language Model Uncertainty
von: Phillips, Edward, et al.
Veröffentlicht: (2026) -
Benchmarking Pathology Foundation Models for Breast Cancer Survival Prediction
von: Gustafsson, Fredrik K., et al.
Veröffentlicht: (2026) -
TECP: Token-Entropy Conformal Prediction for LLMs
von: Xu, Beining, et al.
Veröffentlicht: (2025)