LogitScope: A Framework for Analyzing LLM Uncertainty Through Information Metrics
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ahmed, Farhan, Ong, Yuya Jeremy, DeLuca, Chad |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Analyzing and Improving Chain-of-Thought Monitorability Through Information Theory
von: Anwar, Usman, et al.
Veröffentlicht: (2026)
von: Anwar, Usman, et al.
Veröffentlicht: (2026)
An Information-Theoretic Approach to Analyze NLP Classification Tasks
von: Wang, Luran, et al.
Veröffentlicht: (2024)
von: Wang, Luran, et al.
Veröffentlicht: (2024)
MermaidSeqBench: An Evaluation Benchmark for NL-to-Mermaid Sequence Diagram Generation
von: Shbita, Basel, et al.
Veröffentlicht: (2025)
von: Shbita, Basel, et al.
Veröffentlicht: (2025)
LLMON: An LLM-native Markup Language to Leverage Structure and Semantics at the LLM Interface
von: Hind, Michael, et al.
Veröffentlicht: (2026)
von: Hind, Michael, et al.
Veröffentlicht: (2026)
STaD: Scaffolded Task Design for Identifying Compositional Skill Gaps in LLMs
von: An, Sungeun, et al.
Veröffentlicht: (2026)
von: An, Sungeun, et al.
Veröffentlicht: (2026)
Measuring Uncertainty in Transformer Circuits with Effective Information Consistency
von: Krasnovsky, Anatoly A.
Veröffentlicht: (2025)
von: Krasnovsky, Anatoly A.
Veröffentlicht: (2025)
Backprompting: Leveraging Synthetic Production Data for Health Advice Guardrails
von: Cheng, Kellen Tan, et al.
Veröffentlicht: (2025)
von: Cheng, Kellen Tan, et al.
Veröffentlicht: (2025)
A Communication-Theoretic Framework for LLM Agents: Cost-Aware Adaptive Reliability
von: Omidvar, Hamed, et al.
Veröffentlicht: (2026)
von: Omidvar, Hamed, et al.
Veröffentlicht: (2026)
SearchRAG: Can Search Engines Be Helpful for LLM-based Medical Question Answering?
von: Shi, Yucheng, et al.
Veröffentlicht: (2025)
von: Shi, Yucheng, et al.
Veröffentlicht: (2025)
Measuring Information Distortion in Hierarchical Ultra long Novel Reconstruction:The Optimal Expansion Ratio
von: Shen, Hanwen, et al.
Veröffentlicht: (2025)
von: Shen, Hanwen, et al.
Veröffentlicht: (2025)
Think or Not? Exploring Thinking Efficiency in Large Reasoning Models via an Information-Theoretic Lens
von: Yong, Xixian, et al.
Veröffentlicht: (2025)
von: Yong, Xixian, et al.
Veröffentlicht: (2025)
ELOQ: Resources for Enhancing LLM Detection of Out-of-Scope Questions
von: Peng, Zhiyuan, et al.
Veröffentlicht: (2024)
von: Peng, Zhiyuan, et al.
Veröffentlicht: (2024)
Diff-eRank: A Novel Rank-Based Metric for Evaluating Large Language Models
von: Wei, Lai, et al.
Veröffentlicht: (2024)
von: Wei, Lai, et al.
Veröffentlicht: (2024)
Analyzing the Influence of Knowledge Graph Information on Relation Extraction
von: Möller, Cedric, et al.
Veröffentlicht: (2025)
von: Möller, Cedric, et al.
Veröffentlicht: (2025)
An Information-Theoretic Framework for Comparing Voice and Text Explainability
von: Rajhans, Mona, et al.
Veröffentlicht: (2026)
von: Rajhans, Mona, et al.
Veröffentlicht: (2026)
Learning is Forgetting: LLM Training As Lossy Compression
von: Conklin, Henry C., et al.
Veröffentlicht: (2026)
von: Conklin, Henry C., et al.
Veröffentlicht: (2026)
Iterative Critique-Refine Framework for Enhancing LLM Personalization
von: Maram, Durga Prasad, et al.
Veröffentlicht: (2025)
von: Maram, Durga Prasad, et al.
Veröffentlicht: (2025)
A Training-free Method for LLM Text Attribution
von: Radvand, Tara, et al.
Veröffentlicht: (2025)
von: Radvand, Tara, et al.
Veröffentlicht: (2025)
Adaptive Decoding via Test-Time Policy Learning for Self-Improving Generation
von: Bhardwaj, Asmita, et al.
Veröffentlicht: (2026)
von: Bhardwaj, Asmita, et al.
Veröffentlicht: (2026)
Forgetting-MarI: LLM Unlearning via Marginal Information Regularization
von: Xu, Shizhou, et al.
Veröffentlicht: (2025)
von: Xu, Shizhou, et al.
Veröffentlicht: (2025)
The Information of Large Language Model Geometry
von: Tan, Zhiquan, et al.
Veröffentlicht: (2024)
von: Tan, Zhiquan, et al.
Veröffentlicht: (2024)
From Tokens to Thoughts: How LLMs and Humans Trade Compression for Meaning
von: Shani, Chen, et al.
Veröffentlicht: (2025)
von: Shani, Chen, et al.
Veröffentlicht: (2025)
Test-Time Steering for Lossless Text Compression via Weighted Product of Experts
von: Zhang, Qihang, et al.
Veröffentlicht: (2025)
von: Zhang, Qihang, et al.
Veröffentlicht: (2025)
Surprisal and Metaphor Novelty Judgments: Moderate Correlations and Divergent Scaling Effects Revealed by Corpus-Based and Synthetic Datasets
von: Momen, Omar, et al.
Veröffentlicht: (2026)
von: Momen, Omar, et al.
Veröffentlicht: (2026)
Evaluating Large Language Models for Structured Science Summarization in the Open Research Knowledge Graph
von: Nechakhin, Vladyslav, et al.
Veröffentlicht: (2024)
von: Nechakhin, Vladyslav, et al.
Veröffentlicht: (2024)
On Theoretical Interpretations of Concept-Based In-Context Learning
von: Tang, Huaze, et al.
Veröffentlicht: (2025)
von: Tang, Huaze, et al.
Veröffentlicht: (2025)
Astro-NER -- Astronomy Named Entity Recognition: Is GPT a Good Domain Expert Annotator?
von: Evans, Julia, et al.
Veröffentlicht: (2024)
von: Evans, Julia, et al.
Veröffentlicht: (2024)
Task-Centric Acceleration of Small-Language Models
von: Tsur, Dor, et al.
Veröffentlicht: (2026)
von: Tsur, Dor, et al.
Veröffentlicht: (2026)
Slaves to the Law of Large Numbers: An Asymptotic Equipartition Property for Perplexity in Generative Language Models
von: Bell, Tyler, et al.
Veröffentlicht: (2024)
von: Bell, Tyler, et al.
Veröffentlicht: (2024)
Large Language Models as Evaluators for Scientific Synthesis
von: Evans, Julia, et al.
Veröffentlicht: (2024)
von: Evans, Julia, et al.
Veröffentlicht: (2024)
Graph-Enhanced Retrieval-Augmented Question Answering for E-Commerce Customer Support
von: Patel, Piyushkumar
Veröffentlicht: (2025)
von: Patel, Piyushkumar
Veröffentlicht: (2025)
Complexity Agnostic Recursive Decomposition of Thoughts
von: Qasim, Kaleem Ullah, et al.
Veröffentlicht: (2025)
von: Qasim, Kaleem Ullah, et al.
Veröffentlicht: (2025)
On the Reasoning Capacity of AI Models and How to Quantify It
von: Radha, Santosh Kumar, et al.
Veröffentlicht: (2025)
von: Radha, Santosh Kumar, et al.
Veröffentlicht: (2025)
Conversational Complexity for Assessing Risk in Large Language Models
von: Burden, John, et al.
Veröffentlicht: (2024)
von: Burden, John, et al.
Veröffentlicht: (2024)
Seq2Seq2Seq: Lossless Data Compression via Discrete Latent Transformers and Reinforcement Learning
von: Khodabandeh, Mahdi, et al.
Veröffentlicht: (2026)
von: Khodabandeh, Mahdi, et al.
Veröffentlicht: (2026)
Riddle Quest : The Enigma of Words
von: Parasa, Niharika Sri, et al.
Veröffentlicht: (2026)
von: Parasa, Niharika Sri, et al.
Veröffentlicht: (2026)
Robust AI-Generated Text Detection by Restricted Embeddings
von: Kuznetsov, Kristian, et al.
Veröffentlicht: (2024)
von: Kuznetsov, Kristian, et al.
Veröffentlicht: (2024)
Generalized Measures of Anticipation and Responsivity in Online Language Processing
von: Giulianelli, Mario, et al.
Veröffentlicht: (2024)
von: Giulianelli, Mario, et al.
Veröffentlicht: (2024)
An Information Theoretic Perspective on Agentic System Design
von: He, Shizhe, et al.
Veröffentlicht: (2025)
von: He, Shizhe, et al.
Veröffentlicht: (2025)
OneShield -- the Next Generation of LLM Guardrails
von: DeLuca, Chad, et al.
Veröffentlicht: (2025)
von: DeLuca, Chad, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Analyzing and Improving Chain-of-Thought Monitorability Through Information Theory
von: Anwar, Usman, et al.
Veröffentlicht: (2026) -
An Information-Theoretic Approach to Analyze NLP Classification Tasks
von: Wang, Luran, et al.
Veröffentlicht: (2024) -
MermaidSeqBench: An Evaluation Benchmark for NL-to-Mermaid Sequence Diagram Generation
von: Shbita, Basel, et al.
Veröffentlicht: (2025) -
LLMON: An LLM-native Markup Language to Leverage Structure and Semantics at the LLM Interface
von: Hind, Michael, et al.
Veröffentlicht: (2026) -
STaD: Scaffolded Task Design for Identifying Compositional Skill Gaps in LLMs
von: An, Sungeun, et al.
Veröffentlicht: (2026)