Unused information in token probability distribution of generative LLM: improving LLM reading comprehension through calculation of expected values
Fuente:
arXiv
Enregistré dans:
| Auteur principal: | Zawistowski, Krystian |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Not too long do read: Evaluating LLM-generated extreme scientific summaries
par: Lyu, Zhuoqi, et autres
Publié: (2025)
par: Lyu, Zhuoqi, et autres
Publié: (2025)
Prediction hubs are context-informed frequent tokens in LLMs
par: Nielsen, Beatrix M. G., et autres
Publié: (2025)
par: Nielsen, Beatrix M. G., et autres
Publié: (2025)
AnomaLLMy -- Detecting anomalous tokens in black-box LLMs through low-confidence single-token predictions
par: Witold, Waligóra
Publié: (2024)
par: Witold, Waligóra
Publié: (2024)
LLMzSzŁ: a comprehensive LLM benchmark for Polish
par: Jassem, Krzysztof, et autres
Publié: (2025)
par: Jassem, Krzysztof, et autres
Publié: (2025)
LLM generation novelty through the lens of semantic similarity
par: Davydov, Philipp, et autres
Publié: (2025)
par: Davydov, Philipp, et autres
Publié: (2025)
A comprehensive study of LLM-based argument classification: from Llama through DeepSeek to GPT-5.2
par: Pietroń, Marcin, et autres
Publié: (2026)
par: Pietroń, Marcin, et autres
Publié: (2026)
Not all tokens are created equal: Perplexity Attention Weighted Networks for AI generated text detection
par: Miralles-González, Pablo, et autres
Publié: (2025)
par: Miralles-González, Pablo, et autres
Publié: (2025)
An evaluation of LLM code generation capabilities through graded exercises
par: Jiménez, Álvaro Barbero
Publié: (2024)
par: Jiménez, Álvaro Barbero
Publié: (2024)
A comprehensive study of LLM-based argument classification: from LLAMA through GPT-4o to Deepseek-R1
par: Pietroń, Marcin, et autres
Publié: (2025)
par: Pietroń, Marcin, et autres
Publié: (2025)
Gender Bias in LLM-generated Interview Responses
par: Kong, Haein, et autres
Publié: (2024)
par: Kong, Haein, et autres
Publié: (2024)
Learning to Summarize from LLM-generated Feedback
par: Song, Hwanjun, et autres
Publié: (2024)
par: Song, Hwanjun, et autres
Publié: (2024)
Benchmark of stylistic variation in LLM-generated texts
par: Milička, Jiří, et autres
Publié: (2025)
par: Milička, Jiří, et autres
Publié: (2025)
Semantic uncertainty in advanced decoding methods for LLM generation
par: Foodeei, Darius, et autres
Publié: (2025)
par: Foodeei, Darius, et autres
Publié: (2025)
Multi-Agent LLM Judge: automatic personalized LLM judge design for evaluating natural language generation applications
par: Cao, Hongliu, et autres
Publié: (2025)
par: Cao, Hongliu, et autres
Publié: (2025)
Jacobian Scopes: token-level causal attributions in LLMs
par: Liu, Toni J. B., et autres
Publié: (2026)
par: Liu, Toni J. B., et autres
Publié: (2026)
AutoHarness: improving LLM agents by automatically synthesizing a code harness
par: Lou, Xinghua, et autres
Publié: (2026)
par: Lou, Xinghua, et autres
Publié: (2026)
Pragmatic Reasoning improves LLM Code Generation
par: Cao, Zhuchen, et autres
Publié: (2025)
par: Cao, Zhuchen, et autres
Publié: (2025)
CLAWS:Creativity detection for LLM-generated solutions using Attention Window of Sections
par: Kim, Keuntae, et autres
Publié: (2025)
par: Kim, Keuntae, et autres
Publié: (2025)
LLM-as-a-qualitative-judge: automating error analysis in natural language generation
par: Chirkova, Nadezhda, et autres
Publié: (2025)
par: Chirkova, Nadezhda, et autres
Publié: (2025)
Beyond speculation: Measuring the growing presence of LLM-generated texts in multilingual disinformation
par: Macko, Dominik, et autres
Publié: (2025)
par: Macko, Dominik, et autres
Publié: (2025)
The pitfalls of next-token prediction
par: Bachmann, Gregor, et autres
Publié: (2024)
par: Bachmann, Gregor, et autres
Publié: (2024)
Looking beyond the next token
par: Thankaraj, Abitha, et autres
Publié: (2025)
par: Thankaraj, Abitha, et autres
Publié: (2025)
Strategist: Self-improvement of LLM Decision Making via Bi-Level Tree Search
par: Light, Jonathan, et autres
Publié: (2024)
par: Light, Jonathan, et autres
Publié: (2024)
Cross-cultural Inspiration Detection and Analysis in Real and LLM-generated Social Media Data
par: Ignat, Oana, et autres
Publié: (2024)
par: Ignat, Oana, et autres
Publié: (2024)
The "LLM World of Words" English free association norms generated by large language models
par: Abramski, Katherine, et autres
Publié: (2024)
par: Abramski, Katherine, et autres
Publié: (2024)
Temperature-scaling surprisal estimates improve fit to human reading times -- but does it do so for the "right reasons"?
par: Liu, Tong, et autres
Publié: (2023)
par: Liu, Tong, et autres
Publié: (2023)
Improving LLM Reasoning through Interpretable Role-Playing Steering
par: Wang, Anyi, et autres
Publié: (2025)
par: Wang, Anyi, et autres
Publié: (2025)
Enhancing Debunking Effectiveness through LLM-based Personality Adaptation
par: Dell'Oglio, Pietro, et autres
Publié: (2026)
par: Dell'Oglio, Pietro, et autres
Publié: (2026)
Characterizing Delusional Spirals through Human-LLM Chat Logs
par: Moore, Jared, et autres
Publié: (2026)
par: Moore, Jared, et autres
Publié: (2026)
On the token distance modeling ability of higher RoPE attention dimension
par: Hong, Xiangyu, et autres
Publié: (2024)
par: Hong, Xiangyu, et autres
Publié: (2024)
How good is my story? Towards quantitative metrics for evaluating LLM-generated XAI narratives
par: Ichmoukhamedov, Timour, et autres
Publié: (2024)
par: Ichmoukhamedov, Timour, et autres
Publié: (2024)
MEDSAGE: Enhancing Robustness of Medical Dialogue Summarization to ASR Errors with LLM-generated Synthetic Dialogues
par: Binici, Kuluhan, et autres
Publié: (2024)
par: Binici, Kuluhan, et autres
Publié: (2024)
Calibrating Pre-trained Language Classifiers on LLM-generated Noisy Labels via Iterative Refinement
par: Ye, Liqin, et autres
Publié: (2025)
par: Ye, Liqin, et autres
Publié: (2025)
The Lucie-7B LLM and the Lucie Training Dataset: Open resources for multilingual language generation
par: Gouvert, Olivier, et autres
Publié: (2025)
par: Gouvert, Olivier, et autres
Publié: (2025)
GameArena: Evaluating LLM Reasoning through Live Computer Games
par: Hu, Lanxiang, et autres
Publié: (2024)
par: Hu, Lanxiang, et autres
Publié: (2024)
Memorization or Interpolation ? Detecting LLM Memorization through Input Perturbation Analysis
par: Djiré, Albérick Euraste, et autres
Publié: (2025)
par: Djiré, Albérick Euraste, et autres
Publié: (2025)
LLM CHESS: Benchmarking Reasoning and Instruction-Following in LLMs through Chess
par: Kolasani, Sai, et autres
Publié: (2025)
par: Kolasani, Sai, et autres
Publié: (2025)
Safety Compliance: Rethinking LLM Safety Reasoning through the Lens of Compliance
par: Hu, Wenbin, et autres
Publié: (2025)
par: Hu, Wenbin, et autres
Publié: (2025)
Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context
par: Gemini Team, et autres
Publié: (2024)
par: Gemini Team, et autres
Publié: (2024)
Scaled and Inter-token Relation Enhanced Transformer for Sample-restricted Residential NILM
par: Rahman, Minhajur, et autres
Publié: (2024)
par: Rahman, Minhajur, et autres
Publié: (2024)
Documents similaires
-
Not too long do read: Evaluating LLM-generated extreme scientific summaries
par: Lyu, Zhuoqi, et autres
Publié: (2025) -
Prediction hubs are context-informed frequent tokens in LLMs
par: Nielsen, Beatrix M. G., et autres
Publié: (2025) -
AnomaLLMy -- Detecting anomalous tokens in black-box LLMs through low-confidence single-token predictions
par: Witold, Waligóra
Publié: (2024) -
LLMzSzŁ: a comprehensive LLM benchmark for Polish
par: Jassem, Krzysztof, et autres
Publié: (2025) -
LLM generation novelty through the lens of semantic similarity
par: Davydov, Philipp, et autres
Publié: (2025)