Factual Inconsistency in Data-to-Text Generation Scales Exponentially with LLM Size: A Statistical Validation
Fuente:
arXiv
Salvato in:
| Autori principali: | Mahapatra, Joy, Roy, Soumyajit, Garain, Utpal |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
An Extensive Evaluation of Factual Consistency in Large Language Models for Data-to-Text Generation
di: Mahapatra, Joy, et al.
Pubblicazione: (2024)
di: Mahapatra, Joy, et al.
Pubblicazione: (2024)
Impact of Model Size on Fine-tuned LLM Performance in Data-to-Text Generation: A State-of-the-Art Investigation
di: Mahapatra, Joy, et al.
Pubblicazione: (2024)
di: Mahapatra, Joy, et al.
Pubblicazione: (2024)
Predictable Confabulations: Factual Recall by LLMs Scales with Model Size and Topic Frequency
di: Smith, Matthew L., et al.
Pubblicazione: (2026)
di: Smith, Matthew L., et al.
Pubblicazione: (2026)
Consistency Is the Key: Detecting Hallucinations in LLM Generated Text By Checking Inconsistencies About Key Facts
di: Gupta, Raavi, et al.
Pubblicazione: (2025)
di: Gupta, Raavi, et al.
Pubblicazione: (2025)
Region Mixup
di: Saha, Saptarshi, et al.
Pubblicazione: (2024)
di: Saha, Saptarshi, et al.
Pubblicazione: (2024)
AdaDetectGPT: Adaptive Detection of LLM-Generated Text with Statistical Guarantees
di: Zhou, Hongyi, et al.
Pubblicazione: (2025)
di: Zhou, Hongyi, et al.
Pubblicazione: (2025)
RCStat: A Statistical Framework for using Relative Contextualization in Transformers
di: Mahapatra, Debabrata, et al.
Pubblicazione: (2025)
di: Mahapatra, Debabrata, et al.
Pubblicazione: (2025)
Cyclic Counterfactuals under Shift-Scale Interventions
di: Saha, Saptarshi, et al.
Pubblicazione: (2025)
di: Saha, Saptarshi, et al.
Pubblicazione: (2025)
Leveraging LLM Inconsistency to Boost Pass@k Performance
di: Dalal, Uri, et al.
Pubblicazione: (2025)
di: Dalal, Uri, et al.
Pubblicazione: (2025)
Factual Self-Awareness in Language Models: Representation, Robustness, and Scaling
di: Tamoyan, Hovhannes, et al.
Pubblicazione: (2025)
di: Tamoyan, Hovhannes, et al.
Pubblicazione: (2025)
ALHD: A Large-Scale and Multigenre Benchmark Dataset for Arabic LLM-Generated Text Detection
di: Khairallah, Ali, et al.
Pubblicazione: (2025)
di: Khairallah, Ali, et al.
Pubblicazione: (2025)
Evaluating the Factuality of Large Language Models using Large-Scale Knowledge Graphs
di: Liu, Xiaoze, et al.
Pubblicazione: (2024)
di: Liu, Xiaoze, et al.
Pubblicazione: (2024)
sudoLLM: On Multi-role Alignment of Language Models
di: Saha, Soumadeep, et al.
Pubblicazione: (2025)
di: Saha, Soumadeep, et al.
Pubblicazione: (2025)
On the Detectability of LLM-Generated Text: What Exactly Is LLM-Generated Text?
di: Geng, Mingmeng, et al.
Pubblicazione: (2025)
di: Geng, Mingmeng, et al.
Pubblicazione: (2025)
Power Lines: Scaling Laws for Weight Decay and Batch Size in LLM Pre-training
di: Bergsma, Shane, et al.
Pubblicazione: (2025)
di: Bergsma, Shane, et al.
Pubblicazione: (2025)
Adaptive Conformal Prediction for Improving Factuality of Generations by Large Language Models
di: Rubashevskii, Aleksandr, et al.
Pubblicazione: (2026)
di: Rubashevskii, Aleksandr, et al.
Pubblicazione: (2026)
Language Models with Conformal Factuality Guarantees
di: Mohri, Christopher, et al.
Pubblicazione: (2024)
di: Mohri, Christopher, et al.
Pubblicazione: (2024)
FECT: Factuality Evaluation of Interpretive AI-Generated Claims in Contact Center Conversation Transcripts
di: Shin, Hagyeong, et al.
Pubblicazione: (2025)
di: Shin, Hagyeong, et al.
Pubblicazione: (2025)
Context-Enhanced Contrastive Search for Improved LLM Text Generation
di: Sen, Jaydip, et al.
Pubblicazione: (2025)
di: Sen, Jaydip, et al.
Pubblicazione: (2025)
Learn-to-Distance: Distance Learning for Detecting LLM-Generated Text
di: Zhou, Hongyi, et al.
Pubblicazione: (2026)
di: Zhou, Hongyi, et al.
Pubblicazione: (2026)
HU at SemEval-2024 Task 8A: Can Contrastive Learning Learn Embeddings to Detect Machine-Generated Text?
di: Dipta, Shubhashis Roy, et al.
Pubblicazione: (2024)
di: Dipta, Shubhashis Roy, et al.
Pubblicazione: (2024)
Factuality Challenges in the Era of Large Language Models
di: Augenstein, Isabelle, et al.
Pubblicazione: (2023)
di: Augenstein, Isabelle, et al.
Pubblicazione: (2023)
Prompting Test-Time Scaling Is A Strong LLM Reasoning Data Augmentation
di: Bsharat, Sondos Mahmoud, et al.
Pubblicazione: (2025)
di: Bsharat, Sondos Mahmoud, et al.
Pubblicazione: (2025)
DPIC: Decoupling Prompt and Intrinsic Characteristics for LLM Generated Text Detection
di: Yu, Xiao, et al.
Pubblicazione: (2023)
di: Yu, Xiao, et al.
Pubblicazione: (2023)
Text2Data: Low-Resource Data Generation with Textual Control
di: Wang, Shiyu, et al.
Pubblicazione: (2024)
di: Wang, Shiyu, et al.
Pubblicazione: (2024)
What is Your Data Worth to GPT? LLM-Scale Data Valuation with Influence Functions
di: Choe, Sang Keun, et al.
Pubblicazione: (2024)
di: Choe, Sang Keun, et al.
Pubblicazione: (2024)
BARREL: Boundary-Aware Reasoning for Factual and Reliable LRMs
di: Yang, Junxiao, et al.
Pubblicazione: (2025)
di: Yang, Junxiao, et al.
Pubblicazione: (2025)
Less is More for Improving Automatic Evaluation of Factual Consistency
di: Wang, Tong, et al.
Pubblicazione: (2024)
di: Wang, Tong, et al.
Pubblicazione: (2024)
SIFiD: Reassess Summary Factual Inconsistency Detection with LLM
di: Yang, Jiuding, et al.
Pubblicazione: (2024)
di: Yang, Jiuding, et al.
Pubblicazione: (2024)
RAC: Efficient LLM Factuality Correction with Retrieval Augmentation
di: Li, Changmao, et al.
Pubblicazione: (2024)
di: Li, Changmao, et al.
Pubblicazione: (2024)
Understanding Scaling Laws with Statistical and Approximation Theory for Transformer Neural Networks on Intrinsically Low-dimensional Data
di: Havrilla, Alex, et al.
Pubblicazione: (2024)
di: Havrilla, Alex, et al.
Pubblicazione: (2024)
IPAD: Inverse Prompt for AI Detection - A Robust and Interpretable LLM-Generated Text Detector
di: Chen, Zheng, et al.
Pubblicazione: (2025)
di: Chen, Zheng, et al.
Pubblicazione: (2025)
Beyond LLM-as-a-Judge: Deterministic Metrics for Multilingual Generative Text Evaluation
di: Alam, Firoj, et al.
Pubblicazione: (2026)
di: Alam, Firoj, et al.
Pubblicazione: (2026)
You Can Generate It Again: Data-to-Text Generation with Verification and Correction Prompting
di: Ren, Xuan, et al.
Pubblicazione: (2023)
di: Ren, Xuan, et al.
Pubblicazione: (2023)
Inconsistent Tokenizations Cause Language Models to be Perplexed by Japanese Grammar
di: Gambardella, Andrew, et al.
Pubblicazione: (2025)
di: Gambardella, Andrew, et al.
Pubblicazione: (2025)
Text Detoxification: Data Efficiency, Semantic Preservation and Model Generalization
di: Yu, Jing, et al.
Pubblicazione: (2025)
di: Yu, Jing, et al.
Pubblicazione: (2025)
Generating Pretraining Tokens from Organic Data for Data-Bound Scaling
di: Yu, Zichun, et al.
Pubblicazione: (2026)
di: Yu, Zichun, et al.
Pubblicazione: (2026)
Attention Satisfies: A Constraint-Satisfaction Lens on Factual Errors of Language Models
di: Yuksekgonul, Mert, et al.
Pubblicazione: (2023)
di: Yuksekgonul, Mert, et al.
Pubblicazione: (2023)
Stress Testing Factual Consistency Metrics for Long-Document Summarization
di: Mujahid, Zain Muhammad, et al.
Pubblicazione: (2025)
di: Mujahid, Zain Muhammad, et al.
Pubblicazione: (2025)
How Does Response Length Affect Long-Form Factuality
di: Zhao, James Xu, et al.
Pubblicazione: (2025)
di: Zhao, James Xu, et al.
Pubblicazione: (2025)
Documenti analoghi
-
An Extensive Evaluation of Factual Consistency in Large Language Models for Data-to-Text Generation
di: Mahapatra, Joy, et al.
Pubblicazione: (2024) -
Impact of Model Size on Fine-tuned LLM Performance in Data-to-Text Generation: A State-of-the-Art Investigation
di: Mahapatra, Joy, et al.
Pubblicazione: (2024) -
Predictable Confabulations: Factual Recall by LLMs Scales with Model Size and Topic Frequency
di: Smith, Matthew L., et al.
Pubblicazione: (2026) -
Consistency Is the Key: Detecting Hallucinations in LLM Generated Text By Checking Inconsistencies About Key Facts
di: Gupta, Raavi, et al.
Pubblicazione: (2025) -
Region Mixup
di: Saha, Saptarshi, et al.
Pubblicazione: (2024)