Calibrating Long-form Generations from Large Language Models
Fuente:
arXiv
Salvato in:
| Autori principali: | Huang, Yukun, Liu, Yixin, Thirukovalluru, Raghuveer, Cohan, Arman, Dhingra, Bhuwan |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
GenEOL: Harnessing the Generative Power of LLMs for Training-Free Sentence Embeddings
di: Thirukovalluru, Raghuveer, et al.
Pubblicazione: (2024)
di: Thirukovalluru, Raghuveer, et al.
Pubblicazione: (2024)
Atomic Self-Consistency for Better Long Form Generations
di: Thirukovalluru, Raghuveer, et al.
Pubblicazione: (2024)
di: Thirukovalluru, Raghuveer, et al.
Pubblicazione: (2024)
Cite Pretrain: Retrieval-Free Knowledge Attribution for Large Language Models
di: Huang, Yukun, et al.
Pubblicazione: (2025)
di: Huang, Yukun, et al.
Pubblicazione: (2025)
InData: Towards Secure Multi-Step, Tool-Based Data Analysis
di: K, Karthikeyan, et al.
Pubblicazione: (2025)
di: K, Karthikeyan, et al.
Pubblicazione: (2025)
Document-as-Image Representations Fall Short for Scientific Retrieval
di: Khalighinejad, Ghazal, et al.
Pubblicazione: (2026)
di: Khalighinejad, Ghazal, et al.
Pubblicazione: (2026)
On Evaluating LLM Alignment by Evaluating LLMs as Judges
di: Liu, Yixin, et al.
Pubblicazione: (2025)
di: Liu, Yixin, et al.
Pubblicazione: (2025)
When Greedy Wins: Emergent Exploitation Bias in Meta-Bandit LLM Training
di: Chen, Sanxing, et al.
Pubblicazione: (2025)
di: Chen, Sanxing, et al.
Pubblicazione: (2025)
To Trust or Not to Trust? Enhancing Large Language Models' Situated Faithfulness to External Contexts
di: Huang, Yukun, et al.
Pubblicazione: (2024)
di: Huang, Yukun, et al.
Pubblicazione: (2024)
Additive Large Language Models for Semi-Structured Text
di: K, Karthikeyan, et al.
Pubblicazione: (2025)
di: K, Karthikeyan, et al.
Pubblicazione: (2025)
Real-time Factuality Assessment from Adversarial Feedback
di: Chen, Sanxing, et al.
Pubblicazione: (2024)
di: Chen, Sanxing, et al.
Pubblicazione: (2024)
Hierarchical Multi-Label Classification of Online Vaccine Concerns
di: Zhu, Chloe Qinyu, et al.
Pubblicazione: (2024)
di: Zhu, Chloe Qinyu, et al.
Pubblicazione: (2024)
Atomic Consistency Preference Optimization for Long-Form Question Answering
di: Chen, Jingfeng, et al.
Pubblicazione: (2025)
di: Chen, Jingfeng, et al.
Pubblicazione: (2025)
Graph-based Confidence Calibration for Large Language Models
di: Li, Yukun, et al.
Pubblicazione: (2024)
di: Li, Yukun, et al.
Pubblicazione: (2024)
COMAL: A Convergent Meta-Algorithm for Aligning LLMs with General Preferences
di: Liu, Yixin, et al.
Pubblicazione: (2024)
di: Liu, Yixin, et al.
Pubblicazione: (2024)
Re-evaluating Automatic LLM System Ranking for Alignment with Human Preference
di: Gao, Mingqi, et al.
Pubblicazione: (2024)
di: Gao, Mingqi, et al.
Pubblicazione: (2024)
References Improve LLM Alignment in Non-Verifiable Domains
di: Shi, Kejian, et al.
Pubblicazione: (2026)
di: Shi, Kejian, et al.
Pubblicazione: (2026)
Beware of Calibration Data for Pruning Large Language Models
di: Ji, Yixin, et al.
Pubblicazione: (2024)
di: Ji, Yixin, et al.
Pubblicazione: (2024)
ClinStructor: AI-Powered Structuring of Unstructured Clinical Texts
di: K, Karthikeyan, et al.
Pubblicazione: (2025)
di: K, Karthikeyan, et al.
Pubblicazione: (2025)
Calibrating Large Language Models Using Their Generations Only
di: Ulmer, Dennis, et al.
Pubblicazione: (2024)
di: Ulmer, Dennis, et al.
Pubblicazione: (2024)
ReIFE: Re-evaluating Instruction-Following Evaluation
di: Liu, Yixin, et al.
Pubblicazione: (2024)
di: Liu, Yixin, et al.
Pubblicazione: (2024)
Revisiting Uncertainty Estimation and Calibration of Large Language Models
di: Tao, Linwei, et al.
Pubblicazione: (2025)
di: Tao, Linwei, et al.
Pubblicazione: (2025)
Understanding Reference Policies in Direct Preference Optimization
di: Liu, Yixin, et al.
Pubblicazione: (2024)
di: Liu, Yixin, et al.
Pubblicazione: (2024)
Examining Reasoning LLMs-as-Judges in Non-Verifiable LLM Post-Training
di: Liu, Yixin, et al.
Pubblicazione: (2026)
di: Liu, Yixin, et al.
Pubblicazione: (2026)
Adversarial Math Word Problem Generation
di: Xie, Roy, et al.
Pubblicazione: (2024)
di: Xie, Roy, et al.
Pubblicazione: (2024)
Coding Agents are Effective Long-Context Processors
di: Cao, Weili, et al.
Pubblicazione: (2026)
di: Cao, Weili, et al.
Pubblicazione: (2026)
Linguistic Calibration of Long-Form Generations
di: Band, Neil, et al.
Pubblicazione: (2024)
di: Band, Neil, et al.
Pubblicazione: (2024)
On Calibration of Large Language Models: From Response To Capability
di: Yang, Sin-Han, et al.
Pubblicazione: (2026)
di: Yang, Sin-Han, et al.
Pubblicazione: (2026)
Multi-group Uncertainty Quantification for Long-form Text Generation
di: Liu, Terrance, et al.
Pubblicazione: (2024)
di: Liu, Terrance, et al.
Pubblicazione: (2024)
Graph-based Uncertainty Metrics for Long-form Language Model Outputs
di: Jiang, Mingjian, et al.
Pubblicazione: (2024)
di: Jiang, Mingjian, et al.
Pubblicazione: (2024)
Constrained Entropic Unlearning: A Primal-Dual Framework for Large Language Models
di: Entesari, Taha, et al.
Pubblicazione: (2025)
di: Entesari, Taha, et al.
Pubblicazione: (2025)
IUQ: Interrogative Uncertainty Quantification for Long-Form Large Language Model Generation
di: Fan, Haozhi, et al.
Pubblicazione: (2026)
di: Fan, Haozhi, et al.
Pubblicazione: (2026)
Long-form factuality in large language models
di: Wei, Jerry, et al.
Pubblicazione: (2024)
di: Wei, Jerry, et al.
Pubblicazione: (2024)
LongLoRA: Efficient Fine-tuning of Long-Context Large Language Models
di: Chen, Yukang, et al.
Pubblicazione: (2023)
di: Chen, Yukang, et al.
Pubblicazione: (2023)
Unraveling Text Generation in LLMs: A Stochastic Differential Equation Approach
di: Zhang, Yukun
Pubblicazione: (2024)
di: Zhang, Yukun
Pubblicazione: (2024)
On the Entropy Calibration of Language Models
di: Cao, Steven, et al.
Pubblicazione: (2025)
di: Cao, Steven, et al.
Pubblicazione: (2025)
Enhancing Trust in Large Language Models via Uncertainty-Calibrated Fine-Tuning
di: Krishnan, Ranganath, et al.
Pubblicazione: (2024)
di: Krishnan, Ranganath, et al.
Pubblicazione: (2024)
EfficientLLM: Efficiency in Large Language Models
di: Yuan, Zhengqing, et al.
Pubblicazione: (2025)
di: Yuan, Zhengqing, et al.
Pubblicazione: (2025)
Uncertainty in Language Models: Assessment through Rank-Calibration
di: Huang, Xinmeng, et al.
Pubblicazione: (2024)
di: Huang, Xinmeng, et al.
Pubblicazione: (2024)
Over-Searching in Search-Augmented Large Language Models
di: Xie, Roy, et al.
Pubblicazione: (2026)
di: Xie, Roy, et al.
Pubblicazione: (2026)
Evaluating Morphological Compositional Generalization in Large Language Models
di: Ismayilzada, Mete, et al.
Pubblicazione: (2024)
di: Ismayilzada, Mete, et al.
Pubblicazione: (2024)
Documenti analoghi
-
GenEOL: Harnessing the Generative Power of LLMs for Training-Free Sentence Embeddings
di: Thirukovalluru, Raghuveer, et al.
Pubblicazione: (2024) -
Atomic Self-Consistency for Better Long Form Generations
di: Thirukovalluru, Raghuveer, et al.
Pubblicazione: (2024) -
Cite Pretrain: Retrieval-Free Knowledge Attribution for Large Language Models
di: Huang, Yukun, et al.
Pubblicazione: (2025) -
InData: Towards Secure Multi-Step, Tool-Based Data Analysis
di: K, Karthikeyan, et al.
Pubblicazione: (2025) -
Document-as-Image Representations Fall Short for Scientific Retrieval
di: Khalighinejad, Ghazal, et al.
Pubblicazione: (2026)