Linguistic Calibration of Long-Form Generations
Fuente:
arXiv
Salvato in:
| Autori principali: | Band, Neil, Li, Xuechen, Ma, Tengyu, Hashimoto, Tatsunori |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Reasoning to Learn from Latent Thoughts
di: Ruan, Yangjun, et al.
Pubblicazione: (2025)
di: Ruan, Yangjun, et al.
Pubblicazione: (2025)
Synthetic continued pretraining
di: Yang, Zitong, et al.
Pubblicazione: (2024)
di: Yang, Zitong, et al.
Pubblicazione: (2024)
Synthetic Data for any Differentiable Target
di: Thrush, Tristan, et al.
Pubblicazione: (2026)
di: Thrush, Tristan, et al.
Pubblicazione: (2026)
Language Models with Conformal Factuality Guarantees
di: Mohri, Christopher, et al.
Pubblicazione: (2024)
di: Mohri, Christopher, et al.
Pubblicazione: (2024)
Graph-based Uncertainty Metrics for Long-form Language Model Outputs
di: Jiang, Mingjian, et al.
Pubblicazione: (2024)
di: Jiang, Mingjian, et al.
Pubblicazione: (2024)
Observational Scaling Laws and the Predictability of Language Model Performance
di: Ruan, Yangjun, et al.
Pubblicazione: (2024)
di: Ruan, Yangjun, et al.
Pubblicazione: (2024)
AlpacaFarm: A Simulation Framework for Methods that Learn from Human Feedback
di: Dubois, Yann, et al.
Pubblicazione: (2023)
di: Dubois, Yann, et al.
Pubblicazione: (2023)
The Ideation-Execution Gap: Execution Outcomes of LLM-Generated versus Human Research Ideas
di: Si, Chenglei, et al.
Pubblicazione: (2025)
di: Si, Chenglei, et al.
Pubblicazione: (2025)
Can LLMs Generate Novel Research Ideas? A Large-Scale Human Study with 100+ NLP Researchers
di: Si, Chenglei, et al.
Pubblicazione: (2024)
di: Si, Chenglei, et al.
Pubblicazione: (2024)
Eliciting Language Model Behaviors with Investigator Agents
di: Li, Xiang Lisa, et al.
Pubblicazione: (2025)
di: Li, Xiang Lisa, et al.
Pubblicazione: (2025)
Length-Controlled AlpacaEval: A Simple Way to Debias Automatic Evaluators
di: Dubois, Yann, et al.
Pubblicazione: (2024)
di: Dubois, Yann, et al.
Pubblicazione: (2024)
Agentic Adversarial QA for Improving Domain-Specific LLMs
di: Grari, Vincent, et al.
Pubblicazione: (2026)
di: Grari, Vincent, et al.
Pubblicazione: (2026)
Towards Execution-Grounded Automated AI Research
di: Si, Chenglei, et al.
Pubblicazione: (2026)
di: Si, Chenglei, et al.
Pubblicazione: (2026)
Calibrating Long-form Generations from Large Language Models
di: Huang, Yukun, et al.
Pubblicazione: (2024)
di: Huang, Yukun, et al.
Pubblicazione: (2024)
Real-Time Detection of Hallucinated Entities in Long-Form Generation
di: Obeso, Oscar, et al.
Pubblicazione: (2025)
di: Obeso, Oscar, et al.
Pubblicazione: (2025)
Large Language Models as Tool Makers
di: Cai, Tianle, et al.
Pubblicazione: (2023)
di: Cai, Tianle, et al.
Pubblicazione: (2023)
IUQ: Interrogative Uncertainty Quantification for Long-Form Large Language Model Generation
di: Fan, Haozhi, et al.
Pubblicazione: (2026)
di: Fan, Haozhi, et al.
Pubblicazione: (2026)
s1: Simple test-time scaling
di: Muennighoff, Niklas, et al.
Pubblicazione: (2025)
di: Muennighoff, Niklas, et al.
Pubblicazione: (2025)
When Should LLMs Be Less Specific? Selective Abstraction for Reliable Long-Form Text Generation
di: Goren, Shani, et al.
Pubblicazione: (2026)
di: Goren, Shani, et al.
Pubblicazione: (2026)
Identifying the Risks of LM Agents with an LM-Emulated Sandbox
di: Ruan, Yangjun, et al.
Pubblicazione: (2023)
di: Ruan, Yangjun, et al.
Pubblicazione: (2023)
A comprehensive study of on-device NLP applications -- VQA, automated Form filling, Smart Replies for Linguistic Codeswitching
di: Goyal, Naman
Pubblicazione: (2024)
di: Goyal, Naman
Pubblicazione: (2024)
LongForm: Effective Instruction Tuning with Reverse Instructions
di: Köksal, Abdullatif, et al.
Pubblicazione: (2023)
di: Köksal, Abdullatif, et al.
Pubblicazione: (2023)
Embedding Trust: Semantic Isotropy Predicts Nonfactuality in Long-Form Text Generation
di: Bhardwaj, Dhrupad, et al.
Pubblicazione: (2025)
di: Bhardwaj, Dhrupad, et al.
Pubblicazione: (2025)
Long-Form Information Alignment Evaluation Beyond Atomic Facts
di: Zheng, Danna, et al.
Pubblicazione: (2025)
di: Zheng, Danna, et al.
Pubblicazione: (2025)
How Does Response Length Affect Long-Form Factuality
di: Zhao, James Xu, et al.
Pubblicazione: (2025)
di: Zhao, James Xu, et al.
Pubblicazione: (2025)
Found in the Middle: Calibrating Positional Attention Bias Improves Long Context Utilization
di: Hsieh, Cheng-Yu, et al.
Pubblicazione: (2024)
di: Hsieh, Cheng-Yu, et al.
Pubblicazione: (2024)
Learning to (Learn at Test Time): RNNs with Expressive Hidden States
di: Sun, Yu, et al.
Pubblicazione: (2024)
di: Sun, Yu, et al.
Pubblicazione: (2024)
Knowledge-Level Consistency Reinforcement Learning: Dual-Fact Alignment for Long-Form Factuality
di: Li, Junliang, et al.
Pubblicazione: (2025)
di: Li, Junliang, et al.
Pubblicazione: (2025)
RoleCraft-GLM: Advancing Personalized Role-Playing in Large Language Models
di: Tao, Meiling, et al.
Pubblicazione: (2023)
di: Tao, Meiling, et al.
Pubblicazione: (2023)
FSPO: Few-Shot Optimization of Synthetic Preferences Personalizes to Real Users
di: Singh, Anikait, et al.
Pubblicazione: (2025)
di: Singh, Anikait, et al.
Pubblicazione: (2025)
Fill In The Gaps: Model Calibration and Generalization with Synthetic Data
di: Ba, Yang, et al.
Pubblicazione: (2024)
di: Ba, Yang, et al.
Pubblicazione: (2024)
Calibrating Large Language Models Using Their Generations Only
di: Ulmer, Dennis, et al.
Pubblicazione: (2024)
di: Ulmer, Dennis, et al.
Pubblicazione: (2024)
On Subjective Uncertainty Quantification and Calibration in Natural Language Generation
di: Wang, Ziyu, et al.
Pubblicazione: (2024)
di: Wang, Ziyu, et al.
Pubblicazione: (2024)
Fine-Grained Uncertainty Quantification for Long-Form Language Model Outputs: A Comparative Study
di: Bouchard, Dylan, et al.
Pubblicazione: (2026)
di: Bouchard, Dylan, et al.
Pubblicazione: (2026)
Sectoral Coupling in Linguistic State Space
di: Dumbrava, Sebastian
Pubblicazione: (2025)
di: Dumbrava, Sebastian
Pubblicazione: (2025)
Atomic Calibration of LLMs in Long-Form Generations
di: Zhang, Caiqi, et al.
Pubblicazione: (2024)
di: Zhang, Caiqi, et al.
Pubblicazione: (2024)
LongWriter-Zero: Mastering Ultra-Long Text Generation via Reinforcement Learning
di: Wu, Yuhao, et al.
Pubblicazione: (2025)
di: Wu, Yuhao, et al.
Pubblicazione: (2025)
Calibration Across Layers: Understanding Calibration Evolution in LLMs
di: Joshi, Abhinav, et al.
Pubblicazione: (2025)
di: Joshi, Abhinav, et al.
Pubblicazione: (2025)
RLAC: Reinforcement Learning with Adversarial Critic for Free-Form Generation Tasks
di: Wu, Mian, et al.
Pubblicazione: (2025)
di: Wu, Mian, et al.
Pubblicazione: (2025)
Perceptions of Linguistic Uncertainty by Language Models and Humans
di: Belem, Catarina G, et al.
Pubblicazione: (2024)
di: Belem, Catarina G, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Reasoning to Learn from Latent Thoughts
di: Ruan, Yangjun, et al.
Pubblicazione: (2025) -
Synthetic continued pretraining
di: Yang, Zitong, et al.
Pubblicazione: (2024) -
Synthetic Data for any Differentiable Target
di: Thrush, Tristan, et al.
Pubblicazione: (2026) -
Language Models with Conformal Factuality Guarantees
di: Mohri, Christopher, et al.
Pubblicazione: (2024) -
Graph-based Uncertainty Metrics for Long-form Language Model Outputs
di: Jiang, Mingjian, et al.
Pubblicazione: (2024)