How Does Response Length Affect Long-Form Factuality
Fuente:
arXiv
Salvato in:
| Autori principali: | Zhao, James Xu, Liu, Jimmy Z. J., Hooi, Bryan, Ng, See-Kiong |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Test-Time Scaling in Reasoning Models Is Not Effective for Knowledge-Intensive Tasks Yet
di: Zhao, James Xu, et al.
Pubblicazione: (2025)
di: Zhao, James Xu, et al.
Pubblicazione: (2025)
Uncertainty of Thoughts: Uncertainty-Aware Planning Enhances Information Seeking in Large Language Models
di: Hu, Zhiyuan, et al.
Pubblicazione: (2024)
di: Hu, Zhiyuan, et al.
Pubblicazione: (2024)
Knowledge-Level Consistency Reinforcement Learning: Dual-Fact Alignment for Long-Form Factuality
di: Li, Junliang, et al.
Pubblicazione: (2025)
di: Li, Junliang, et al.
Pubblicazione: (2025)
Prompt Optimization with EASE? Efficient Ordering-aware Automated Selection of Exemplars
di: Wu, Zhaoxuan, et al.
Pubblicazione: (2024)
di: Wu, Zhaoxuan, et al.
Pubblicazione: (2024)
Use Your INSTINCT: INSTruction optimization for LLMs usIng Neural bandits Coupled with Transformers
di: Lin, Xiaoqiang, et al.
Pubblicazione: (2023)
di: Lin, Xiaoqiang, et al.
Pubblicazione: (2023)
Dipper: Diversity in Prompts for Producing Large Language Model Ensembles in Reasoning tasks
di: Lau, Gregory Kang Ruey, et al.
Pubblicazione: (2024)
di: Lau, Gregory Kang Ruey, et al.
Pubblicazione: (2024)
Geneshift: Impact of different scenario shift on Jailbreaking LLM
di: Wu, Tianyi, et al.
Pubblicazione: (2025)
di: Wu, Tianyi, et al.
Pubblicazione: (2025)
Integrating Time Series into LLMs via Multi-layer Steerable Embedding Fusion for Enhanced Forecasting
di: Chen, Zhuomin, et al.
Pubblicazione: (2025)
di: Chen, Zhuomin, et al.
Pubblicazione: (2025)
KnowHalu: Hallucination Detection via Multi-Form Knowledge Based Factual Checking
di: Zhang, Jiawei, et al.
Pubblicazione: (2024)
di: Zhang, Jiawei, et al.
Pubblicazione: (2024)
Beyond Prompt-Induced Lies: Investigating LLM Deception on Benign Prompts
di: Wu, Zhaomin, et al.
Pubblicazione: (2025)
di: Wu, Zhaomin, et al.
Pubblicazione: (2025)
Long-Form Information Alignment Evaluation Beyond Atomic Facts
di: Zheng, Danna, et al.
Pubblicazione: (2025)
di: Zheng, Danna, et al.
Pubblicazione: (2025)
Stress Testing Factual Consistency Metrics for Long-Document Summarization
di: Mujahid, Zain Muhammad, et al.
Pubblicazione: (2025)
di: Mujahid, Zain Muhammad, et al.
Pubblicazione: (2025)
Ferret: Federated Full-Parameter Tuning at Scale for Large Language Models
di: Shu, Yao, et al.
Pubblicazione: (2024)
di: Shu, Yao, et al.
Pubblicazione: (2024)
Towards a Holistic Evaluation of LLMs on Factual Knowledge Recall
di: Yuan, Jiaqing, et al.
Pubblicazione: (2024)
di: Yuan, Jiaqing, et al.
Pubblicazione: (2024)
Balancing Truthfulness and Informativeness with Uncertainty-Aware Instruction Fine-Tuning
di: Wu, Tianyi, et al.
Pubblicazione: (2025)
di: Wu, Tianyi, et al.
Pubblicazione: (2025)
Contextual Drag: How Errors in the Context Affect LLM Reasoning
di: Cheng, Yun, et al.
Pubblicazione: (2026)
di: Cheng, Yun, et al.
Pubblicazione: (2026)
ActiveDPO: Active Direct Preference Optimization for Sample-Efficient Alignment
di: Lin, Xiaoqiang, et al.
Pubblicazione: (2025)
di: Lin, Xiaoqiang, et al.
Pubblicazione: (2025)
On Newton's Method to Unlearn Neural Networks
di: Bui, Nhung, et al.
Pubblicazione: (2024)
di: Bui, Nhung, et al.
Pubblicazione: (2024)
Prompt Optimization with Human Feedback
di: Lin, Xiaoqiang, et al.
Pubblicazione: (2024)
di: Lin, Xiaoqiang, et al.
Pubblicazione: (2024)
Towards A Unified View of Answer Calibration for Multi-Step Reasoning
di: Deng, Shumin, et al.
Pubblicazione: (2023)
di: Deng, Shumin, et al.
Pubblicazione: (2023)
Multi-Modal One-Shot Federated Ensemble Learning for Medical Data with Vision Large Language Model
di: Wang, Naibo, et al.
Pubblicazione: (2025)
di: Wang, Naibo, et al.
Pubblicazione: (2025)
Evaluating the Paperclip Maximizer: Are RL-Based Language Models More Likely to Pursue Instrumental Goals?
di: He, Yufei, et al.
Pubblicazione: (2025)
di: He, Yufei, et al.
Pubblicazione: (2025)
Promote, Suppress, Iterate: How Language Models Answer One-to-Many Factual Queries
di: Yan, Tianyi Lorena, et al.
Pubblicazione: (2025)
di: Yan, Tianyi Lorena, et al.
Pubblicazione: (2025)
Source Attribution for Large Language Model-Generated Data
di: Wang, Jingtan, et al.
Pubblicazione: (2023)
di: Wang, Jingtan, et al.
Pubblicazione: (2023)
Spatio-Temporal Foundation Models: Vision, Challenges, and Opportunities
di: Goodge, Adam, et al.
Pubblicazione: (2025)
di: Goodge, Adam, et al.
Pubblicazione: (2025)
Linguistic Calibration of Long-Form Generations
di: Band, Neil, et al.
Pubblicazione: (2024)
di: Band, Neil, et al.
Pubblicazione: (2024)
IUQ: Interrogative Uncertainty Quantification for Long-Form Large Language Model Generation
di: Fan, Haozhi, et al.
Pubblicazione: (2026)
di: Fan, Haozhi, et al.
Pubblicazione: (2026)
Evaluating the Factuality of Large Language Models using Large-Scale Knowledge Graphs
di: Liu, Xiaoze, et al.
Pubblicazione: (2024)
di: Liu, Xiaoze, et al.
Pubblicazione: (2024)
ThinkTank-ME: A Multi-Expert Framework for Middle East Event Forecasting
di: Li, Haoxuan, et al.
Pubblicazione: (2026)
di: Li, Haoxuan, et al.
Pubblicazione: (2026)
Integrative Decoding: Improve Factuality via Implicit Self-consistency
di: Cheng, Yi, et al.
Pubblicazione: (2024)
di: Cheng, Yi, et al.
Pubblicazione: (2024)
Seeing is Believing: Mitigating Hallucination in Large Vision-Language Models via CLIP-Guided Decoding
di: Deng, Ailin, et al.
Pubblicazione: (2024)
di: Deng, Ailin, et al.
Pubblicazione: (2024)
LongDocFACTScore: Evaluating the Factuality of Long Document Abstractive Summarisation
di: Bishop, Jennifer A, et al.
Pubblicazione: (2023)
di: Bishop, Jennifer A, et al.
Pubblicazione: (2023)
Information Extraction in Low-Resource Scenarios: Survey and Perspective
di: Deng, Shumin, et al.
Pubblicazione: (2022)
di: Deng, Shumin, et al.
Pubblicazione: (2022)
Language Models with Conformal Factuality Guarantees
di: Mohri, Christopher, et al.
Pubblicazione: (2024)
di: Mohri, Christopher, et al.
Pubblicazione: (2024)
How Numerical Precision Affects Arithmetical Reasoning Capabilities of LLMs
di: Feng, Guhao, et al.
Pubblicazione: (2024)
di: Feng, Guhao, et al.
Pubblicazione: (2024)
Uncovering Scaling Laws for Large Language Models via Inverse Problems
di: Verma, Arun, et al.
Pubblicazione: (2025)
di: Verma, Arun, et al.
Pubblicazione: (2025)
DoLa: Decoding by Contrasting Layers Improves Factuality in Large Language Models
di: Chuang, Yung-Sung, et al.
Pubblicazione: (2023)
di: Chuang, Yung-Sung, et al.
Pubblicazione: (2023)
BARREL: Boundary-Aware Reasoning for Factual and Reliable LRMs
di: Yang, Junxiao, et al.
Pubblicazione: (2025)
di: Yang, Junxiao, et al.
Pubblicazione: (2025)
Rating Quality of Diverse Time Series Data by Meta-learning from LLM Judgment
di: Wu, Shunyu, et al.
Pubblicazione: (2025)
di: Wu, Shunyu, et al.
Pubblicazione: (2025)
How Abilities in Large Language Models are Affected by Supervised Fine-tuning Data Composition
di: Dong, Guanting, et al.
Pubblicazione: (2023)
di: Dong, Guanting, et al.
Pubblicazione: (2023)
Documenti analoghi
-
Test-Time Scaling in Reasoning Models Is Not Effective for Knowledge-Intensive Tasks Yet
di: Zhao, James Xu, et al.
Pubblicazione: (2025) -
Uncertainty of Thoughts: Uncertainty-Aware Planning Enhances Information Seeking in Large Language Models
di: Hu, Zhiyuan, et al.
Pubblicazione: (2024) -
Knowledge-Level Consistency Reinforcement Learning: Dual-Fact Alignment for Long-Form Factuality
di: Li, Junliang, et al.
Pubblicazione: (2025) -
Prompt Optimization with EASE? Efficient Ordering-aware Automated Selection of Exemplars
di: Wu, Zhaoxuan, et al.
Pubblicazione: (2024) -
Use Your INSTINCT: INSTruction optimization for LLMs usIng Neural bandits Coupled with Transformers
di: Lin, Xiaoqiang, et al.
Pubblicazione: (2023)