Do LLMs estimate uncertainty well in instruction-following?
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Heo, Juyeon, Xiong, Miao, Heinze-Deml, Christina, Narain, Jaya |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Do LLMs "know" internally when they follow instructions?
von: Heo, Juyeon, et al.
Veröffentlicht: (2024)
von: Heo, Juyeon, et al.
Veröffentlicht: (2024)
Reinforcement learning fine-tuning of language model for instruction following and math reasoning
von: Han, Yifu, et al.
Veröffentlicht: (2025)
von: Han, Yifu, et al.
Veröffentlicht: (2025)
Trustful LLMs: Customizing and Grounding Text Generation with Knowledge Bases and Dual Decoders
von: Zhu, Xiaofeng, et al.
Veröffentlicht: (2024)
von: Zhu, Xiaofeng, et al.
Veröffentlicht: (2024)
WizardLM: Empowering large pre-trained language models to follow complex instructions
von: Xu, Can, et al.
Veröffentlicht: (2023)
von: Xu, Can, et al.
Veröffentlicht: (2023)
Estimation of Concept Explanations Should be Uncertainty Aware
von: Piratla, Vihari, et al.
Veröffentlicht: (2023)
von: Piratla, Vihari, et al.
Veröffentlicht: (2023)
BLEUBERI: BLEU is a surprisingly effective reward for instruction following
von: Chang, Yapei, et al.
Veröffentlicht: (2025)
von: Chang, Yapei, et al.
Veröffentlicht: (2025)
Discerning minds or generic tutors? Evaluating instructional guidance capabilities in Socratic LLMs
von: Liu, Ying, et al.
Veröffentlicht: (2025)
von: Liu, Ying, et al.
Veröffentlicht: (2025)
Analysis of instruction-based LLMs' capabilities to score and judge text-input problems in an academic setting
von: Ramirez-Garcia, Valeria, et al.
Veröffentlicht: (2025)
von: Ramirez-Garcia, Valeria, et al.
Veröffentlicht: (2025)
Enabling robots to follow abstract instructions and complete complex dynamic tasks
von: Mon-Williams, Ruaridh, et al.
Veröffentlicht: (2024)
von: Mon-Williams, Ruaridh, et al.
Veröffentlicht: (2024)
How well can LLMs Grade Essays in Arabic?
von: Ghazawi, Rayed, et al.
Veröffentlicht: (2025)
von: Ghazawi, Rayed, et al.
Veröffentlicht: (2025)
Do LLMs Dream of Ontologies?
von: Bombieri, Marco, et al.
Veröffentlicht: (2024)
von: Bombieri, Marco, et al.
Veröffentlicht: (2024)
Do LLMs have Consistent Values?
von: Rozen, Naama, et al.
Veröffentlicht: (2024)
von: Rozen, Naama, et al.
Veröffentlicht: (2024)
How Do LLMs Use Their Depth?
von: Gupta, Akshat, et al.
Veröffentlicht: (2025)
von: Gupta, Akshat, et al.
Veröffentlicht: (2025)
Do as I Say, Not as I Do: Instruction-Induction Conflict in LLMs
von: Camassa, Carolina, et al.
Veröffentlicht: (2026)
von: Camassa, Carolina, et al.
Veröffentlicht: (2026)
Agent-based Automated Claim Matching with Instruction-following LLMs
von: Pisarevskaya, Dina, et al.
Veröffentlicht: (2025)
von: Pisarevskaya, Dina, et al.
Veröffentlicht: (2025)
LLMs Do Not Grade Essays Like Humans
von: Mathew, Jerin George, et al.
Veröffentlicht: (2026)
von: Mathew, Jerin George, et al.
Veröffentlicht: (2026)
Do LLMs Benefit From Their Own Words?
von: Huang, Jenny Y., et al.
Veröffentlicht: (2026)
von: Huang, Jenny Y., et al.
Veröffentlicht: (2026)
How well do LLMs cite relevant medical references? An evaluation framework and analyses
von: Wu, Kevin, et al.
Veröffentlicht: (2024)
von: Wu, Kevin, et al.
Veröffentlicht: (2024)
Do LLMs Agree on the Creativity Evaluation of Alternative Uses?
von: Rabeyah, Abdullah Al, et al.
Veröffentlicht: (2024)
von: Rabeyah, Abdullah Al, et al.
Veröffentlicht: (2024)
Do LLMs "Feel"? Emotion Circuits Discovery and Control
von: Wang, Chenxi, et al.
Veröffentlicht: (2025)
von: Wang, Chenxi, et al.
Veröffentlicht: (2025)
How Well Do LLMs Understand Tunisian Arabic?
von: Mahdi, Mohamed
Veröffentlicht: (2025)
von: Mahdi, Mohamed
Veröffentlicht: (2025)
ReadCtrl: Personalizing text generation with readability-controlled instruction learning
von: Tran, Hieu, et al.
Veröffentlicht: (2024)
von: Tran, Hieu, et al.
Veröffentlicht: (2024)
Do LLMs Have Distinct and Consistent Personality? TRAIT: Personality Testset designed for LLMs with Psychometrics
von: Lee, Seungbeen, et al.
Veröffentlicht: (2024)
von: Lee, Seungbeen, et al.
Veröffentlicht: (2024)
URL: Universal Referential Knowledge Linking via Task-instructed Representation Compression
von: Li, Zhuoqun, et al.
Veröffentlicht: (2024)
von: Li, Zhuoqun, et al.
Veröffentlicht: (2024)
Zero-shot cross-lingual transfer in instruction tuning of large language models
von: Chirkova, Nadezhda, et al.
Veröffentlicht: (2024)
von: Chirkova, Nadezhda, et al.
Veröffentlicht: (2024)
Do LLMs Really Adapt to Domains? An Ontology Learning Perspective
von: Mai, Huu Tan, et al.
Veröffentlicht: (2024)
von: Mai, Huu Tan, et al.
Veröffentlicht: (2024)
Do LLMs Really Think Step-by-step In Implicit Reasoning?
von: Yu, Yijiong
Veröffentlicht: (2024)
von: Yu, Yijiong
Veröffentlicht: (2024)
How Do LLMs Perform Two-Hop Reasoning in Context?
von: Guo, Tianyu, et al.
Veröffentlicht: (2025)
von: Guo, Tianyu, et al.
Veröffentlicht: (2025)
KGMEL: Knowledge Graph-Enhanced Multimodal Entity Linking
von: Kim, Juyeon, et al.
Veröffentlicht: (2025)
von: Kim, Juyeon, et al.
Veröffentlicht: (2025)
Do LLMs Exhibit Human-Like Reasoning? Evaluating Theory of Mind in LLMs for Open-Ended Responses
von: Amirizaniani, Maryam, et al.
Veröffentlicht: (2024)
von: Amirizaniani, Maryam, et al.
Veröffentlicht: (2024)
Re-Ex: Revising after Explanation Reduces the Factual Errors in LLM Responses
von: Kim, Juyeon, et al.
Veröffentlicht: (2024)
von: Kim, Juyeon, et al.
Veröffentlicht: (2024)
MedReadCtrl: Personalizing medical text generation with readability-controlled instruction learning
von: Tran, Hieu, et al.
Veröffentlicht: (2025)
von: Tran, Hieu, et al.
Veröffentlicht: (2025)
Do Compressed LLMs Forget Knowledge? An Experimental Study with Practical Implications
von: Hoang, Duc N. M, et al.
Veröffentlicht: (2023)
von: Hoang, Duc N. M, et al.
Veröffentlicht: (2023)
Caught in the Web of Words: Do LLMs Fall for Spin in Medical Literature?
von: Yun, Hye Sun, et al.
Veröffentlicht: (2025)
von: Yun, Hye Sun, et al.
Veröffentlicht: (2025)
How Context Shapes Truth: Geometric Transformations of Statement-level Truth Representations in LLMs
von: Adarsh, Shivam, et al.
Veröffentlicht: (2026)
von: Adarsh, Shivam, et al.
Veröffentlicht: (2026)
Do Large Language Models Mirror Cognitive Language Processing?
von: Ren, Yuqi, et al.
Veröffentlicht: (2024)
von: Ren, Yuqi, et al.
Veröffentlicht: (2024)
Not All Personas Are Worth It: Culture-Reflective Persona Data Augmentation
von: Han, Ji-Eun, et al.
Veröffentlicht: (2025)
von: Han, Ji-Eun, et al.
Veröffentlicht: (2025)
Zero-shot and Few-shot Learning with Instruction-following LLMs for Claim Matching in Automated Fact-checking
von: Pisarevskaya, Dina, et al.
Veröffentlicht: (2025)
von: Pisarevskaya, Dina, et al.
Veröffentlicht: (2025)
Do Hallucination Neurons Generalize? Evidence from Cross-Domain Transfer in LLMs
von: Vaddi, Snehit, et al.
Veröffentlicht: (2026)
von: Vaddi, Snehit, et al.
Veröffentlicht: (2026)
Do LLMs Triage Like Clinicians? A Dynamic Study of Outpatient Referral
von: Liu, Xiaoxiao, et al.
Veröffentlicht: (2025)
von: Liu, Xiaoxiao, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Do LLMs "know" internally when they follow instructions?
von: Heo, Juyeon, et al.
Veröffentlicht: (2024) -
Reinforcement learning fine-tuning of language model for instruction following and math reasoning
von: Han, Yifu, et al.
Veröffentlicht: (2025) -
Trustful LLMs: Customizing and Grounding Text Generation with Knowledge Bases and Dual Decoders
von: Zhu, Xiaofeng, et al.
Veröffentlicht: (2024) -
WizardLM: Empowering large pre-trained language models to follow complex instructions
von: Xu, Can, et al.
Veröffentlicht: (2023) -
Estimation of Concept Explanations Should be Uncertainty Aware
von: Piratla, Vihari, et al.
Veröffentlicht: (2023)