Capability Self-Assessment: Teaching LLMs to Know Their Limits
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Yang, Haoyan, Shirkavand, Reza, Jin, Yukai, Zhou, Jiawei, Gao, Shangqian, Huang, Heng |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Privacy-Preserving LLMs Routing
par: Wu, Xidong, et autres
Publié: (2026)
par: Wu, Xidong, et autres
Publié: (2026)
Cost-Aware Contrastive Routing for LLMs
par: Shirkavand, Reza, et autres
Publié: (2025)
par: Shirkavand, Reza, et autres
Publié: (2025)
Not All Prompts Are Made Equal: Prompt-based Pruning of Text-to-Image Diffusion Models
par: Ganjdanesh, Alireza, et autres
Publié: (2024)
par: Ganjdanesh, Alireza, et autres
Publié: (2024)
Learning to Plan Before Answering: Self-Teaching LLMs to Learn Abstract Plans for Problem Solving
par: Zhang, Jin, et autres
Publié: (2025)
par: Zhang, Jin, et autres
Publié: (2025)
GeoMotionGPT: Geometry-Aligned Motion Understanding with Large Language Models
par: Ye, Zhankai, et autres
Publié: (2026)
par: Ye, Zhankai, et autres
Publié: (2026)
Efficient Fine-Tuning and Concept Suppression for Pruned Diffusion Models
par: Shirkavand, Reza, et autres
Publié: (2024)
par: Shirkavand, Reza, et autres
Publié: (2024)
KnowRL: Teaching Language Models to Know What They Know
par: Kale, Sahil, et autres
Publié: (2025)
par: Kale, Sahil, et autres
Publié: (2025)
LLM Self Defense: By Self Examination, LLMs Know They Are Being Tricked
par: Phute, Mansi, et autres
Publié: (2023)
par: Phute, Mansi, et autres
Publié: (2023)
Knowing What LLMs DO NOT Know: A Simple Yet Effective Self-Detection Method
par: Zhao, Yukun, et autres
Publié: (2023)
par: Zhao, Yukun, et autres
Publié: (2023)
Evolving-RL: End-to-End Optimization of Experience-Driven Self-Evolving Capability within Agents
par: Fan, Zhiyuan, et autres
Publié: (2026)
par: Fan, Zhiyuan, et autres
Publié: (2026)
KnowCoder-A1: Incentivizing Agentic Reasoning Capability with Outcome Supervision for KBQA
par: Chen, Zhuo, et autres
Publié: (2025)
par: Chen, Zhuo, et autres
Publié: (2025)
SaySelf: Teaching LLMs to Express Confidence with Self-Reflective Rationales
par: Xu, Tianyang, et autres
Publié: (2024)
par: Xu, Tianyang, et autres
Publié: (2024)
MedReflect: Teaching Medical LLMs to Self-Improve via Reflective Correction
par: Huang, Yue, et autres
Publié: (2025)
par: Huang, Yue, et autres
Publié: (2025)
Do Large Language Models Know What They Are Capable Of?
par: Barkan, Casey O., et autres
Publié: (2025)
par: Barkan, Casey O., et autres
Publié: (2025)
Spilling the Beans: Teaching LLMs to Self-Report Their Hidden Objectives
par: Li, Chloe, et autres
Publié: (2025)
par: Li, Chloe, et autres
Publié: (2025)
KnowBias: Mitigating Social Bias in LLMs via Know-Bias Neuron Enhancement
par: Pan, Jinhao, et autres
Publié: (2026)
par: Pan, Jinhao, et autres
Publié: (2026)
SMART: Self-Generating and Self-Validating Multi-Dimensional Assessment for LLMs' Mathematical Problem Solving
par: Hou, Yujie, et autres
Publié: (2025)
par: Hou, Yujie, et autres
Publié: (2025)
Evaluating the Capabilities of LLMs for Supporting Anticipatory Impact Assessment
par: Allaham, Mowafak, et autres
Publié: (2024)
par: Allaham, Mowafak, et autres
Publié: (2024)
Are the Hidden States Hiding Something? Testing the Limits of Factuality-Encoding Capabilities in LLMs
par: Servedio, Giovanni, et autres
Publié: (2025)
par: Servedio, Giovanni, et autres
Publié: (2025)
Tracking the Limits of Knowledge Propagation: How LLMs Fail at Multi-Step Reasoning with Conflicting Knowledge
par: Feng, Yiyang, et autres
Publié: (2026)
par: Feng, Yiyang, et autres
Publié: (2026)
CaRT: Teaching LLM Agents to Know When They Know Enough
par: Liu, Grace, et autres
Publié: (2025)
par: Liu, Grace, et autres
Publié: (2025)
Do LLMs Know When to Flip a Coin? Strategic Randomization through Reasoning and Experience
par: Yang, Lingyu
Publié: (2025)
par: Yang, Lingyu
Publié: (2025)
LLMs Don't Know Their Own Decision Boundaries: The Unreliability of Self-Generated Counterfactual Explanations
par: Mayne, Harry, et autres
Publié: (2025)
par: Mayne, Harry, et autres
Publié: (2025)
Do Retrieval Augmented Language Models Know When They Don't Know?
par: Zhou, Youchao, et autres
Publié: (2025)
par: Zhou, Youchao, et autres
Publié: (2025)
On the Limitations and Capabilities of Position Embeddings for Length Generalization
par: Chen, Yang, et autres
Publié: (2025)
par: Chen, Yang, et autres
Publié: (2025)
PRO: Enabling Precise and Robust Text Watermark for Open-Source LLMs
par: Xue, Jiaqi, et autres
Publié: (2025)
par: Xue, Jiaqi, et autres
Publié: (2025)
Are Your LLMs Capable of Stable Reasoning?
par: Liu, Junnan, et autres
Publié: (2024)
par: Liu, Junnan, et autres
Publié: (2024)
InfiAlign: A Scalable and Sample-Efficient Framework for Aligning LLMs to Enhance Reasoning Capabilities
par: Cai, Shuo, et autres
Publié: (2025)
par: Cai, Shuo, et autres
Publié: (2025)
DeepInnovator: Triggering the Innovative Capabilities of LLMs
par: Fan, Tianyu, et autres
Publié: (2026)
par: Fan, Tianyu, et autres
Publié: (2026)
Evaluating LLMs When They Do Not Know the Answer: Statistical Evaluation of Mathematical Reasoning via Comparative Signals
par: Dong, Zihan, et autres
Publié: (2026)
par: Dong, Zihan, et autres
Publié: (2026)
Know When to Abstain: Optimal Selective Classification with Likelihood Ratios
par: Heng, Alvin, et autres
Publié: (2025)
par: Heng, Alvin, et autres
Publié: (2025)
KnowCoder: Coding Structured Knowledge into LLMs for Universal Information Extraction
par: Li, Zixuan, et autres
Publié: (2024)
par: Li, Zixuan, et autres
Publié: (2024)
All-in-One Tuning and Structural Pruning for Domain-Specific LLMs
par: Lu, Lei, et autres
Publié: (2024)
par: Lu, Lei, et autres
Publié: (2024)
Evaluating Developmental Cognition Capabilities of LLMs
par: Xiao, Xiao, et autres
Publié: (2026)
par: Xiao, Xiao, et autres
Publié: (2026)
Unleashing the Denoising Capability of Diffusion Prior for Solving Inverse Problems
par: Zhang, Jiawei, et autres
Publié: (2024)
par: Zhang, Jiawei, et autres
Publié: (2024)
Teaching LLMs to Ask: Self-Querying Category-Theoretic Planning for Under-Specified Reasoning
par: Qu, Shuhui
Publié: (2026)
par: Qu, Shuhui
Publié: (2026)
Observations on LLMs for Telecom Domain: Capabilities and Limitations
par: Soman, Sumit, et autres
Publié: (2023)
par: Soman, Sumit, et autres
Publié: (2023)
Tell Me What You Don't Know: Enhancing Refusal Capabilities of Role-Playing Agents via Representation Space Analysis and Editing
par: Liu, Wenhao, et autres
Publié: (2024)
par: Liu, Wenhao, et autres
Publié: (2024)
From Understanding to Utilization: A Survey on Explainability for Large Language Models
par: Luo, Haoyan, et autres
Publié: (2024)
par: Luo, Haoyan, et autres
Publié: (2024)
Tuning Language Models by Mixture-of-Depths Ensemble
par: Luo, Haoyan, et autres
Publié: (2024)
par: Luo, Haoyan, et autres
Publié: (2024)
Documents similaires
-
Privacy-Preserving LLMs Routing
par: Wu, Xidong, et autres
Publié: (2026) -
Cost-Aware Contrastive Routing for LLMs
par: Shirkavand, Reza, et autres
Publié: (2025) -
Not All Prompts Are Made Equal: Prompt-based Pruning of Text-to-Image Diffusion Models
par: Ganjdanesh, Alireza, et autres
Publié: (2024) -
Learning to Plan Before Answering: Self-Teaching LLMs to Learn Abstract Plans for Problem Solving
par: Zhang, Jin, et autres
Publié: (2025) -
GeoMotionGPT: Geometry-Aligned Motion Understanding with Large Language Models
par: Ye, Zhankai, et autres
Publié: (2026)