Do not Abstain! Identify and Solve the Uncertainty
Fuente:
arXiv
Saved in:
| Main Authors: | Liu, Jingyu, Peng, Jingquan, Wu, xiaopeng, Li, Xubin, Ge, Tiezheng, Zheng, Bo, Liu, Yong |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Gradient Coupling: The Hidden Barrier to Generalization in Agentic Reinforcement Learning
by: Liu, Jingyu, et al.
Published: (2025)
by: Liu, Jingyu, et al.
Published: (2025)
Knowing When to Abstain: Medical LLMs Under Clinical Uncertainty
by: Machcha, Sravanthi, et al.
Published: (2026)
by: Machcha, Sravanthi, et al.
Published: (2026)
Abstaining Machine Learning -- Philosophical Considerations
by: Schuster, Daniela
Published: (2024)
by: Schuster, Daniela
Published: (2024)
Interpretable and Fair Mechanisms for Abstaining Classifiers
by: Lenders, Daphne, et al.
Published: (2025)
by: Lenders, Daphne, et al.
Published: (2025)
The Confidence Gate Theorem: When Should Ranked Decision Systems Abstain?
by: Doku, Ronald
Published: (2026)
by: Doku, Ronald
Published: (2026)
MT-Bench-101: A Fine-Grained Benchmark for Evaluating Large Language Models in Multi-Turn Dialogues
by: Bai, Ge, et al.
Published: (2024)
by: Bai, Ge, et al.
Published: (2024)
Know When to Abstain: Optimal Selective Classification with Likelihood Ratios
by: Heng, Alvin, et al.
Published: (2025)
by: Heng, Alvin, et al.
Published: (2025)
Logic-of-Thought: Empowering Large Language Models with Logic Programs for Solving Puzzles in Natural Language
by: Li, Naiqi, et al.
Published: (2025)
by: Li, Naiqi, et al.
Published: (2025)
Detecting Multimodal Situations with Insufficient Context and Abstaining from Baseless Predictions
by: Liu, Junzhang, et al.
Published: (2024)
by: Liu, Junzhang, et al.
Published: (2024)
How Much Can RAG Help the Reasoning of LLM?
by: Liu, Jingyu, et al.
Published: (2024)
by: Liu, Jingyu, et al.
Published: (2024)
Perfect Alignment May be Poisonous to Graph Contrastive Learning
by: Liu, Jingyu, et al.
Published: (2023)
by: Liu, Jingyu, et al.
Published: (2023)
Tackling the Inherent Difficulty of Noise Filtering in RAG
by: Liu, Jingyu, et al.
Published: (2026)
by: Liu, Jingyu, et al.
Published: (2026)
CausalAbstain: Enhancing Multilingual LLMs with Causal Reasoning for Trustworthy Abstention
by: Sun, Yuxi, et al.
Published: (2025)
by: Sun, Yuxi, et al.
Published: (2025)
Teaching LLMs to Abstain via Fine-Grained Semantic Confidence Reward
by: An, Hao, et al.
Published: (2025)
by: An, Hao, et al.
Published: (2025)
Clarify, Abstain or Answer? Strategising in Conversation with Belief-Augmented Generation
by: Baan, Joris, et al.
Published: (2026)
by: Baan, Joris, et al.
Published: (2026)
When Silence Is Golden: Can LLMs Learn to Abstain in Temporal QA and Beyond?
by: Zhou, Xinyu, et al.
Published: (2026)
by: Zhou, Xinyu, et al.
Published: (2026)
Do We Really Need to Approach the Entire Pareto Front in Many-Objective Bayesian Optimisation?
by: Jiang, Chao, et al.
Published: (2026)
by: Jiang, Chao, et al.
Published: (2026)
Influencing LLM Multi-Agent Dialogue via Policy-Parameterized Prompts
by: Bo, Hongbo, et al.
Published: (2026)
by: Bo, Hongbo, et al.
Published: (2026)
DeePoly: A High-Order Accuracy Scientific Machine Learning Framework for Function Approximation and Solving PDEs
by: Liu, Li, et al.
Published: (2025)
by: Liu, Li, et al.
Published: (2025)
Abstain-R1: Calibrated Abstention and Post-Refusal Clarification via Verifiable RL
by: Zhai, Skylar, et al.
Published: (2026)
by: Zhai, Skylar, et al.
Published: (2026)
ConceptMath: A Bilingual Concept-wise Benchmark for Measuring Mathematical Reasoning of Large Language Models
by: Wu, Yanan, et al.
Published: (2024)
by: Wu, Yanan, et al.
Published: (2024)
Enhancing Text Annotation through Rationale-Driven Collaborative Few-Shot Prompting
by: Wu, Jianfei, et al.
Published: (2024)
by: Wu, Jianfei, et al.
Published: (2024)
AdvDMD: Adversarial Reward Meets DMD For High-Quality Few-Step Generation
by: Wang, Xu, et al.
Published: (2026)
by: Wang, Xu, et al.
Published: (2026)
Prompt Stability in Code LLMs: Measuring Sensitivity across Emotion- and Personality-Driven Variations
by: Ma, Wei, et al.
Published: (2025)
by: Ma, Wei, et al.
Published: (2025)
AutoHealth: An Uncertainty-Aware Multi-Agent System for Autonomous Health Data Modeling
by: Xia, Tong, et al.
Published: (2026)
by: Xia, Tong, et al.
Published: (2026)
Open-Source AI-based SE Tools: Opportunities and Challenges of Collaborative Software Learning
by: Lin, Zhihao, et al.
Published: (2024)
by: Lin, Zhihao, et al.
Published: (2024)
VC4VG: Optimizing Video Captions for Text-to-Video Generation
by: Du, Yang, et al.
Published: (2025)
by: Du, Yang, et al.
Published: (2025)
Identifying and Solving Conditional Image Leakage in Image-to-Video Diffusion Model
by: Zhao, Min, et al.
Published: (2024)
by: Zhao, Min, et al.
Published: (2024)
EasyRec: Simple yet Effective Language Models for Recommendation
by: Ren, Xubin, et al.
Published: (2024)
by: Ren, Xubin, et al.
Published: (2024)
Cognitive Edge Computing: A Comprehensive Survey on Optimizing Large Models and AI Agents for Pervasive Deployment
by: Wang, Xubin, et al.
Published: (2025)
by: Wang, Xubin, et al.
Published: (2025)
Geometric Manifold Rectification for Imbalanced Learning
by: Wang, Xubin, et al.
Published: (2026)
by: Wang, Xubin, et al.
Published: (2026)
Employee Turnover Prediction: A Cross-component Attention Transformer with Consideration of Competitor Influence and Contagious Effect
by: Liu, Hao, et al.
Published: (2025)
by: Liu, Hao, et al.
Published: (2025)
How Well Do LLMs Perform on the Simplest Long-Chain Reasoning Tasks: An Empirical Study on the Equivalence Class Problem
by: Zheng, Chun, et al.
Published: (2026)
by: Zheng, Chun, et al.
Published: (2026)
E^2-LLM: Efficient and Extreme Length Extension of Large Language Models
by: Liu, Jiaheng, et al.
Published: (2024)
by: Liu, Jiaheng, et al.
Published: (2024)
Abstain and Validate: A Dual-LLM Policy for Reducing Noise in Agentic Program Repair
by: Cambronero, José, et al.
Published: (2025)
by: Cambronero, José, et al.
Published: (2025)
Failing on Bias Mitigation: A Case Study on the Challenges of Fairness in Government Data
by: Bo, Hongbo, et al.
Published: (2026)
by: Bo, Hongbo, et al.
Published: (2026)
From Detection to Diagnosis: Advancing Hallucination Analysis with Automated Data Synthesis
by: Liu, Yanyi, et al.
Published: (2025)
by: Liu, Yanyi, et al.
Published: (2025)
Do Large Language Models have Problem-Solving Capability under Incomplete Information Scenarios?
by: Chen, Yuyan, et al.
Published: (2024)
by: Chen, Yuyan, et al.
Published: (2024)
DDA-Thinker: Decoupled Dual-Atomic Reinforcement Learning for Reasoning-Driven Image Editing
by: Yang, Hanqing, et al.
Published: (2026)
by: Yang, Hanqing, et al.
Published: (2026)
Do Transformers Have the Ability for Periodicity Generalization?
by: Liu, Huanyu, et al.
Published: (2026)
by: Liu, Huanyu, et al.
Published: (2026)
Similar Items
-
Gradient Coupling: The Hidden Barrier to Generalization in Agentic Reinforcement Learning
by: Liu, Jingyu, et al.
Published: (2025) -
Knowing When to Abstain: Medical LLMs Under Clinical Uncertainty
by: Machcha, Sravanthi, et al.
Published: (2026) -
Abstaining Machine Learning -- Philosophical Considerations
by: Schuster, Daniela
Published: (2024) -
Interpretable and Fair Mechanisms for Abstaining Classifiers
by: Lenders, Daphne, et al.
Published: (2025) -
The Confidence Gate Theorem: When Should Ranked Decision Systems Abstain?
by: Doku, Ronald
Published: (2026)