Do not Abstain! Identify and Solve the Uncertainty
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Liu, Jingyu, Peng, Jingquan, Wu, xiaopeng, Li, Xubin, Ge, Tiezheng, Zheng, Bo, Liu, Yong |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Gradient Coupling: The Hidden Barrier to Generalization in Agentic Reinforcement Learning
von: Liu, Jingyu, et al.
Veröffentlicht: (2025)
von: Liu, Jingyu, et al.
Veröffentlicht: (2025)
Knowing When to Abstain: Medical LLMs Under Clinical Uncertainty
von: Machcha, Sravanthi, et al.
Veröffentlicht: (2026)
von: Machcha, Sravanthi, et al.
Veröffentlicht: (2026)
Abstaining Machine Learning -- Philosophical Considerations
von: Schuster, Daniela
Veröffentlicht: (2024)
von: Schuster, Daniela
Veröffentlicht: (2024)
Interpretable and Fair Mechanisms for Abstaining Classifiers
von: Lenders, Daphne, et al.
Veröffentlicht: (2025)
von: Lenders, Daphne, et al.
Veröffentlicht: (2025)
The Confidence Gate Theorem: When Should Ranked Decision Systems Abstain?
von: Doku, Ronald
Veröffentlicht: (2026)
von: Doku, Ronald
Veröffentlicht: (2026)
MT-Bench-101: A Fine-Grained Benchmark for Evaluating Large Language Models in Multi-Turn Dialogues
von: Bai, Ge, et al.
Veröffentlicht: (2024)
von: Bai, Ge, et al.
Veröffentlicht: (2024)
Know When to Abstain: Optimal Selective Classification with Likelihood Ratios
von: Heng, Alvin, et al.
Veröffentlicht: (2025)
von: Heng, Alvin, et al.
Veröffentlicht: (2025)
Logic-of-Thought: Empowering Large Language Models with Logic Programs for Solving Puzzles in Natural Language
von: Li, Naiqi, et al.
Veröffentlicht: (2025)
von: Li, Naiqi, et al.
Veröffentlicht: (2025)
Detecting Multimodal Situations with Insufficient Context and Abstaining from Baseless Predictions
von: Liu, Junzhang, et al.
Veröffentlicht: (2024)
von: Liu, Junzhang, et al.
Veröffentlicht: (2024)
How Much Can RAG Help the Reasoning of LLM?
von: Liu, Jingyu, et al.
Veröffentlicht: (2024)
von: Liu, Jingyu, et al.
Veröffentlicht: (2024)
Perfect Alignment May be Poisonous to Graph Contrastive Learning
von: Liu, Jingyu, et al.
Veröffentlicht: (2023)
von: Liu, Jingyu, et al.
Veröffentlicht: (2023)
Tackling the Inherent Difficulty of Noise Filtering in RAG
von: Liu, Jingyu, et al.
Veröffentlicht: (2026)
von: Liu, Jingyu, et al.
Veröffentlicht: (2026)
CausalAbstain: Enhancing Multilingual LLMs with Causal Reasoning for Trustworthy Abstention
von: Sun, Yuxi, et al.
Veröffentlicht: (2025)
von: Sun, Yuxi, et al.
Veröffentlicht: (2025)
Teaching LLMs to Abstain via Fine-Grained Semantic Confidence Reward
von: An, Hao, et al.
Veröffentlicht: (2025)
von: An, Hao, et al.
Veröffentlicht: (2025)
Clarify, Abstain or Answer? Strategising in Conversation with Belief-Augmented Generation
von: Baan, Joris, et al.
Veröffentlicht: (2026)
von: Baan, Joris, et al.
Veröffentlicht: (2026)
When Silence Is Golden: Can LLMs Learn to Abstain in Temporal QA and Beyond?
von: Zhou, Xinyu, et al.
Veröffentlicht: (2026)
von: Zhou, Xinyu, et al.
Veröffentlicht: (2026)
Do We Really Need to Approach the Entire Pareto Front in Many-Objective Bayesian Optimisation?
von: Jiang, Chao, et al.
Veröffentlicht: (2026)
von: Jiang, Chao, et al.
Veröffentlicht: (2026)
Influencing LLM Multi-Agent Dialogue via Policy-Parameterized Prompts
von: Bo, Hongbo, et al.
Veröffentlicht: (2026)
von: Bo, Hongbo, et al.
Veröffentlicht: (2026)
DeePoly: A High-Order Accuracy Scientific Machine Learning Framework for Function Approximation and Solving PDEs
von: Liu, Li, et al.
Veröffentlicht: (2025)
von: Liu, Li, et al.
Veröffentlicht: (2025)
Abstain-R1: Calibrated Abstention and Post-Refusal Clarification via Verifiable RL
von: Zhai, Skylar, et al.
Veröffentlicht: (2026)
von: Zhai, Skylar, et al.
Veröffentlicht: (2026)
ConceptMath: A Bilingual Concept-wise Benchmark for Measuring Mathematical Reasoning of Large Language Models
von: Wu, Yanan, et al.
Veröffentlicht: (2024)
von: Wu, Yanan, et al.
Veröffentlicht: (2024)
Enhancing Text Annotation through Rationale-Driven Collaborative Few-Shot Prompting
von: Wu, Jianfei, et al.
Veröffentlicht: (2024)
von: Wu, Jianfei, et al.
Veröffentlicht: (2024)
AdvDMD: Adversarial Reward Meets DMD For High-Quality Few-Step Generation
von: Wang, Xu, et al.
Veröffentlicht: (2026)
von: Wang, Xu, et al.
Veröffentlicht: (2026)
Prompt Stability in Code LLMs: Measuring Sensitivity across Emotion- and Personality-Driven Variations
von: Ma, Wei, et al.
Veröffentlicht: (2025)
von: Ma, Wei, et al.
Veröffentlicht: (2025)
AutoHealth: An Uncertainty-Aware Multi-Agent System for Autonomous Health Data Modeling
von: Xia, Tong, et al.
Veröffentlicht: (2026)
von: Xia, Tong, et al.
Veröffentlicht: (2026)
Open-Source AI-based SE Tools: Opportunities and Challenges of Collaborative Software Learning
von: Lin, Zhihao, et al.
Veröffentlicht: (2024)
von: Lin, Zhihao, et al.
Veröffentlicht: (2024)
VC4VG: Optimizing Video Captions for Text-to-Video Generation
von: Du, Yang, et al.
Veröffentlicht: (2025)
von: Du, Yang, et al.
Veröffentlicht: (2025)
Identifying and Solving Conditional Image Leakage in Image-to-Video Diffusion Model
von: Zhao, Min, et al.
Veröffentlicht: (2024)
von: Zhao, Min, et al.
Veröffentlicht: (2024)
EasyRec: Simple yet Effective Language Models for Recommendation
von: Ren, Xubin, et al.
Veröffentlicht: (2024)
von: Ren, Xubin, et al.
Veröffentlicht: (2024)
Cognitive Edge Computing: A Comprehensive Survey on Optimizing Large Models and AI Agents for Pervasive Deployment
von: Wang, Xubin, et al.
Veröffentlicht: (2025)
von: Wang, Xubin, et al.
Veröffentlicht: (2025)
Geometric Manifold Rectification for Imbalanced Learning
von: Wang, Xubin, et al.
Veröffentlicht: (2026)
von: Wang, Xubin, et al.
Veröffentlicht: (2026)
Employee Turnover Prediction: A Cross-component Attention Transformer with Consideration of Competitor Influence and Contagious Effect
von: Liu, Hao, et al.
Veröffentlicht: (2025)
von: Liu, Hao, et al.
Veröffentlicht: (2025)
How Well Do LLMs Perform on the Simplest Long-Chain Reasoning Tasks: An Empirical Study on the Equivalence Class Problem
von: Zheng, Chun, et al.
Veröffentlicht: (2026)
von: Zheng, Chun, et al.
Veröffentlicht: (2026)
E^2-LLM: Efficient and Extreme Length Extension of Large Language Models
von: Liu, Jiaheng, et al.
Veröffentlicht: (2024)
von: Liu, Jiaheng, et al.
Veröffentlicht: (2024)
Abstain and Validate: A Dual-LLM Policy for Reducing Noise in Agentic Program Repair
von: Cambronero, José, et al.
Veröffentlicht: (2025)
von: Cambronero, José, et al.
Veröffentlicht: (2025)
Failing on Bias Mitigation: A Case Study on the Challenges of Fairness in Government Data
von: Bo, Hongbo, et al.
Veröffentlicht: (2026)
von: Bo, Hongbo, et al.
Veröffentlicht: (2026)
From Detection to Diagnosis: Advancing Hallucination Analysis with Automated Data Synthesis
von: Liu, Yanyi, et al.
Veröffentlicht: (2025)
von: Liu, Yanyi, et al.
Veröffentlicht: (2025)
Do Large Language Models have Problem-Solving Capability under Incomplete Information Scenarios?
von: Chen, Yuyan, et al.
Veröffentlicht: (2024)
von: Chen, Yuyan, et al.
Veröffentlicht: (2024)
DDA-Thinker: Decoupled Dual-Atomic Reinforcement Learning for Reasoning-Driven Image Editing
von: Yang, Hanqing, et al.
Veröffentlicht: (2026)
von: Yang, Hanqing, et al.
Veröffentlicht: (2026)
Do Transformers Have the Ability for Periodicity Generalization?
von: Liu, Huanyu, et al.
Veröffentlicht: (2026)
von: Liu, Huanyu, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Gradient Coupling: The Hidden Barrier to Generalization in Agentic Reinforcement Learning
von: Liu, Jingyu, et al.
Veröffentlicht: (2025) -
Knowing When to Abstain: Medical LLMs Under Clinical Uncertainty
von: Machcha, Sravanthi, et al.
Veröffentlicht: (2026) -
Abstaining Machine Learning -- Philosophical Considerations
von: Schuster, Daniela
Veröffentlicht: (2024) -
Interpretable and Fair Mechanisms for Abstaining Classifiers
von: Lenders, Daphne, et al.
Veröffentlicht: (2025) -
The Confidence Gate Theorem: When Should Ranked Decision Systems Abstain?
von: Doku, Ronald
Veröffentlicht: (2026)