Anytime Safe PAC Efficient Reasoning
Fuente:
arXiv
Saved in:
| Main Authors: | Yu, Chengyao, Zeng, Hao, Zhu, Youxin, Huang, Jianguo, Zeng, Huajun, Jing, Bingyi |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
HyPAC: Cost-Efficient LLMs-Human Hybrid Annotation with PAC Error Guarantees
by: Zeng, Hao, et al.
Published: (2026)
by: Zeng, Hao, et al.
Published: (2026)
On the Provable Performance Guarantee of Efficient Reasoning Models
by: Zeng, Hao, et al.
Published: (2025)
by: Zeng, Hao, et al.
Published: (2025)
Conditional Performance Guarantee for Large Reasoning Models
by: Huang, Jianguo, et al.
Published: (2026)
by: Huang, Jianguo, et al.
Published: (2026)
RACER: Risk-Aware Calibrated Efficient Routing for Large Language Models
by: Hao, Sai, et al.
Published: (2026)
by: Hao, Sai, et al.
Published: (2026)
A note on the impossibility of conditional PAC-efficient reasoning in large language models
by: Zeng, Hao
Published: (2025)
by: Zeng, Hao
Published: (2025)
FAIR-Pruner: A Flexible Framework for Automatic Layer-Wise Pruning via Tolerance of Difference
by: Lin, Chenqing, et al.
Published: (2025)
by: Lin, Chenqing, et al.
Published: (2025)
Done Is Better than Perfect: Unlocking Efficient Reasoning by Structured Multi-Turn Decomposition
by: Zeng, Zihao, et al.
Published: (2025)
by: Zeng, Zihao, et al.
Published: (2025)
Multi-Condition Conformal Selection
by: Hao, Qingyang, et al.
Published: (2025)
by: Hao, Qingyang, et al.
Published: (2025)
A Lightweight Traffic Map for Efficient Anytime LaCAM*
by: Shen, Bojie, et al.
Published: (2026)
by: Shen, Bojie, et al.
Published: (2026)
Model-agnostic Selective Labeling with Provable Statistical Guarantees
by: Huang, Huipeng, et al.
Published: (2025)
by: Huang, Huipeng, et al.
Published: (2025)
Anytime Single-Step MAPF Planning with Anytime PIBT
by: Gandotra, Nayesha, et al.
Published: (2025)
by: Gandotra, Nayesha, et al.
Published: (2025)
An Optimistic Algorithm for online CMDPS with Anytime Adversarial Constraints
by: Zhu, Jiahui, et al.
Published: (2025)
by: Zhu, Jiahui, et al.
Published: (2025)
Reflective Confidence: Correcting Reasoning Flaws via Online Self-Correction
by: Zeng, Qinglin, et al.
Published: (2025)
by: Zeng, Qinglin, et al.
Published: (2025)
Anytime-Constrained Reinforcement Learning
by: McMahan, Jeremy, et al.
Published: (2023)
by: McMahan, Jeremy, et al.
Published: (2023)
Towards Anytime-Valid Statistical Watermarking
by: Huang, Baihe, et al.
Published: (2026)
by: Huang, Baihe, et al.
Published: (2026)
RealSafe-R1: Safety-Aligned DeepSeek-R1 without Compromising Reasoning Capability
by: Zhang, Yichi, et al.
Published: (2025)
by: Zhang, Yichi, et al.
Published: (2025)
From Accuracy to Robustness: A Study of Rule- and Model-based Verifiers in Mathematical Reasoning
by: Huang, Yuzhen, et al.
Published: (2025)
by: Huang, Yuzhen, et al.
Published: (2025)
Optimizing Anytime Reasoning via Budget Relative Policy Optimization
by: Qi, Penghui, et al.
Published: (2025)
by: Qi, Penghui, et al.
Published: (2025)
Tail-Risk-Safe Monte Carlo Tree Search under PAC-Level Guarantees
by: Zhang, Zuyuan, et al.
Published: (2025)
by: Zhang, Zuyuan, et al.
Published: (2025)
When to Continue Thinking: Adaptive Thinking Mode Switching for Efficient Reasoning
by: Zhang, Xiaoyun, et al.
Published: (2025)
by: Zhang, Xiaoyun, et al.
Published: (2025)
Anytime Cooperative Implicit Hitting Set Solving
by: Rollón, Emma, et al.
Published: (2025)
by: Rollón, Emma, et al.
Published: (2025)
Hi-Drive: Hierarchical POMDP Planning for Safe Autonomous Driving in Diverse Urban Environments
by: Jin, Xuanjin, et al.
Published: (2024)
by: Jin, Xuanjin, et al.
Published: (2024)
PAC Privacy Preserving Diffusion Models
by: Xu, Qipan, et al.
Published: (2023)
by: Xu, Qipan, et al.
Published: (2023)
TabDSR: Decompose, Sanitize, and Reason for Complex Numerical Reasoning in Tabular Data
by: Jiang, Changjiang, et al.
Published: (2025)
by: Jiang, Changjiang, et al.
Published: (2025)
Is Efficient PAC Learning Possible with an Oracle That Responds 'Yes' or 'No'?
by: Daskalakis, Constantinos, et al.
Published: (2024)
by: Daskalakis, Constantinos, et al.
Published: (2024)
From Bias Mitigation to Bias Negotiation: Governing Identity and Sociocultural Reasoning in Generative AI
by: Dunivin, Zackary Okun, et al.
Published: (2026)
by: Dunivin, Zackary Okun, et al.
Published: (2026)
DIANOIA: Diagnostic Decomposition and Joint Optimization for Multi-Agent Reasoning
by: Yang, Yiming, et al.
Published: (2026)
by: Yang, Yiming, et al.
Published: (2026)
A-MHA*: Anytime Multi-Heuristic A*
by: Natarajan, Ramkumar, et al.
Published: (2025)
by: Natarajan, Ramkumar, et al.
Published: (2025)
FloydNet: A Learning Paradigm for Global Relational Reasoning
by: Yu, Jingcheng, et al.
Published: (2026)
by: Yu, Jingcheng, et al.
Published: (2026)
Enhancing Safe and Controllable Protein Generation via Knowledge Preference Optimization
by: Wang, Yuhao, et al.
Published: (2025)
by: Wang, Yuhao, et al.
Published: (2025)
SafeRBench: Dissecting the Reasoning Safety of Large Language Models
by: Gao, Xin, et al.
Published: (2025)
by: Gao, Xin, et al.
Published: (2025)
LogicLens: Visual-Logical Co-Reasoning for Text-Centric Forgery Analysis
by: Zeng, Fanwei, et al.
Published: (2025)
by: Zeng, Fanwei, et al.
Published: (2025)
IFDNS: An Iterative Feedback-Driven Neuro-Symbolic Method for Faithful Logical Reasoning
by: Wang, Xiaoheng, et al.
Published: (2026)
by: Wang, Xiaoheng, et al.
Published: (2026)
SafeSteer: Localized On-Policy Distillation for Efficient Safety Alignment
by: Li, Hao, et al.
Published: (2026)
by: Li, Hao, et al.
Published: (2026)
Online POMDP Planning with Anytime Deterministic Optimality Guarantees
by: Barenboim, Moran, et al.
Published: (2023)
by: Barenboim, Moran, et al.
Published: (2023)
Constant-time Motion Planning with Anytime Refinement for Manipulation
by: Mishani, Itamar, et al.
Published: (2023)
by: Mishani, Itamar, et al.
Published: (2023)
StatLLaMA: Multi-Stage training for domain-optimized statistical large language models
by: Zeng, Jing-Yi, et al.
Published: (2025)
by: Zeng, Jing-Yi, et al.
Published: (2025)
Data-Efficient On-Policy Distillation for Automatic Speech Recognition
by: Lin, Yu, et al.
Published: (2026)
by: Lin, Yu, et al.
Published: (2026)
InfoCom: Kilobyte-Scale Communication-Efficient Collaborative Perception with Information Bottleneck
by: Wei, Quanmin, et al.
Published: (2025)
by: Wei, Quanmin, et al.
Published: (2025)
SIFT: Grounding LLM Reasoning in Contexts via Stickers
by: Zeng, Zihao, et al.
Published: (2025)
by: Zeng, Zihao, et al.
Published: (2025)
Similar Items
-
HyPAC: Cost-Efficient LLMs-Human Hybrid Annotation with PAC Error Guarantees
by: Zeng, Hao, et al.
Published: (2026) -
On the Provable Performance Guarantee of Efficient Reasoning Models
by: Zeng, Hao, et al.
Published: (2025) -
Conditional Performance Guarantee for Large Reasoning Models
by: Huang, Jianguo, et al.
Published: (2026) -
RACER: Risk-Aware Calibrated Efficient Routing for Large Language Models
by: Hao, Sai, et al.
Published: (2026) -
A note on the impossibility of conditional PAC-efficient reasoning in large language models
by: Zeng, Hao
Published: (2025)