Mitigating Legibility Tax with Decoupled Prover-Verifier Games
Fuente:
arXiv
Saved in:
| Main Authors: | Kim, Yegon, Lee, Juho |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Model-Free Universal AI
by: Kim, Yegon, et al.
Published: (2026)
by: Kim, Yegon, et al.
Published: (2026)
Parallel Test-Time Scaling with Multi-Sequence Verifiers
by: Kim, Yegon, et al.
Published: (2026)
by: Kim, Yegon, et al.
Published: (2026)
Variational Partial Group Convolutions for Input-Aware Partial Equivariance of Rotations and Color-Shifts
by: Kim, Hyunsu, et al.
Published: (2024)
by: Kim, Hyunsu, et al.
Published: (2024)
Verbalized Confidence Triggers Self-Verification: Emergent Behavior Without Explicit Reasoning Supervision
by: Jang, Chaeyun, et al.
Published: (2025)
by: Jang, Chaeyun, et al.
Published: (2025)
Neural Concept Verifier: Scaling Prover-Verifier Games via Concept Encodings
by: Turan, Berkant, et al.
Published: (2025)
by: Turan, Berkant, et al.
Published: (2025)
PokerKit: A Comprehensive Python Library for Fine-Grained Multi-Variant Poker Game Simulations
by: Kim, Juho
Published: (2023)
by: Kim, Juho
Published: (2023)
StepFun-Prover Preview: Let's Think and Verify Step by Step
by: Shang, Shijie, et al.
Published: (2025)
by: Shang, Shijie, et al.
Published: (2025)
Active Legibility in Multiagent Reinforcement Learning
by: Liu, Yanyu, et al.
Published: (2024)
by: Liu, Yanyu, et al.
Published: (2024)
Leanabell-Prover-V2: Verifier-integrated Reasoning for Formal Theorem Proving via Reinforcement Learning
by: Ji, Xingguang, et al.
Published: (2025)
by: Ji, Xingguang, et al.
Published: (2025)
Domain-Independent Game Abstraction using Word Embedding Techniques
by: Kim, Juho, et al.
Published: (2026)
by: Kim, Juho, et al.
Published: (2026)
Recording and Describing Poker Hands
by: Kim, Juho
Published: (2023)
by: Kim, Juho
Published: (2023)
Watermarking Game-Playing Agents in Perfect-Information Extensive-Form Games
by: Kim, Juho, et al.
Published: (2026)
by: Kim, Juho, et al.
Published: (2026)
Leanabell-Prover: Posttraining Scaling in Formal Reasoning
by: Zhang, Jingyuan, et al.
Published: (2025)
by: Zhang, Jingyuan, et al.
Published: (2025)
"Teammates, Am I Clear?": Analysing Legible Behaviours in Teams
by: Faria, Miguel, et al.
Published: (2025)
by: Faria, Miguel, et al.
Published: (2025)
ProofCompass: Enhancing Specialized Provers with LLM Guidance
by: Wischermann, Nicolas, et al.
Published: (2025)
by: Wischermann, Nicolas, et al.
Published: (2025)
Theorem Prover as a Judge for Synthetic Data Generation
by: Leang, Joshua Ong Jun, et al.
Published: (2025)
by: Leang, Joshua Ong Jun, et al.
Published: (2025)
REAL-Prover: Retrieval Augmented Lean Prover for Mathematical Reasoning
by: Shen, Ziju, et al.
Published: (2025)
by: Shen, Ziju, et al.
Published: (2025)
Fast Ensembling with Diffusion Schrödinger Bridge
by: Kim, Hyunsu, et al.
Published: (2024)
by: Kim, Hyunsu, et al.
Published: (2024)
Learning Infinitesimal Generators of Continuous Symmetries from Data
by: Ko, Gyeonghoon, et al.
Published: (2024)
by: Ko, Gyeonghoon, et al.
Published: (2024)
Mitigating Safety Tax via Distribution-Grounded Refinement in Large Reasoning Models
by: Xie, Yingsha, et al.
Published: (2026)
by: Xie, Yingsha, et al.
Published: (2026)
Clarifying Before Reasoning: A Coq Prover with Structural Context
by: Lu, Yanzhen, et al.
Published: (2025)
by: Lu, Yanzhen, et al.
Published: (2025)
Model Fusion through Bayesian Optimization in Language Model Fine-Tuning
by: Jang, Chaeyun, et al.
Published: (2024)
by: Jang, Chaeyun, et al.
Published: (2024)
Measuring Reasoning Trace Legibility: Can Those Who Understand Teach?
by: Roytburg, Dani, et al.
Published: (2026)
by: Roytburg, Dani, et al.
Published: (2026)
Vibe Coding an LLM-powered Theorem Prover
by: Hou, Zhe
Published: (2026)
by: Hou, Zhe
Published: (2026)
PhysProver: Advancing Automatic Theorem Proving for Physics
by: Zhang, Hanning, et al.
Published: (2026)
by: Zhang, Hanning, et al.
Published: (2026)
Active Learning with Selective Time-Step Acquisition for PDEs
by: Kim, Yegon, et al.
Published: (2025)
by: Kim, Yegon, et al.
Published: (2025)
Kimina-Prover Preview: Towards Large Formal Reasoning Models with Reinforcement Learning
by: Wang, Haiming, et al.
Published: (2025)
by: Wang, Haiming, et al.
Published: (2025)
Avoiding Obfuscation with Prover-Estimator Debate
by: Brown-Cohen, Jonah, et al.
Published: (2025)
by: Brown-Cohen, Jonah, et al.
Published: (2025)
Reward Hacking Mitigation using Verifiable Composite Rewards
by: Tarek, Mirza Farhan Bin, et al.
Published: (2025)
by: Tarek, Mirza Farhan Bin, et al.
Published: (2025)
Seed-Prover: Deep and Broad Reasoning for Automated Theorem Proving
by: Chen, Luoxin, et al.
Published: (2025)
by: Chen, Luoxin, et al.
Published: (2025)
Prover Agent: An Agent-Based Framework for Formal Mathematical Proofs
by: Baba, Kaito, et al.
Published: (2025)
by: Baba, Kaito, et al.
Published: (2025)
Proof Recommendation System for the HOL4 Theorem Prover
by: Dekhil, Nour, et al.
Published: (2024)
by: Dekhil, Nour, et al.
Published: (2024)
Inference-Time Diversity in RL-Trained Lean Theorem Provers: A Diagnostic Study
by: Burton, Zachary
Published: (2026)
by: Burton, Zachary
Published: (2026)
MPS-Prover: Advancing Stepwise Theorem Proving by Multi-Perspective Search and Data Curation
by: Liang, Zhenwen, et al.
Published: (2025)
by: Liang, Zhenwen, et al.
Published: (2025)
Can AI Models Appreciate Document Aesthetics? An Exploration of Legibility and Layout Quality in Relation to Prediction Confidence
by: Yang, Hsiu-Wei, et al.
Published: (2024)
by: Yang, Hsiu-Wei, et al.
Published: (2024)
TaxCalcBench: Evaluating Frontier Models on the Tax Calculation Task
by: Bock, Michael R., et al.
Published: (2025)
by: Bock, Michael R., et al.
Published: (2025)
LLMs Gaming Verifiers: RLVR can Lead to Reward Hacking
by: Helff, Lukas, et al.
Published: (2026)
by: Helff, Lukas, et al.
Published: (2026)
BFS-Prover: Scalable Best-First Tree Search for LLM-based Automatic Theorem Proving
by: Xin, Ran, et al.
Published: (2025)
by: Xin, Ran, et al.
Published: (2025)
EvolProver: Advancing Automated Theorem Proving by Evolving Formalized Problems via Symmetry and Difficulty
by: Tian, Yuchen, et al.
Published: (2025)
by: Tian, Yuchen, et al.
Published: (2025)
DeepSeek-Prover: Advancing Theorem Proving in LLMs through Large-Scale Synthetic Data
by: Xin, Huajian, et al.
Published: (2024)
by: Xin, Huajian, et al.
Published: (2024)
Similar Items
-
A Model-Free Universal AI
by: Kim, Yegon, et al.
Published: (2026) -
Parallel Test-Time Scaling with Multi-Sequence Verifiers
by: Kim, Yegon, et al.
Published: (2026) -
Variational Partial Group Convolutions for Input-Aware Partial Equivariance of Rotations and Color-Shifts
by: Kim, Hyunsu, et al.
Published: (2024) -
Verbalized Confidence Triggers Self-Verification: Emergent Behavior Without Explicit Reasoning Supervision
by: Jang, Chaeyun, et al.
Published: (2025) -
Neural Concept Verifier: Scaling Prover-Verifier Games via Concept Encodings
by: Turan, Berkant, et al.
Published: (2025)