Taming Imperfect Process Verifiers: A Sampling Perspective on Backtracking
Fuente:
arXiv
Salvato in:
| Autori principali: | Rohatgi, Dhruv, Shetty, Abhishek, Saless, Donya, Li, Yuchen, Moitra, Ankur, Risteski, Andrej, Foster, Dylan J. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
The tractability landscape of diffusion alignment: regularization, rewards, and computational primitives
di: Moitra, Ankur, et al.
Pubblicazione: (2026)
di: Moitra, Ankur, et al.
Pubblicazione: (2026)
Steering diffusion models with quadratic rewards: a fine-grained analysis
di: Moitra, Ankur, et al.
Pubblicazione: (2026)
di: Moitra, Ankur, et al.
Pubblicazione: (2026)
Exploration is Harder than Prediction: Cryptographically Separating Reinforcement Learning from Supervised Learning
di: Golowich, Noah, et al.
Pubblicazione: (2024)
di: Golowich, Noah, et al.
Pubblicazione: (2024)
Regularized Robustly Reliable Learners and Instance Targeted Attacks
di: Blum, Avrim, et al.
Pubblicazione: (2024)
di: Blum, Avrim, et al.
Pubblicazione: (2024)
Fit Like You Sample: Sample-Efficient Generalized Score Matching from Fast Mixing Diffusions
di: Qin, Yilong, et al.
Pubblicazione: (2023)
di: Qin, Yilong, et al.
Pubblicazione: (2023)
On Learning Parities with Dependent Noise
di: Golowich, Noah, et al.
Pubblicazione: (2024)
di: Golowich, Noah, et al.
Pubblicazione: (2024)
Computational-Statistical Tradeoffs at the Next-Token Prediction Barrier: Autoregressive and Imitation Learning under Misspecification
di: Rohatgi, Dhruv, et al.
Pubblicazione: (2025)
di: Rohatgi, Dhruv, et al.
Pubblicazione: (2025)
Prior Knowledge Makes It Possible: From Sublinear Graph Algorithms to LLM Test-Time Methods
di: Blum, Avrim, et al.
Pubblicazione: (2025)
di: Blum, Avrim, et al.
Pubblicazione: (2025)
Better Models and Algorithms for Learning Ising Models from Dynamics
di: Gaitonde, Jason, et al.
Pubblicazione: (2025)
di: Gaitonde, Jason, et al.
Pubblicazione: (2025)
Bypassing the Noisy Parity Barrier: Learning Higher-Order Markov Random Fields from Dynamics
di: Gaitonde, Jason, et al.
Pubblicazione: (2024)
di: Gaitonde, Jason, et al.
Pubblicazione: (2024)
Overcomplete Tensor Decomposition via Koszul-Young Flattenings
di: Kothari, Pravesh K., et al.
Pubblicazione: (2024)
di: Kothari, Pravesh K., et al.
Pubblicazione: (2024)
Model Stealing for Any Low-Rank Language Model
di: Liu, Allen, et al.
Pubblicazione: (2024)
di: Liu, Allen, et al.
Pubblicazione: (2024)
A computational phase transition for learning-to-sample from Ising models
di: Risteski, Andrej, et al.
Pubblicazione: (2026)
di: Risteski, Andrej, et al.
Pubblicazione: (2026)
Learning $\mathsf{AC}^0$ Under Graphical Models
di: Chandrasekaran, Gautam, et al.
Pubblicazione: (2026)
di: Chandrasekaran, Gautam, et al.
Pubblicazione: (2026)
Structure learning of Hamiltonians from real-time evolution
di: Bakshi, Ainesh, et al.
Pubblicazione: (2024)
di: Bakshi, Ainesh, et al.
Pubblicazione: (2024)
Learning quantum Hamiltonians at any temperature in polynomial time
di: Bakshi, Ainesh, et al.
Pubblicazione: (2023)
di: Bakshi, Ainesh, et al.
Pubblicazione: (2023)
Learning with Monotone Adversarial Corruptions
di: Larsen, Kasper Green, et al.
Pubblicazione: (2026)
di: Larsen, Kasper Green, et al.
Pubblicazione: (2026)
Adversarial Resilience in Sequential Prediction via Abstention
di: Goel, Surbhi, et al.
Pubblicazione: (2023)
di: Goel, Surbhi, et al.
Pubblicazione: (2023)
Tolerant Algorithms for Learning with Arbitrary Covariate Shift
di: Goel, Surbhi, et al.
Pubblicazione: (2024)
di: Goel, Surbhi, et al.
Pubblicazione: (2024)
Lasso with Latents: Efficient Estimation, Covariate Rescaling, and Computational-Statistical Gaps
di: Kelner, Jonathan, et al.
Pubblicazione: (2024)
di: Kelner, Jonathan, et al.
Pubblicazione: (2024)
Provably Learning from Modern Language Models via Low Logit Rank
di: Golowich, Noah, et al.
Pubblicazione: (2025)
di: Golowich, Noah, et al.
Pubblicazione: (2025)
Theoretically Grounded Pruning of Large Ground Sets for Constrained, Discrete Optimization
di: Nath, Ankur, et al.
Pubblicazione: (2024)
di: Nath, Ankur, et al.
Pubblicazione: (2024)
Omnipredictors for Regression and the Approximate Rank of Convex Functions
di: Gopalan, Parikshit, et al.
Pubblicazione: (2024)
di: Gopalan, Parikshit, et al.
Pubblicazione: (2024)
$O(\sqrt{T})$ Static Regret and Instance Dependent Constraint Violation for Constrained Online Convex Optimization
di: Vaze, Rahul, et al.
Pubblicazione: (2025)
di: Vaze, Rahul, et al.
Pubblicazione: (2025)
On the query complexity of sampling from non-log-concave distributions
di: He, Yuchen, et al.
Pubblicazione: (2025)
di: He, Yuchen, et al.
Pubblicazione: (2025)
Efficient Algorithms for Verifying Kruskal Rank in Sparse Linear Regression and Related Applications
di: Zhou, Fengqin
Pubblicazione: (2025)
di: Zhou, Fengqin
Pubblicazione: (2025)
Sample-Adaptivity Tradeoff in On-Demand Sampling
di: Haghtalab, Nika, et al.
Pubblicazione: (2025)
di: Haghtalab, Nika, et al.
Pubblicazione: (2025)
Understanding Memory-Regret Trade-Off for Streaming Stochastic Multi-Armed Bandits
di: He, Yuchen, et al.
Pubblicazione: (2024)
di: He, Yuchen, et al.
Pubblicazione: (2024)
On the Problem of Best Arm Retention
di: Chen, Houshuang, et al.
Pubblicazione: (2025)
di: Chen, Houshuang, et al.
Pubblicazione: (2025)
Smooth Nash Equilibria: Algorithms and Complexity
di: Daskalakis, Constantinos, et al.
Pubblicazione: (2023)
di: Daskalakis, Constantinos, et al.
Pubblicazione: (2023)
Dimension-Free Correlated Sampling for the Hypersimplex
di: Joseph, et al.
Pubblicazione: (2025)
di: Joseph, et al.
Pubblicazione: (2025)
Thompson Sampling Itself is Differentially Private
di: Ou, Tingting, et al.
Pubblicazione: (2024)
di: Ou, Tingting, et al.
Pubblicazione: (2024)
Efficient Sample-optimal Learning of Gaussian Tree Models via Sample-optimal Testing of Gaussian Mutual Information
di: Gayen, Sutanu, et al.
Pubblicazione: (2024)
di: Gayen, Sutanu, et al.
Pubblicazione: (2024)
Finite Sample Bounds for Learning with Score Matching
di: Smedira, Devin, et al.
Pubblicazione: (2026)
di: Smedira, Devin, et al.
Pubblicazione: (2026)
Optimal Dimension-Free Sampling for Regularized Classification
di: Alishahi, Meysam, et al.
Pubblicazione: (2026)
di: Alishahi, Meysam, et al.
Pubblicazione: (2026)
Metalearning with Very Few Samples Per Task
di: Aliakbarpour, Maryam, et al.
Pubblicazione: (2023)
di: Aliakbarpour, Maryam, et al.
Pubblicazione: (2023)
Sharper Bounds for $\ell_p$ Sensitivity Sampling
di: Woodruff, David P., et al.
Pubblicazione: (2023)
di: Woodruff, David P., et al.
Pubblicazione: (2023)
Distribution Learning Meets Graph Structure Sampling
di: Bhattacharyya, Arnab, et al.
Pubblicazione: (2024)
di: Bhattacharyya, Arnab, et al.
Pubblicazione: (2024)
Sample-efficient Multiclass Calibration under $\ell_{p}$ Error
di: Bairaktari, Konstantina, et al.
Pubblicazione: (2025)
di: Bairaktari, Konstantina, et al.
Pubblicazione: (2025)
Structure-Aware Spectral Sparsification via Uniform Edge Sampling
di: He, Kaiwen, et al.
Pubblicazione: (2025)
di: He, Kaiwen, et al.
Pubblicazione: (2025)
Documenti analoghi
-
The tractability landscape of diffusion alignment: regularization, rewards, and computational primitives
di: Moitra, Ankur, et al.
Pubblicazione: (2026) -
Steering diffusion models with quadratic rewards: a fine-grained analysis
di: Moitra, Ankur, et al.
Pubblicazione: (2026) -
Exploration is Harder than Prediction: Cryptographically Separating Reinforcement Learning from Supervised Learning
di: Golowich, Noah, et al.
Pubblicazione: (2024) -
Regularized Robustly Reliable Learners and Instance Targeted Attacks
di: Blum, Avrim, et al.
Pubblicazione: (2024) -
Fit Like You Sample: Sample-Efficient Generalized Score Matching from Fast Mixing Diffusions
di: Qin, Yilong, et al.
Pubblicazione: (2023)