FraPPE: Fast and Efficient Preference-based Pure Exploration
Fuente:
arXiv
Saved in:
| Main Authors: | Das, Udvas, Shukla, Apurv, Basu, Debabrota |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Performative Policy Gradient: Optimality in Performative Reinforcement Learning
by: Basu, Debabrota, et al.
Published: (2025)
by: Basu, Debabrota, et al.
Published: (2025)
Preference-based Pure Exploration
by: Shukla, Apurv, et al.
Published: (2024)
by: Shukla, Apurv, et al.
Published: (2024)
Learning to Explore with Lagrangians for Bandits under Unknown Linear Constraints
by: Das, Udvas, et al.
Published: (2024)
by: Das, Udvas, et al.
Published: (2024)
Straight-Through meets Sparse Recovery: the Support Exploration Algorithm
by: Mohamed, Mimoun, et al.
Published: (2023)
by: Mohamed, Mimoun, et al.
Published: (2023)
Inverse Mixed-Integer Programming: Learning Constraints then Objective Functions
by: Kitaoka, Akira
Published: (2025)
by: Kitaoka, Akira
Published: (2025)
Precise gradient descent training dynamics for finite-width multi-layer neural networks
by: Han, Qiyang, et al.
Published: (2025)
by: Han, Qiyang, et al.
Published: (2025)
Learning to Fuse Temporal Proximity Networks: A Case Study in Chimpanzee Social Interactions
by: He, Yixuan, et al.
Published: (2025)
by: He, Yixuan, et al.
Published: (2025)
Statistical and Algorithmic Foundations of Reinforcement Learning
by: Chi, Yuejie, et al.
Published: (2025)
by: Chi, Yuejie, et al.
Published: (2025)
Decoupled Continuous-Time Reinforcement Learning via Hamiltonian Flow
by: Nguyen, Minh
Published: (2026)
by: Nguyen, Minh
Published: (2026)
Byzantine Machine Learning: MultiKrum and an optimal notion of robustness
by: Bareilles, Gilles, et al.
Published: (2026)
by: Bareilles, Gilles, et al.
Published: (2026)
Sinkhorn Based Associative Memory Retrieval Using Spherical Hellinger Kantorovich Dynamics
by: Mustafi, Aratrika, et al.
Published: (2026)
by: Mustafi, Aratrika, et al.
Published: (2026)
A Differential and Pointwise Control Approach to Reinforcement Learning
by: Nguyen, Minh, et al.
Published: (2024)
by: Nguyen, Minh, et al.
Published: (2024)
Training Dynamics of Multi-Head Softmax Attention for In-Context Learning: Emergence, Convergence, and Optimality
by: Chen, Siyu, et al.
Published: (2024)
by: Chen, Siyu, et al.
Published: (2024)
Sail into the Headwind: Alignment via Robust Rewards and Dynamic Labels against Reward Hacking
by: Rashidinejad, Paria, et al.
Published: (2024)
by: Rashidinejad, Paria, et al.
Published: (2024)
Optimism Stabilizes Thompson Sampling for Adaptive Inference
by: Yan, Shunxing, et al.
Published: (2026)
by: Yan, Shunxing, et al.
Published: (2026)
Smooth Non-Stationary Bandits
by: Jia, Su, et al.
Published: (2023)
by: Jia, Su, et al.
Published: (2023)
Piecewise Polynomial Regression of Tame Functions via Integer Programming
by: Bareilles, Gilles, et al.
Published: (2023)
by: Bareilles, Gilles, et al.
Published: (2023)
The Fair Game: Auditing & Debiasing AI Algorithms Over Time
by: Basu, Debabrota, et al.
Published: (2025)
by: Basu, Debabrota, et al.
Published: (2025)
Differentially Private High Dimensional Bandits
by: Shukla, Apurv
Published: (2024)
by: Shukla, Apurv
Published: (2024)
Geometry-induced Regularization in Deep ReLU Neural Networks
by: Bona-Pellissier, Joachim, et al.
Published: (2024)
by: Bona-Pellissier, Joachim, et al.
Published: (2024)
Efficient Group Lasso Regularized Rank Regression with Data-Driven Parameter Determination
by: Lin, Meixia, et al.
Published: (2025)
by: Lin, Meixia, et al.
Published: (2025)
Data-Efficient Non-Gaussian Semi-Nonparametric Density Estimation for Nonlinear Dynamical Systems
by: Liao, Aaron R., et al.
Published: (2026)
by: Liao, Aaron R., et al.
Published: (2026)
Fast convergence of the Expectation Maximization algorithm under a logarithmic Sobolev inequality
by: Caprio, Rocco, et al.
Published: (2024)
by: Caprio, Rocco, et al.
Published: (2024)
Convergence of flow-based generative models via proximal gradient descent in Wasserstein space
by: Cheng, Xiuyuan, et al.
Published: (2023)
by: Cheng, Xiuyuan, et al.
Published: (2023)
Comparing Few to Rank Many: Active Human Preference Learning using Randomized Frank-Wolfe
by: Thekumparampil, Kiran Koshy, et al.
Published: (2024)
by: Thekumparampil, Kiran Koshy, et al.
Published: (2024)
Fast Spawn\&Prune (FS\&P): Global convergence of stochastic conic particle gradient descent via birth/death process
by: De Castro, Yohann, et al.
Published: (2026)
by: De Castro, Yohann, et al.
Published: (2026)
Analytic Bridge Diffusions for Controlled Path Generation
by: Chertkov, Michael
Published: (2026)
by: Chertkov, Michael
Published: (2026)
Causal Invariance Learning via Efficient Nonconvex Optimization
by: Wang, Zhenyu, et al.
Published: (2024)
by: Wang, Zhenyu, et al.
Published: (2024)
A Graphical Global Optimization Framework for Parameter Estimation of Statistical Models with Nonconvex Regularization Functions
by: Davarnia, Danial, et al.
Published: (2025)
by: Davarnia, Danial, et al.
Published: (2025)
ECPv2: Fast, Efficient, and Scalable Global Optimization of Lipschitz Functions
by: Fourati, Fares, et al.
Published: (2025)
by: Fourati, Fares, et al.
Published: (2025)
Delightful Exploration
by: Osband, Ian
Published: (2026)
by: Osband, Ian
Published: (2026)
Stochastic Optimization with Optimal Importance Sampling
by: Aolaritei, Liviu, et al.
Published: (2025)
by: Aolaritei, Liviu, et al.
Published: (2025)
Joint learning of a network of linear dynamical systems via total variation penalization
by: Donnat, Claire, et al.
Published: (2025)
by: Donnat, Claire, et al.
Published: (2025)
A review of NMF, PLSA, LBA, EMA, and LCA with a focus on the identifiability issue
by: Qi, Qianqian, et al.
Published: (2025)
by: Qi, Qianqian, et al.
Published: (2025)
Error Analysis of Triangular Optimal Transport Maps for Filtering
by: Al-Jarrah, Mohammad, et al.
Published: (2025)
by: Al-Jarrah, Mohammad, et al.
Published: (2025)
Online Inference of Constrained Optimization: Primal-Dual Optimality and Sequential Quadratic Programming
by: Gao, Yihang, et al.
Published: (2025)
by: Gao, Yihang, et al.
Published: (2025)
Mixing Times and Privacy Analysis for the Projected Langevin Algorithm under a Modulus of Continuity
by: Bravo, Mario, et al.
Published: (2025)
by: Bravo, Mario, et al.
Published: (2025)
Extreme mass distributions for quasi-copulas
by: Omladič, Matjaž, et al.
Published: (2025)
by: Omladič, Matjaž, et al.
Published: (2025)
Learning an Optimal Assortment Policy under Observational Data
by: Han, Yuxuan, et al.
Published: (2025)
by: Han, Yuxuan, et al.
Published: (2025)
An Elementary Proof of the Near Optimality of LogSumExp Smoothing
by: Samakhoana, Thabo, et al.
Published: (2025)
by: Samakhoana, Thabo, et al.
Published: (2025)
Similar Items
-
Performative Policy Gradient: Optimality in Performative Reinforcement Learning
by: Basu, Debabrota, et al.
Published: (2025) -
Preference-based Pure Exploration
by: Shukla, Apurv, et al.
Published: (2024) -
Learning to Explore with Lagrangians for Bandits under Unknown Linear Constraints
by: Das, Udvas, et al.
Published: (2024) -
Straight-Through meets Sparse Recovery: the Support Exploration Algorithm
by: Mohamed, Mimoun, et al.
Published: (2023) -
Inverse Mixed-Integer Programming: Learning Constraints then Objective Functions
by: Kitaoka, Akira
Published: (2025)