Guardado en:
| Autores principales: | Xu, Ziping, Xu, Zifan, Jiang, Runxuan, Stone, Peter, Tewari, Ambuj |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2403.01636 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
A Primal-Dual Algorithm for Offline Constrained Reinforcement Learning with Linear MDPs
por: Hong, Kihyuk, et al.
Publicado: (2024)
por: Hong, Kihyuk, et al.
Publicado: (2024)
Near Optimal Pure Exploration in Logistic Bandits
por: Rivera, Eduardo Ochoa, et al.
Publicado: (2024)
por: Rivera, Eduardo Ochoa, et al.
Publicado: (2024)
If generative AI is the answer, what is the question?
por: Tewari, Ambuj
Publicado: (2025)
por: Tewari, Ambuj
Publicado: (2025)
Offline Constrained Reinforcement Learning under Partial Data Coverage
por: Ko, Seokmin, et al.
Publicado: (2025)
por: Ko, Seokmin, et al.
Publicado: (2025)
Operator Learning: A Statistical Perspective
por: Subedi, Unique, et al.
Publicado: (2025)
por: Subedi, Unique, et al.
Publicado: (2025)
On the Benefits of Active Data Collection in Operator Learning
por: Subedi, Unique, et al.
Publicado: (2024)
por: Subedi, Unique, et al.
Publicado: (2024)
A Computationally Efficient Algorithm for Infinite-Horizon Average-Reward Linear MDPs
por: Hong, Kihyuk, et al.
Publicado: (2025)
por: Hong, Kihyuk, et al.
Publicado: (2025)
Learning to Partially Defer for Sequences
por: Rayan, Sahana, et al.
Publicado: (2025)
por: Rayan, Sahana, et al.
Publicado: (2025)
Is Zero-Shot Super-Resolution Possible in Operator Learning?
por: Subedi, Unique, et al.
Publicado: (2026)
por: Subedi, Unique, et al.
Publicado: (2026)
Controlling Statistical, Discretization, and Truncation Errors in Learning Fourier Linear Operators
por: Subedi, Unique, et al.
Publicado: (2024)
por: Subedi, Unique, et al.
Publicado: (2024)
Dexterous Legged Locomotion in Confined 3D Spaces with Reinforcement Learning
por: Xu, Zifan, et al.
Publicado: (2024)
por: Xu, Zifan, et al.
Publicado: (2024)
Quantum Learning Theory Beyond Batch Binary Classification
por: Mohan, Preetham, et al.
Publicado: (2023)
por: Mohan, Preetham, et al.
Publicado: (2023)
On the Minimax Regret in Online Ranking with Top-k Feedback
por: Zhang, Mingyuan, et al.
Publicado: (2023)
por: Zhang, Mingyuan, et al.
Publicado: (2023)
An Asymptotically Optimal Algorithm for the Convex Hull Membership Problem
por: Qiao, Gang, et al.
Publicado: (2023)
por: Qiao, Gang, et al.
Publicado: (2023)
Distribution-Free Robust Predict-Then-Optimize in Function Spaces
por: Patel, Yash, et al.
Publicado: (2026)
por: Patel, Yash, et al.
Publicado: (2026)
Online Learning with Set-Valued Feedback
por: Raman, Vinod, et al.
Publicado: (2023)
por: Raman, Vinod, et al.
Publicado: (2023)
Online Infinite-Dimensional Regression: Learning Linear Operators
por: Raman, Vinod, et al.
Publicado: (2023)
por: Raman, Vinod, et al.
Publicado: (2023)
Continuum Transformers Perform In-Context Learning by Operator Gradient Descent
por: Mishra, Abhiti, et al.
Publicado: (2025)
por: Mishra, Abhiti, et al.
Publicado: (2025)
Operator Learning for Schrödinger Equation: Unitarity, Error Bounds, and Time Generalization
por: Patel, Yash, et al.
Publicado: (2025)
por: Patel, Yash, et al.
Publicado: (2025)
Compute Aligned Training: Optimizing for Test Time Inference
por: Ousherovitch, Adam, et al.
Publicado: (2026)
por: Ousherovitch, Adam, et al.
Publicado: (2026)
On Next-Token Prediction in LLMs: How End Goals Determine the Consistency of Decoding Algorithms
por: Trauger, Jacob, et al.
Publicado: (2025)
por: Trauger, Jacob, et al.
Publicado: (2025)
Online Classification with Predictions
por: Raman, Vinod, et al.
Publicado: (2024)
por: Raman, Vinod, et al.
Publicado: (2024)
Optimal Thresholding Linear Bandit
por: Rivera, Eduardo Ochoa, et al.
Publicado: (2024)
por: Rivera, Eduardo Ochoa, et al.
Publicado: (2024)
Online Conformal Prediction: Enforcing monotonicity via Online Optimization
por: Rivera, Eduardo Ochoa, et al.
Publicado: (2026)
por: Rivera, Eduardo Ochoa, et al.
Publicado: (2026)
Generator-Mediated Bandits: Thompson Sampling for GenAI-Powered Adaptive Interventions
por: Brooks, Marc, et al.
Publicado: (2025)
por: Brooks, Marc, et al.
Publicado: (2025)
Reinforcement Learning for Infinite-Horizon Average-Reward Linear MDPs via Approximation by Discounted-Reward MDPs
por: Hong, Kihyuk, et al.
Publicado: (2024)
por: Hong, Kihyuk, et al.
Publicado: (2024)
Generation through the lens of learning theory
por: Li, Jiaxun, et al.
Publicado: (2024)
por: Li, Jiaxun, et al.
Publicado: (2024)
The Complexity of Sequential Prediction in Dynamical Systems
por: Raman, Vinod, et al.
Publicado: (2024)
por: Raman, Vinod, et al.
Publicado: (2024)
Smoothed Online Classification can be Harder than Batch Classification
por: Raman, Vinod, et al.
Publicado: (2024)
por: Raman, Vinod, et al.
Publicado: (2024)
A Combinatorial Characterization of Supervised Online Learnability
por: Raman, Vinod, et al.
Publicado: (2023)
por: Raman, Vinod, et al.
Publicado: (2023)
A Characterization of Multioutput Learnability
por: Raman, Vinod, et al.
Publicado: (2023)
por: Raman, Vinod, et al.
Publicado: (2023)
On Generation in Metric Spaces
por: Li, Jiaxun, et al.
Publicado: (2026)
por: Li, Jiaxun, et al.
Publicado: (2026)
Characterizing the Multiclass Learnability of Forgiving 0-1 Loss Functions
por: Trauger, Jacob, et al.
Publicado: (2025)
por: Trauger, Jacob, et al.
Publicado: (2025)
A Greedy PDE Router for Blending Neural Operators and Classical Methods
por: Rayan, Sahana, et al.
Publicado: (2025)
por: Rayan, Sahana, et al.
Publicado: (2025)
Leveraging Offline Data in Linear Latent Contextual Bandits
por: Kausik, Chinmaya, et al.
Publicado: (2024)
por: Kausik, Chinmaya, et al.
Publicado: (2024)
On the Computational Complexity of Private High-dimensional Model Selection
por: Roy, Saptarshi, et al.
Publicado: (2023)
por: Roy, Saptarshi, et al.
Publicado: (2023)
Understanding Best Subset Selection: A Tale of Two C(omplex)ities
por: Roy, Saptarshi, et al.
Publicado: (2023)
por: Roy, Saptarshi, et al.
Publicado: (2023)
Latency-Aware Contextual Bandit: Application to Cryo-EM Data Collection
por: Wei, Lai, et al.
Publicado: (2024)
por: Wei, Lai, et al.
Publicado: (2024)
Conformal Prediction for Ensembles: Improving Efficiency via Score-Based Aggregation
por: Rivera, Eduardo Ochoa, et al.
Publicado: (2024)
por: Rivera, Eduardo Ochoa, et al.
Publicado: (2024)
Bandit Social Learning: Exploration under Myopic Behavior
por: Banihashem, Kiarash, et al.
Publicado: (2023)
por: Banihashem, Kiarash, et al.
Publicado: (2023)
Ejemplares similares
-
A Primal-Dual Algorithm for Offline Constrained Reinforcement Learning with Linear MDPs
por: Hong, Kihyuk, et al.
Publicado: (2024) -
Near Optimal Pure Exploration in Logistic Bandits
por: Rivera, Eduardo Ochoa, et al.
Publicado: (2024) -
If generative AI is the answer, what is the question?
por: Tewari, Ambuj
Publicado: (2025) -
Offline Constrained Reinforcement Learning under Partial Data Coverage
por: Ko, Seokmin, et al.
Publicado: (2025) -
Operator Learning: A Statistical Perspective
por: Subedi, Unique, et al.
Publicado: (2025)