Off-policy Evaluation with Deeply-abstracted States
Fuente:
arXiv
Salvato in:
| Autori principali: | Hao, Meiling, Su, Pingfan, Hu, Liyuan, Szabo, Zoltan, Zhao, Qingyuan, Shi, Chengchun |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Scalable and Interpretable Scientific Discovery via Sparse Variational Gaussian Process Kolmogorov-Arnold Networks (SVGP KAN)
di: Ju, Y. Sungtaek
Pubblicazione: (2025)
di: Ju, Y. Sungtaek
Pubblicazione: (2025)
Uncertainty Quantification for Scientific Machine Learning using Sparse Variational Gaussian Process Kolmogorov-Arnold Networks (SVGP KAN)
di: Ju, Y. Sungtaek
Pubblicazione: (2025)
di: Ju, Y. Sungtaek
Pubblicazione: (2025)
Deep Learning for Solving and Estimating Dynamic Models in Economics and Finance
di: Scheidegger, Simon
Pubblicazione: (2026)
di: Scheidegger, Simon
Pubblicazione: (2026)
Memory-Efficient Training with In-Place FFT Implementation
di: Ding, Xinyu, et al.
Pubblicazione: (2025)
di: Ding, Xinyu, et al.
Pubblicazione: (2025)
Non-convex cost functionals in boosting algorithms and methods for panel selection
di: Visentin, Marco
Pubblicazione: (2001)
di: Visentin, Marco
Pubblicazione: (2001)
Distribution-Free Uncertainty Quantification for Continuous AI Agent Evaluation
di: Gao, Yuxuan, et al.
Pubblicazione: (2026)
di: Gao, Yuxuan, et al.
Pubblicazione: (2026)
Front-door Reducibility: Reducing ADMGs to the Standard Front-door Setting via a Graphical Criterion
di: Mao, Jianqiao, et al.
Pubblicazione: (2025)
di: Mao, Jianqiao, et al.
Pubblicazione: (2025)
An objective function for order preserving hierarchical clustering
di: Bakkelund, Daniel
Pubblicazione: (2021)
di: Bakkelund, Daniel
Pubblicazione: (2021)
Constant-Target Energy Matching: A Unified Framework for Continuous and Discrete Density Estimation
di: Zeng, Zhijun, et al.
Pubblicazione: (2026)
di: Zeng, Zhijun, et al.
Pubblicazione: (2026)
PhAIL: A Real-Robot VLA Benchmark and Distributional Methodology
di: Arkhangelskiy, Sergey
Pubblicazione: (2026)
di: Arkhangelskiy, Sergey
Pubblicazione: (2026)
Variational Search Distributions
di: Steinberg, Daniel M., et al.
Pubblicazione: (2024)
di: Steinberg, Daniel M., et al.
Pubblicazione: (2024)
Bayesian Optimization in Linear Time
di: Schneider, Jesse, et al.
Pubblicazione: (2026)
di: Schneider, Jesse, et al.
Pubblicazione: (2026)
Learning Nonlinear Regime Transitions via Semi-Parametric State-Space Models
di: Hiremath, Prakul Sunil
Pubblicazione: (2026)
di: Hiremath, Prakul Sunil
Pubblicazione: (2026)
CDFlow: Building Invertible Layers with Circulant and Diagonal Matrices
di: Feng, Xuchen, et al.
Pubblicazione: (2025)
di: Feng, Xuchen, et al.
Pubblicazione: (2025)
Evaluating Temporal Observation-Based Causal Discovery Techniques Applied to Road Driver Behaviour
di: Howard, Rhys, et al.
Pubblicazione: (2023)
di: Howard, Rhys, et al.
Pubblicazione: (2023)
Enhanced Random Subspace Local Projections for High-Dimensional Time Series Analysis
di: Khalid, Eman, et al.
Pubblicazione: (2026)
di: Khalid, Eman, et al.
Pubblicazione: (2026)
HAPEns: Hardware-Aware Post-Hoc Ensembling for Tabular Data
di: Maier, Jannis, et al.
Pubblicazione: (2026)
di: Maier, Jannis, et al.
Pubblicazione: (2026)
Blind-Spot Mass: A Good-Turing Framework for Quantifying Deployment Coverage Risk in Machine Learning Systems
di: Pal, Biplab, et al.
Pubblicazione: (2026)
di: Pal, Biplab, et al.
Pubblicazione: (2026)
Uncertainty Propagation Networks for Neural Ordinary Differential Equations
di: Jahanshahi, Hadi, et al.
Pubblicazione: (2025)
di: Jahanshahi, Hadi, et al.
Pubblicazione: (2025)
SmallML: Bayesian Transfer Learning for Small-Data Predictive Analytics
di: Leontev, Semen
Pubblicazione: (2025)
di: Leontev, Semen
Pubblicazione: (2025)
How Does the Pretraining Distribution Shape In-Context Learning? Task Selection, Generalization, and Robustness
di: Azizian, Waïss, et al.
Pubblicazione: (2025)
di: Azizian, Waïss, et al.
Pubblicazione: (2025)
The Role of Causal Features in Strategic Classification for Robustness and Alignment
di: Gois, Antonio, et al.
Pubblicazione: (2026)
di: Gois, Antonio, et al.
Pubblicazione: (2026)
Nyström Kernel Stein Discrepancy
di: Kalinke, Florian, et al.
Pubblicazione: (2024)
di: Kalinke, Florian, et al.
Pubblicazione: (2024)
Digging deeper: deep joint species distribution modeling reveals environmental drivers of Earthworm Communities
di: Si-moussi, Sara, et al.
Pubblicazione: (2025)
di: Si-moussi, Sara, et al.
Pubblicazione: (2025)
From Theory to Practice with RAVEN-UCB: Addressing Non-Stationarity in Multi-Armed Bandits through Variance Adaptation
di: Fang, Junyi, et al.
Pubblicazione: (2025)
di: Fang, Junyi, et al.
Pubblicazione: (2025)
Improving ML Training Data with Gold-Standard Quality Metrics
di: Barrett, Leslie, et al.
Pubblicazione: (2025)
di: Barrett, Leslie, et al.
Pubblicazione: (2025)
Randomized Approach to Matrix Completion: Applications in Recommendation Systems and Image Inpainting
di: Krajewska, Antonina, et al.
Pubblicazione: (2024)
di: Krajewska, Antonina, et al.
Pubblicazione: (2024)
Shallow ReLU$^s$ Networks in $L^p$-Type and Sobolev Spaces: Approximation and Path-Norm Controlled Generalization
di: Li, Weizhao, et al.
Pubblicazione: (2026)
di: Li, Weizhao, et al.
Pubblicazione: (2026)
TrajGPT: Controlled Synthetic Trajectory Generation Using a Multitask Transformer-Based Spatiotemporal Model
di: Hsu, Shang-Ling, et al.
Pubblicazione: (2024)
di: Hsu, Shang-Ling, et al.
Pubblicazione: (2024)
Divergence-free Linearized Neural Networks: Integral Representation and Optimal Approximation Rates
di: He, Juncai, et al.
Pubblicazione: (2026)
di: He, Juncai, et al.
Pubblicazione: (2026)
A Fast Convoluted Story: Scaling Probabilistic Inference for Integer Arithmetic
di: De Smet, Lennert, et al.
Pubblicazione: (2024)
di: De Smet, Lennert, et al.
Pubblicazione: (2024)
Proof-Carrying No-Arbitrage Surfaces: Constructive PCA-Smolyak Meets Chain-Consistent Diffusion with c-EMOT Certificates
di: Zhang, Jian'an
Pubblicazione: (2025)
di: Zhang, Jian'an
Pubblicazione: (2025)
Manifold-Adapted Sparse RBF-SINDy: Unbiased Library Construction and Unsupervised Discovery of Dynamical States in Turbulent Wall Flows
di: Perez-Cuadrado, Miguel, et al.
Pubblicazione: (2026)
di: Perez-Cuadrado, Miguel, et al.
Pubblicazione: (2026)
Learning What Evaluators Value: A Reliable Approach to Modeling Evaluator Preferences
di: Kitch, Madeline Celi, et al.
Pubblicazione: (2026)
di: Kitch, Madeline Celi, et al.
Pubblicazione: (2026)
Sinusoidal Initialization, Time for a New Start
di: Fernández-Hernández, Alberto, et al.
Pubblicazione: (2025)
di: Fernández-Hernández, Alberto, et al.
Pubblicazione: (2025)
Closing the Curvature Gap: Full Transformer Hessians and Their Implications for Scaling Laws
di: Petrov, Egor, et al.
Pubblicazione: (2025)
di: Petrov, Egor, et al.
Pubblicazione: (2025)
Unified Unbiased Variance Estimation for Maximum Mean Discrepancy: Robust Finite-Sample Performance with Imbalanced Data and Exact Acceleration under Null and Alternative Hypotheses
di: Zhong, Shijie, et al.
Pubblicazione: (2026)
di: Zhong, Shijie, et al.
Pubblicazione: (2026)
Bayesian Hierarchical Probabilistic Forecasting of Intraday Electricity Prices
di: Nickelsen, Daniel, et al.
Pubblicazione: (2024)
di: Nickelsen, Daniel, et al.
Pubblicazione: (2024)
The Reasoning-Creativity Trade-off: Toward Creativity-Driven Problem Solving
di: Luyten, Max Ruiz, et al.
Pubblicazione: (2026)
di: Luyten, Max Ruiz, et al.
Pubblicazione: (2026)
Actor-Critic Model Predictive Control: Differentiable Optimization meets Reinforcement Learning for Agile Flight
di: Romero, Angel, et al.
Pubblicazione: (2023)
di: Romero, Angel, et al.
Pubblicazione: (2023)
Documenti analoghi
-
Scalable and Interpretable Scientific Discovery via Sparse Variational Gaussian Process Kolmogorov-Arnold Networks (SVGP KAN)
di: Ju, Y. Sungtaek
Pubblicazione: (2025) -
Uncertainty Quantification for Scientific Machine Learning using Sparse Variational Gaussian Process Kolmogorov-Arnold Networks (SVGP KAN)
di: Ju, Y. Sungtaek
Pubblicazione: (2025) -
Deep Learning for Solving and Estimating Dynamic Models in Economics and Finance
di: Scheidegger, Simon
Pubblicazione: (2026) -
Memory-Efficient Training with In-Place FFT Implementation
di: Ding, Xinyu, et al.
Pubblicazione: (2025) -
Non-convex cost functionals in boosting algorithms and methods for panel selection
di: Visentin, Marco
Pubblicazione: (2001)