Guardado en:
| Autores principales: | Wang, Jimmy, Che, Ethan, Jiang, Daniel R., Namkoong, Hongseok |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2408.04531 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Optimization-Driven Adaptive Experimentation
por: Che, Ethan, et al.
Publicado: (2024)
por: Che, Ethan, et al.
Publicado: (2024)
Differentiable Discrete Event Simulation for Queuing Network Control
por: Che, Ethan, et al.
Publicado: (2024)
por: Che, Ethan, et al.
Publicado: (2024)
Adaptive Elicitation of Latent Information Using Natural Language
por: Wang, Jimmy, et al.
Publicado: (2025)
por: Wang, Jimmy, et al.
Publicado: (2025)
Empirical Likelihood for Nonsmooth Functionals
por: Namkoong, Hongseok
Publicado: (2026)
por: Namkoong, Hongseok
Publicado: (2026)
QGym: Scalable Simulation and Benchmarking of Queuing Network Controllers
por: Chen, Haozhe, et al.
Publicado: (2024)
por: Chen, Haozhe, et al.
Publicado: (2024)
Exchangeable Sequence Models Quantify Uncertainty Over Latent Concepts
por: Ye, Naimeng, et al.
Publicado: (2024)
por: Ye, Naimeng, et al.
Publicado: (2024)
A Planning Framework for Adaptive Labeling
por: Mittal, Daksh, et al.
Publicado: (2025)
por: Mittal, Daksh, et al.
Publicado: (2025)
Benchmarking In-context Experiential Learning Through Repeated Product Recommendations
por: Yang, Gilbert, et al.
Publicado: (2025)
por: Yang, Gilbert, et al.
Publicado: (2025)
A Broader View of Thompson Sampling
por: Qu, Yanlin, et al.
Publicado: (2025)
por: Qu, Yanlin, et al.
Publicado: (2025)
A Sensitivity Approach to Causal Inference Under Limited Overlap
por: Ma, Yuanzhe, et al.
Publicado: (2025)
por: Ma, Yuanzhe, et al.
Publicado: (2025)
Design and Scheduling of an AI-based Queueing System
por: Lee, Jiung, et al.
Publicado: (2024)
por: Lee, Jiung, et al.
Publicado: (2024)
Distilled Thompson Sampling: Practical and Efficient Thompson Sampling via Imitation Learning
por: Namkoong, Hongseok, et al.
Publicado: (2020)
por: Namkoong, Hongseok, et al.
Publicado: (2020)
Minimax Optimal Estimation of Stability Under Distribution Shift
por: Namkoong, Hongseok, et al.
Publicado: (2022)
por: Namkoong, Hongseok, et al.
Publicado: (2022)
Rethinking Distribution Shifts: Empirical Analysis and Inductive Modeling for Tabular Data
por: Wang, Tianyu, et al.
Publicado: (2023)
por: Wang, Tianyu, et al.
Publicado: (2023)
Data-Driven Stochastic Modeling Using Autoregressive Sequence Models: Translating Event Tables to Queueing Dynamics
por: Mittal, Daksh, et al.
Publicado: (2025)
por: Mittal, Daksh, et al.
Publicado: (2025)
Active Exploration via Autoregressive Generation of Missing Data
por: Cai, Tiffany Tianhui, et al.
Publicado: (2024)
por: Cai, Tiffany Tianhui, et al.
Publicado: (2024)
Evaluating Model Performance Under Worst-case Subpopulations
por: Li, Mike, et al.
Publicado: (2024)
por: Li, Mike, et al.
Publicado: (2024)
Contextual Thompson Sampling via Generation of Missing Data
por: Zhang, Kelly W., et al.
Publicado: (2025)
por: Zhang, Kelly W., et al.
Publicado: (2025)
C-Learner: Constrained Learning for Causal Inference
por: Cai, Tiffany Tianhui, et al.
Publicado: (2024)
por: Cai, Tiffany Tianhui, et al.
Publicado: (2024)
Architectural and Inferential Inductive Biases For Exchangeable Sequence Modeling
por: Mittal, Daksh, et al.
Publicado: (2025)
por: Mittal, Daksh, et al.
Publicado: (2025)
LLM Embeddings Improve Test-time Adaptation to Tabular $Y|X$-Shifts
por: Zeng, Yibo, et al.
Publicado: (2024)
por: Zeng, Yibo, et al.
Publicado: (2024)
DRO: A Python Library for Distributionally Robust Optimization in Machine Learning
por: Liu, Jiashuo, et al.
Publicado: (2025)
por: Liu, Jiashuo, et al.
Publicado: (2025)
LLM Swiss Round: Aggregating Multi-Benchmark Performance via Competitive Swiss-System Dynamics
por: Liu, Jiashuo, et al.
Publicado: (2025)
por: Liu, Jiashuo, et al.
Publicado: (2025)
Data Mixture Optimization: A Multi-fidelity Multi-scale Bayesian Framework
por: Yen, Thomson, et al.
Publicado: (2025)
por: Yen, Thomson, et al.
Publicado: (2025)
BoxingGym: Benchmarking Progress in Automated Experimental Design and Model Discovery
por: Gandhi, Kanishk, et al.
Publicado: (2025)
por: Gandhi, Kanishk, et al.
Publicado: (2025)
OceanGym: A Benchmark Environment for Underwater Embodied Agents
por: Xue, Yida, et al.
Publicado: (2025)
por: Xue, Yida, et al.
Publicado: (2025)
SynthTools: A Framework for Scaling Synthetic Tools for Agent Development
por: Castellani, Tommaso, et al.
Publicado: (2025)
por: Castellani, Tommaso, et al.
Publicado: (2025)
Mining--Gym: A Configurable RL Benchmarking Environment for Truck Dispatch Scheduling
por: Banerjee, Chayan, et al.
Publicado: (2025)
por: Banerjee, Chayan, et al.
Publicado: (2025)
Stochastic Gradient Descent with Adaptive Data
por: Che, Ethan, et al.
Publicado: (2024)
por: Che, Ethan, et al.
Publicado: (2024)
PersonalLLM: Tailoring LLMs to Individual Preferences
por: Zollo, Thomas P., et al.
Publicado: (2024)
por: Zollo, Thomas P., et al.
Publicado: (2024)
SeekerGym: A Benchmark for Reliable Information Seeking
por: Kim, Remy, et al.
Publicado: (2026)
por: Kim, Remy, et al.
Publicado: (2026)
Gym-Anything: Turn any Software into an Agent Environment
por: Aggarwal, Pranjal, et al.
Publicado: (2026)
por: Aggarwal, Pranjal, et al.
Publicado: (2026)
BlenderGym: Benchmarking Foundational Model Systems for Graphics Editing
por: Gu, Yunqi, et al.
Publicado: (2025)
por: Gu, Yunqi, et al.
Publicado: (2025)
CUA-Gym: Scaling Verifiable Training Environments and Tasks for Computer-Use Agents
por: Wang, Bowen, et al.
Publicado: (2026)
por: Wang, Bowen, et al.
Publicado: (2026)
UserBench: An Interactive Gym Environment for User-Centric Agents
por: Qian, Cheng, et al.
Publicado: (2025)
por: Qian, Cheng, et al.
Publicado: (2025)
Memory Gym: Towards Endless Tasks to Benchmark Memory Capabilities of Agents
por: Pleines, Marco, et al.
Publicado: (2023)
por: Pleines, Marco, et al.
Publicado: (2023)
PeersimGym: An Environment for Solving the Task Offloading Problem with Reinforcement Learning
por: Metelo, Frederico, et al.
Publicado: (2024)
por: Metelo, Frederico, et al.
Publicado: (2024)
AbideGym: Turning Static RL Worlds into Adaptive Challenges
por: Aryan, Abi, et al.
Publicado: (2025)
por: Aryan, Abi, et al.
Publicado: (2025)
EduGym: An Environment and Notebook Suite for Reinforcement Learning Education
por: Moerland, Thomas M., et al.
Publicado: (2023)
por: Moerland, Thomas M., et al.
Publicado: (2023)
CrystalGym: A New Benchmark for Materials Discovery Using Reinforcement Learning
por: Govindarajan, Prashant, et al.
Publicado: (2025)
por: Govindarajan, Prashant, et al.
Publicado: (2025)
Ejemplares similares
-
Optimization-Driven Adaptive Experimentation
por: Che, Ethan, et al.
Publicado: (2024) -
Differentiable Discrete Event Simulation for Queuing Network Control
por: Che, Ethan, et al.
Publicado: (2024) -
Adaptive Elicitation of Latent Information Using Natural Language
por: Wang, Jimmy, et al.
Publicado: (2025) -
Empirical Likelihood for Nonsmooth Functionals
por: Namkoong, Hongseok
Publicado: (2026) -
QGym: Scalable Simulation and Benchmarking of Queuing Network Controllers
por: Chen, Haozhe, et al.
Publicado: (2024)