Guardado en:
| Autores principales: | Ballentine, Alex E., Bapat, Nachiket U., Cowlagi, Raghvendra V. |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2604.02438 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Synthetic Data Generation for Minimum-Exposure Navigation in a Time-Varying Environment using Generative AI Models
por: Bapat, Nachiket U., et al.
Publicado: (2025)
por: Bapat, Nachiket U., et al.
Publicado: (2025)
Inverse Reinforcement Learning for Minimum-Exposure Paths in Spatiotemporally Varying Scalar Fields
por: Ballentine, Alexandra E., et al.
Publicado: (2025)
por: Ballentine, Alexandra E., et al.
Publicado: (2025)
Case Studies of Generative Machine Learning Models for Dynamical Systems
por: Bapat, Nachiket U., et al.
Publicado: (2025)
por: Bapat, Nachiket U., et al.
Publicado: (2025)
Trajectory Optimization for Minimum Threat Exposure using Physics-Informed Neural Networks
por: Ballentine, Alexandra E., et al.
Publicado: (2025)
por: Ballentine, Alexandra E., et al.
Publicado: (2025)
Guided Data Augmentation for Offline Reinforcement Learning and Imitation Learning
por: Corrado, Nicholas E., et al.
Publicado: (2023)
por: Corrado, Nicholas E., et al.
Publicado: (2023)
Hierarchical Reinforcement Learning Framework for Stochastic Spaceflight Campaign Design
por: Takubo, Yuji, et al.
Publicado: (2021)
por: Takubo, Yuji, et al.
Publicado: (2021)
Optimal Coupled Sensor Placement and Path-Planning in Unknown Time-Varying Environments
por: Poudel, Prakash, et al.
Publicado: (2025)
por: Poudel, Prakash, et al.
Publicado: (2025)
Using Offline Data to Speed Up Reinforcement Learning in Procedurally Generated Environments
por: Andres, Alain, et al.
Publicado: (2023)
por: Andres, Alain, et al.
Publicado: (2023)
Mitigating Data Scarcity in Time Series Analysis: A Foundation Model with Series-Symbol Data Generation
por: Wang, Wenxuan, et al.
Publicado: (2025)
por: Wang, Wenxuan, et al.
Publicado: (2025)
Unsupervised Data Generation for Offline Reinforcement Learning: A Perspective from Model
por: He, Shuncheng, et al.
Publicado: (2025)
por: He, Shuncheng, et al.
Publicado: (2025)
An Investigation of Offline Reinforcement Learning in Factorisable Action Spaces
por: Beeson, Alex, et al.
Publicado: (2024)
por: Beeson, Alex, et al.
Publicado: (2024)
Probabilistic Forecasting of Radiation Exposure for Spaceflight
por: Gurav, Rutuja, et al.
Publicado: (2024)
por: Gurav, Rutuja, et al.
Publicado: (2024)
BeamVQ: Beam Search with Vector Quantization to Mitigate Data Scarcity in Physical Spatiotemporal Forecasting
por: Wang, Weiyan, et al.
Publicado: (2025)
por: Wang, Weiyan, et al.
Publicado: (2025)
Model-Free Robust $ϕ$-Divergence Reinforcement Learning Using Both Offline and Online Data
por: Panaganti, Kishan, et al.
Publicado: (2024)
por: Panaganti, Kishan, et al.
Publicado: (2024)
A Survey of Reinforcement Learning for Large Language Models under Data Scarcity: Challenges and Solutions
por: Yu, Zhiyin, et al.
Publicado: (2026)
por: Yu, Zhiyin, et al.
Publicado: (2026)
Temporal Abstraction in Reinforcement Learning with Offline Data
por: Ayyagari, Ranga Shaarad, et al.
Publicado: (2024)
por: Ayyagari, Ranga Shaarad, et al.
Publicado: (2024)
Offline Reinforcement Learning with Domain-Unlabeled Data
por: Nishimori, Soichiro, et al.
Publicado: (2024)
por: Nishimori, Soichiro, et al.
Publicado: (2024)
The Generalization Gap in Offline Reinforcement Learning
por: Mediratta, Ishita, et al.
Publicado: (2023)
por: Mediratta, Ishita, et al.
Publicado: (2023)
Minimum-Time Sequential Traversal by a Team of Small Unmanned Aerial Vehicles in an Unknown Environment with Winds
por: DesRoches, Jeffrey A., et al.
Publicado: (2025)
por: DesRoches, Jeffrey A., et al.
Publicado: (2025)
Deep Transfer $Q$-Learning for Offline Non-Stationary Reinforcement Learning
por: Chai, Jinhang, et al.
Publicado: (2025)
por: Chai, Jinhang, et al.
Publicado: (2025)
Model-Based Offline Reinforcement Learning with Adversarial Data Augmentation
por: Cao, Hongye, et al.
Publicado: (2025)
por: Cao, Hongye, et al.
Publicado: (2025)
Hints-In-Browser: Benchmarking Language Models for Programming Feedback Generation
por: Kotalwar, Nachiket, et al.
Publicado: (2024)
por: Kotalwar, Nachiket, et al.
Publicado: (2024)
Beyond the Generative Learning Trilemma: Generative Model Assessment in Data Scarcity Domains
por: Salmè, Marco, et al.
Publicado: (2025)
por: Salmè, Marco, et al.
Publicado: (2025)
Time Series Generation Under Data Scarcity: A Unified Generative Modeling Approach
por: Gonen, Tal, et al.
Publicado: (2025)
por: Gonen, Tal, et al.
Publicado: (2025)
Data Center Cooling System Optimization Using Offline Reinforcement Learning
por: Zhan, Xianyuan, et al.
Publicado: (2025)
por: Zhan, Xianyuan, et al.
Publicado: (2025)
Jump Diffusion-Informed Neural Networks with Transfer Learning for Accurate American Option Pricing under Data Scarcity
por: Sun, Qiguo, et al.
Publicado: (2024)
por: Sun, Qiguo, et al.
Publicado: (2024)
A Primer on Variational Inference for Physics-Informed Deep Generative Modelling
por: Glyn-Davies, Alex, et al.
Publicado: (2024)
por: Glyn-Davies, Alex, et al.
Publicado: (2024)
Trajectory-Level Data Augmentation for Offline Reinforcement Learning
por: Schmähling, Tobias, et al.
Publicado: (2026)
por: Schmähling, Tobias, et al.
Publicado: (2026)
Leveraging Large Language Models to Address Data Scarcity in Machine Learning: Applications in Graphene Synthesis
por: Biswajeet, Devi Dutta, et al.
Publicado: (2025)
por: Biswajeet, Devi Dutta, et al.
Publicado: (2025)
Offline Reinforcement Learning with Generative Trajectory Policies
por: Feng, Xinsong, et al.
Publicado: (2025)
por: Feng, Xinsong, et al.
Publicado: (2025)
Doubly Mild Generalization for Offline Reinforcement Learning
por: Mao, Yixiu, et al.
Publicado: (2024)
por: Mao, Yixiu, et al.
Publicado: (2024)
Data-Incremental Continual Offline Reinforcement Learning
por: Gai, Sibo, et al.
Publicado: (2024)
por: Gai, Sibo, et al.
Publicado: (2024)
Tackling Data Corruption in Offline Reinforcement Learning via Sequence Modeling
por: Xu, Jiawei, et al.
Publicado: (2024)
por: Xu, Jiawei, et al.
Publicado: (2024)
Using Non-Expert Data to Robustify Imitation Learning via Offline Reinforcement Learning
por: Huang, Kevin, et al.
Publicado: (2025)
por: Huang, Kevin, et al.
Publicado: (2025)
Switching the Loss Reduces the Cost in Batch (Offline) Reinforcement Learning
por: Ayoub, Alex, et al.
Publicado: (2024)
por: Ayoub, Alex, et al.
Publicado: (2024)
Benchmarks for Reinforcement Learning with Biased Offline Data and Imperfect Simulators
por: Linial, Ori, et al.
Publicado: (2024)
por: Linial, Ori, et al.
Publicado: (2024)
Active Advantage-Aligned Online Reinforcement Learning with Offline Data
por: Liu, Xuefeng, et al.
Publicado: (2025)
por: Liu, Xuefeng, et al.
Publicado: (2025)
Offline Constrained Reinforcement Learning under Partial Data Coverage
por: Ko, Seokmin, et al.
Publicado: (2025)
por: Ko, Seokmin, et al.
Publicado: (2025)
FAWAC: Feasibility Informed Advantage Weighted Regression for Persistent Safety in Offline Reinforcement Learning
por: Koirala, Prajwal, et al.
Publicado: (2024)
por: Koirala, Prajwal, et al.
Publicado: (2024)
Offline Trajectory Optimization for Offline Reinforcement Learning
por: Zhao, Ziqi, et al.
Publicado: (2024)
por: Zhao, Ziqi, et al.
Publicado: (2024)
Ejemplares similares
-
Synthetic Data Generation for Minimum-Exposure Navigation in a Time-Varying Environment using Generative AI Models
por: Bapat, Nachiket U., et al.
Publicado: (2025) -
Inverse Reinforcement Learning for Minimum-Exposure Paths in Spatiotemporally Varying Scalar Fields
por: Ballentine, Alexandra E., et al.
Publicado: (2025) -
Case Studies of Generative Machine Learning Models for Dynamical Systems
por: Bapat, Nachiket U., et al.
Publicado: (2025) -
Trajectory Optimization for Minimum Threat Exposure using Physics-Informed Neural Networks
por: Ballentine, Alexandra E., et al.
Publicado: (2025) -
Guided Data Augmentation for Offline Reinforcement Learning and Imitation Learning
por: Corrado, Nicholas E., et al.
Publicado: (2023)