Algorithmic Guarantees for Distilling Supervised and Offline RL Datasets
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Gupta, Aaryan, Saket, Rishi, Raghuveer, Aravindan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Dense and Diverse Goal Coverage in Multi Goal Reinforcement Learning
von: Singh, Sagalpreet, et al.
Veröffentlicht: (2025)
von: Singh, Sagalpreet, et al.
Veröffentlicht: (2025)
LLP-Bench: A Large Scale Tabular Benchmark for Learning from Label Proportions
von: Brahmbhatt, Anand, et al.
Veröffentlicht: (2023)
von: Brahmbhatt, Anand, et al.
Veröffentlicht: (2023)
Aggregating Data for Optimal and Private Learning
von: Agarwal, Sushant, et al.
Veröffentlicht: (2024)
von: Agarwal, Sushant, et al.
Veröffentlicht: (2024)
Learning from Label Proportions and Covariate-shifted Instances
von: Singh, Sagalpreet, et al.
Veröffentlicht: (2024)
von: Singh, Sagalpreet, et al.
Veröffentlicht: (2024)
FRACTAL: Fine-Grained Scoring from Aggregate Text Labels
von: Makhija, Yukti, et al.
Veröffentlicht: (2024)
von: Makhija, Yukti, et al.
Veröffentlicht: (2024)
Learning from Label Proportions: Bootstrapping Supervised Learners via Belief Propagation
von: Havaldar, Shreyas, et al.
Veröffentlicht: (2023)
von: Havaldar, Shreyas, et al.
Veröffentlicht: (2023)
Preserving Expert-Level Privacy in Offline Reinforcement Learning
von: Sharma, Navodita, et al.
Veröffentlicht: (2024)
von: Sharma, Navodita, et al.
Veröffentlicht: (2024)
Weak to Strong Learning from Aggregate Labels
von: Makhija, Yukti, et al.
Veröffentlicht: (2024)
von: Makhija, Yukti, et al.
Veröffentlicht: (2024)
Hardness of Learning Boolean Functions from Label Proportions
von: Guruswami, Venkatesan, et al.
Veröffentlicht: (2024)
von: Guruswami, Venkatesan, et al.
Veröffentlicht: (2024)
Fairness under Covariate Shift: Improving Fairness-Accuracy tradeoff with few Unlabeled Test Samples
von: Havaldar, Shreyas, et al.
Veröffentlicht: (2023)
von: Havaldar, Shreyas, et al.
Veröffentlicht: (2023)
Location Aware Embedding for Geotargeting in Sponsored Search Advertising
von: Gligorijevic, Jelena, et al.
Veröffentlicht: (2026)
von: Gligorijevic, Jelena, et al.
Veröffentlicht: (2026)
Dataset Distillation for Offline Reinforcement Learning
von: Light, Jonathan, et al.
Veröffentlicht: (2024)
von: Light, Jonathan, et al.
Veröffentlicht: (2024)
RoiRL: Efficient, Self-Supervised Reasoning with Offline Iterative Reinforcement Learning
von: Arzhantsev, Aleksei, et al.
Veröffentlicht: (2025)
von: Arzhantsev, Aleksei, et al.
Veröffentlicht: (2025)
Offline Reinforcement Learning in Large State Spaces: Algorithms and Guarantees
von: Jiang, Nan, et al.
Veröffentlicht: (2025)
von: Jiang, Nan, et al.
Veröffentlicht: (2025)
Offline Inverse RL: New Solution Concepts and Provably Efficient Algorithms
von: Lazzati, Filippo, et al.
Veröffentlicht: (2024)
von: Lazzati, Filippo, et al.
Veröffentlicht: (2024)
Towards Distillation Guarantees under Algorithmic Alignment for Combinatorial Optimization
von: Le, Thien, et al.
Veröffentlicht: (2026)
von: Le, Thien, et al.
Veröffentlicht: (2026)
Augmenting Online RL with Offline Data is All You Need: A Unified Hybrid RL Algorithm Design and Analysis
von: Huang, Ruiquan, et al.
Veröffentlicht: (2025)
von: Huang, Ruiquan, et al.
Veröffentlicht: (2025)
Accelerating Diffusion Planners in Offline RL via Reward-Aware Consistency Trajectory Distillation
von: Duan, Xintong, et al.
Veröffentlicht: (2025)
von: Duan, Xintong, et al.
Veröffentlicht: (2025)
Self-Supervised Dataset Distillation for Transfer Learning
von: Lee, Dong Bok, et al.
Veröffentlicht: (2023)
von: Lee, Dong Bok, et al.
Veröffentlicht: (2023)
Improving Offline RL by Blending Heuristics
von: Geng, Sinong, et al.
Veröffentlicht: (2023)
von: Geng, Sinong, et al.
Veröffentlicht: (2023)
Statistical Guarantees for Offline Domain Randomization
von: Fickinger, Arnaud, et al.
Veröffentlicht: (2025)
von: Fickinger, Arnaud, et al.
Veröffentlicht: (2025)
Budgeting Counterfactual for Offline RL
von: Liu, Yao, et al.
Veröffentlicht: (2023)
von: Liu, Yao, et al.
Veröffentlicht: (2023)
When Are RL Hyperparameters Benign? A Study in Offline Goal-Conditioned RL
von: Töpperwien, Jan Malte, et al.
Veröffentlicht: (2026)
von: Töpperwien, Jan Malte, et al.
Veröffentlicht: (2026)
Distributional Offline Policy Evaluation with Predictive Error Guarantees
von: Wu, Runzhe, et al.
Veröffentlicht: (2023)
von: Wu, Runzhe, et al.
Veröffentlicht: (2023)
Offline Behavior Distillation
von: Lei, Shiye, et al.
Veröffentlicht: (2024)
von: Lei, Shiye, et al.
Veröffentlicht: (2024)
Making Offline RL Online: Collaborative World Models for Offline Visual Reinforcement Learning
von: Wang, Qi, et al.
Veröffentlicht: (2023)
von: Wang, Qi, et al.
Veröffentlicht: (2023)
Selective Uncertainty Propagation in Offline RL
von: Krishnamurthy, Sanath Kumar, et al.
Veröffentlicht: (2023)
von: Krishnamurthy, Sanath Kumar, et al.
Veröffentlicht: (2023)
Decoupled Prioritized Resampling for Offline RL
von: Yue, Yang, et al.
Veröffentlicht: (2023)
von: Yue, Yang, et al.
Veröffentlicht: (2023)
Augmenting Offline RL with Unlabeled Data
von: Wang, Zhao, et al.
Veröffentlicht: (2024)
von: Wang, Zhao, et al.
Veröffentlicht: (2024)
AD4RL: Autonomous Driving Benchmarks for Offline Reinforcement Learning with Value-based Dataset
von: Lee, Dongsu, et al.
Veröffentlicht: (2024)
von: Lee, Dongsu, et al.
Veröffentlicht: (2024)
Offline RL via Feature-Occupancy Gradient Ascent
von: Neu, Gergely, et al.
Veröffentlicht: (2024)
von: Neu, Gergely, et al.
Veröffentlicht: (2024)
Reinformer: Max-Return Sequence Modeling for Offline RL
von: Zhuang, Zifeng, et al.
Veröffentlicht: (2024)
von: Zhuang, Zifeng, et al.
Veröffentlicht: (2024)
FuzzDistill: Intelligent Fuzzing Target Selection using Compile-Time Analysis and Machine Learning
von: Upadhyay, Saket
Veröffentlicht: (2024)
von: Upadhyay, Saket
Veröffentlicht: (2024)
Tuning Algorithmic and Architectural Hyperparameters in Graph-Based Semi-Supervised Learning with Provable Guarantees
von: Du, Ally Yalei, et al.
Veröffentlicht: (2025)
von: Du, Ally Yalei, et al.
Veröffentlicht: (2025)
An Empirical Study on the Effectiveness of Incorporating Offline RL As Online RL Subroutines
von: Su, Jianhai, et al.
Veröffentlicht: (2025)
von: Su, Jianhai, et al.
Veröffentlicht: (2025)
Policy Agnostic RL: Offline RL and Online RL Fine-Tuning of Any Class and Backbone
von: Mark, Max Sobol, et al.
Veröffentlicht: (2024)
von: Mark, Max Sobol, et al.
Veröffentlicht: (2024)
Dataset Distillation via Knowledge Distillation: Towards Efficient Self-Supervised Pre-Training of Deep Networks
von: Joshi, Siddharth, et al.
Veröffentlicht: (2024)
von: Joshi, Siddharth, et al.
Veröffentlicht: (2024)
Merge Now, Regret Later: The Hidden Cost of Model Merging is Adversarial Transferability
von: Gangwal, Ankit, et al.
Veröffentlicht: (2025)
von: Gangwal, Ankit, et al.
Veröffentlicht: (2025)
Design Considerations in Offline Preference-based RL
von: Agarwal, Alekh, et al.
Veröffentlicht: (2025)
von: Agarwal, Alekh, et al.
Veröffentlicht: (2025)
A Tractable Inference Perspective of Offline RL
von: Liu, Xuejie, et al.
Veröffentlicht: (2023)
von: Liu, Xuejie, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Dense and Diverse Goal Coverage in Multi Goal Reinforcement Learning
von: Singh, Sagalpreet, et al.
Veröffentlicht: (2025) -
LLP-Bench: A Large Scale Tabular Benchmark for Learning from Label Proportions
von: Brahmbhatt, Anand, et al.
Veröffentlicht: (2023) -
Aggregating Data for Optimal and Private Learning
von: Agarwal, Sushant, et al.
Veröffentlicht: (2024) -
Learning from Label Proportions and Covariate-shifted Instances
von: Singh, Sagalpreet, et al.
Veröffentlicht: (2024) -
FRACTAL: Fine-Grained Scoring from Aggregate Text Labels
von: Makhija, Yukti, et al.
Veröffentlicht: (2024)