Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments
Fuente:
arXiv
Saved in:
| Main Authors: | Qu, Yun, Wang, Qi Cheems, Mao, Yixiu, Lv, Yiqin, Ji, Xiangyang |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Robust Fast Adaptation from Adversarially Explicit Task Distribution Generation
by: Wang, Cheems, et al.
Published: (2024)
by: Wang, Cheems, et al.
Published: (2024)
Adaptive Neighborhood-Constrained Q Learning for Offline Reinforcement Learning
by: Mao, Yixiu, et al.
Published: (2025)
by: Mao, Yixiu, et al.
Published: (2025)
Dynamics-Predictive Sampling for Active RL Finetuning of Large Reasoning Models
by: Mao, Yixiu, et al.
Published: (2026)
by: Mao, Yixiu, et al.
Published: (2026)
Latent Reward: LLM-Empowered Credit Assignment in Episodic Reinforcement Learning
by: Qu, Yun, et al.
Published: (2024)
by: Qu, Yun, et al.
Published: (2024)
RLVR without Ineffective Samples: Group Prioritized Off-Policy Optimization for LLM Reasoning
by: Mao, Yixiu, et al.
Published: (2026)
by: Mao, Yixiu, et al.
Published: (2026)
Model Predictive Task Sampling for Efficient and Robust Adaptation
by: Wang, Qi, et al.
Published: (2025)
by: Wang, Qi, et al.
Published: (2025)
Utility-Diversity Aware Online Batch Selection for LLM Supervised Fine-tuning
by: Zou, Heming, et al.
Published: (2025)
by: Zou, Heming, et al.
Published: (2025)
Offline Reinforcement Learning with OOD State Correction and OOD Action Suppression
by: Mao, Yixiu, et al.
Published: (2024)
by: Mao, Yixiu, et al.
Published: (2024)
Doubly Mild Generalization for Offline Reinforcement Learning
by: Mao, Yixiu, et al.
Published: (2024)
by: Mao, Yixiu, et al.
Published: (2024)
VAO: Validation-Aligned Optimization for Cross-Task Generative Auto-Bidding
by: Lv, Yiqin, et al.
Published: (2025)
by: Lv, Yiqin, et al.
Published: (2025)
Can Prompt Difficulty be Online Predicted for Accelerating RL Finetuning of Reasoning Models?
by: Qu, Yun, et al.
Published: (2025)
by: Qu, Yun, et al.
Published: (2025)
Stop Wandering, Find the Keys: LLMs Discriminate Key States for Efficient Multi-Agent Exploration
by: Qu, Yun, et al.
Published: (2024)
by: Qu, Yun, et al.
Published: (2024)
Small Generalizable Prompt Predictive Models Can Steer Efficient RL Post-Training of Large Reasoning Models
by: Qu, Yun, et al.
Published: (2026)
by: Qu, Yun, et al.
Published: (2026)
Fast and Robust Likelihood-Guided Diffusion Posterior Sampling with Amortized Variational Inference
by: Zheng, Léon, et al.
Published: (2026)
by: Zheng, Léon, et al.
Published: (2026)
Listwise Policy Optimization: Group-based RLVR as Target-Projection on the LLM Response Simplex
by: Qu, Yun, et al.
Published: (2026)
by: Qu, Yun, et al.
Published: (2026)
Enhancing Generative Auto-bidding with Offline Reward Evaluation and Policy Search
by: Mou, Zhiyu, et al.
Published: (2025)
by: Mou, Zhiyu, et al.
Published: (2025)
On Sample-Efficient Offline Reinforcement Learning: Data Diversity, Posterior Sampling, and Beyond
by: Nguyen-Tang, Thanh, et al.
Published: (2024)
by: Nguyen-Tang, Thanh, et al.
Published: (2024)
Tiny, On-Device Decision Makers with the MiniConv Library
by: Purves, Carlos
Published: (2025)
by: Purves, Carlos
Published: (2025)
2-Step Agent: A Framework for the Interaction of a Decision Maker with AI Decision Support
by: Nyberg, Otto, et al.
Published: (2026)
by: Nyberg, Otto, et al.
Published: (2026)
Statistical Tests for Replacing Human Decision Makers with Algorithms
by: Feng, Kai, et al.
Published: (2023)
by: Feng, Kai, et al.
Published: (2023)
FM-EAC: Feature Model-based Enhanced Actor-Critic for Multi-Task Control in Dynamic Environments
by: Zhou, Quanxi, et al.
Published: (2025)
by: Zhou, Quanxi, et al.
Published: (2025)
(Un)certainty of (Un)fairness: Preference-Based Selection of Certainly Fair Decision-Makers
by: Duong, Manh Khoi, et al.
Published: (2024)
by: Duong, Manh Khoi, et al.
Published: (2024)
Synergy and Diversity in CLIP: Enhancing Performance Through Adaptive Backbone Ensembling
by: Rodriguez-Opazo, Cristian, et al.
Published: (2024)
by: Rodriguez-Opazo, Cristian, et al.
Published: (2024)
Gains: Fine-grained Federated Domain Adaptation in Open Set
by: Zhong, Zhengyi, et al.
Published: (2025)
by: Zhong, Zhengyi, et al.
Published: (2025)
Solving Diffusion Inverse Problems with Restart Posterior Sampling
by: Ahmed, Bilal, et al.
Published: (2025)
by: Ahmed, Bilal, et al.
Published: (2025)
OrchMoE: Efficient Multi-Adapter Learning with Task-Skill Synergy
by: Wang, Haowen, et al.
Published: (2024)
by: Wang, Haowen, et al.
Published: (2024)
Reducing Fine-Tuning Memory Overhead by Approximate and Memory-Sharing Backpropagation
by: Yang, Yuchen, et al.
Published: (2024)
by: Yang, Yuchen, et al.
Published: (2024)
WLFM: A Well-Logs Foundation Model for Multi-Task and Cross-Well Geological Interpretation
by: Qi, Zhenyu, et al.
Published: (2025)
by: Qi, Zhenyu, et al.
Published: (2025)
Symmetric Reinforcement Learning Loss for Robust Learning on Diverse Tasks and Model Scales
by: Byun, Ju-Seung, et al.
Published: (2024)
by: Byun, Ju-Seung, et al.
Published: (2024)
Diffusion Posterior Sampling is Computationally Intractable
by: Gupta, Shivam, et al.
Published: (2024)
by: Gupta, Shivam, et al.
Published: (2024)
TASER: Temporal Adaptive Sampling for Fast and Accurate Dynamic Graph Representation Learning
by: Deng, Gangda, et al.
Published: (2024)
by: Deng, Gangda, et al.
Published: (2024)
Efficient Approximate Posterior Sampling with Annealed Langevin Monte Carlo
by: Parulekar, Advait, et al.
Published: (2025)
by: Parulekar, Advait, et al.
Published: (2025)
HybridLinker: Topology-Guided Posterior Sampling for Enhanced Diversity and Validity in 3D Molecular Linker Generation
by: Hwang, Minyeong, et al.
Published: (2025)
by: Hwang, Minyeong, et al.
Published: (2025)
Task-agnostic Decision Transformer for Multi-type Agent Control with Federated Split Training
by: Wang, Zhiyuan, et al.
Published: (2024)
by: Wang, Zhiyuan, et al.
Published: (2024)
Leveraging Ensemble Diversity for Robust Self-Training in the Presence of Sample Selection Bias
by: Odonnat, Ambroise, et al.
Published: (2023)
by: Odonnat, Ambroise, et al.
Published: (2023)
AMS-QUANT: Adaptive Mantissa Sharing for Floating-point Quantization
by: Lv, Mengtao, et al.
Published: (2025)
by: Lv, Mengtao, et al.
Published: (2025)
Coupled Data and Measurement Space Dynamics for Enhanced Diffusion Posterior Sampling
by: Hamidi, Shayan Mohajer, et al.
Published: (2025)
by: Hamidi, Shayan Mohajer, et al.
Published: (2025)
Active Attacks: Red-teaming LLMs via Adaptive Environments
by: Yun, Taeyoung, et al.
Published: (2025)
by: Yun, Taeyoung, et al.
Published: (2025)
Variational Randomized Smoothing for Sample-Wise Adversarial Robustness
by: Hase, Ryo, et al.
Published: (2024)
by: Hase, Ryo, et al.
Published: (2024)
FlexCare: Leveraging Cross-Task Synergy for Flexible Multimodal Healthcare Prediction
by: Xu, Muhao, et al.
Published: (2024)
by: Xu, Muhao, et al.
Published: (2024)
Similar Items
-
Robust Fast Adaptation from Adversarially Explicit Task Distribution Generation
by: Wang, Cheems, et al.
Published: (2024) -
Adaptive Neighborhood-Constrained Q Learning for Offline Reinforcement Learning
by: Mao, Yixiu, et al.
Published: (2025) -
Dynamics-Predictive Sampling for Active RL Finetuning of Large Reasoning Models
by: Mao, Yixiu, et al.
Published: (2026) -
Latent Reward: LLM-Empowered Credit Assignment in Episodic Reinforcement Learning
by: Qu, Yun, et al.
Published: (2024) -
RLVR without Ineffective Samples: Group Prioritized Off-Policy Optimization for LLM Reasoning
by: Mao, Yixiu, et al.
Published: (2026)