Action Shapley: A Training Data Selection Metric for World Model in Reinforcement Learning
Fuente:
arXiv
Salvato in:
| Autori principali: | Ghosh, Rajat, Dutta, Debojyoti |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Efficient Alignment of Large Language Models via Data Sampling
di: Khera, Amrit, et al.
Pubblicazione: (2024)
di: Khera, Amrit, et al.
Pubblicazione: (2024)
Predictive Scaling Laws for Efficient GRPO Training of Large Reasoning Models
di: Nimmaturi, Datta, et al.
Pubblicazione: (2025)
di: Nimmaturi, Datta, et al.
Pubblicazione: (2025)
CPP-UT-Bench: Can LLMs Write Complex Unit Tests in C++?
di: Bhargava, Vaishnavi, et al.
Pubblicazione: (2024)
di: Bhargava, Vaishnavi, et al.
Pubblicazione: (2024)
SWE-Tester: Training Open-Source LLMs for Issue Reproduction in Real-World Repositories
di: Soni, Aditya Bharat, et al.
Pubblicazione: (2026)
di: Soni, Aditya Bharat, et al.
Pubblicazione: (2026)
Shapley Curves: A Smoothing Perspective
di: Miftachov, Ratmir, et al.
Pubblicazione: (2022)
di: Miftachov, Ratmir, et al.
Pubblicazione: (2022)
Reinforcement Learning with Continuous Actions Under Unmeasured Confounding
di: Li, Yuhan, et al.
Pubblicazione: (2025)
di: Li, Yuhan, et al.
Pubblicazione: (2025)
Shapley-PC: Constraint-based Causal Structure Learning with a Shapley Inspired Framework
di: Russo, Fabrizio, et al.
Pubblicazione: (2023)
di: Russo, Fabrizio, et al.
Pubblicazione: (2023)
Group Shapley Value and Counterfactual Simulations in a Structural Model
di: Kwon, Yongchan, et al.
Pubblicazione: (2024)
di: Kwon, Yongchan, et al.
Pubblicazione: (2024)
DR-VIDAL -- Doubly Robust Variational Information-theoretic Deep Adversarial Learning for Counterfactual Prediction and Treatment Effect Estimation on Real World Data
di: Ghosh, Shantanu, et al.
Pubblicazione: (2023)
di: Ghosh, Shantanu, et al.
Pubblicazione: (2023)
Multivariate outlier explanations using Shapley values and Mahalanobis distances
di: Mayrhofer, Marcus, et al.
Pubblicazione: (2022)
di: Mayrhofer, Marcus, et al.
Pubblicazione: (2022)
Efficient nonparametric statistical inference on population feature importance using Shapley values
di: Williamson, Brian D., et al.
Pubblicazione: (2020)
di: Williamson, Brian D., et al.
Pubblicazione: (2020)
RANGER -- Repository-Level Agent for Graph-Enhanced Retrieval
di: Shah, Pratik, et al.
Pubblicazione: (2025)
di: Shah, Pratik, et al.
Pubblicazione: (2025)
Overview and practical recommendations on using Shapley Values for identifying predictive biomarkers via CATE modeling
di: Svensson, David, et al.
Pubblicazione: (2025)
di: Svensson, David, et al.
Pubblicazione: (2025)
i-IF-Learn: Iterative Feature Selection and Unsupervised Learning for High-Dimensional Complex Data
di: Ma, Chen, et al.
Pubblicazione: (2026)
di: Ma, Chen, et al.
Pubblicazione: (2026)
BAR Conjecture: the Feasibility of Inference Budget-Constrained LLM Services with Authenticity and Reasoning
di: Zhou, Jinan, et al.
Pubblicazione: (2025)
di: Zhou, Jinan, et al.
Pubblicazione: (2025)
A Multi-Agent Framework for Stateful Inference-Time Search
di: Lalan, Arshika, et al.
Pubblicazione: (2025)
di: Lalan, Arshika, et al.
Pubblicazione: (2025)
Go-UT-Bench: A Fine-Tuning Dataset for LLM-Based Unit Test Generation in Go
di: Pipalani, Yashshi, et al.
Pubblicazione: (2025)
di: Pipalani, Yashshi, et al.
Pubblicazione: (2025)
Bayesian Models for Joint Selection of Features and Auto-Regressive Lags: Theory and Applications in Environmental and Financial Forecasting
di: Manna, Alokesh, et al.
Pubblicazione: (2025)
di: Manna, Alokesh, et al.
Pubblicazione: (2025)
Joint Distribution-Informed Shapley Values for Sparse Counterfactual Explanations
di: You, Lei, et al.
Pubblicazione: (2024)
di: You, Lei, et al.
Pubblicazione: (2024)
Characterization and Learning of Causal Graphs with Latent Confounders and Post-treatment Selection from Interventional Data
di: Luo, Gongxu, et al.
Pubblicazione: (2025)
di: Luo, Gongxu, et al.
Pubblicazione: (2025)
Fast approximative estimation of conditional Shapley values when using a linear regression model or a polynomial regression model
di: Aanes, Fredrik Lohne
Pubblicazione: (2025)
di: Aanes, Fredrik Lohne
Pubblicazione: (2025)
DFNN: A Deep Fréchet Neural Network Framework for Learning Metric-Space-Valued Responses
di: Kim, Kyum, et al.
Pubblicazione: (2025)
di: Kim, Kyum, et al.
Pubblicazione: (2025)
Federated Offline Reinforcement Learning
di: Zhou, Doudou, et al.
Pubblicazione: (2022)
di: Zhou, Doudou, et al.
Pubblicazione: (2022)
No $D_{\text{train}}$: Model-Agnostic Counterfactual Explanations Using Reinforcement Learning
di: Sun, Xiangyu, et al.
Pubblicazione: (2024)
di: Sun, Xiangyu, et al.
Pubblicazione: (2024)
A feature selection method based on Shapley values robust to concept shift in regression
di: Sebastián, Carlos, et al.
Pubblicazione: (2023)
di: Sebastián, Carlos, et al.
Pubblicazione: (2023)
Inference on Variable Importance for Treatment Effect Heterogeneity: Shapley Values and Beyond
di: Morzywolek, Pawel, et al.
Pubblicazione: (2025)
di: Morzywolek, Pawel, et al.
Pubblicazione: (2025)
Model Class Selection
di: Cecil, Ryan, et al.
Pubblicazione: (2025)
di: Cecil, Ryan, et al.
Pubblicazione: (2025)
High-dimensional Functional Graphical Model Structure Learning via Neighborhood Selection Approach
di: Zhao, Boxin, et al.
Pubblicazione: (2021)
di: Zhao, Boxin, et al.
Pubblicazione: (2021)
Variable Selection for Comparing High-dimensional Time-Series Data
di: Mitsuzawa, Kensuke, et al.
Pubblicazione: (2024)
di: Mitsuzawa, Kensuke, et al.
Pubblicazione: (2024)
Assessing Surrogate Heterogeneity in Real World Data Using Meta-Learners
di: Knowlton, Rebecca, et al.
Pubblicazione: (2025)
di: Knowlton, Rebecca, et al.
Pubblicazione: (2025)
Position: Benchmarking is Limited in Reinforcement Learning Research
di: Jordan, Scott M., et al.
Pubblicazione: (2024)
di: Jordan, Scott M., et al.
Pubblicazione: (2024)
A General Control-Theoretic Approach for Reinforcement Learning: Theory and Algorithms
di: Chen, Weiqin, et al.
Pubblicazione: (2024)
di: Chen, Weiqin, et al.
Pubblicazione: (2024)
Riemannian Laplace Approximation with the Fisher Metric
di: Yu, Hanlin, et al.
Pubblicazione: (2023)
di: Yu, Hanlin, et al.
Pubblicazione: (2023)
Tackling Interference Induced by Data Training Loops in A/B Tests: A Weighted Training Approach
di: Si, Nian
Pubblicazione: (2023)
di: Si, Nian
Pubblicazione: (2023)
Causally-Aware Unsupervised Feature Selection Learning
di: Shen, Zongxin, et al.
Pubblicazione: (2024)
di: Shen, Zongxin, et al.
Pubblicazione: (2024)
A Graphical Approach to State Variable Selection in Off-policy Learning
di: Andersen, Joakim Blach, et al.
Pubblicazione: (2025)
di: Andersen, Joakim Blach, et al.
Pubblicazione: (2025)
Counterfactually Fair Reinforcement Learning via Sequential Data Preprocessing
di: Wang, Jitao, et al.
Pubblicazione: (2025)
di: Wang, Jitao, et al.
Pubblicazione: (2025)
The Landscape of Causal Discovery Data: Grounding Causal Discovery in Real-World Applications
di: Brouillard, Philippe, et al.
Pubblicazione: (2024)
di: Brouillard, Philippe, et al.
Pubblicazione: (2024)
Designing Time Series Experiments in A/B Testing with Transformer Reinforcement Learning
di: Wu, Xiangkun, et al.
Pubblicazione: (2026)
di: Wu, Xiangkun, et al.
Pubblicazione: (2026)
Robust Classification of High-Dimensional Data using Data-Adaptive Energy Distance
di: Choudhury, Jyotishka Ray, et al.
Pubblicazione: (2023)
di: Choudhury, Jyotishka Ray, et al.
Pubblicazione: (2023)
Documenti analoghi
-
Efficient Alignment of Large Language Models via Data Sampling
di: Khera, Amrit, et al.
Pubblicazione: (2024) -
Predictive Scaling Laws for Efficient GRPO Training of Large Reasoning Models
di: Nimmaturi, Datta, et al.
Pubblicazione: (2025) -
CPP-UT-Bench: Can LLMs Write Complex Unit Tests in C++?
di: Bhargava, Vaishnavi, et al.
Pubblicazione: (2024) -
SWE-Tester: Training Open-Source LLMs for Issue Reproduction in Real-World Repositories
di: Soni, Aditya Bharat, et al.
Pubblicazione: (2026) -
Shapley Curves: A Smoothing Perspective
di: Miftachov, Ratmir, et al.
Pubblicazione: (2022)