Salvato in:
| Autore principale: | Shelopugin, Andrei |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2406.00814 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Player-Team Heterogeneous Interaction Graph Transformer for Soccer Outcome Prediction
di: Wang, Lintao, et al.
Pubblicazione: (2025)
di: Wang, Lintao, et al.
Pubblicazione: (2025)
Unveiling Hidden Pivotal Players with GoalNet: A GNN-Based Soccer Player Evaluation System
di: Jiang, Jacky Hao, et al.
Pubblicazione: (2025)
di: Jiang, Jacky Hao, et al.
Pubblicazione: (2025)
Online Clustering of Dueling Bandits
di: Wang, Zhiyong, et al.
Pubblicazione: (2025)
di: Wang, Zhiyong, et al.
Pubblicazione: (2025)
Fusing Reward and Dueling Feedback in Stochastic Bandits
di: Wang, Xuchuang, et al.
Pubblicazione: (2025)
di: Wang, Xuchuang, et al.
Pubblicazione: (2025)
Linear and Neural Dueling Bandits with Delayed Feedback
di: Wang, Xiangyi, et al.
Pubblicazione: (2026)
di: Wang, Xiangyi, et al.
Pubblicazione: (2026)
Learning Agile Soccer Skills for a Bipedal Robot with Deep Reinforcement Learning
di: Haarnoja, Tuomas, et al.
Pubblicazione: (2023)
di: Haarnoja, Tuomas, et al.
Pubblicazione: (2023)
A Controlled Study of Double DQN and Dueling DQN Under Cross-Environment Transfer
di: Nasir, Azkaa, et al.
Pubblicazione: (2026)
di: Nasir, Azkaa, et al.
Pubblicazione: (2026)
Ranking Policy Learning via Marketplace Expected Value Estimation From Observational Data
di: Ebrahimzadeh, Ehsan, et al.
Pubblicazione: (2024)
di: Ebrahimzadeh, Ehsan, et al.
Pubblicazione: (2024)
Neural Dueling Bandits: Preference-Based Optimization with Human Feedback
di: Verma, Arun, et al.
Pubblicazione: (2024)
di: Verma, Arun, et al.
Pubblicazione: (2024)
Learning to Play 7 Wonders Duel Without Human Supervision
di: Paolini, Giovanni, et al.
Pubblicazione: (2024)
di: Paolini, Giovanni, et al.
Pubblicazione: (2024)
Adaptive Action Chunking via Multi-Chunk Q Value Estimation
di: Shin, Yongjae, et al.
Pubblicazione: (2026)
di: Shin, Yongjae, et al.
Pubblicazione: (2026)
Do We Need Large VLMs for Spotting Soccer Actions?
di: Chakraborty, Ritabrata, et al.
Pubblicazione: (2025)
di: Chakraborty, Ritabrata, et al.
Pubblicazione: (2025)
Learning To Play Atari Games Using Dueling Q-Learning and Hebbian Plasticity
di: Salehin, Md Ashfaq
Pubblicazione: (2024)
di: Salehin, Md Ashfaq
Pubblicazione: (2024)
Improving Dribbling, Passing, and Marking Actions in Soccer Simulation 2D Games Using Machine Learning
di: Zare, Nader, et al.
Pubblicazione: (2024)
di: Zare, Nader, et al.
Pubblicazione: (2024)
SoccerDiffusion: Toward Learning End-to-End Humanoid Robot Soccer from Gameplay Recordings
di: Vahl, Florian, et al.
Pubblicazione: (2025)
di: Vahl, Florian, et al.
Pubblicazione: (2025)
Learning Tennis Strategy Through Curriculum-Based Dueling Double Deep Q-Networks
di: Mohan, Vishnu
Pubblicazione: (2025)
di: Mohan, Vishnu
Pubblicazione: (2025)
Beyond Numeric Rewards: In-Context Dueling Bandits with LLM Agents
di: Xia, Fanzeng, et al.
Pubblicazione: (2024)
di: Xia, Fanzeng, et al.
Pubblicazione: (2024)
PathCRF: Ball-Free Soccer Event Detection via Possession Path Inference from Player Trajectories
di: Kim, Hyunsung, et al.
Pubblicazione: (2026)
di: Kim, Hyunsung, et al.
Pubblicazione: (2026)
An Analysis of Action-Value Temporal-Difference Methods That Learn State Values
di: Daley, Brett, et al.
Pubblicazione: (2025)
di: Daley, Brett, et al.
Pubblicazione: (2025)
Multi-Player Approaches for Dueling Bandits
di: Raveh, Or, et al.
Pubblicazione: (2024)
di: Raveh, Or, et al.
Pubblicazione: (2024)
Survey of Action Recognition, Spotting and Spatio-Temporal Localization in Soccer -- Current Trends and Research Perspectives
di: Seweryn, Karolina, et al.
Pubblicazione: (2023)
di: Seweryn, Karolina, et al.
Pubblicazione: (2023)
Duel-Evolve: Reward-Free Test-Time Scaling via LLM Self-Preferences
di: Karlekar, Sweta, et al.
Pubblicazione: (2026)
di: Karlekar, Sweta, et al.
Pubblicazione: (2026)
An Odd Estimator for Shapley Values
di: Fumagalli, Fabian, et al.
Pubblicazione: (2026)
di: Fumagalli, Fabian, et al.
Pubblicazione: (2026)
Enhancing Two-Player Performance Through Single-Player Knowledge Transfer: An Empirical Study on Atari 2600 Games
di: Saadat, Kimiya, et al.
Pubblicazione: (2024)
di: Saadat, Kimiya, et al.
Pubblicazione: (2024)
Regression-adjusted Monte Carlo Estimators for Shapley Values and Probabilistic Values
di: Witter, R. Teal, et al.
Pubblicazione: (2025)
di: Witter, R. Teal, et al.
Pubblicazione: (2025)
Expected Coordinate Improvement for High-Dimensional Bayesian Optimization
di: Zhan, Dawei
Pubblicazione: (2024)
di: Zhan, Dawei
Pubblicazione: (2024)
The Strain of Success: A Predictive Model for Injury Risk Mitigation and Team Success in Soccer
di: Everett, Gregory, et al.
Pubblicazione: (2024)
di: Everett, Gregory, et al.
Pubblicazione: (2024)
SkillTree: Explainable Skill-Based Deep Reinforcement Learning for Long-Horizon Control Tasks
di: Wen, Yongyan, et al.
Pubblicazione: (2024)
di: Wen, Yongyan, et al.
Pubblicazione: (2024)
Uncertainty-aware Evaluation of Auxiliary Anomalies with the Expected Anomaly Posterior
di: Perini, Lorenzo, et al.
Pubblicazione: (2024)
di: Perini, Lorenzo, et al.
Pubblicazione: (2024)
Engineering Features to Improve Pass Prediction in Soccer Simulation 2D Games
di: Zare, Nader, et al.
Pubblicazione: (2024)
di: Zare, Nader, et al.
Pubblicazione: (2024)
Mixture of Masters: Sparse Chess Language Models with Player Routing
di: Frisoni, Giacomo, et al.
Pubblicazione: (2026)
di: Frisoni, Giacomo, et al.
Pubblicazione: (2026)
Reinforcement Learning Within the Classical Robotics Stack: A Case Study in Robot Soccer
di: Labiosa, Adam, et al.
Pubblicazione: (2024)
di: Labiosa, Adam, et al.
Pubblicazione: (2024)
GroupSegment-SHAP: Shapley Value Explanations with Group-Segment Players for Multivariate Time Series
di: Kim, Jinwoong, et al.
Pubblicazione: (2026)
di: Kim, Jinwoong, et al.
Pubblicazione: (2026)
Guided Flow Policy: Learning from High-Value Actions in Offline Reinforcement Learning
di: Tiofack, Franki Nguimatsia, et al.
Pubblicazione: (2025)
di: Tiofack, Franki Nguimatsia, et al.
Pubblicazione: (2025)
Learning to Move Like Professional Counter-Strike Players
di: Durst, David, et al.
Pubblicazione: (2024)
di: Durst, David, et al.
Pubblicazione: (2024)
Kernel Banzhaf: A Fast and Robust Estimator for Banzhaf Values
di: Liu, Yurong, et al.
Pubblicazione: (2024)
di: Liu, Yurong, et al.
Pubblicazione: (2024)
Action Controlled Paraphrasing
di: Shi, Ning, et al.
Pubblicazione: (2024)
di: Shi, Ning, et al.
Pubblicazione: (2024)
Offline Imitation of Badminton Player Behavior via Experiential Contexts and Brownian Motion
di: Wang, Kuang-Da, et al.
Pubblicazione: (2024)
di: Wang, Kuang-Da, et al.
Pubblicazione: (2024)
State-Action Inpainting Diffuser for Continuous Control with Delay
di: Han, Dongqi, et al.
Pubblicazione: (2026)
di: Han, Dongqi, et al.
Pubblicazione: (2026)
Revisiting the Test-Time Scaling of o1-like Models: Do they Truly Possess Test-Time Scaling Capabilities?
di: Zeng, Zhiyuan, et al.
Pubblicazione: (2025)
di: Zeng, Zhiyuan, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Player-Team Heterogeneous Interaction Graph Transformer for Soccer Outcome Prediction
di: Wang, Lintao, et al.
Pubblicazione: (2025) -
Unveiling Hidden Pivotal Players with GoalNet: A GNN-Based Soccer Player Evaluation System
di: Jiang, Jacky Hao, et al.
Pubblicazione: (2025) -
Online Clustering of Dueling Bandits
di: Wang, Zhiyong, et al.
Pubblicazione: (2025) -
Fusing Reward and Dueling Feedback in Stochastic Bandits
di: Wang, Xuchuang, et al.
Pubblicazione: (2025) -
Linear and Neural Dueling Bandits with Delayed Feedback
di: Wang, Xiangyi, et al.
Pubblicazione: (2026)