Agentick: A Unified Benchmark for General Sequential Decision-Making Agents
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Castanyer, Roger Creus, Castro, Pablo Samuel, Berseth, Glen |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Improving Intrinsic Exploration by Creating Stationary Objectives
von: Castanyer, Roger Creus, et al.
Veröffentlicht: (2023)
von: Castanyer, Roger Creus, et al.
Veröffentlicht: (2023)
ARM-FM: Automated Reward Machines via Foundation Models for Compositional Reinforcement Learning
von: Castanyer, Roger Creus, et al.
Veröffentlicht: (2025)
von: Castanyer, Roger Creus, et al.
Veröffentlicht: (2025)
Surprise-Adaptive Intrinsic Motivation for Unsupervised Reinforcement Learning
von: Hugessen, Adriana, et al.
Veröffentlicht: (2024)
von: Hugessen, Adriana, et al.
Veröffentlicht: (2024)
Align and Filter: Improving Performance in Asynchronous On-Policy RL
von: Honari, Homayoun, et al.
Veröffentlicht: (2026)
von: Honari, Homayoun, et al.
Veröffentlicht: (2026)
Is Exploration or Optimization the Problem for Deep Reinforcement Learning?
von: Berseth, Glen
Veröffentlicht: (2025)
von: Berseth, Glen
Veröffentlicht: (2025)
BuilderBench: The Building Blocks of Intelligent Agents
von: Ghugare, Raj, et al.
Veröffentlicht: (2025)
von: Ghugare, Raj, et al.
Veröffentlicht: (2025)
xgenius: LLM-Oriented Autonomous Research Platform for SLURM Clusters
von: Creus Castanyer, Roger
Veröffentlicht: (2026)
von: Creus Castanyer, Roger
Veröffentlicht: (2026)
Improving Deep Reinforcement Learning by Reducing the Chain Effect of Value and Policy Churn
von: Tang, Hongyao, et al.
Veröffentlicht: (2024)
von: Tang, Hongyao, et al.
Veröffentlicht: (2024)
SegDAC: Visual Generalization in Reinforcement Learning via Dynamic Object Tokens
von: Brown, Alexandre, et al.
Veröffentlicht: (2025)
von: Brown, Alexandre, et al.
Veröffentlicht: (2025)
Towards a Unified Framework for Sequential Decision Making
von: Núñez-Molina, Carlos, et al.
Veröffentlicht: (2023)
von: Núñez-Molina, Carlos, et al.
Veröffentlicht: (2023)
Mitigating Plasticity Loss in Continual Reinforcement Learning by Reducing Churn
von: Tang, Hongyao, et al.
Veröffentlicht: (2025)
von: Tang, Hongyao, et al.
Veröffentlicht: (2025)
DLM: Unified Decision Language Models for Offline Multi-Agent Sequential Decision Making
von: Zhang, Zhuohui, et al.
Veröffentlicht: (2026)
von: Zhang, Zhuohui, et al.
Veröffentlicht: (2026)
RLeXplore: Accelerating Research in Intrinsically-Motivated Reinforcement Learning
von: Yuan, Mingqi, et al.
Veröffentlicht: (2024)
von: Yuan, Mingqi, et al.
Veröffentlicht: (2024)
Stable Gradients for Stable Learning at Scale in Deep Reinforcement Learning
von: Castanyer, Roger Creus, et al.
Veröffentlicht: (2025)
von: Castanyer, Roger Creus, et al.
Veröffentlicht: (2025)
PopuLoRA: Co-Evolving LLM Populations for Reasoning Self-Play
von: Castanyer, Roger Creus, et al.
Veröffentlicht: (2026)
von: Castanyer, Roger Creus, et al.
Veröffentlicht: (2026)
KellyBench: A Benchmark for Long-Horizon Sequential Decision Making
von: Grady, Thomas, et al.
Veröffentlicht: (2026)
von: Grady, Thomas, et al.
Veröffentlicht: (2026)
Unifying and Optimizing Data Values for Selection via Sequential Decision-Making
von: Chi, Hongliang, et al.
Veröffentlicht: (2025)
von: Chi, Hongliang, et al.
Veröffentlicht: (2025)
MoralityGym: A Benchmark for Evaluating Hierarchical Moral Alignment in Sequential Decision-Making Agents
von: Rosen, Simon, et al.
Veröffentlicht: (2026)
von: Rosen, Simon, et al.
Veröffentlicht: (2026)
unix-ctf: Procedural Environments for Unix-Competence Reinforcement Learning
von: Bradway, Geoffrey, et al.
Veröffentlicht: (2026)
von: Bradway, Geoffrey, et al.
Veröffentlicht: (2026)
MazeEval: A Benchmark for Testing Sequential Decision-Making in Language Models
von: Einarsson, Hafsteinn
Veröffentlicht: (2025)
von: Einarsson, Hafsteinn
Veröffentlicht: (2025)
MTBBench: A Multimodal Sequential Clinical Decision-Making Benchmark in Oncology
von: Vasilev, Kiril, et al.
Veröffentlicht: (2025)
von: Vasilev, Kiril, et al.
Veröffentlicht: (2025)
Self-Predictive Representations for Combinatorial Generalization in Behavioral Cloning
von: Lawson, Daniel, et al.
Veröffentlicht: (2025)
von: Lawson, Daniel, et al.
Veröffentlicht: (2025)
Counterfactual Effect Decomposition in Multi-Agent Sequential Decision Making
von: Triantafyllou, Stelios, et al.
Veröffentlicht: (2024)
von: Triantafyllou, Stelios, et al.
Veröffentlicht: (2024)
Temporal Representations for Exploration: Learning Complex Exploratory Behavior without Extrinsic Rewards
von: Mohamed, Faisal, et al.
Veröffentlicht: (2026)
von: Mohamed, Faisal, et al.
Veröffentlicht: (2026)
Intelligent Switching for Reset-Free RL
von: Patil, Darshan, et al.
Veröffentlicht: (2024)
von: Patil, Darshan, et al.
Veröffentlicht: (2024)
Enabling Realtime Reinforcement Learning at Scale with Staggered Asynchronous Inference
von: Riemer, Matthew, et al.
Veröffentlicht: (2024)
von: Riemer, Matthew, et al.
Veröffentlicht: (2024)
Enhancing Agent Learning through World Dynamics Modeling
von: Sun, Zhiyuan, et al.
Veröffentlicht: (2024)
von: Sun, Zhiyuan, et al.
Veröffentlicht: (2024)
A Sequential Decision-Making Model for Perimeter Identification
von: Taitler, Ayal
Veröffentlicht: (2024)
von: Taitler, Ayal
Veröffentlicht: (2024)
Causal Explanations for Sequential Decision-Making in Multi-Agent Systems
von: Gyevnar, Balint, et al.
Veröffentlicht: (2023)
von: Gyevnar, Balint, et al.
Veröffentlicht: (2023)
Understanding the Training and Generalization of Pretrained Transformer for Sequential Decision Making
von: Wang, Hanzhao, et al.
Veröffentlicht: (2024)
von: Wang, Hanzhao, et al.
Veröffentlicht: (2024)
M^4olGen: Multi-Agent, Multi-Stage Molecular Generation under Precise Multi-Property Constraints
von: Li, Yizhan, et al.
Veröffentlicht: (2026)
von: Li, Yizhan, et al.
Veröffentlicht: (2026)
PRISM: Perception Reasoning Interleaved for Sequential Decision Making
von: Aissi, Mohamed Salim, et al.
Veröffentlicht: (2026)
von: Aissi, Mohamed Salim, et al.
Veröffentlicht: (2026)
Responsibility Gap and Diffusion in Sequential Decision-Making Mechanisms
von: Jiang, Junli, et al.
Veröffentlicht: (2025)
von: Jiang, Junli, et al.
Veröffentlicht: (2025)
Online Sequential Decision-Making with Unknown Delays
von: Wu, Ping, et al.
Veröffentlicht: (2024)
von: Wu, Ping, et al.
Veröffentlicht: (2024)
Reward Design for Justifiable Sequential Decision-Making
von: Sukovic, Aleksa, et al.
Veröffentlicht: (2024)
von: Sukovic, Aleksa, et al.
Veröffentlicht: (2024)
On the Modeling Capabilities of Large Language Models for Sequential Decision Making
von: Klissarov, Martin, et al.
Veröffentlicht: (2024)
von: Klissarov, Martin, et al.
Veröffentlicht: (2024)
Improving Pre-Trained Vision-Language-Action Policies with Model-Based Search
von: Neary, Cyrus, et al.
Veröffentlicht: (2025)
von: Neary, Cyrus, et al.
Veröffentlicht: (2025)
Bayesian Orchestration of Multi-LLM Agents for Cost-Aware Sequential Decision-Making
von: Amin, Danial
Veröffentlicht: (2026)
von: Amin, Danial
Veröffentlicht: (2026)
Translating the Rashomon Effect to Sequential Decision-Making Tasks
von: Gross, Dennis, et al.
Veröffentlicht: (2025)
von: Gross, Dennis, et al.
Veröffentlicht: (2025)
SkillGen: Learning Domain Skills for In-Context Sequential Decision Making
von: Ding, Ruomeng, et al.
Veröffentlicht: (2025)
von: Ding, Ruomeng, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Improving Intrinsic Exploration by Creating Stationary Objectives
von: Castanyer, Roger Creus, et al.
Veröffentlicht: (2023) -
ARM-FM: Automated Reward Machines via Foundation Models for Compositional Reinforcement Learning
von: Castanyer, Roger Creus, et al.
Veröffentlicht: (2025) -
Surprise-Adaptive Intrinsic Motivation for Unsupervised Reinforcement Learning
von: Hugessen, Adriana, et al.
Veröffentlicht: (2024) -
Align and Filter: Improving Performance in Asynchronous On-Policy RL
von: Honari, Homayoun, et al.
Veröffentlicht: (2026) -
Is Exploration or Optimization the Problem for Deep Reinforcement Learning?
von: Berseth, Glen
Veröffentlicht: (2025)