Rollout Cards: A Reproducibility Standard for Agent Research
Fuente:
arXiv
Saved in:
| Main Authors: | Masters, Charlie, Liu, Ziyuan, Albrecht, Stefano V. |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ARCANE: A Multi-Agent Framework for Interpretable and Configurable Alignment
by: Masters, Charlie, et al.
Published: (2025)
by: Masters, Charlie, et al.
Published: (2025)
Orchestrating Human-AI Teams: The Manager Agent as a Unifying Research Challenge
by: Masters, Charlie, et al.
Published: (2025)
by: Masters, Charlie, et al.
Published: (2025)
CAPF: Guiding Search-Agent Rollouts with Credit-Attenuated Privileged Feedback
by: Chen, Bin, et al.
Published: (2026)
by: Chen, Bin, et al.
Published: (2026)
Multi-Agent Reinforcement Learning for Energy Networks: Computational Challenges, Progress and Open Problems
by: Keren, Sarah, et al.
Published: (2024)
by: Keren, Sarah, et al.
Published: (2024)
Valet: A Standardized Testbed of Traditional Imperfect-Information Card Games
by: Goadrich, Mark, et al.
Published: (2026)
by: Goadrich, Mark, et al.
Published: (2026)
ProRL Agent: Rollout-as-a-Service for RL Training of Multi-Turn LLM Agents
by: Zhang, Hao, et al.
Published: (2026)
by: Zhang, Hao, et al.
Published: (2026)
Integrating Counterfactual Simulations with Language Models for Explaining Multi-Agent Behaviour
by: Gyevnár, Bálint, et al.
Published: (2025)
by: Gyevnár, Bálint, et al.
Published: (2025)
Towards Ethical Multi-Agent Systems of Large Language Models: A Mechanistic Interpretability Perspective
by: Lee, Jae Hee, et al.
Published: (2025)
by: Lee, Jae Hee, et al.
Published: (2025)
EvalCards: A Framework for Standardized Evaluation Reporting
by: Dhar, Ruchira, et al.
Published: (2025)
by: Dhar, Ruchira, et al.
Published: (2025)
Maximum Entropy Exploration Without the Rollouts
by: Adamczyk, Jacob, et al.
Published: (2026)
by: Adamczyk, Jacob, et al.
Published: (2026)
SAJA: A State-Action Joint Attack Framework on Multi-Agent Deep Reinforcement Learning
by: Guo, Weiqi, et al.
Published: (2025)
by: Guo, Weiqi, et al.
Published: (2025)
Interaction as Intelligence Part II: Asynchronous Human-Agent Rollout for Long-Horizon Task Training
by: Fu, Dayuan, et al.
Published: (2025)
by: Fu, Dayuan, et al.
Published: (2025)
Multi-Agent Reinforcement Learning: Methods, Applications, Visionary Prospects, and Challenges
by: Zhou, Ziyuan, et al.
Published: (2023)
by: Zhou, Ziyuan, et al.
Published: (2023)
Partially Observable Mean Field Multi-Agent Reinforcement Learning Based on Graph-Attention
by: Yang, Min, et al.
Published: (2023)
by: Yang, Min, et al.
Published: (2023)
DiffusionRollout: Uncertainty-Aware Rollout Planning in Long-Horizon PDE Solving
by: Yoo, Seungwoo, et al.
Published: (2026)
by: Yoo, Seungwoo, et al.
Published: (2026)
HyperMARL: Adaptive Hypernetworks for Multi-Agent RL
by: Tessera, Kale-ab Abebe, et al.
Published: (2024)
by: Tessera, Kale-ab Abebe, et al.
Published: (2024)
Portfolio Reinforcement Learning with Scenario-Context Rollout
by: Bendatu, Vanya Priscillia, et al.
Published: (2026)
by: Bendatu, Vanya Priscillia, et al.
Published: (2026)
Learning to Beat ByteRL: Exploitability of Collectible Card Game Agents
by: Haluska, Radovan, et al.
Published: (2024)
by: Haluska, Radovan, et al.
Published: (2024)
Not All Rollouts are Useful: Down-Sampling Rollouts in LLM Reinforcement Learning
by: Xu, Yixuan Even, et al.
Published: (2025)
by: Xu, Yixuan Even, et al.
Published: (2025)
LMR-BENCH: Evaluating LLM Agent's Ability on Reproducing Language Modeling Research
by: Yan, Shuo, et al.
Published: (2025)
by: Yan, Shuo, et al.
Published: (2025)
TSR: Trajectory-Search Rollouts for Multi-Turn RL of LLM Agents
by: Djuhera, Aladin, et al.
Published: (2026)
by: Djuhera, Aladin, et al.
Published: (2026)
PTCG-Bench: Can LLM Agents Master Pokémon Trading Card Game?
by: Hua, Dongdong, et al.
Published: (2026)
by: Hua, Dongdong, et al.
Published: (2026)
Generalized Nested Rollout Policy Adaptation with Limited Repetitions
by: Cazenave, Tristan
Published: (2024)
by: Cazenave, Tristan
Published: (2024)
Catching the Infection Before It Spreads: Foresight-Guided Defense in Multi-Agent Systems
by: Ma, Yue, et al.
Published: (2026)
by: Ma, Yue, et al.
Published: (2026)
AgentHub: A Registry for Discoverable, Verifiable, and Reproducible AI Agents
by: Pautsch, Erik, et al.
Published: (2025)
by: Pautsch, Erik, et al.
Published: (2025)
Acceptance Cards:A Four-Diagnostic Standard for Safe Fine-Tuning Defense Claims
by: Konrad, Phongsakon Mark, et al.
Published: (2026)
by: Konrad, Phongsakon Mark, et al.
Published: (2026)
Knowledgeable Agents by Offline Reinforcement Learning from Large Language Model Rollouts
by: Pang, Jing-Cheng, et al.
Published: (2024)
by: Pang, Jing-Cheng, et al.
Published: (2024)
PrismAgent: Illuminating Harm in Memes via a Zero-Shot Interpretable Multi-Agent Framework
by: Ding, Zihan, et al.
Published: (2026)
by: Ding, Zihan, et al.
Published: (2026)
CuSearch: Curriculum Rollout Sampling via Search Depth for Agentic RAG
by: Shen, Jianghan, et al.
Published: (2026)
by: Shen, Jianghan, et al.
Published: (2026)
A Trembling House of Cards? Mapping Adversarial Attacks against Language Agents
by: Mo, Lingbo, et al.
Published: (2024)
by: Mo, Lingbo, et al.
Published: (2024)
Bridging the Reproducibility Divide: Open Source Software's Role in Standardizing Healthcare AI
by: Wu, John, et al.
Published: (2026)
by: Wu, John, et al.
Published: (2026)
AblateCell: A Reproduce-then-Ablate Agent for Virtual Cell Repositories
by: Xia, Xue, et al.
Published: (2026)
by: Xia, Xue, et al.
Published: (2026)
Causal Explanations for Sequential Decision-Making in Multi-Agent Systems
by: Gyevnar, Balint, et al.
Published: (2023)
by: Gyevnar, Balint, et al.
Published: (2023)
Agentic Flow Steering and Parallel Rollout Search for Spatially Grounded Text-to-Image Generation
by: Chen, Ping, et al.
Published: (2026)
by: Chen, Ping, et al.
Published: (2026)
CUBE: A Standard for Unifying Agent Benchmarks
by: Lacoste, Alexandre, et al.
Published: (2026)
by: Lacoste, Alexandre, et al.
Published: (2026)
AI Identity: Standards, Gaps, and Research Directions for AI Agents
by: Otsuka, Takumi, et al.
Published: (2026)
by: Otsuka, Takumi, et al.
Published: (2026)
DR$^{3}$-Eval: Towards Realistic and Reproducible Deep Research Evaluation
by: Xie, Qianqian, et al.
Published: (2026)
by: Xie, Qianqian, et al.
Published: (2026)
CardAIc-Agents: A Multimodal Framework with Hierarchical Adaptation for Cardiac Care Support
by: Zhang, Yuting, et al.
Published: (2025)
by: Zhang, Yuting, et al.
Published: (2025)
From Static Analysis to Audience Dissemination: A Training-Free Multimodal Controversy Detection Multi-Agent Framework
by: Ding, Zihan, et al.
Published: (2026)
by: Ding, Zihan, et al.
Published: (2026)
UrzaGPT: LoRA-Tuned Large Language Models for Card Selection in Collectible Card Games
by: Bertram, Timo
Published: (2025)
by: Bertram, Timo
Published: (2025)
Similar Items
-
ARCANE: A Multi-Agent Framework for Interpretable and Configurable Alignment
by: Masters, Charlie, et al.
Published: (2025) -
Orchestrating Human-AI Teams: The Manager Agent as a Unifying Research Challenge
by: Masters, Charlie, et al.
Published: (2025) -
CAPF: Guiding Search-Agent Rollouts with Credit-Attenuated Privileged Feedback
by: Chen, Bin, et al.
Published: (2026) -
Multi-Agent Reinforcement Learning for Energy Networks: Computational Challenges, Progress and Open Problems
by: Keren, Sarah, et al.
Published: (2024) -
Valet: A Standardized Testbed of Traditional Imperfect-Information Card Games
by: Goadrich, Mark, et al.
Published: (2026)