The PokeAgent Challenge: Competitive and Long-Context Learning at Scale
Fuente:
arXiv
Saved in:
| Main Authors: | Karten, Seth, Grigsby, Jake, Upaa Jr, Tersoo, Bae, Junik, Hong, Seonghun, Jeong, Hyunyoung, Jung, Jaeyoon, Kerdthaisong, Kun, Kim, Gyungbo, Kim, Hyeokgi, Kim, Yujin, Kwon, Eunju, Liu, Dongyu, Mariglia, Patrick, Park, Sangyeon, Schink, Benedikt, Shi, Xianwei, Sistilli, Anthony, Twin, Joseph, Urdu, Arian, Urdu, Matin, Wang, Qiao, Wu, Ling, Zhang, Wenli, Zhou, Kunsheng, Milani, Stephanie, Vodrahalli, Kiran, Zhang, Amy, Fang, Fei, Zhu, Yuke, Jin, Chi |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Continual Harness: Online Adaptation for Self-Improving Foundation Agents
by: Karten, Seth, et al.
Published: (2026)
by: Karten, Seth, et al.
Published: (2026)
PokéChamp: an Expert-level Minimax Language Agent
by: Karten, Seth, et al.
Published: (2025)
by: Karten, Seth, et al.
Published: (2025)
Understanding Political Communication and Political Communicators on Twitch
by: Kim, Sangyeon
Published: (2024)
by: Kim, Sangyeon
Published: (2024)
A Humanoid Visual-Tactile-Action Dataset for Contact-Rich Manipulation
by: Kwon, Eunju, et al.
Published: (2025)
by: Kwon, Eunju, et al.
Published: (2025)
ESREAL: Exploiting Semantic Reconstruction to Mitigate Hallucinations in Vision-Language Models
by: Kim, Minchan, et al.
Published: (2024)
by: Kim, Minchan, et al.
Published: (2024)
KOFFVQA: An Objectively Evaluated Free-form VQA Benchmark for Large Vision-Language Models in the Korean Language
by: Kim, Yoonshik, et al.
Published: (2025)
by: Kim, Yoonshik, et al.
Published: (2025)
The exchange rate pass‐through to domestic prices: A meta‐analysis
by: Tersoo David Iorngurum
Published: (2024)
by: Tersoo David Iorngurum
Published: (2024)
A Framework for Portrait Stylization with Skin-Tone Awareness and Nudity Identification
by: Kim, Seungkwon, et al.
Published: (2024)
by: Kim, Seungkwon, et al.
Published: (2024)
Effects of Multisensory Feedback on the Perception and Performance of Virtual Reality Hand-Retargeted Interaction
by: Jang, Hyunyoung, et al.
Published: (2024)
by: Jang, Hyunyoung, et al.
Published: (2024)
Beyond Static Scoring: Enhancing Assessment Validity via AI-Generated Interactive Verification
by: Lee, Tom, et al.
Published: (2025)
by: Lee, Tom, et al.
Published: (2025)
FENCE: A Financial and Multimodal Jailbreak Detection Dataset
by: Kim, Mirae, et al.
Published: (2026)
by: Kim, Mirae, et al.
Published: (2026)
Tourists' Carbon‐Neutral Intentions and Willingness to Pay Tourism Carbon Tax: Integrating the Climate Change Risk Perception Model and Attitude Framework
by: Eunju Woo, et al.
Published: (2026)
by: Eunju Woo, et al.
Published: (2026)
Synergy-CLIP: Extending CLIP with Multi-modal Integration for Robust Representation Learning
by: Cho, Sangyeon, et al.
Published: (2025)
by: Cho, Sangyeon, et al.
Published: (2025)
A cavity-mediated reconfigurable coupling scheme for superconducting qubits
by: Hwang, Shinyoung, et al.
Published: (2026)
by: Hwang, Shinyoung, et al.
Published: (2026)
The Persistence of Contrarianism on Twitter: Mapping users' sharing habits for the Ukraine war, COVID-19 vaccination, and the 2022 Midterm Elections
by: Axelrod, David, et al.
Published: (2024)
by: Axelrod, David, et al.
Published: (2024)
MoFE: Mixture of Frozen Experts Architecture
by: Seo, Jean, et al.
Published: (2025)
by: Seo, Jean, et al.
Published: (2025)
TetraSDF: Precise Mesh Extraction with Multi-resolution Tetrahedral Grid
by: Oh, Seonghun, et al.
Published: (2025)
by: Oh, Seonghun, et al.
Published: (2025)
Enhancing Visual Re-ranking through Denoising Nearest Neighbor Graph via Continuous CRF
by: Kim, Jaeyoon, et al.
Published: (2024)
by: Kim, Jaeyoon, et al.
Published: (2024)
Exploring Fine-Tuning of Large Audio Language Models for Spoken Language Understanding under Limited Speech Data
by: Choi, Youngwon, et al.
Published: (2025)
by: Choi, Youngwon, et al.
Published: (2025)
Structural analysis of the peptidoglycan DL‐endopeptidase CwlO complexed with its inhibitory protein IseA
by: Sudarshan Tandukar, et al.
Published: (2024)
by: Sudarshan Tandukar, et al.
Published: (2024)
Overwintering Ticks as a Reservoir for Severe Fever With Thrombocytopenia Syndrome Virus in the Republic of Korea
by: Hyunyoung Yoon, et al.
Published: (2026)
by: Hyunyoung Yoon, et al.
Published: (2026)
VisAgent: Narrative-Preserving Story Visualization Framework
by: Kim, Seungkwon, et al.
Published: (2025)
by: Kim, Seungkwon, et al.
Published: (2025)
Towards Test-time Efficient Visual Place Recognition via Asymmetric Query Processing
by: Kim, Jaeyoon, et al.
Published: (2025)
by: Kim, Jaeyoon, et al.
Published: (2025)
2D Materials: From Design and Synthesis to Applications in Electrical and Electrochemical Biosensors
by: Masud, et al.
Published: (2025)
by: Masud, et al.
Published: (2025)
Comparative Genomics and Virulence Mechanisms to Identify Genes Related to Mucin O‐Glycan Degradation and Pathogenicity in a Potentially Multidrug‐Resistant Clostridium tertium Strain
by: Seonghun Kim, et al.
Published: (2025)
by: Seonghun Kim, et al.
Published: (2025)
Gag Gifting: The Joke and the Poke
by: Robert M. Schindler, et al.
Published: (2025)
by: Robert M. Schindler, et al.
Published: (2025)
PokeRRT: Poking as a Skill and Failure Recovery Tactic for Planar Non-Prehensile Manipulation
by: Pasricha, Anuj, et al.
Published: (2022)
by: Pasricha, Anuj, et al.
Published: (2022)
A scalable quantum-neural hybrid variational algorithm for ground state estimation
by: Kim, Minwoo, et al.
Published: (2025)
by: Kim, Minwoo, et al.
Published: (2025)
EL POETA MILITAR CHUCHÚ MARTÍNEZ
by: William Grigsby
Published: (2011)
by: William Grigsby
Published: (2011)
Knowledge Beyond Language: Bridging the Gap in Multilingual Machine Unlearning Evaluation
by: Hwang, Kyomin, et al.
Published: (2026)
by: Hwang, Kyomin, et al.
Published: (2026)
ConCSE: Unified Contrastive Learning and Augmentation for Code-Switched Embeddings
by: Jeon, Jangyeong, et al.
Published: (2024)
by: Jeon, Jangyeong, et al.
Published: (2024)
Unraveling reaction discrepancy and electrolyte stabilizing effects of auto‐oxygenated porphyrin catalysts in lithium–oxygen and lithium–air cells
by: Boran Kim, et al.
Published: (2024)
by: Boran Kim, et al.
Published: (2024)
Simulation of integrated nonlinear quantum optics: from nonlinear interferometer to temporal walk-off compensator
by: Kim, Seonghun, et al.
Published: (2024)
by: Kim, Seonghun, et al.
Published: (2024)
Rank-O-ToM: Unlocking Emotional Nuance Ranking to Enhance Affective Theory-of-Mind
by: Kim, JiHyun, et al.
Published: (2025)
by: Kim, JiHyun, et al.
Published: (2025)
Dual Recursive Feedback on Generation and Appearance Latents for Pose-Robust Text-to-Image Diffusion
by: Kim, Jiwon, et al.
Published: (2025)
by: Kim, Jiwon, et al.
Published: (2025)
StableSketcher: Enhancing Diffusion Model for Pixel-based Sketch Generation via Visual Question Answering Feedback
by: Park, Jiho, et al.
Published: (2025)
by: Park, Jiho, et al.
Published: (2025)
Online Learning with Bounded Recall
by: Schneider, Jon, et al.
Published: (2022)
by: Schneider, Jon, et al.
Published: (2022)
ArtWhisperer: A Dataset for Characterizing Human-AI Interactions in Artistic Creations
by: Vodrahalli, Kailas, et al.
Published: (2023)
by: Vodrahalli, Kailas, et al.
Published: (2023)
An Exposure Model Framework for Signal Detection based on Electronic Healthcare Data
by: Dijkstra, Louis, et al.
Published: (2024)
by: Dijkstra, Louis, et al.
Published: (2024)
TLDR: Unsupervised Goal-Conditioned RL via Temporal Distance-Aware Representations
by: Bae, Junik, et al.
Published: (2024)
by: Bae, Junik, et al.
Published: (2024)
Similar Items
-
Continual Harness: Online Adaptation for Self-Improving Foundation Agents
by: Karten, Seth, et al.
Published: (2026) -
PokéChamp: an Expert-level Minimax Language Agent
by: Karten, Seth, et al.
Published: (2025) -
Understanding Political Communication and Political Communicators on Twitch
by: Kim, Sangyeon
Published: (2024) -
A Humanoid Visual-Tactile-Action Dataset for Contact-Rich Manipulation
by: Kwon, Eunju, et al.
Published: (2025) -
ESREAL: Exploiting Semantic Reconstruction to Mitigate Hallucinations in Vision-Language Models
by: Kim, Minchan, et al.
Published: (2024)