AI Gamestore: Scalable, Open-Ended Evaluation of Machine General Intelligence with Human Games
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ying, Lance, Truong, Ryan, Sharma, Prafull, Zhao, Kaiya Ivy, Cloos, Nathan, Allen, Kelsey R., Griffiths, Thomas L., Collins, Katherine M., Hernández-Orallo, José, Isola, Phillip, Gershman, Samuel J., Tenenbaum, Joshua B. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Assessing Adaptive World Models in Machines with Novel Games
von: Ying, Lance, et al.
Veröffentlicht: (2025)
von: Ying, Lance, et al.
Veröffentlicht: (2025)
Adaptive Social Learning using Theory of Mind
von: Ying, Lance, et al.
Veröffentlicht: (2025)
von: Ying, Lance, et al.
Veröffentlicht: (2025)
Evaluating Language Models' Evaluations of Games
von: Collins, Katherine M., et al.
Veröffentlicht: (2025)
von: Collins, Katherine M., et al.
Veröffentlicht: (2025)
On Benchmarking Human-Like Intelligence in Machines
von: Ying, Lance, et al.
Veröffentlicht: (2025)
von: Ying, Lance, et al.
Veröffentlicht: (2025)
Digital Red Queen: Adversarial Program Evolution in Core War with LLMs
von: Kumar, Akarsh, et al.
Veröffentlicht: (2026)
von: Kumar, Akarsh, et al.
Veröffentlicht: (2026)
Language-Informed Synthesis of Rational Agent Models for Grounded Theory-of-Mind Reasoning On-The-Fly
von: Ying, Lance, et al.
Veröffentlicht: (2025)
von: Ying, Lance, et al.
Veröffentlicht: (2025)
Belief Attribution as Mental Explanation: The Role of Accuracy, Informativity, and Causality
von: Ying, Lance, et al.
Veröffentlicht: (2025)
von: Ying, Lance, et al.
Veröffentlicht: (2025)
What's in the Box? Reasoning about Unseen Objects from Multimodal Cues
von: Ying, Lance, et al.
Veröffentlicht: (2025)
von: Ying, Lance, et al.
Veröffentlicht: (2025)
"Just in Time" World Modeling Supports Human Planning and Reasoning
von: Chen, Tony, et al.
Veröffentlicht: (2026)
von: Chen, Tony, et al.
Veröffentlicht: (2026)
Generation and Evaluation in the Human Invention Process through the Lens of Game Design
von: Collins, Katherine M., et al.
Veröffentlicht: (2025)
von: Collins, Katherine M., et al.
Veröffentlicht: (2025)
Synthesizing world models for bilevel planning
von: Ahmed, Zergham, et al.
Veröffentlicht: (2025)
von: Ahmed, Zergham, et al.
Veröffentlicht: (2025)
Neural Thickets: Diverse Task Experts Are Dense Around Pretrained Weights
von: Gan, Yulu, et al.
Veröffentlicht: (2026)
von: Gan, Yulu, et al.
Veröffentlicht: (2026)
Improved Representation of Asymmetrical Distances with Interval Quasimetric Embeddings
von: Wang, Tongzhou, et al.
Veröffentlicht: (2022)
von: Wang, Tongzhou, et al.
Veröffentlicht: (2022)
Learning Abstractions for Hierarchical Planning in Program-Synthesis Agents
von: Ahmed, Zergham, et al.
Veröffentlicht: (2026)
von: Ahmed, Zergham, et al.
Veröffentlicht: (2026)
Scalable Optimization in the Modular Norm
von: Large, Tim, et al.
Veröffentlicht: (2024)
von: Large, Tim, et al.
Veröffentlicht: (2024)
Infinite Ends from Finite Samples: Open-Ended Goal Inference as Top-Down Bayesian Filtering of Bottom-Up Proposals
von: Zhi-Xuan, Tan, et al.
Veröffentlicht: (2024)
von: Zhi-Xuan, Tan, et al.
Veröffentlicht: (2024)
Atypical Hodge Loci
von: Griffiths, Phillip
Veröffentlicht: (2025)
von: Griffiths, Phillip
Veröffentlicht: (2025)
Pragmatic Instruction Following and Goal Assistance via Cooperative Language-Guided Inverse Planning
von: Zhi-Xuan, Tan, et al.
Veröffentlicht: (2024)
von: Zhi-Xuan, Tan, et al.
Veröffentlicht: (2024)
Scalable Real2Sim: Physics-Aware Asset Generation Via Robotic Pick-and-Place Setups
von: Pfaff, Nicholas, et al.
Veröffentlicht: (2025)
von: Pfaff, Nicholas, et al.
Veröffentlicht: (2025)
Under the Influence: Quantifying Persuasion and Vigilance in Large Language Models
von: Robinson, Sasha, et al.
Veröffentlicht: (2026)
von: Robinson, Sasha, et al.
Veröffentlicht: (2026)
Empathy in Explanation
von: Collins, Katherine M., et al.
Veröffentlicht: (2025)
von: Collins, Katherine M., et al.
Veröffentlicht: (2025)
Grounding Language about Belief in a Bayesian Theory-of-Mind
von: Ying, Lance, et al.
Veröffentlicht: (2024)
von: Ying, Lance, et al.
Veröffentlicht: (2024)
Understanding Epistemic Language with a Language-augmented Bayesian Theory of Mind
von: Ying, Lance, et al.
Veröffentlicht: (2024)
von: Ying, Lance, et al.
Veröffentlicht: (2024)
The Truth Lies Somewhere in the Middle (of the Generated Tokens)
von: Wang, Sophie L., et al.
Veröffentlicht: (2026)
von: Wang, Sophie L., et al.
Veröffentlicht: (2026)
Words That Make Language Models Perceive
von: Wang, Sophie L., et al.
Veröffentlicht: (2025)
von: Wang, Sophie L., et al.
Veröffentlicht: (2025)
A Framework for Standardizing Similarity Measures in a Rapidly Evolving Field
von: Cloos, Nathan, et al.
Veröffentlicht: (2024)
von: Cloos, Nathan, et al.
Veröffentlicht: (2024)
The Role of Fat in Osteoarthritis
von: Kelsey H. Collins
Veröffentlicht: (2026)
von: Kelsey H. Collins
Veröffentlicht: (2026)
People use fast, goal-directed simulation to reason about novel games
von: Zhang, Cedegao E., et al.
Veröffentlicht: (2024)
von: Zhang, Cedegao E., et al.
Veröffentlicht: (2024)
Interspecific Conformity and Asymmetric Behavioral Convergence in Drosophila
von: Kaiya Hamamichi, et al.
Veröffentlicht: (2026)
von: Kaiya Hamamichi, et al.
Veröffentlicht: (2026)
Identification, chemical synthesis, and receptor binding of a reptilian gecko ghrelin
von: Hidekazu Katayama, et al.
Veröffentlicht: (2024)
von: Hidekazu Katayama, et al.
Veröffentlicht: (2024)
GOMA: Proactive Embodied Cooperative Communication via Goal-Oriented Mental Alignment
von: Ying, Lance, et al.
Veröffentlicht: (2024)
von: Ying, Lance, et al.
Veröffentlicht: (2024)
Characterizing Security and Privacy Teaching Standards for Schools in the United States
von: Limes, Katherine, et al.
Veröffentlicht: (2025)
von: Limes, Katherine, et al.
Veröffentlicht: (2025)
MARBLE: Material Recomposition and Blending in CLIP-Space
von: Cheng, Ta-Ying, et al.
Veröffentlicht: (2025)
von: Cheng, Ta-Ying, et al.
Veröffentlicht: (2025)
Limiting mixed Hodge structures associated to I-surfaces with simple elliptic singularities
von: Friedman, Robert, et al.
Veröffentlicht: (2024)
von: Friedman, Robert, et al.
Veröffentlicht: (2024)
Deformations of I-surfaces with elliptic singularities
von: Friedman, Robert, et al.
Veröffentlicht: (2024)
von: Friedman, Robert, et al.
Veröffentlicht: (2024)
Introducción a la minería de datos / José Hern ndez Orallo, Ma. José Ramírez Quintana, César Ferri Ramírez
von: Hernández Orallo, José
von: Hernández Orallo, José
Aprendizaje Automático de Programas Lógico-Funcionales
von: José Hernández Orallo
Veröffentlicht: (2000)
von: José Hernández Orallo
Veröffentlicht: (2000)
The Platonic Representation Hypothesis
von: Huh, Minyoung, et al.
Veröffentlicht: (2024)
von: Huh, Minyoung, et al.
Veröffentlicht: (2024)
Cycle Consistency as Reward: Learning Image-Text Alignment without Human Preferences
von: Bahng, Hyojin, et al.
Veröffentlicht: (2025)
von: Bahng, Hyojin, et al.
Veröffentlicht: (2025)
Adaptive Length Image Tokenization via Recurrent Allocation
von: Duggal, Shivam, et al.
Veröffentlicht: (2024)
von: Duggal, Shivam, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Assessing Adaptive World Models in Machines with Novel Games
von: Ying, Lance, et al.
Veröffentlicht: (2025) -
Adaptive Social Learning using Theory of Mind
von: Ying, Lance, et al.
Veröffentlicht: (2025) -
Evaluating Language Models' Evaluations of Games
von: Collins, Katherine M., et al.
Veröffentlicht: (2025) -
On Benchmarking Human-Like Intelligence in Machines
von: Ying, Lance, et al.
Veröffentlicht: (2025) -
Digital Red Queen: Adversarial Program Evolution in Core War with LLMs
von: Kumar, Akarsh, et al.
Veröffentlicht: (2026)