Missed Connections: Lateral Thinking Puzzles for Large Language Models
Fuente:
arXiv
Guardado en:
| Autores principales: | Todd, Graham, Merino, Tim, Earle, Sam, Togelius, Julian |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Making New Connections: LLMs as Puzzle Generators for The New York Times' Connections Word Game
por: Merino, Tim, et al.
Publicado: (2024)
por: Merino, Tim, et al.
Publicado: (2024)
Large Language Models and Games: A Survey and Roadmap
por: Gallotta, Roberto, et al.
Publicado: (2024)
por: Gallotta, Roberto, et al.
Publicado: (2024)
ScriptDoctor: Automatic Generation of PuzzleScript Games via Large Language Models and Tree Search
por: Earle, Sam, et al.
Publicado: (2025)
por: Earle, Sam, et al.
Publicado: (2025)
LLMatic: Neural Architecture Search via Large Language Models and Quality Diversity Optimization
por: Nasir, Muhammad U., et al.
Publicado: (2023)
por: Nasir, Muhammad U., et al.
Publicado: (2023)
In Search of the Ingredients of Open-Endedness: Replicating Picbreeder with Large Vision-Language Models
por: Earle, Sam, et al.
Publicado: (2026)
por: Earle, Sam, et al.
Publicado: (2026)
Autoverse: An Evolvable Game Language for Learning Robust Embodied Agents
por: Earle, Sam, et al.
Publicado: (2024)
por: Earle, Sam, et al.
Publicado: (2024)
DreamCraft: Text-Guided Generation of Functional 3D Environments in Minecraft
por: Earle, Sam, et al.
Publicado: (2024)
por: Earle, Sam, et al.
Publicado: (2024)
Word2World: Generating Stories and Worlds through Large Language Models
por: Nasir, Muhammad U., et al.
Publicado: (2024)
por: Nasir, Muhammad U., et al.
Publicado: (2024)
GameTraversalBenchmark: Evaluating Planning Abilities Of Large Language Models Through Traversing 2D Game Maps
por: Nasir, Muhammad Umair, et al.
Publicado: (2024)
por: Nasir, Muhammad Umair, et al.
Publicado: (2024)
PuzzleJAX: A Benchmark for Reasoning and Learning
por: Earle, Sam, et al.
Publicado: (2025)
por: Earle, Sam, et al.
Publicado: (2025)
AILS-NTUA at SemEval-2024 Task 9: Cracking Brain Teasers: Transformer Models for Lateral Thinking Puzzles
por: Panagiotopoulos, Ioannis, et al.
Publicado: (2024)
por: Panagiotopoulos, Ioannis, et al.
Publicado: (2024)
PCGRL+: Scaling, Control and Generalization in Reinforcement Learning Level Generators
por: Earle, Sam, et al.
Publicado: (2024)
por: Earle, Sam, et al.
Publicado: (2024)
Weak-eval-Strong: Evaluating and Eliciting Lateral Thinking of LLMs with Situation Puzzles
por: Chen, Qi, et al.
Publicado: (2024)
por: Chen, Qi, et al.
Publicado: (2024)
Video Game Level Design as a Multi-Agent Reinforcement Learning Problem
por: Earle, Sam, et al.
Publicado: (2025)
por: Earle, Sam, et al.
Publicado: (2025)
FoodPuzzle: Developing Large Language Model Agents as Flavor Scientists
por: Huang, Tenghao, et al.
Publicado: (2024)
por: Huang, Tenghao, et al.
Publicado: (2024)
Puzzle Solving using Reasoning of Large Language Models: A Survey
por: Giadikiaroglou, Panagiotis, et al.
Publicado: (2024)
por: Giadikiaroglou, Panagiotis, et al.
Publicado: (2024)
Enigmata: Scaling Logical Reasoning in Large Language Models with Synthetic Verifiable Puzzles
por: Chen, Jiangjie, et al.
Publicado: (2025)
por: Chen, Jiangjie, et al.
Publicado: (2025)
What Is Missing: Interpretable Ratings for Large Language Model Outputs
por: Stranges, Nicholas, et al.
Publicado: (2026)
por: Stranges, Nicholas, et al.
Publicado: (2026)
"Lost-in-the-Later": Framework for Quantifying Contextual Grounding in Large Language Models
por: Tao, Yufei, et al.
Publicado: (2025)
por: Tao, Yufei, et al.
Publicado: (2025)
DreamGarden: A Designer Assistant for Growing Games from a Single Prompt
por: Earle, Sam, et al.
Publicado: (2024)
por: Earle, Sam, et al.
Publicado: (2024)
PuzzlePlex: Benchmarking Foundation Models on Reasoning and Planning with Puzzles
por: Long, Yitao, et al.
Publicado: (2025)
por: Long, Yitao, et al.
Publicado: (2025)
Cognitive Decision Routing in Large Language Models: When to Think Fast, When to Think Slow
por: Du, Y., et al.
Publicado: (2025)
por: Du, Y., et al.
Publicado: (2025)
The Cost of Thinking: Increased Jailbreak Risk in Large Language Models
por: Yang, Fan
Publicado: (2025)
por: Yang, Fan
Publicado: (2025)
Multi-Persona Thinking for Bias Mitigation in Large Language Models
por: Chen, Yuxing, et al.
Publicado: (2026)
por: Chen, Yuxing, et al.
Publicado: (2026)
THiNK: Can Large Language Models Think-aloud?
por: Yu, Yongan, et al.
Publicado: (2025)
por: Yu, Yongan, et al.
Publicado: (2025)
Think$^{2}$: Grounded Metacognitive Reasoning in Large Language Models
por: Elenjical, Abraham Paul, et al.
Publicado: (2026)
por: Elenjical, Abraham Paul, et al.
Publicado: (2026)
The Garden of Forking Paths: Narrative Arc-Conditioned Gameplay Planning
por: Wen, Yunge, et al.
Publicado: (2026)
por: Wen, Yunge, et al.
Publicado: (2026)
Enhancing Creativity in Large Language Models through Associative Thinking Strategies
por: Mehrotra, Pronita, et al.
Publicado: (2024)
por: Mehrotra, Pronita, et al.
Publicado: (2024)
Fast-Slow-Thinking: Complex Task Solving with Large Language Models
por: Sun, Yiliu, et al.
Publicado: (2025)
por: Sun, Yiliu, et al.
Publicado: (2025)
Incentivizing Dual Process Thinking for Efficient Large Language Model Reasoning
por: Cheng, Xiaoxue, et al.
Publicado: (2025)
por: Cheng, Xiaoxue, et al.
Publicado: (2025)
Dream-Cubed: Controllable Generative Modeling in Minecraft by Training on Billions of Cubes
por: Merino, Tim, et al.
Publicado: (2026)
por: Merino, Tim, et al.
Publicado: (2026)
PCGRLLM: Large Language Model-Driven Reward Design for Procedural Content Generation Reinforcement Learning
por: Baek, In-Chang, et al.
Publicado: (2025)
por: Baek, In-Chang, et al.
Publicado: (2025)
Augmenting Lateral Thinking in Language Models with Humor and Riddle Data for the BRAINTEASER Task
por: Ghashami, Mina, et al.
Publicado: (2024)
por: Ghashami, Mina, et al.
Publicado: (2024)
TangramPuzzle: Evaluating Multimodal Large Language Models with Compositional Spatial Reasoning
por: Liu, Daixian, et al.
Publicado: (2026)
por: Liu, Daixian, et al.
Publicado: (2026)
Thinking Tokens for Language Modeling
por: Herel, David, et al.
Publicado: (2024)
por: Herel, David, et al.
Publicado: (2024)
Thinking Before Constraining: A Unified Decoding Framework for Large Language Models
por: Nguyen, Ngoc Trinh Hung, et al.
Publicado: (2026)
por: Nguyen, Ngoc Trinh Hung, et al.
Publicado: (2026)
DynamicMind: A Tri-Mode Thinking System for Large Language Models
por: Li, Wei, et al.
Publicado: (2025)
por: Li, Wei, et al.
Publicado: (2025)
Brain in a Vat: On Missing Pieces Towards Artificial General Intelligence in Large Language Models
por: Ma, Yuxi, et al.
Publicado: (2023)
por: Ma, Yuxi, et al.
Publicado: (2023)
To Think or Not To Think, That is The Question for Large Reasoning Models in Theory of Mind Tasks
por: Gong, Nanxu, et al.
Publicado: (2026)
por: Gong, Nanxu, et al.
Publicado: (2026)
Modeling Hierarchical Thinking in Large Reasoning Models
por: Shahariar, G M, et al.
Publicado: (2025)
por: Shahariar, G M, et al.
Publicado: (2025)
Ejemplares similares
-
Making New Connections: LLMs as Puzzle Generators for The New York Times' Connections Word Game
por: Merino, Tim, et al.
Publicado: (2024) -
Large Language Models and Games: A Survey and Roadmap
por: Gallotta, Roberto, et al.
Publicado: (2024) -
ScriptDoctor: Automatic Generation of PuzzleScript Games via Large Language Models and Tree Search
por: Earle, Sam, et al.
Publicado: (2025) -
LLMatic: Neural Architecture Search via Large Language Models and Quality Diversity Optimization
por: Nasir, Muhammad U., et al.
Publicado: (2023) -
In Search of the Ingredients of Open-Endedness: Replicating Picbreeder with Large Vision-Language Models
por: Earle, Sam, et al.
Publicado: (2026)