Connecting the Dots: Evaluating Abstract Reasoning Capabilities of LLMs Using the New York Times Connections Word Game
Fuente:
arXiv
Salvato in:
| Autori principali: | Samadarshi, Prisha, Mustafa, Mariam, Kulkarni, Anushka, Rothkopf, Raven, Chakrabarty, Tuhin, Muresan, Smaranda |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Understanding Figurative Meaning through Explainable Visual Entailment
di: Saakyan, Arkadiy, et al.
Pubblicazione: (2024)
di: Saakyan, Arkadiy, et al.
Pubblicazione: (2024)
Death of the Novel(ty): Beyond n-Gram Novelty as a Metric for Textual Creativity
di: Saakyan, Arkadiy, et al.
Pubblicazione: (2025)
di: Saakyan, Arkadiy, et al.
Pubblicazione: (2025)
Creativity Support in the Age of Large Language Models: An Empirical Study Involving Emerging Writers
di: Chakrabarty, Tuhin, et al.
Pubblicazione: (2023)
di: Chakrabarty, Tuhin, et al.
Pubblicazione: (2023)
Making New Connections: LLMs as Puzzle Generators for The New York Times' Connections Word Game
di: Merino, Tim, et al.
Pubblicazione: (2024)
di: Merino, Tim, et al.
Pubblicazione: (2024)
LiveNewsBench: Evaluating LLM Web Search Capabilities with Freshly Curated News
di: Zhang, Yunfan, et al.
Pubblicazione: (2026)
di: Zhang, Yunfan, et al.
Pubblicazione: (2026)
"Is ChatGPT a Better Explainer than My Professor?": Evaluating the Explanation Capabilities of LLMs in Conversation Compared to a Human Baseline
di: Li, Grace, et al.
Pubblicazione: (2024)
di: Li, Grace, et al.
Pubblicazione: (2024)
Art or Artifice? Large Language Models and the False Promise of Creativity
di: Chakrabarty, Tuhin, et al.
Pubblicazione: (2023)
di: Chakrabarty, Tuhin, et al.
Pubblicazione: (2023)
Designing and Evaluating Chain-of-Hints for Scientific Question Answering
di: Jangra, Anubhav, et al.
Pubblicazione: (2025)
di: Jangra, Anubhav, et al.
Pubblicazione: (2025)
ICLEF: In-Context Learning with Expert Feedback for Explainable Style Transfer
di: Saakyan, Arkadiy, et al.
Pubblicazione: (2023)
di: Saakyan, Arkadiy, et al.
Pubblicazione: (2023)
Identifying Self-Disclosures of Use, Misuse and Addiction in Community-based Social Media Posts
di: Yang, Chenghao, et al.
Pubblicazione: (2023)
di: Yang, Chenghao, et al.
Pubblicazione: (2023)
Exploring Chain-of-Thought Reasoning for Steerable Pluralistic Alignment
di: Zhang, Yunfan, et al.
Pubblicazione: (2025)
di: Zhang, Yunfan, et al.
Pubblicazione: (2025)
Large Language Models are Few-Shot Training Example Generators: A Case Study in Fallacy Recognition
di: Alhindi, Tariq, et al.
Pubblicazione: (2023)
di: Alhindi, Tariq, et al.
Pubblicazione: (2023)
NormSAGE: Multi-Lingual Multi-Cultural Norm Discovery from Conversations On-the-Fly
di: Fung, Yi R., et al.
Pubblicazione: (2022)
di: Fung, Yi R., et al.
Pubblicazione: (2022)
Forecasting Conversation Derailments Through Generation
di: Zhang, Yunfan, et al.
Pubblicazione: (2025)
di: Zhang, Yunfan, et al.
Pubblicazione: (2025)
Browsing Lost Unformed Recollections: A Benchmark for Tip-of-the-Tongue Search and Reasoning
di: CH-Wang, Sky, et al.
Pubblicazione: (2025)
di: CH-Wang, Sky, et al.
Pubblicazione: (2025)
Navigating the Landscape of Hint Generation Research: From the Past to the Future
di: Jangra, Anubhav, et al.
Pubblicazione: (2024)
di: Jangra, Anubhav, et al.
Pubblicazione: (2024)
LLMs as Science Journalists: Supporting Early-stage Researchers in Communicating Their Science to the Public
di: Alshomary, Milad, et al.
Pubblicazione: (2026)
di: Alshomary, Milad, et al.
Pubblicazione: (2026)
Can Good Writing Be Generative? Expert-Level AI Writing Emerges through Fine-Tuning on High-Quality Books
di: Chakrabarty, Tuhin, et al.
Pubblicazione: (2026)
di: Chakrabarty, Tuhin, et al.
Pubblicazione: (2026)
Connecting the Contamination Dots
di: Ken Sansone, et al.
Pubblicazione: (2024)
di: Ken Sansone, et al.
Pubblicazione: (2024)
Procedural Adherence and Interpretability Through Neuro-Symbolic Generative Agents
di: Rothkopf, Raven, et al.
Pubblicazione: (2024)
di: Rothkopf, Raven, et al.
Pubblicazione: (2024)
Why and How LLMs Hallucinate: Connecting the Dots with Subsequence Associations
di: Sun, Yiyou, et al.
Pubblicazione: (2025)
di: Sun, Yiyou, et al.
Pubblicazione: (2025)
Hypergraph Connectivity Augmentation in Strongly Polynomial Time
di: Bérczi, Kristóf, et al.
Pubblicazione: (2024)
di: Bérczi, Kristóf, et al.
Pubblicazione: (2024)
CrossWordBench: Evaluating the Reasoning Capabilities of LLMs and LVLMs with Controllable Puzzle Generation
di: Leng, Jixuan, et al.
Pubblicazione: (2025)
di: Leng, Jixuan, et al.
Pubblicazione: (2025)
PRAGyan -- Connecting the Dots in Tweets
di: Ravi, Rahul, et al.
Pubblicazione: (2024)
di: Ravi, Rahul, et al.
Pubblicazione: (2024)
Rose: Composable Autodiff for the Interactive Web
di: Estep, Sam, et al.
Pubblicazione: (2024)
di: Estep, Sam, et al.
Pubblicazione: (2024)
Connecting the Dots: Training-Free Visual Grounding via Agentic Reasoning
di: Luo, Liqin, et al.
Pubblicazione: (2025)
di: Luo, Liqin, et al.
Pubblicazione: (2025)
Lightweight Connective Detection Using Gradient Boosting
di: Er, Mustafa Erolcan, et al.
Pubblicazione: (2024)
di: Er, Mustafa Erolcan, et al.
Pubblicazione: (2024)
Anatomy Connected 2024 Scientific Abstracts
Pubblicazione: (2024)
Pubblicazione: (2024)
Towards Connected Smart Work Zones: Advancing Work Zone Management through Improved Connectivity
di: Nour, Mariam, et al.
Pubblicazione: (2025)
di: Nour, Mariam, et al.
Pubblicazione: (2025)
Graphs of Reduced Words and Some Connections
di: Adeyemo, Praise
Pubblicazione: (2024)
di: Adeyemo, Praise
Pubblicazione: (2024)
VideoNorms: Benchmarking Cultural Awareness of Video Language Models
di: Varimalla, Nikhil Reddy, et al.
Pubblicazione: (2025)
di: Varimalla, Nikhil Reddy, et al.
Pubblicazione: (2025)
Layered Insights: Generalizable Analysis of Authorial Style by Leveraging All Transformer Layers
di: Alshomary, Milad, et al.
Pubblicazione: (2025)
di: Alshomary, Milad, et al.
Pubblicazione: (2025)
SNGR: Selective Non-Gaussian Refinement for Ambiguous SLAM Factor Graphs
di: Kulkarni, Anushka, et al.
Pubblicazione: (2026)
di: Kulkarni, Anushka, et al.
Pubblicazione: (2026)
Resting-State Functional Connectivity Correlates of Emotional Memory Control under Cognitive load in Subclinical Anxiety
di: Kinger, Shruti, et al.
Pubblicazione: (2026)
di: Kinger, Shruti, et al.
Pubblicazione: (2026)
Dense Optical Tracking: Connecting the Dots
di: Moing, Guillaume Le, et al.
Pubblicazione: (2023)
di: Moing, Guillaume Le, et al.
Pubblicazione: (2023)
Cardiotoxicity and Neurotoxicity in Breast Cancer Patients: Is It Time for MRI to Connect the Dots?
di: Jose de Arimateia Batista Araujo‐Filho, et al.
Pubblicazione: (2025)
di: Jose de Arimateia Batista Araujo‐Filho, et al.
Pubblicazione: (2025)
HiLDe: Intentional Code Generation via Human-in-the-Loop Decoding
di: González, Emmanuel Anaya, et al.
Pubblicazione: (2025)
di: González, Emmanuel Anaya, et al.
Pubblicazione: (2025)
XAM: Interactive Explainability for Authorship Attribution Models
di: Alshomary, Milad, et al.
Pubblicazione: (2025)
di: Alshomary, Milad, et al.
Pubblicazione: (2025)
Latent Space Interpretation for Stylistic Analysis and Explainable Authorship Attribution
di: Alshomary, Milad, et al.
Pubblicazione: (2024)
di: Alshomary, Milad, et al.
Pubblicazione: (2024)
Effect of the Dot-Dot Interaction Strength on the Conductance of Side-Connected Quantum Dots
di: V. M. Apel
Pubblicazione: (2006)
di: V. M. Apel
Pubblicazione: (2006)
Documenti analoghi
-
Understanding Figurative Meaning through Explainable Visual Entailment
di: Saakyan, Arkadiy, et al.
Pubblicazione: (2024) -
Death of the Novel(ty): Beyond n-Gram Novelty as a Metric for Textual Creativity
di: Saakyan, Arkadiy, et al.
Pubblicazione: (2025) -
Creativity Support in the Age of Large Language Models: An Empirical Study Involving Emerging Writers
di: Chakrabarty, Tuhin, et al.
Pubblicazione: (2023) -
Making New Connections: LLMs as Puzzle Generators for The New York Times' Connections Word Game
di: Merino, Tim, et al.
Pubblicazione: (2024) -
LiveNewsBench: Evaluating LLM Web Search Capabilities with Freshly Curated News
di: Zhang, Yunfan, et al.
Pubblicazione: (2026)