ReasoningWeekly: A General Knowledge and Verbal Reasoning Challenge for Large Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wu, Zixuan, Lucchetti, Francesca, Boruch-Gruszecki, Aleksander, Zhao, Jingmiao, Anderson, Carolyn Jane, Biswas, Joydeep, Cassano, Federico, Guha, Arjun |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Agnostics: Learning to Code in Any Programming Language via Reinforcement with a Universal Learning Environment
von: Boruch-Gruszecki, Aleksander, et al.
Veröffentlicht: (2025)
von: Boruch-Gruszecki, Aleksander, et al.
Veröffentlicht: (2025)
AgentPack: A Dataset of Code Changes, Co-Authored by Agents and Humans
von: Zi, Yangtian, et al.
Veröffentlicht: (2025)
von: Zi, Yangtian, et al.
Veröffentlicht: (2025)
Substance Beats Style: Why Beginning Students Fail to Code with LLMs
von: Lucchetti, Francesca, et al.
Veröffentlicht: (2024)
von: Lucchetti, Francesca, et al.
Veröffentlicht: (2024)
Knowledge Transfer from High-Resource to Low-Resource Programming Languages for Code LLMs
von: Cassano, Federico, et al.
Veröffentlicht: (2023)
von: Cassano, Federico, et al.
Veröffentlicht: (2023)
Understanding How CodeLLMs (Mis)Predict Types with Activation Steering
von: Lucchetti, Francesca, et al.
Veröffentlicht: (2024)
von: Lucchetti, Francesca, et al.
Veröffentlicht: (2024)
Creating and Repairing Robot Programs in Open-World Domains
von: Schlesinger, Claire, et al.
Veröffentlicht: (2024)
von: Schlesinger, Claire, et al.
Veröffentlicht: (2024)
Deploying and Evaluating LLMs to Program Service Mobile Robots
von: Hu, Zichao, et al.
Veröffentlicht: (2023)
von: Hu, Zichao, et al.
Veröffentlicht: (2023)
Can It Edit? Evaluating the Ability of Large Language Models to Follow Code Editing Instructions
von: Cassano, Federico, et al.
Veröffentlicht: (2023)
von: Cassano, Federico, et al.
Veröffentlicht: (2023)
Robo-Instruct: Simulator-Augmented Instruction Alignment For Finetuning Code LLMs
von: Hu, Zichao, et al.
Veröffentlicht: (2024)
von: Hu, Zichao, et al.
Veröffentlicht: (2024)
Knowledge Vector of Logical Reasoning in Large Language Models
von: Wang, Zixuan, et al.
Veröffentlicht: (2026)
von: Wang, Zixuan, et al.
Veröffentlicht: (2026)
"I Would Have Written My Code Differently'': Beginners Struggle to Understand LLM-Generated Code
von: Zi, Yangtian, et al.
Veröffentlicht: (2025)
von: Zi, Yangtian, et al.
Veröffentlicht: (2025)
GlyphPattern: An Abstract Pattern Recognition Benchmark for Vision-Language Models
von: Wu, Zixuan, et al.
Veröffentlicht: (2024)
von: Wu, Zixuan, et al.
Veröffentlicht: (2024)
Follow the Path: Reasoning over Knowledge Graph Paths to Improve Large Language Model Factuality
von: Zhang, Mike, et al.
Veröffentlicht: (2025)
von: Zhang, Mike, et al.
Veröffentlicht: (2025)
Process Supervision via Verbal Critique Improves Reasoning in Large Language Models
von: Chen, Hao-Yuan
Veröffentlicht: (2026)
von: Chen, Hao-Yuan
Veröffentlicht: (2026)
Evaluating Computational Accuracy of Large Language Models in Numerical Reasoning Tasks for Healthcare Applications
von: Malghan, Arjun R.
Veröffentlicht: (2025)
von: Malghan, Arjun R.
Veröffentlicht: (2025)
How Beginning Programmers and Code LLMs (Mis)read Each Other
von: Nguyen, Sydney, et al.
Veröffentlicht: (2024)
von: Nguyen, Sydney, et al.
Veröffentlicht: (2024)
ReMEmbR: Building and Reasoning Over Long-Horizon Spatio-Temporal Memory for Robot Navigation
von: Anwar, Abrar, et al.
Veröffentlicht: (2024)
von: Anwar, Abrar, et al.
Veröffentlicht: (2024)
Dictionary Insertion Prompting for Multilingual Reasoning on Multilingual Large Language Models
von: Lu, Hongyuan, et al.
Veröffentlicht: (2024)
von: Lu, Hongyuan, et al.
Veröffentlicht: (2024)
La colaboración campbell y la Práctica Basada en la evidencia
von: Robert F. Boruch
Veröffentlicht: (2002)
von: Robert F. Boruch
Veröffentlicht: (2002)
Are Retrials All You Need? Enhancing Large Language Model Reasoning Without Verbalized Feedback
von: Potamitis, Nearchos, et al.
Veröffentlicht: (2025)
von: Potamitis, Nearchos, et al.
Veröffentlicht: (2025)
Large Language Models Still Face Challenges in Multi-Hop Reasoning with External Knowledge
von: Zhang, Haotong
Veröffentlicht: (2024)
von: Zhang, Haotong
Veröffentlicht: (2024)
Verbal-R3: Verbal Reranker as the Missing Bridge between Retrieval and Reasoning
von: Park, Sangkwon, et al.
Veröffentlicht: (2026)
von: Park, Sangkwon, et al.
Veröffentlicht: (2026)
Knowledge Crosswords: Geometric Knowledge Reasoning with Large Language Models
von: Ding, Wenxuan, et al.
Veröffentlicht: (2023)
von: Ding, Wenxuan, et al.
Veröffentlicht: (2023)
SelfCodeAlign: Self-Alignment for Code Generation
von: Wei, Yuxiang, et al.
Veröffentlicht: (2024)
von: Wei, Yuxiang, et al.
Veröffentlicht: (2024)
Graph-constrained Reasoning: Faithful Reasoning on Knowledge Graphs with Large Language Models
von: Luo, Linhao, et al.
Veröffentlicht: (2024)
von: Luo, Linhao, et al.
Veröffentlicht: (2024)
Selective Temporal Knowledge Graph Reasoning
von: Hou, Zhongni, et al.
Veröffentlicht: (2024)
von: Hou, Zhongni, et al.
Veröffentlicht: (2024)
The Geometry of Thought: How Scale Restructures Reasoning In Large Language Models
von: Anderson, Samuel Cyrenius
Veröffentlicht: (2026)
von: Anderson, Samuel Cyrenius
Veröffentlicht: (2026)
Agricultural Reason in the Shadow of Subsistence Capitalism
von: Appadurai, Arjun
Veröffentlicht: (2024)
von: Appadurai, Arjun
Veröffentlicht: (2024)
Knowledge Reasoning Language Model: Unifying Knowledge and Language for Inductive Knowledge Graph Reasoning
von: Zhuo, Xingrui, et al.
Veröffentlicht: (2025)
von: Zhuo, Xingrui, et al.
Veröffentlicht: (2025)
OWL: Geometry-Aware Spatial Reasoning for Audio Large Language Models
von: Biswas, Subrata, et al.
Veröffentlicht: (2025)
von: Biswas, Subrata, et al.
Veröffentlicht: (2025)
Large Language Models for Mathematical Reasoning: Progresses and Challenges
von: Ahn, Janice, et al.
Veröffentlicht: (2024)
von: Ahn, Janice, et al.
Veröffentlicht: (2024)
Disentangling Reasoning and Knowledge in Medical Large Language Models
von: Thapa, Rahul, et al.
Veröffentlicht: (2025)
von: Thapa, Rahul, et al.
Veröffentlicht: (2025)
Large Language Models are In-context Teachers for Knowledge Reasoning
von: Zhao, Jiachen, et al.
Veröffentlicht: (2023)
von: Zhao, Jiachen, et al.
Veröffentlicht: (2023)
Celebrating 2002-Children's Book Week and the Newbery and Caldecott Awards.
von: Brodie, Carolyn S.
Veröffentlicht: (2002)
von: Brodie, Carolyn S.
Veröffentlicht: (2002)
Physics Reasoner: Knowledge-Augmented Reasoning for Solving Physics Problems with Large Language Models
von: Pang, Xinyu, et al.
Veröffentlicht: (2024)
von: Pang, Xinyu, et al.
Veröffentlicht: (2024)
Evaluating Computational Representations of Character: An Austen Character Similarity Benchmark
von: Yang, Funing, et al.
Veröffentlicht: (2024)
von: Yang, Funing, et al.
Veröffentlicht: (2024)
Knowledge Distillation for Temporal Knowledge Graph Reasoning with Large Language Models
von: Xing, Wang, et al.
Veröffentlicht: (2026)
von: Xing, Wang, et al.
Veröffentlicht: (2026)
Reasoning About Reasoning: Towards Informed and Reflective Use of LLM Reasoning in HCI
von: Mothilal, Ramaravind Kommiya, et al.
Veröffentlicht: (2025)
von: Mothilal, Ramaravind Kommiya, et al.
Veröffentlicht: (2025)
General365: Benchmarking General Reasoning in Large Language Models Across Diverse and Challenging Tasks
von: Liu, Junlin, et al.
Veröffentlicht: (2026)
von: Liu, Junlin, et al.
Veröffentlicht: (2026)
Strategic Facility Location with Limited Liars
von: Gruszecki, Yue, et al.
Veröffentlicht: (2026)
von: Gruszecki, Yue, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Agnostics: Learning to Code in Any Programming Language via Reinforcement with a Universal Learning Environment
von: Boruch-Gruszecki, Aleksander, et al.
Veröffentlicht: (2025) -
AgentPack: A Dataset of Code Changes, Co-Authored by Agents and Humans
von: Zi, Yangtian, et al.
Veröffentlicht: (2025) -
Substance Beats Style: Why Beginning Students Fail to Code with LLMs
von: Lucchetti, Francesca, et al.
Veröffentlicht: (2024) -
Knowledge Transfer from High-Resource to Low-Resource Programming Languages for Code LLMs
von: Cassano, Federico, et al.
Veröffentlicht: (2023) -
Understanding How CodeLLMs (Mis)Predict Types with Activation Steering
von: Lucchetti, Francesca, et al.
Veröffentlicht: (2024)