LLMs for Relational Reasoning: How Far are We?
Fuente:
arXiv
Salvato in:
| Autori principali: | Li, Zhiming, Cao, Yushi, Xu, Xiufeng, Jiang, Junzhe, Liu, Xu, Teo, Yon Shin, Lin, Shang-wei, Liu, Yang |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
How Far Are We from Optimal Reasoning Efficiency?
di: Gao, Jiaxuan, et al.
Pubblicazione: (2025)
di: Gao, Jiaxuan, et al.
Pubblicazione: (2025)
How Far Are We From AGI: Are LLMs All We Need?
di: Feng, Tao, et al.
Pubblicazione: (2024)
di: Feng, Tao, et al.
Pubblicazione: (2024)
How Reliable are LLMs for Reasoning on the Re-ranking task?
di: Islam, Nafis Tanveer, et al.
Pubblicazione: (2025)
di: Islam, Nafis Tanveer, et al.
Pubblicazione: (2025)
How Far Are We on the Decision-Making of LLMs? Evaluating LLMs' Gaming Ability in Multi-Agent Environments
di: Huang, Jen-tse, et al.
Pubblicazione: (2024)
di: Huang, Jen-tse, et al.
Pubblicazione: (2024)
Logic-Q: Improving Deep Reinforcement Learning-based Quantitative Trading via Program Sketch-based Tuning
di: Li, Zhiming, et al.
Pubblicazione: (2023)
di: Li, Zhiming, et al.
Pubblicazione: (2023)
How Far Are We from Intelligent Visual Deductive Reasoning?
di: Zhang, Yizhe, et al.
Pubblicazione: (2024)
di: Zhang, Yizhe, et al.
Pubblicazione: (2024)
MotiveBench: How Far Are We From Human-Like Motivational Reasoning in Large Language Models?
di: Yong, Xixian, et al.
Pubblicazione: (2025)
di: Yong, Xixian, et al.
Pubblicazione: (2025)
VisualWebBench: How Far Have Multimodal LLMs Evolved in Web Page Understanding and Grounding?
di: Liu, Junpeng, et al.
Pubblicazione: (2024)
di: Liu, Junpeng, et al.
Pubblicazione: (2024)
Enabling and Analyzing How to Efficiently Extract Information from Hybrid Long Documents with LLMs
di: Yue, Chongjian, et al.
Pubblicazione: (2023)
di: Yue, Chongjian, et al.
Pubblicazione: (2023)
Recall, Retrieve and Reason: Towards Better In-Context Relation Extraction
di: Li, Guozheng, et al.
Pubblicazione: (2024)
di: Li, Guozheng, et al.
Pubblicazione: (2024)
Nondeterministic Polynomial-time Problem Challenge: An Ever-Scaling Reasoning Benchmark for LLMs
di: Yang, Chang, et al.
Pubblicazione: (2025)
di: Yang, Chang, et al.
Pubblicazione: (2025)
A Primer in Post-Training Reasoning Data: What We Know About How It Works
di: Li, Yaoming, et al.
Pubblicazione: (2026)
di: Li, Yaoming, et al.
Pubblicazione: (2026)
Teaching LLMs According to Their Aptitude: Adaptive Reasoning for Mathematical Problem Solving
di: Xu, Xin, et al.
Pubblicazione: (2025)
di: Xu, Xin, et al.
Pubblicazione: (2025)
Emergent Hierarchical Reasoning in LLMs through Reinforcement Learning
di: Wang, Haozhe, et al.
Pubblicazione: (2025)
di: Wang, Haozhe, et al.
Pubblicazione: (2025)
Unveiling Project-Specific Bias in Neural Code Models
di: Li, Zhiming, et al.
Pubblicazione: (2022)
di: Li, Zhiming, et al.
Pubblicazione: (2022)
Chain of History: Learning and Forecasting with LLMs for Temporal Knowledge Graph Completion
di: Luo, Ruilin, et al.
Pubblicazione: (2024)
di: Luo, Ruilin, et al.
Pubblicazione: (2024)
R-Horizon: How Far Can Your Large Reasoning Model Really Go in Breadth and Depth?
di: Lu, Yi, et al.
Pubblicazione: (2025)
di: Lu, Yi, et al.
Pubblicazione: (2025)
MedReason: Eliciting Factual Medical Reasoning Steps in LLMs via Knowledge Graphs
di: Wu, Juncheng, et al.
Pubblicazione: (2025)
di: Wu, Juncheng, et al.
Pubblicazione: (2025)
Under the Shadow of Babel: How Language Shapes Reasoning in LLMs
di: Wang, Chenxi, et al.
Pubblicazione: (2025)
di: Wang, Chenxi, et al.
Pubblicazione: (2025)
Deceptive Semantic Shortcuts on Reasoning Chains: How Far Can Models Go without Hallucination?
di: Li, Bangzheng, et al.
Pubblicazione: (2023)
di: Li, Bangzheng, et al.
Pubblicazione: (2023)
How Much Can RAG Help the Reasoning of LLM?
di: Liu, Jingyu, et al.
Pubblicazione: (2024)
di: Liu, Jingyu, et al.
Pubblicazione: (2024)
The AI Hippocampus: How Far are We From Human Memory?
di: Jia, Zixia, et al.
Pubblicazione: (2026)
di: Jia, Zixia, et al.
Pubblicazione: (2026)
Reinforcement Learning with Verifiable Rewards Implicitly Incentivizes Correct Reasoning in Base LLMs
di: Wen, Xumeng, et al.
Pubblicazione: (2025)
di: Wen, Xumeng, et al.
Pubblicazione: (2025)
How Far Are Vision-Language Models from Constructing the Real World? A Benchmark for Physical Generative Reasoning
di: Yang, Luyu, et al.
Pubblicazione: (2026)
di: Yang, Luyu, et al.
Pubblicazione: (2026)
Deep Learning-Based Identification of Inconsistent Method Names: How Far Are We?
di: Wang, Taiming, et al.
Pubblicazione: (2025)
di: Wang, Taiming, et al.
Pubblicazione: (2025)
Tokenization Constraints in LLMs: A Study of Symbolic and Arithmetic Reasoning Limits
di: Zhang, Xiang, et al.
Pubblicazione: (2025)
di: Zhang, Xiang, et al.
Pubblicazione: (2025)
Empirical Analysis of Dialogue Relation Extraction with Large Language Models
di: Li, Guozheng, et al.
Pubblicazione: (2024)
di: Li, Guozheng, et al.
Pubblicazione: (2024)
Route-and-Reason: Scaling Large Language Model Reasoning with Reinforced Model Router
di: Shao, Chenyang, et al.
Pubblicazione: (2025)
di: Shao, Chenyang, et al.
Pubblicazione: (2025)
How Far Can In-Context Alignment Go? Exploring the State of In-Context Alignment
di: Huang, Heyan, et al.
Pubblicazione: (2024)
di: Huang, Heyan, et al.
Pubblicazione: (2024)
Select2Reason: Efficient Instruction-Tuning Data Selection for Long-CoT Reasoning
di: Yang, Cehao, et al.
Pubblicazione: (2025)
di: Yang, Cehao, et al.
Pubblicazione: (2025)
How Emotion Shapes the Behavior of LLMs and Agents: A Mechanistic Study
di: Sun, Moran, et al.
Pubblicazione: (2026)
di: Sun, Moran, et al.
Pubblicazione: (2026)
MedMT-Bench: Can LLMs Memorize and Understand Long Multi-Turn Conversations in Medical Scenarios?
di: Yang, Lin, et al.
Pubblicazione: (2026)
di: Yang, Lin, et al.
Pubblicazione: (2026)
GSM-Infinite: How Do Your LLMs Behave over Infinitely Increasing Context Length and Reasoning Complexity?
di: Zhou, Yang, et al.
Pubblicazione: (2025)
di: Zhou, Yang, et al.
Pubblicazione: (2025)
Empowering LLMs with Logical Reasoning: A Comprehensive Survey
di: Cheng, Fengxiang, et al.
Pubblicazione: (2025)
di: Cheng, Fengxiang, et al.
Pubblicazione: (2025)
How Do Answer Tokens Read Reasoning Traces? Self-Reading Patterns in Thinking LLMs for Quantitative Reasoning
di: Chen, Haoyang, et al.
Pubblicazione: (2026)
di: Chen, Haoyang, et al.
Pubblicazione: (2026)
Spider2-V: How Far Are Multimodal Agents From Automating Data Science and Engineering Workflows?
di: Cao, Ruisheng, et al.
Pubblicazione: (2024)
di: Cao, Ruisheng, et al.
Pubblicazione: (2024)
Reasoning or Not? A Comprehensive Evaluation of Reasoning LLMs for Dialogue Summarization
di: Jin, Keyan, et al.
Pubblicazione: (2025)
di: Jin, Keyan, et al.
Pubblicazione: (2025)
An Online Reference-Free Evaluation Framework for Flowchart Image-to-Code Generation
di: Nguyen, Giang Son, et al.
Pubblicazione: (2026)
di: Nguyen, Giang Son, et al.
Pubblicazione: (2026)
Is Depth All You Need? An Exploration of Iterative Reasoning in LLMs
di: Wu, Zongqian, et al.
Pubblicazione: (2025)
di: Wu, Zongqian, et al.
Pubblicazione: (2025)
Retrieval-Augmented Test Generation: How Far Are We?
di: Shin, Jiho, et al.
Pubblicazione: (2024)
di: Shin, Jiho, et al.
Pubblicazione: (2024)
Documenti analoghi
-
How Far Are We from Optimal Reasoning Efficiency?
di: Gao, Jiaxuan, et al.
Pubblicazione: (2025) -
How Far Are We From AGI: Are LLMs All We Need?
di: Feng, Tao, et al.
Pubblicazione: (2024) -
How Reliable are LLMs for Reasoning on the Re-ranking task?
di: Islam, Nafis Tanveer, et al.
Pubblicazione: (2025) -
How Far Are We on the Decision-Making of LLMs? Evaluating LLMs' Gaming Ability in Multi-Agent Environments
di: Huang, Jen-tse, et al.
Pubblicazione: (2024) -
Logic-Q: Improving Deep Reinforcement Learning-based Quantitative Trading via Program Sketch-based Tuning
di: Li, Zhiming, et al.
Pubblicazione: (2023)