Disentangling Logic: The Role of Context in Large Language Model Reasoning Capabilities
Fuente:
arXiv
Salvato in:
| Autori principali: | Hua, Wenyue, Zhu, Kaijie, Li, Lingyao, Fan, Lizhou, Lin, Shuhang, Jin, Mingyu, Xue, Haochen, Li, Zelong, Wang, JinDong, Zhang, Yongfeng |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
NPHardEval: Dynamic Benchmark on Reasoning Ability of Large Language Models via Complexity Classes
di: Fan, Lizhou, et al.
Pubblicazione: (2023)
di: Fan, Lizhou, et al.
Pubblicazione: (2023)
BattleAgent: Multi-modal Dynamic Emulation on Historical Battles to Complement Historical Analysis
di: Lin, Shuhang, et al.
Pubblicazione: (2024)
di: Lin, Shuhang, et al.
Pubblicazione: (2024)
NPHardEval4V: Dynamic Evaluation of Large Vision-Language Models with Effects of Vision
di: Li, Xiang, et al.
Pubblicazione: (2024)
di: Li, Xiang, et al.
Pubblicazione: (2024)
Disentangling Memory and Reasoning Ability in Large Language Models
di: Jin, Mingyu, et al.
Pubblicazione: (2024)
di: Jin, Mingyu, et al.
Pubblicazione: (2024)
EmojiPrompt: Generative Prompt Obfuscation for Privacy-Preserving Communication with Cloud-based LLMs
di: Lin, Sam, et al.
Pubblicazione: (2024)
di: Lin, Sam, et al.
Pubblicazione: (2024)
War and Peace (WarAgent): Large Language Model-based Multi-Agent Simulation of World Wars
di: Hua, Wenyue, et al.
Pubblicazione: (2023)
di: Hua, Wenyue, et al.
Pubblicazione: (2023)
Large Language Models in Biomedical and Health Informatics: A Review with Bibliometric Analysis
di: Yu, Huizi, et al.
Pubblicazione: (2024)
di: Yu, Huizi, et al.
Pubblicazione: (2024)
Formal-LLM: Integrating Formal Language and Natural Language for Controllable LLM-based Agents
di: Li, Zelong, et al.
Pubblicazione: (2024)
di: Li, Zelong, et al.
Pubblicazione: (2024)
The Impact of Reasoning Step Length on Large Language Models
di: Jin, Mingyu, et al.
Pubblicazione: (2024)
di: Jin, Mingyu, et al.
Pubblicazione: (2024)
TrustAgent: Towards Safe and Trustworthy LLM-based Agents
di: Hua, Wenyue, et al.
Pubblicazione: (2024)
di: Hua, Wenyue, et al.
Pubblicazione: (2024)
AttackEval: How to Evaluate the Effectiveness of Jailbreak Attacking on Large Language Models
di: Shu, Dong, et al.
Pubblicazione: (2024)
di: Shu, Dong, et al.
Pubblicazione: (2024)
PAP-REC: Personalized Automatic Prompt for Recommendation Language Model
di: Li, Zelong, et al.
Pubblicazione: (2024)
di: Li, Zelong, et al.
Pubblicazione: (2024)
When AI Meets Finance (StockAgent): Large Language Model-based Stock Trading in Simulated Real-world Environments
di: Zhang, Chong, et al.
Pubblicazione: (2024)
di: Zhang, Chong, et al.
Pubblicazione: (2024)
Game-theoretic LLM: Agent Workflow for Negotiation Games
di: Hua, Wenyue, et al.
Pubblicazione: (2024)
di: Hua, Wenyue, et al.
Pubblicazione: (2024)
Toward Equitable Access: Leveraging Crowdsourced Reviews to Investigate Public Perceptions of Health Resource Accessibility
di: Xue, Zhaoqian, et al.
Pubblicazione: (2025)
di: Xue, Zhaoqian, et al.
Pubblicazione: (2025)
A scoping review of using Large Language Models (LLMs) to investigate Electronic Health Records (EHRs)
di: Li, Lingyao, et al.
Pubblicazione: (2024)
di: Li, Lingyao, et al.
Pubblicazione: (2024)
Know the Ropes: A Heuristic Strategy for LLM-based Multi-Agent System Design
di: Li, Zhenkun, et al.
Pubblicazione: (2025)
di: Li, Zhenkun, et al.
Pubblicazione: (2025)
ADO: Automatic Data Optimization for Inputs in LLM Prompts
di: Lin, Sam, et al.
Pubblicazione: (2025)
di: Lin, Sam, et al.
Pubblicazione: (2025)
Health-LLM: Personalized Retrieval-Augmented Disease Prediction System
di: Yu, Qinkai, et al.
Pubblicazione: (2024)
di: Yu, Qinkai, et al.
Pubblicazione: (2024)
Characterizing Online Toxicity During the 2022 Mpox Outbreak: A Computational Analysis of Topical and Network Dynamics
di: Fan, Lizhou, et al.
Pubblicazione: (2024)
di: Fan, Lizhou, et al.
Pubblicazione: (2024)
Invisible Prompts, Visible Threats: Malicious Font Injection in External Resources for Large Language Models
di: Xiong, Junjie, et al.
Pubblicazione: (2025)
di: Xiong, Junjie, et al.
Pubblicazione: (2025)
AIOS: LLM Agent Operating System
di: Mei, Kai, et al.
Pubblicazione: (2024)
di: Mei, Kai, et al.
Pubblicazione: (2024)
AutoFlow: Automated Workflow Generation for Large Language Model Agents
di: Li, Zelong, et al.
Pubblicazione: (2024)
di: Li, Zelong, et al.
Pubblicazione: (2024)
IDGenRec: LLM-RecSys Alignment with Textual ID Learning
di: Tan, Juntao, et al.
Pubblicazione: (2024)
di: Tan, Juntao, et al.
Pubblicazione: (2024)
MoralBench: Moral Evaluation of LLMs
di: Ji, Jianchao, et al.
Pubblicazione: (2024)
di: Ji, Jianchao, et al.
Pubblicazione: (2024)
Cache Mechanism for Agent RAG Systems
di: Lin, Shuhang, et al.
Pubblicazione: (2025)
di: Lin, Shuhang, et al.
Pubblicazione: (2025)
"HOT" ChatGPT: The promise of ChatGPT in detecting and discriminating hateful, offensive, and toxic comments on social media
di: Li, Lingyao, et al.
Pubblicazione: (2023)
di: Li, Lingyao, et al.
Pubblicazione: (2023)
Goal-guided Generative Prompt Injection Attack on Large Language Models
di: Zhang, Chong, et al.
Pubblicazione: (2024)
di: Zhang, Chong, et al.
Pubblicazione: (2024)
Evaluating the Formal Reasoning Capabilities of Large Language Models through Chomsky Hierarchy
di: Dong, Yihong, et al.
Pubblicazione: (2026)
di: Dong, Yihong, et al.
Pubblicazione: (2026)
ReaGAN: Node-as-Agent-Reasoning Graph Agentic Network
di: Guo, Minghao, et al.
Pubblicazione: (2025)
di: Guo, Minghao, et al.
Pubblicazione: (2025)
Counterfactual Explainable Incremental Prompt Attack Analysis on Large Language Models
di: Shu, Dong, et al.
Pubblicazione: (2024)
di: Shu, Dong, et al.
Pubblicazione: (2024)
RAGRouter-Bench: A Dataset and Benchmark for Adaptive RAG Routing
di: Wang, Ziqi, et al.
Pubblicazione: (2026)
di: Wang, Ziqi, et al.
Pubblicazione: (2026)
Time Series Forecasting with LLMs: Understanding and Enhancing Model Capabilities
di: Tang, Hua, et al.
Pubblicazione: (2024)
di: Tang, Hua, et al.
Pubblicazione: (2024)
OpenP5: An Open-Source Platform for Developing, Training, and Evaluating LLM-based Recommender Systems
di: Xu, Shuyuan, et al.
Pubblicazione: (2023)
di: Xu, Shuyuan, et al.
Pubblicazione: (2023)
Knowledge Graph Large Language Model (KG-LLM) for Link Prediction
di: Shu, Dong, et al.
Pubblicazione: (2024)
di: Shu, Dong, et al.
Pubblicazione: (2024)
Socrates or Smartypants: Testing Logic Reasoning Capabilities of Large Language Models with Logic Programming-based Test Oracles
di: Xu, Zihao, et al.
Pubblicazione: (2025)
di: Xu, Zihao, et al.
Pubblicazione: (2025)
Exploring Concept Depth: How Large Language Models Acquire Knowledge and Concept at Different Layers?
di: Jin, Mingyu, et al.
Pubblicazione: (2024)
di: Jin, Mingyu, et al.
Pubblicazione: (2024)
Can LLM Agents Really Debate? A Controlled Study of Multi-Agent Debate in Logical Reasoning
di: Wu, Haolun, et al.
Pubblicazione: (2025)
di: Wu, Haolun, et al.
Pubblicazione: (2025)
Logic-of-Thought: Injecting Logic into Contexts for Full Reasoning in Large Language Models
di: Liu, Tongxuan, et al.
Pubblicazione: (2024)
di: Liu, Tongxuan, et al.
Pubblicazione: (2024)
AIOS Compiler: LLM as Interpreter for Natural Language Programming and Flow Programming of AI Agents
di: Xu, Shuyuan, et al.
Pubblicazione: (2024)
di: Xu, Shuyuan, et al.
Pubblicazione: (2024)
Documenti analoghi
-
NPHardEval: Dynamic Benchmark on Reasoning Ability of Large Language Models via Complexity Classes
di: Fan, Lizhou, et al.
Pubblicazione: (2023) -
BattleAgent: Multi-modal Dynamic Emulation on Historical Battles to Complement Historical Analysis
di: Lin, Shuhang, et al.
Pubblicazione: (2024) -
NPHardEval4V: Dynamic Evaluation of Large Vision-Language Models with Effects of Vision
di: Li, Xiang, et al.
Pubblicazione: (2024) -
Disentangling Memory and Reasoning Ability in Large Language Models
di: Jin, Mingyu, et al.
Pubblicazione: (2024) -
EmojiPrompt: Generative Prompt Obfuscation for Privacy-Preserving Communication with Cloud-based LLMs
di: Lin, Sam, et al.
Pubblicazione: (2024)