LLM for Complex Reasoning Task: An Exploratory Study in Fermi Problems
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Liu, Zishuo, Villarreal, Carlos Rabat, Rahgouy, Mostafa, Das, Amit, Zhang, Zheng, Ren, Chang, Feng, Dongji |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
LLM for Comparative Narrative Analysis
von: Kampen, Leo, et al.
Veröffentlicht: (2025)
von: Kampen, Leo, et al.
Veröffentlicht: (2025)
OffensiveLang: A Community Based Implicit Offensive Language Dataset
von: Das, Amit, et al.
Veröffentlicht: (2024)
von: Das, Amit, et al.
Veröffentlicht: (2024)
Prompting a Weighting Mechanism into LLM-as-a-Judge in Two-Step: A Case Study
von: Xie, Wenwen, et al.
Veröffentlicht: (2025)
von: Xie, Wenwen, et al.
Veröffentlicht: (2025)
Investigating Annotator Bias in Large Language Models for Hate Speech Detection
von: Das, Amit, et al.
Veröffentlicht: (2024)
von: Das, Amit, et al.
Veröffentlicht: (2024)
Homa at SemEval-2025 Task 5: Aligning Librarian Records with OntoAligner for Subject Tagging
von: Tekanlou, Hadi Bayrami Asl, et al.
Veröffentlicht: (2025)
von: Tekanlou, Hadi Bayrami Asl, et al.
Veröffentlicht: (2025)
Textualized Agent-Style Reasoning for Complex Tasks by Multiple Round LLM Generation
von: Liang, Chen, et al.
Veröffentlicht: (2024)
von: Liang, Chen, et al.
Veröffentlicht: (2024)
Investigating Hallucination in Conversations for Low Resource Languages
von: Das, Amit, et al.
Veröffentlicht: (2025)
von: Das, Amit, et al.
Veröffentlicht: (2025)
CNMBERT: A Model for Converting Hanyu Pinyin Abbreviations to Chinese Characters
von: Feng, Zishuo, et al.
Veröffentlicht: (2024)
von: Feng, Zishuo, et al.
Veröffentlicht: (2024)
Advances in LLM Reasoning Enable Flexibility in Clinical Problem-Solving
von: Shidara, Kie, et al.
Veröffentlicht: (2026)
von: Shidara, Kie, et al.
Veröffentlicht: (2026)
Towards Effective Authorship Attribution: Integrating Class-Incremental Learning
von: Rahgouy, Mostafa, et al.
Veröffentlicht: (2024)
von: Rahgouy, Mostafa, et al.
Veröffentlicht: (2024)
Reasoning Up the Instruction Ladder for Controllable Language Models
von: Zheng, Zishuo, et al.
Veröffentlicht: (2025)
von: Zheng, Zishuo, et al.
Veröffentlicht: (2025)
Enrich-on-Graph: Query-Graph Alignment for Complex Reasoning with LLM Enriching
von: Li, Songze, et al.
Veröffentlicht: (2025)
von: Li, Songze, et al.
Veröffentlicht: (2025)
Unbiased Reasoning for Knowledge-Intensive Tasks in Large Language Models via Conditional Front-Door Adjustment
von: Zhao, Bo, et al.
Veröffentlicht: (2025)
von: Zhao, Bo, et al.
Veröffentlicht: (2025)
Neural-Symbolic Collaborative Distillation: Advancing Small Language Models for Complex Reasoning Tasks
von: Liao, Huanxuan, et al.
Veröffentlicht: (2024)
von: Liao, Huanxuan, et al.
Veröffentlicht: (2024)
MMESGBench: Pioneering Multimodal Understanding and Complex Reasoning Benchmark for ESG Tasks
von: Zhang, Lei, et al.
Veröffentlicht: (2025)
von: Zhang, Lei, et al.
Veröffentlicht: (2025)
Reasoning Depth and Environment Complexity: A Controlled Study of RLVR Data Allocation across Logical Reasoning Tasks
von: Zhu, Yihua, et al.
Veröffentlicht: (2026)
von: Zhu, Yihua, et al.
Veröffentlicht: (2026)
The Behavior Gap: Evaluating Zero-shot LLM Agents in Complex Task-Oriented Dialogs
von: Baidya, Avinash, et al.
Veröffentlicht: (2025)
von: Baidya, Avinash, et al.
Veröffentlicht: (2025)
Towards Multi-Agent Reasoning Systems for Collaborative Expertise Delegation: An Exploratory Design Study
von: Xu, Baixuan, et al.
Veröffentlicht: (2025)
von: Xu, Baixuan, et al.
Veröffentlicht: (2025)
Multimodal Sentiment Analysis Based on Causal Reasoning
von: Chen, Fuhai, et al.
Veröffentlicht: (2024)
von: Chen, Fuhai, et al.
Veröffentlicht: (2024)
Case Study: Testing Model Capabilities in Some Reasoning Tasks
von: Zhang, Min, et al.
Veröffentlicht: (2024)
von: Zhang, Min, et al.
Veröffentlicht: (2024)
Reasoning Topology Matters: Network-of-Thought for Complex Reasoning Tasks
von: Huang, Fan
Veröffentlicht: (2026)
von: Huang, Fan
Veröffentlicht: (2026)
Can LLM Reasoning Be Trusted? A Comparative Study: Using Human Benchmarking on Statistical Tasks
von: Nagarkar, Crish, et al.
Veröffentlicht: (2026)
von: Nagarkar, Crish, et al.
Veröffentlicht: (2026)
Teaching LLM to Reason: Reinforcement Learning from Algorithmic Problems without Code
von: Bao, Keqin, et al.
Veröffentlicht: (2025)
von: Bao, Keqin, et al.
Veröffentlicht: (2025)
RATT: A Thought Structure for Coherent and Correct LLM Reasoning
von: Zhang, Jinghan, et al.
Veröffentlicht: (2024)
von: Zhang, Jinghan, et al.
Veröffentlicht: (2024)
Modeling LLM Unlearning as an Asymmetric Two-Task Learning Problem
von: Xiao, Zeguan, et al.
Veröffentlicht: (2026)
von: Xiao, Zeguan, et al.
Veröffentlicht: (2026)
MCP-Bench: Benchmarking Tool-Using LLM Agents with Complex Real-World Tasks via MCP Servers
von: Wang, Zhenting, et al.
Veröffentlicht: (2025)
von: Wang, Zhenting, et al.
Veröffentlicht: (2025)
D-CORE: Incentivizing Task Decomposition in Large Reasoning Models for Complex Tool Use
von: Xu, Bowen, et al.
Veröffentlicht: (2026)
von: Xu, Bowen, et al.
Veröffentlicht: (2026)
Question-Analysis Prompting Improves LLM Performance in Reasoning Tasks
von: Yugeswardeenoo, Dharunish, et al.
Veröffentlicht: (2024)
von: Yugeswardeenoo, Dharunish, et al.
Veröffentlicht: (2024)
BAR: A Backward Reasoning based Agent for Complex Minecraft Tasks
von: Du, Weihong, et al.
Veröffentlicht: (2025)
von: Du, Weihong, et al.
Veröffentlicht: (2025)
XFinBench: Benchmarking LLMs in Complex Financial Problem Solving and Reasoning
von: Zhang, Zhihan, et al.
Veröffentlicht: (2025)
von: Zhang, Zhihan, et al.
Veröffentlicht: (2025)
CoMM: Collaborative Multi-Agent, Multi-Reasoning-Path Prompting for Complex Problem Solving
von: Chen, Pei, et al.
Veröffentlicht: (2024)
von: Chen, Pei, et al.
Veröffentlicht: (2024)
AI Can Be Cognitively Biased: An Exploratory Study on Threshold Priming in LLM-Based Batch Relevance Assessment
von: Chen, Nuo, et al.
Veröffentlicht: (2024)
von: Chen, Nuo, et al.
Veröffentlicht: (2024)
A Training-free LLM Framework with Interaction between Contextually Related Subtasks in Solving Complex Tasks
von: Liu, Hongjia, et al.
Veröffentlicht: (2025)
von: Liu, Hongjia, et al.
Veröffentlicht: (2025)
ACADREASON: Exploring the Limits of Reasoning Models with Academic Research Problems
von: Gui, Xin, et al.
Veröffentlicht: (2025)
von: Gui, Xin, et al.
Veröffentlicht: (2025)
Bayesian Optimization for Enhanced Language Models: Optimizing Acquisition Functions
von: Bao, Zishuo, et al.
Veröffentlicht: (2025)
von: Bao, Zishuo, et al.
Veröffentlicht: (2025)
InternBootcamp Technical Report: Boosting LLM Reasoning with Verifiable Task Scaling
von: Li, Peiji, et al.
Veröffentlicht: (2025)
von: Li, Peiji, et al.
Veröffentlicht: (2025)
Augmenting LLM Reasoning with Dynamic Notes Writing for Complex QA
von: Maheshwary, Rishabh, et al.
Veröffentlicht: (2025)
von: Maheshwary, Rishabh, et al.
Veröffentlicht: (2025)
Lemmatization as a Classification Task: Results from Arabic across Multiple Genres
von: Saeed, Mostafa, et al.
Veröffentlicht: (2025)
von: Saeed, Mostafa, et al.
Veröffentlicht: (2025)
Navigating the Unknown: A Chat-Based Collaborative Interface for Personalized Exploratory Tasks
von: Peng, Yingzhe, et al.
Veröffentlicht: (2024)
von: Peng, Yingzhe, et al.
Veröffentlicht: (2024)
Multi-Task GRPO: Reliable LLM Reasoning Across Tasks
von: Ramesh, Shyam Sundhar, et al.
Veröffentlicht: (2026)
von: Ramesh, Shyam Sundhar, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
LLM for Comparative Narrative Analysis
von: Kampen, Leo, et al.
Veröffentlicht: (2025) -
OffensiveLang: A Community Based Implicit Offensive Language Dataset
von: Das, Amit, et al.
Veröffentlicht: (2024) -
Prompting a Weighting Mechanism into LLM-as-a-Judge in Two-Step: A Case Study
von: Xie, Wenwen, et al.
Veröffentlicht: (2025) -
Investigating Annotator Bias in Large Language Models for Hate Speech Detection
von: Das, Amit, et al.
Veröffentlicht: (2024) -
Homa at SemEval-2025 Task 5: Aligning Librarian Records with OntoAligner for Subject Tagging
von: Tekanlou, Hadi Bayrami Asl, et al.
Veröffentlicht: (2025)