Evaluating Large Language Models on Solved and Unsolved Problems in Graph Theory: Implications for Computing Education
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kulkarni, Adithya, Chakraborty, Mohna, Bagga, Jay |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Is Large Language Model Performance on Reasoning Tasks Impacted by Different Ways Questions Are Asked?
von: Song, Seok Hwan, et al.
Veröffentlicht: (2025)
von: Song, Seok Hwan, et al.
Veröffentlicht: (2025)
Modeling Data Diversity for Joint Instance and Verbalizer Selection in Cold-Start Scenarios
von: Chakraborty, Mohna, et al.
Veröffentlicht: (2025)
von: Chakraborty, Mohna, et al.
Veröffentlicht: (2025)
UQ: Assessing Language Models on Unsolved Questions
von: Nie, Fan, et al.
Veröffentlicht: (2025)
von: Nie, Fan, et al.
Veröffentlicht: (2025)
GENUINE: Graph Enhanced Multi-level Uncertainty Estimation for Large Language Models
von: Wang, Tuo, et al.
Veröffentlicht: (2025)
von: Wang, Tuo, et al.
Veröffentlicht: (2025)
Graph-of-Thought: Utilizing Large Language Models to Solve Complex and Dynamic Business Problems
von: Li, Ye
Veröffentlicht: (2024)
von: Li, Ye
Veröffentlicht: (2024)
Graph of Thoughts: Solving Elaborate Problems with Large Language Models
von: Besta, Maciej, et al.
Veröffentlicht: (2023)
von: Besta, Maciej, et al.
Veröffentlicht: (2023)
Mathify: Evaluating Large Language Models on Mathematical Problem Solving Tasks
von: Anand, Avinash, et al.
Veröffentlicht: (2024)
von: Anand, Avinash, et al.
Veröffentlicht: (2024)
EngiBench: A Benchmark for Evaluating Large Language Models on Engineering Problem Solving
von: Zhou, Xiyuan, et al.
Veröffentlicht: (2025)
von: Zhou, Xiyuan, et al.
Veröffentlicht: (2025)
Can Language Models Solve Graph Problems in Natural Language?
von: Wang, Heng, et al.
Veröffentlicht: (2023)
von: Wang, Heng, et al.
Veröffentlicht: (2023)
Evaluating Large Language Models on the GMAT: Implications for the Future of Business Education
von: Ashrafimoghari, Vahid, et al.
Veröffentlicht: (2024)
von: Ashrafimoghari, Vahid, et al.
Veröffentlicht: (2024)
Reinforcement Learning Problem Solving with Large Language Models
von: Gholamian, Sina, et al.
Veröffentlicht: (2024)
von: Gholamian, Sina, et al.
Veröffentlicht: (2024)
Graph of States: Solving Abductive Tasks with Large Language Models
von: Luo, Yu, et al.
Veröffentlicht: (2026)
von: Luo, Yu, et al.
Veröffentlicht: (2026)
GraphArena: Evaluating and Exploring Large Language Models on Graph Computation
von: Tang, Jianheng, et al.
Veröffentlicht: (2024)
von: Tang, Jianheng, et al.
Veröffentlicht: (2024)
Evaluating GPT- and Reasoning-based Large Language Models on Physics Olympiad Problems: Surpassing Human Performance and Implications for Educational Assessment
von: Tschisgale, Paul, et al.
Veröffentlicht: (2025)
von: Tschisgale, Paul, et al.
Veröffentlicht: (2025)
The Effect of Sampling Temperature on Problem Solving in Large Language Models
von: Renze, Matthew, et al.
Veröffentlicht: (2024)
von: Renze, Matthew, et al.
Veröffentlicht: (2024)
Is Mathematical Problem-Solving Expertise in Large Language Models Associated with Assessment Performance?
von: Zhang, Liang, et al.
Veröffentlicht: (2026)
von: Zhang, Liang, et al.
Veröffentlicht: (2026)
Digital Epidemiology: Leveraging Social Media for Insight into Epilepsy and Mental Health
von: Dahiya, Liza, et al.
Veröffentlicht: (2024)
von: Dahiya, Liza, et al.
Veröffentlicht: (2024)
Inference Scaling Laws: An Empirical Analysis of Compute-Optimal Inference for Problem-Solving with Language Models
von: Wu, Yangzhen, et al.
Veröffentlicht: (2024)
von: Wu, Yangzhen, et al.
Veröffentlicht: (2024)
Multi-hop Question Answering over Knowledge Graphs using Large Language Models
von: Chakraborty, Abir
Veröffentlicht: (2024)
von: Chakraborty, Abir
Veröffentlicht: (2024)
SciBench: Evaluating College-Level Scientific Problem-Solving Abilities of Large Language Models
von: Wang, Xiaoxuan, et al.
Veröffentlicht: (2023)
von: Wang, Xiaoxuan, et al.
Veröffentlicht: (2023)
OptiMUS-0.3: Using Large Language Models to Model and Solve Optimization Problems at Scale
von: AhmadiTeshnizi, Ali, et al.
Veröffentlicht: (2024)
von: AhmadiTeshnizi, Ali, et al.
Veröffentlicht: (2024)
Learn to Relax with Large Language Models: Solving Constraint Optimization Problems via Bidirectional Coevolution
von: Liu, Beidan, et al.
Veröffentlicht: (2025)
von: Liu, Beidan, et al.
Veröffentlicht: (2025)
Eyeballing Combinatorial Problems: A Case Study of Using Multimodal Large Language Models to Solve Traveling Salesman Problems
von: Elhenawy, Mohammed, et al.
Veröffentlicht: (2024)
von: Elhenawy, Mohammed, et al.
Veröffentlicht: (2024)
LLM-ProS: Analyzing Large Language Models' Performance in Competitive Problem Solving
von: Hossain, Md Sifat, et al.
Veröffentlicht: (2025)
von: Hossain, Md Sifat, et al.
Veröffentlicht: (2025)
Exploring Solution Divergence and Its Effect on Large Language Model Problem Solving
von: Li, Hang, et al.
Veröffentlicht: (2025)
von: Li, Hang, et al.
Veröffentlicht: (2025)
Creative Problem Solving in Large Language and Vision Models -- What Would it Take?
von: Nair, Lakshmi, et al.
Veröffentlicht: (2024)
von: Nair, Lakshmi, et al.
Veröffentlicht: (2024)
The Benefits of a Concise Chain of Thought on Problem-Solving in Large Language Models
von: Renze, Matthew, et al.
Veröffentlicht: (2024)
von: Renze, Matthew, et al.
Veröffentlicht: (2024)
OR-Toolformer: Modeling and Solving Operations Research Problems with Tool Augmented Large Language Models
von: Zhang, Jianzhang, et al.
Veröffentlicht: (2025)
von: Zhang, Jianzhang, et al.
Veröffentlicht: (2025)
Problem-Solving in Language Model Networks
von: Regan, Ciaran, et al.
Veröffentlicht: (2024)
von: Regan, Ciaran, et al.
Veröffentlicht: (2024)
Knowledge Graph Modeling-Driven Large Language Model Operating System (LLM OS) for Task Automation in Process Engineering Problem-Solving
von: Srinivas, Sakhinana Sagar, et al.
Veröffentlicht: (2024)
von: Srinivas, Sakhinana Sagar, et al.
Veröffentlicht: (2024)
Computational Blueprints: Generating Isomorphic Mathematics Problems with Large Language Models
von: Kim, Jeong-Hoon, et al.
Veröffentlicht: (2025)
von: Kim, Jeong-Hoon, et al.
Veröffentlicht: (2025)
Unsolved Problems in Spectral Graph Theory
von: Liu, Lele, et al.
Veröffentlicht: (2023)
von: Liu, Lele, et al.
Veröffentlicht: (2023)
Boosting of Thoughts: Trial-and-Error Problem Solving with Large Language Models
von: Chen, Sijia, et al.
Veröffentlicht: (2024)
von: Chen, Sijia, et al.
Veröffentlicht: (2024)
Rethinking the Uncertainty: A Critical Review and Analysis in the Era of Large Language Models
von: Beigi, Mohammad, et al.
Veröffentlicht: (2024)
von: Beigi, Mohammad, et al.
Veröffentlicht: (2024)
Can Graph Descriptive Order Affect Solving Graph Problems with LLMs?
von: Ge, Yuyao, et al.
Veröffentlicht: (2024)
von: Ge, Yuyao, et al.
Veröffentlicht: (2024)
Token-Supervised Value Models for Enhancing Mathematical Problem-Solving Capabilities of Large Language Models
von: Lee, Jung Hyun, et al.
Veröffentlicht: (2024)
von: Lee, Jung Hyun, et al.
Veröffentlicht: (2024)
Limitations of Large Language Models in Clinical Problem-Solving Arising from Inflexible Reasoning
von: Kim, Jonathan, et al.
Veröffentlicht: (2025)
von: Kim, Jonathan, et al.
Veröffentlicht: (2025)
TeleMath: A Benchmark for Large Language Models in Telecom Mathematical Problem Solving
von: Colle, Vincenzo, et al.
Veröffentlicht: (2025)
von: Colle, Vincenzo, et al.
Veröffentlicht: (2025)
MathGLM-Vision: Solving Mathematical Problems with Multi-Modal Large Language Model
von: Yang, Zhen, et al.
Veröffentlicht: (2024)
von: Yang, Zhen, et al.
Veröffentlicht: (2024)
Evaluating Large Language Models for Detecting Antisemitism
von: Patel, Jay, et al.
Veröffentlicht: (2025)
von: Patel, Jay, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Is Large Language Model Performance on Reasoning Tasks Impacted by Different Ways Questions Are Asked?
von: Song, Seok Hwan, et al.
Veröffentlicht: (2025) -
Modeling Data Diversity for Joint Instance and Verbalizer Selection in Cold-Start Scenarios
von: Chakraborty, Mohna, et al.
Veröffentlicht: (2025) -
UQ: Assessing Language Models on Unsolved Questions
von: Nie, Fan, et al.
Veröffentlicht: (2025) -
GENUINE: Graph Enhanced Multi-level Uncertainty Estimation for Large Language Models
von: Wang, Tuo, et al.
Veröffentlicht: (2025) -
Graph-of-Thought: Utilizing Large Language Models to Solve Complex and Dynamic Business Problems
von: Li, Ye
Veröffentlicht: (2024)