EduBot -- Can LLMs Solve Personalized Learning and Programming Assignments?
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Yibin, Xie, Jiaxi, Subramanian, Lakshminarayanan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Can LLMs Enable Verification in Mainstream Programming?
von: Shefer, Aleksandr, et al.
Veröffentlicht: (2025)
von: Shefer, Aleksandr, et al.
Veröffentlicht: (2025)
Prompt Stability in Code LLMs: Measuring Sensitivity across Emotion- and Personality-Driven Variations
von: Ma, Wei, et al.
Veröffentlicht: (2025)
von: Ma, Wei, et al.
Veröffentlicht: (2025)
DroidBot-GPT: GPT-powered UI Automation for Android
von: Wen, Hao, et al.
Veröffentlicht: (2023)
von: Wen, Hao, et al.
Veröffentlicht: (2023)
The Last Dependency Crusade: Solving Python Dependency Conflicts with LLMs
von: Bartlett, Antony, et al.
Veröffentlicht: (2025)
von: Bartlett, Antony, et al.
Veröffentlicht: (2025)
ProgramBench: Can Language Models Rebuild Programs From Scratch?
von: Yang, John, et al.
Veröffentlicht: (2026)
von: Yang, John, et al.
Veröffentlicht: (2026)
MutaBot: A Mutation Testing Approach for Chatbots
von: Urrico, Michael Ferdinando, et al.
Veröffentlicht: (2024)
von: Urrico, Michael Ferdinando, et al.
Veröffentlicht: (2024)
LeafTutor: An AI Agent for Programming Assignment Tutoring
von: Bochard, Madison, et al.
Veröffentlicht: (2025)
von: Bochard, Madison, et al.
Veröffentlicht: (2025)
Can LLMs Replace Humans During Code Chunking?
von: Glasz, Christopher, et al.
Veröffentlicht: (2025)
von: Glasz, Christopher, et al.
Veröffentlicht: (2025)
Can LLMs Generate User Stories and Assess Their Quality?
von: Quattrocchi, Giovanni, et al.
Veröffentlicht: (2025)
von: Quattrocchi, Giovanni, et al.
Veröffentlicht: (2025)
Understanding the Limits of Automated Evaluation for Code Review Bots in Practice
von: Karakaya, Veli, et al.
Veröffentlicht: (2026)
von: Karakaya, Veli, et al.
Veröffentlicht: (2026)
Can LLMs Reason About Program Semantics? A Comprehensive Evaluation of LLMs on Formal Specification Inference
von: Le-Cong, Thanh, et al.
Veröffentlicht: (2025)
von: Le-Cong, Thanh, et al.
Veröffentlicht: (2025)
Evaluating the Generalizability of LLMs in Automated Program Repair
von: Li, Fengjie, et al.
Veröffentlicht: (2025)
von: Li, Fengjie, et al.
Veröffentlicht: (2025)
Past, Present and Future: Exploring Adaptive AI in Software Development Bots
von: Elsisi, Omar, et al.
Veröffentlicht: (2025)
von: Elsisi, Omar, et al.
Veröffentlicht: (2025)
Can LLMs Reason Like Automated Theorem Provers for Rust Verification? VCoT-Bench: Evaluating via Verification Chain of Thought
von: Xie, Zichen, et al.
Veröffentlicht: (2026)
von: Xie, Zichen, et al.
Veröffentlicht: (2026)
A Study of LLMs' Preferences for Libraries and Programming Languages
von: Twist, Lukas, et al.
Veröffentlicht: (2025)
von: Twist, Lukas, et al.
Veröffentlicht: (2025)
RuleFlow : Generating Reusable Program Optimizations with LLMs
von: Singh, Avaljot, et al.
Veröffentlicht: (2026)
von: Singh, Avaljot, et al.
Veröffentlicht: (2026)
MAS-Algorithm: A Workflow for Solving Algorithmic Programming Problems with a Multi-Agent System
von: Xu, Yuliang, et al.
Veröffentlicht: (2026)
von: Xu, Yuliang, et al.
Veröffentlicht: (2026)
Can LLMs Replace Human Evaluators? An Empirical Study of LLM-as-a-Judge in Software Engineering
von: Wang, Ruiqi, et al.
Veröffentlicht: (2025)
von: Wang, Ruiqi, et al.
Veröffentlicht: (2025)
Empowering AI to Generate Better AI Code: Guided Generation of Deep Learning Projects with LLMs
von: Xie, Chen, et al.
Veröffentlicht: (2025)
von: Xie, Chen, et al.
Veröffentlicht: (2025)
Can LLMs Generate Reliable Test Case Generators? A Study on Competition-Level Programming Problems
von: Cao, Yuhan, et al.
Veröffentlicht: (2025)
von: Cao, Yuhan, et al.
Veröffentlicht: (2025)
Accuracy, Stability, and Repeated-Run Reliability of Large Language Models on Deterministic Programming Tasks
von: Zhou, Yongxi, et al.
Veröffentlicht: (2026)
von: Zhou, Yongxi, et al.
Veröffentlicht: (2026)
The Hitchhiker's Guide to Program Analysis, Part II: Deep Thoughts by LLMs
von: Li, Haonan, et al.
Veröffentlicht: (2025)
von: Li, Haonan, et al.
Veröffentlicht: (2025)
Logic Error Localization in Student Programming Assignments Using Pseudocode and Graph Neural Networks
von: Xu, Zhenyu, et al.
Veröffentlicht: (2024)
von: Xu, Zhenyu, et al.
Veröffentlicht: (2024)
Assessing Large Language Models for Automated Feedback Generation in Learning Programming Problem Solving
von: Silva, Priscylla, et al.
Veröffentlicht: (2025)
von: Silva, Priscylla, et al.
Veröffentlicht: (2025)
Terminus-4B: Can a Smaller Model Replace Frontier LLMs at Agentic Execution Tasks?
von: Garg, Spandan, et al.
Veröffentlicht: (2026)
von: Garg, Spandan, et al.
Veröffentlicht: (2026)
Can GPT-O1 Kill All Bugs? An Evaluation of GPT-Family LLMs on QuixBugs
von: Hu, Haichuan, et al.
Veröffentlicht: (2024)
von: Hu, Haichuan, et al.
Veröffentlicht: (2024)
RAG-Verus: Repository-Level Program Verification with LLMs using Retrieval Augmented Generation
von: Zhong, Sicheng, et al.
Veröffentlicht: (2025)
von: Zhong, Sicheng, et al.
Veröffentlicht: (2025)
Utilizing LLMs for Industrial Process Automation: A Case Study on Modifying RAPID Programs
von: Fares, Salim, et al.
Veröffentlicht: (2025)
von: Fares, Salim, et al.
Veröffentlicht: (2025)
LIA: Supervised Fine-Tuning of Large Language Models for Automatic Issue Assignment
von: Khosravani, Arsham, et al.
Veröffentlicht: (2026)
von: Khosravani, Arsham, et al.
Veröffentlicht: (2026)
Learning to Debug: LLM-Organized Knowledge Trees for Solving RTL Assertion Failures
von: Bai, Yunsheng, et al.
Veröffentlicht: (2025)
von: Bai, Yunsheng, et al.
Veröffentlicht: (2025)
CoRe: Benchmarking LLMs Code Reasoning Capabilities through Static Analysis Tasks
von: Xie, Danning, et al.
Veröffentlicht: (2025)
von: Xie, Danning, et al.
Veröffentlicht: (2025)
Programming with Data: Test-Driven Data Engineering for Self-Improving LLMs from Raw Corpora
von: Pan, Chenkai, et al.
Veröffentlicht: (2026)
von: Pan, Chenkai, et al.
Veröffentlicht: (2026)
Peer-aided Repairer: Empowering Large Language Models to Repair Advanced Student Assignments
von: Zhao, Qianhui, et al.
Veröffentlicht: (2024)
von: Zhao, Qianhui, et al.
Veröffentlicht: (2024)
Can We Enhance Bug Report Quality Using LLMs?: An Empirical Study of LLM-Based Bug Report Generation
von: Acharya, Jagrit, et al.
Veröffentlicht: (2025)
von: Acharya, Jagrit, et al.
Veröffentlicht: (2025)
MergeRepair: An Exploratory Study on Merging Task-Specific Adapters in Code LLMs for Automated Program Repair
von: Dehghan, Meghdad, et al.
Veröffentlicht: (2024)
von: Dehghan, Meghdad, et al.
Veröffentlicht: (2024)
R-Log: Incentivizing Log Analysis Capability in LLMs via Reasoning-based Reinforcement Learning
von: Liu, Yilun, et al.
Veröffentlicht: (2025)
von: Liu, Yilun, et al.
Veröffentlicht: (2025)
Teaching LLMs to Learn Tool Trialing and Execution through Environment Interaction
von: Gao, Xingjie, et al.
Veröffentlicht: (2026)
von: Gao, Xingjie, et al.
Veröffentlicht: (2026)
Energy-Aware Code Generation with LLMs: Benchmarking Small vs. Large Language Models for Sustainable AI Programming
von: Ashraf, Humza, et al.
Veröffentlicht: (2025)
von: Ashraf, Humza, et al.
Veröffentlicht: (2025)
Self-Bootstrapping Automated Program Repair: Using LLMs to Generate and Evaluate Synthetic Training Data for Bug Repair
von: de-Fitero-Dominguez, David, et al.
Veröffentlicht: (2025)
von: de-Fitero-Dominguez, David, et al.
Veröffentlicht: (2025)
InvAASTCluster: On Applying Invariant-Based Program Clustering to Introductory Programming Assignments
von: Orvalho, Pedro, et al.
Veröffentlicht: (2022)
von: Orvalho, Pedro, et al.
Veröffentlicht: (2022)
Ähnliche Einträge
-
Can LLMs Enable Verification in Mainstream Programming?
von: Shefer, Aleksandr, et al.
Veröffentlicht: (2025) -
Prompt Stability in Code LLMs: Measuring Sensitivity across Emotion- and Personality-Driven Variations
von: Ma, Wei, et al.
Veröffentlicht: (2025) -
DroidBot-GPT: GPT-powered UI Automation for Android
von: Wen, Hao, et al.
Veröffentlicht: (2023) -
The Last Dependency Crusade: Solving Python Dependency Conflicts with LLMs
von: Bartlett, Antony, et al.
Veröffentlicht: (2025) -
ProgramBench: Can Language Models Rebuild Programs From Scratch?
von: Yang, John, et al.
Veröffentlicht: (2026)