arXiv · 2403.09744
Evaluating the Application of Large Language Models to Generate Feedback in Programming Education
Abstract
This study investigates the application of large language models, specifically GPT-4, to enhance programming education. The research outlines the design of a web application that uses GPT-4 to provide feedback on programming tasks, without giving away the solution. A web application for working on programming tasks was developed for the study and evaluated with 51 students over the course of one semester. The results show that most of the feedback generated by GPT-4 effectively addressed code errors. However, challenges with incorrect suggestions and hallucinated issues indicate the need for further improvements.
Explore related subjects
Keep this discovery
Sven Jacobs, Steffen Jaschke. 2024-03-13. Evaluating the Application of Large Language Models to Generate Feedback in Programming Education. https://doi.org/10.1109/educon60312.2024.10578838
Cite the original work for its findings. Save a collection to share your selection of sources.