This repository contains the evaluation of the experiment "Examination of Code generated by Large Language Models", conducted in September 2023. A paper based on this study will be published soon.
Results from conducted tests and scripts used to evaluate this data are included here, alongside all study outcomes as images. The official source code for the study is available in a separate repository.
As the experiment is based on the bachelor thesis "Code Correctness and Quality in the Era of AI Code Generation" by Emilia Hannson & Oliwer Ellréus from 2023, the evaluation is partly similar to its repository.
- Robin Beer (@BoLer1)
- Alexander Feix (@alexanderfeix)
- Tim Guttzeit (@tguttzeit)
- Vincent Müller (@vincent-mueller)
- Tamara Muras (@t-muras)
- Maurice Rauscher (@mrm021)
- Florian Schäffler (@schaefflerf)