PaperScope
LIVE · 2026-09-17 05:40 UTC

Finding Common Mistakes In Modelling With Mathematical Formalisms Using LLMs

Lilian Killich, Marko Schmellenkamp, Fabian Vehlken, Thomas Zeume

Latestcs.CLcs.LGcs.AIcs.CV
arXiv ID
2609.17111 v1
Category
Submitted
2026-09-15

Abstract

Modelling with mathematical formalisms like logical formulas, mathematical equations, or regular expressions is an important yet challenging task for students of computer science and other STEM disciplines. Identifying common mistakes occurring in this context is an important step towards helping struggling students by providing targeted high-quality feedback, e.g. in interactive learning systems. We present a tool-supported workflow that allows to (1) identify candidates for common mistakes that explain many student mistakes in large educational data sets, (2) cluster candidates according to similarities, and (3) visualize resulting clusters for instructors and CS education researchers. The visualization is designed to help researchers to identify common modelling mistakes. The candidates for common mistakes are represented by bug fixing transformations that translate incorrect formalizations into correct formalizations; they are generated by an LLM and validated algorithmically. We show that this approach works well by reproducing common mistakes in propositional logic modelling that were identified by hand in the literature; showing that, unlike other algorithmic approaches, the LLM-based approach is suitable for very large sets of data; and applying it to multiple other formalisms to showcase it generalizes beyond propositional logic.

arXiv abs page · PDF · same-day batch