Balancing AI responsibility with privacy, safety, and utility: Unlearning in large language models for mathematics education
Abstract
Online mathematics learning platforms are increasingly adopting large language models (LLMs) to provide scalable, on-demand support, but these models may reproduce private information from training data or generate harmful language. This raises concerns about responsibility in educational settings regarding the use of pre-trained models. LLM unlearning is an emerging area for reducing a model's ability to produce specific unwanted content and remains underexplored in educational research. This study aims to investigate how LLM unlearning reduces the model's reliance on personally identifiable information (PII) and inappropriate content in the math tutoring context, while maintaining the model's utility on both single-label and multi-label downstream math tasks. We applied a gradient-based LLM unlearning approach to three different models, which were pre-trained on approximately 3 million data points from an Algebra I online discussion forum between students and professional tutors. PII and harmful content were detected on this training data and used for unlearning in two different orders (PII and harmful content unlearning). Then, the generated outputs from these two unlearning models were compared with those of the pre-trained model in terms of PII-containing output rate and harmful rate. Moreover, unlearned models were evaluated on two different math classification tasks. The results showed that the rates of PII-containing output rate and harmfulness substantially decreased compared to the pre-trained models, and the utility of the unlearned model was still maintained. These findings demonstrate how LLM unlearning can be applied to pre-trained models to behave them more responsibly, while maintaining strong model performance on math-related tasks.
Identifier Metadata
| Identifier | 110.0975/CON.2026.00946 |
| Canonical | mdoi:110.0975/CON.2026.00946 |
| Resolver URL | https://mdoi.org/110.0975/CON.2026.00946 |
| Resource URL | Open resource |
| Document URL | Open document |
| Content Type | Article |
| Authors | Chenglu Li, Gokhan Gülfidan, Yinqi Zhang-Kopf |
| Year | 2026 |
| Depositor | Convergence Chronicles Organisation |
| Prefix | 110.0975 |
| Registered | Aug. 3, 2026 |
| Updated | Aug. 3, 2026 |
| Status | Active |
| Visibility | Public |
Cite This Identifier
APA 7th Edition
Click to copy
MLA 9th Edition
Click to copy
Chicago 17th Edition
Click to copy
BibTeX
Click to copy
Persistent Identifier
mdoi:110.0975/CON.2026.00946Click to copy