Högskolan i Skövde

his.sePublications
Change search
CiteExportLink to record
Permanent link

Direct link
Cite
Citation style
  • apa
  • apa-cv
  • ieee
  • modern-language-association-8th-edition
  • vancouver
  • Other style
More styles
Language
  • de-DE
  • en-GB
  • en-US
  • fi-FI
  • nn-NO
  • nn-NB
  • sv-SE
  • Other locale
More languages
Output format
  • html
  • text
  • asciidoc
  • rtf
Understanding Learner Confusion Across Languages with Explainable AI
University of Skövde, School of Informatics.
2025 (English)Independent thesis Advanced level (degree of Master (One Year)), 10 credits / 15 HE creditsStudent thesis
Abstract [en]

This thesis investigates automated confusion detection in student discussion posts, focusing on both English and German data. Building on the pretrained EduDistilBERT model, its performance is evaluated in an exploratory cross-lingual application: German posts are translated into English through a machine translation pipeline and then classified using the English-trained model. To address the opacity of predictions, LIME is employed to analyze feature attributions and gain insights into decision-making. Results show that the model achieves high recall in both languages, with systematic differences in precision. In English, confusion is frequently overpredicted, while in German, exploratory analysis suggests that polite or formal expressions are often misclassified as confusion, and subtle indicators are overlooked. LIME analyses reveal language-specific lexical cues, such as pronouns, task-related words, and politeness markers, as key drivers of predictions. The study highlights the feasibility of applying educational NLP models across languages via translation, while emphasizing that findings on German data remain exploratory due to the smaller dataset size.

Place, publisher, year, edition, pages
2025. , p. 30
Keywords [en]
Confusion Detection, Online Learning, Explainable AI (XAI), LIME, Educational NLP, Cross-lingual Application, Machine Translation
National Category
Natural Language Processing
Identifiers
URN: urn:nbn:se:his:diva-25890OAI: oai:DiVA.org:his-25890DiVA, id: diva2:2003166
Subject / course
Informationsteknologi
Educational program
Data Science - Master’s Programme
Supervisors
Examiners
Available from: 2025-10-03 Created: 2025-10-03 Last updated: 2025-10-03Bibliographically approved

Open Access in DiVA

fulltext(586 kB)192 downloads
File information
File name FULLTEXT01.pdfFile size 586 kBChecksum SHA-512
53e63d8efee6e1db0f662fe11d4edf2a5786711ff5342b58e139580ad9e1219faa51b0395e7c8cd9943264745c2df773f69df75c1533e343e9a1dd466ccb0a1f
Type fulltextMimetype application/pdf

By organisation
School of Informatics
Natural Language Processing

Search outside of DiVA

GoogleGoogle Scholar
The number of downloads is the sum of all downloads of full texts. It may include eg previous versions that are now no longer available

urn-nbn

Altmetric score

urn-nbn
Total: 2894 hits
CiteExportLink to record
Permanent link

Direct link
Cite
Citation style
  • apa
  • apa-cv
  • ieee
  • modern-language-association-8th-edition
  • vancouver
  • Other style
More styles
Language
  • de-DE
  • en-GB
  • en-US
  • fi-FI
  • nn-NO
  • nn-NB
  • sv-SE
  • Other locale
More languages
Output format
  • html
  • text
  • asciidoc
  • rtf