Evidence map›Paper›PMID 42594116›Full record

ArticlePloS one2026

A comparative analysis of topic modelling techniques for the thematic analysis of student feedback.

Neha Kardam, Denise Wilson

Abstract readComparative Study
In one paragraph

Article in PloS one, 2026. The graph could read no effect estimate from its abstract, so it casts no vote on the map. Not yet cited in PubMed.

0numbers the graph read from it
0cells of the map it votes in
0citing papers in PubMed
–field-weighted citation impact
1 · What the graph read from it

What it found

Each row is one number read from the abstract, on the scale the paper reported it, with its interval. Left of the dashed line favours the treatment, right favours the comparator. Under each row is the sentence it came from. New to these charts? A ten-minute tutorial.

The abstract states no effect estimate the extractor could read, or names no intervention and outcome on the map, so this paper lights no cell and moves no belief. It is still indexed, cited and linked below.

2 · The registry

The trial behind it

Trials whose registry record cites this paper, or whose number appears in the abstract. A trial that started after this paper was published is citing it as background, not reporting it.

Neither the registry nor the abstract names a trial number. If this is a trial report, that itself is worth knowing.

3 · Its place in the literature

Who cites it

0 citing papers in PubMed.

No citing paper in PubMed yet.

4 · The record

Corrections and comments

PubMed lists nothing against this paper. Absence here is not a guarantee, only a check that was made.

5 · Who and what money

Authors and funding

2 authors.

Neha KardamElectrical and Computer Engineering, University of Washington, Seattle, Washington, United States of America.ORCID https://orcid.org/0009-0006-8767-8552
Denise WilsonElectrical and Computer Engineering, University of Washington, Seattle, Washington, United States of America.

Funding

No grant is acknowledged in the PubMed record.

6 · The paper itself

Abstract

This study seeks best practices for when and how to apply short text topic modelling (STTM) techniques using natural language processing to semi-structured data collected as part of education research in order to provide accurate guidance for interventions and avoid misguided improvements in education practice. Student feedback was collected using short answer questions that resulted in 1,667, 1,592, and 1,376 expectations for faculty support, teaching assistant (TA) support, and peer support respectively as part of a larger survey conducted via convenience sampling in over 40 engineering courses offered at a single large university between 2016 and 2023. After cleaning and preprocessing the data, short text responses were analyzed using five unsupervised topic models implemented in Python: traditional models Latent Dirichlet Allocation (LDA), Latent Semantic Analysis (LSA), Non-Negative Matrix Factorization (NMF), and k-means and one deep learning model (BERTopic). Model performance was evaluated using topic coherence and exse than that expected by chance (internal performance metrics. Addressing a methodological gap in prior comparative studies that rely predominantly on machine-led evaluation, two approaches to establishing ground truth were evaluated: (a) keywords from each topic model guided manual (human) coding of the data (a machine-led approach); and (b) themes in the data were extracted and coded independently by a domain expert (a human-led approach). NMF achieved the highest average performance in two of the three datasets, reaching 75.6% accuracy, 75.7% F1-Score, and 0.63 interrater reliability for the peer support dataset and 72.6% accuracy, 72.0% F1-Score, and 0.57 interrater reliability for the TA support dataset. The human-led approach yielded higher accuracy and F1-scores for faculty and peer support but failed for TA support when the topics extracted by topic models did not align with themes identified by a domain expert. These findings highlight the need for humans to be involved in the analysis of short text data in contexts like education research where high performance is necessary to achieve appropriate rigor. Domain expert intervention also enables strategic use of topic models to optimize their use in qualitative data analysis.

Indexed as

StudentsFeedbackHumansNatural Language ProcessingUniversities

Identifiers

PMID42594116
PMCPMC13472439

What Socratic holds

Textmetadata
Read underepoch 390

Registered trials

None linked

Read under generation 80e0d062 · epoch 390. Bibliography from PubMed, PubMed Central and OpenAlex; grants from NIH RePORTER; trial links from ClinicalTrials.gov; estimates, votes and beliefs from the Socratic graph.