Publicação em atas de evento científico
BERT-Based Classification of Cross-Linguistic MCQs According to the Bloom Taxonomy
João Botas (Botas, J.); Ana Rita Peixoto (Peixoto, A.); Eugénio Ribeiro (Ribeiro, E.);
15th Symposium on Languages, Applications and Technologies (SLATE 2026)
Ano (publicação definitiva)
2026
Língua
Inglês
País
--
Mais Informação
Web of Science®

Esta publicação não está indexada na Web of Science®

Scopus

Esta publicação não está indexada na Scopus

Google Scholar

N.º de citações: 0

(Última verificação: 2026-08-20 15:47)

Ver o registo no Google Scholar

Esta publicação não está indexada no Overton

Abstract/Resumo
The Bloom Taxonomy provides a useful framework for describing the cognitive demands of Multiple Choice Questions (MCQs), yet automatically assigning Bloom levels remains challenging due to category overlap and the limited evidence of higher-order thinking in MCQs. This study aims to evaluate transformer-based models, including BERT and ModernBERT variants, to enhance their performance on a four-level Bloom classification task across multiple scenarios, while also conducting model interpretability and error analysis. Across 40 settings tested, BERT models consistently outperform a keyword-based baseline, highlighting the contextual and semantic representations for this task. Performance is broadly similar between BERT and ModernBERT, with only minor differences across architectures, while the inclusion of answer options yields only small gains, suggesting that the question stem alone contains most of the discriminative signal. Besides, cross-linguistic experiments show comparable results between the original English data and Portuguese, although translated data exhibits a slight performance degradation, likely due to direct translation noise and subtle syntactic shifts. Errors are concentrated between adjacent Bloom levels, and models perform best on the Remembering level. In contrast, the Applying level remains the most difficult class to predict, largely because of class imbalance and conceptual overlap. The findings indicate that our approach constitutes a stable pipeline for Bloom classification, but its performance remains constrained by the intrinsic ambiguity of the taxonomy and the structural characteristics of the available data.
Agradecimentos/Acknowledgements
--
Palavras-chave
  • Ciências da Computação e da Informação - Ciências Naturais