PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
February 8, 2026Information1 citationsOpen Access

Evaluating ChatGPT’s Cognitive Performance in Chemical Engineering Education

View Full Paper
SSSalman ShahidUniversity of ManchesterSWShaun WalmsleyUniversity of Manchester

Key Points

  • The study aims to assess ChatGPT's cognitive performance in solving Chemical Engineering problems based on Bloom's Taxonomy.
  • Evaluated a diverse dataset of undergraduate-level Chemical Engineering problems.
  • Mapped problems to six cognitive domains of Bloom's Taxonomy.
  • Categorized responses into five types of errors based on accuracy.
  • Significant differences in ChatGPT's performance across cognitive levels were observed.
  • Strong performance at lower cognitive levels (Remember–Apply).
  • Substantial decline in performance at higher levels (Analyze, Evaluate, Create).

Abstract

Large Language Models (LLMs) now occupy a prominent role in science, engineering, and higher education. Their capacity to generate step-wise solutions, conceptual explanations, and problem-solving pathways creates new opportunities—but also new risks—for Chemical Engineering learners. Despite widespread informal use, few empirical studies have evaluated LLM performance using a systematically designed dataset mapped directly to Bloom’s Taxonomy. This study evaluates the competency of ChatGPT in solving Chemical Engineering problems mapped to Bloom’s Taxonomy. A diverse dataset of undergraduate-level problems spanning six cognitive domains was used to assess the model’s reasoning across ascending levels of cognitive complexity. Each response was evaluated for accuracy and categorized into five error types. Although ChatGPT demonstrated considerable potential across a range of topics, the analysis also revealed important challenges and limitations that inform best practices for integrating LLMs into Chemical Engineering education. Results show significant differences in ChatGPT performance across Bloom levels, revealing three distinct tiers of capability. Strong performance was observed at lower cognitive levels (Remember–Apply), while substantial degradation occurred at Analyze, Evaluate, and specially Create. The findings provide a nuanced, empirically grounded understanding of current LLM capability limits, with practical recommendations for educators integrating LLMs into engineering curricula.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Shahid et al. (2026) studied this question.

synapsesocial.com/papers/698828410fc35cd7a8847a1bhttps://doi.org/10.3390/info17020162
Ask AI
Helpful
Bookmark
Share
View Full Paper