PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
October 3, 2025Digital2 citationsOpen Access

Using LLM to Identify Pillars of the Mind Within Physics Learning Materials

View Full Paper
DČDaša ČerveňováPDPeter Demkanin

Key Points

  • Using large language models can help identify cognitive pillars in physics learning materials, enhancing educational frameworks.
  • The case study analyzed eight pages of materials aimed at 12- to 14-year-olds, comparing AI results with manual assessments.
  • MAXQDA AI Assist achieved the highest precision of 1.00, while both OpenAI models exhibited significant identification errors.
  • Future applications of this method may extend to analyzing students' written work and video activities for deeper educational insights.

Abstract

Artificial intelligence tools are quickly being applied in many areas of science, including learning sciences. Learning requires various types of thinking, sustained by distinct sets of neural networks in the brain. Labelling these systems gives us tools to manage them. This paper presents a pilot application of Large Language Models (LLMs) to physics textbook analysis, grounded in a well-developed neural network theory known as the Five Pillars of the Mind. The domain-specific networks, innate sense, and the five pillars provide a framework with which to examine how physics is learnt. For example, one can identify which pillars are active when discussing a physics concept. Identifying which pillars belong to which physics concept may be significantly influenced by the bias of the author and could be too time-consuming for longer, more complex texts involving physics concepts. Therefore, using LLMs to identify pillars could enhance the application of this framework to physics education. This article presents a case study in which we used selected Large Language Models to identify pillars within eight pages of learning material concerning forces aimed at 12- to 14-year-old pupils. We used GPT-4o and o4-mini, as well as MAXQDA AI Assist. Results from these models were compared with the authors’ manual analysis. Precision, recall, and F1-Score were used to evaluate the results quantitatively. MAXQDA AI Assist obtained the best results with 1.00 precision, 0.67 recall, and an F1-Score of 0.80. Both products by OpenAI hallucinated and falsely identified several concepts, resulting in low precision and, consequently, low F1-Score. As predicted, ChatGPT o4-mini scored twice as high as ChatGPT 4o. The method proved to be promising, and its future development has the potential to provide research teams with analysis not only of written learning material, but also of pupils’ written work and their video-recorded activities.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Červeňová et al. (2025) studied this question.

synapsesocial.com/papers/68e040f7a99c246f578b3cd1https://doi.org/10.3390/digital5040047
Ask AI
Helpful
Bookmark
Share
View Full Paper

Also Consider

Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context:

  1. 1The Vision of University Students from the Educational Field in the Integration of ChatGPT2024 · 12 citations
  2. 2Systematic Review of Artificial Intelligence in Education: Trends, Benefits, and Challenges2025 · 204 citations
  3. 3Empowering student self‐regulated learning and science education through ChatGPT : A pioneering pilot study2024 · 323 citations
  4. 4What Is the Impact of ChatGPT on Education? A Rapid Review of the Literature2023 · 1,917 citations
  5. 5Artificial Intelligence in Education (AIEd): Publication Patterns, Keywords, and Research Focuses2025 · 15 citations