PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
March 16, 2026Frontiers in Medicine2 citationsOpen Access

Improving TCM question answering through tree-organized self-reflective retrieval with LLMs

CLChang LiuYCYing ChangJLJu Li

Key Points

  • This research aims to evaluate a novel framework to improve LLM performance in TCM question answering through innovative knowledge organization.
  • Developed a hierarchical knowledge representation system using subject-predicate-object-text (SPO-T) units.
  • Implemented a tree-like architecture to capture multi-dimensional relationships in TCM knowledge.
  • Utilized iterative self-reflection for dynamic knowledge retrieval and validation across TCM resources.
  • Evaluated performance using questions from the TCM Medical Licensing Examination and Classics Course Exam.
  • Achieved a 19.85% improvement in accuracy on the TCM MLE benchmark.
  • Increased recall accuracy from 27% to 38% on CCE datasets.
  • Expert evaluations indicated significant improvements in safety, consistency, and explainability.
  • Demonstrated superior knowledge utilization and retrieval precision over standard methods.

Abstract

Background Large language models (LLMs) offer significant potential for intelligent question answering (QA) in healthcare, yet traditional knowledge representation methods fail to capture the complex, hierarchical nature of Traditional Chinese Medicine (TCM) knowledge systems. The lack of effective retrieval-augmented generation (RAG) frameworks specifically tailored for TCM’s unique epistemology limits applications. Objectives This study aims to evaluate the effectiveness of a novel Tree-Organized Self-Reflective Retrieval (TOSRR) framework in enhancing LLM performance on TCM QA tasks through innovative knowledge organization and dynamic self-correction mechanisms. Methods We developed a hierarchical knowledge representation system that structures TCM knowledge as subject-predicate-object-text (SPO-T) units within a tree-like architecture, enabling multi-dimensional relationships while preserving semantic context. Our iterative self-reflection mechanism implements dynamic knowledge retrieval and validation across textbook chapters and disciplines. Performance was evaluated using randomly selected questions from the TCM Medical Licensing Examination (MLE) and college Classics Course Exam (CCE), representing both standardized clinical knowledge and classical theory assessment. Results When integrated with GPT-4, the TOSRR framework demonstrated a 19.85% improvement in absolute accuracy on the TCM MLE benchmark and increased recall accuracy from 27 to 38% on CCE datasets. Expert manual evaluation revealed substantial enhancements across critical dimensions: safety, consistency, explainability, compliance, and coherence, with a comprehensive improvement of 18.64 points. Retrieval-Augmented Generation Assessment (RAGAs) metrics confirmed the framework’s superior knowledge utilization, retrieval precision, and resistance to information noise compared to standard RAG approaches. Conclusion The TOSRR framework enhances LLM performance in TCM knowledge tasks through its hierarchical knowledge representation and self-reflective retrieval approach. And the framework has potential for application in teaching.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Liu et al. (2026) studied this question.

synapsesocial.com/papers/69b79d538166e15b153aac02https://doi.org/10.3389/fmed.2026.1752778
Ask AI
Helpful
Bookmark
Share
View Full Paper