Awesome-TCM-LLM · Catalog
TCM LLMs
A living catalog of Traditional Chinese Medicine large language models (TCM LLMs): which models exist, plus papers, surveys, benchmarks, datasets, and patents. This page is the web edition of Awesome-TCM-LLM, maintained by Yang Tan and synced with the GitHub README.
What is a TCM LLM?
A TCM LLM (Traditional Chinese Medicine large language model) is a large language model trained or fine-tuned for the domain of Traditional Chinese Medicine. These models are typically built on general-purpose Chinese foundation models (such as Qwen, ChatGLM, Baichuan, or LLaMA) and further adapted with domain corpora — classical TCM canons, renowned physicians' case records, herbal formula knowledge, clinical guidelines, and consultation dialogues — acquiring capabilities such as syndrome differentiation (辨证论治), herbal formula recommendation, TCM knowledge question answering, and understanding of the four diagnostic methods (inspection, listening/smelling, inquiry, and palpation).
Since the first wave of TCM LLMs in 2023 — BianQue, HuaTuoGPT, and ShenNong-TCM-LLM — the field has grown rapidly, now spanning intelligent consultation, clinical decision support, TCM education, licensing-exam evaluation, and knowledge-graph construction, including multimodal systems such as ShizhenGPT that handle tongue and facial imagery. This page tracks model releases, academic papers, benchmarks, open datasets, and related patents in the TCM LLM space.
Representative TCM LLMs
- BianQue (扁鹊): proactive-health large language model for Chinese living spaces (2023)
- HuaTuoGPT: large language model trained on Chinese medical corpora (2023)
- ShenNong-TCM-LLM (神农): large-scale TCM language model with instruction data and open weights (2023)
- ZhongJing (仲景): expert-knowledge-guided TCM LLM with vertical-domain fine-tuning (2025)
- Qibo (岐伯): continued pretraining + SFT for syndrome differentiation and QA, with the Qibo Benchmark (2025)
- ShizhenGPT: multimodal TCM LLM supporting the four diagnostic methods (2025)
- Baize-TCM-LLM (白泽): Qwen3-based TCM QA model series from the Institute of Chinese Materia Medica, CACMS (2025)
- TCMChat: generative TCM LLM with a 600k-sample herbal-knowledge dialogue dataset (2025)
- Xinghe (杏核): Xinghe-TCM Neijing reasoning model on Qwen3.5-9B with open weights (2026)
- ZhiFangDanTai (知方丹台): GraphRAG + fine-tuning formula model from Capital Normal University and UQ, with open weights (2025)
- XuanHuGPT (悬壶): parameter-efficient fine-tuned TCM domain LLM (2025)
- MedChatZH: first-author work — a Baichuan-7B model fine-tuned for Chinese medical / TCM consultation
See the full catalog below — filterable by type and year.
TCM LLM benchmarks & datasets
- TCMBench: comprehensive benchmark for evaluating LLMs in the TCM domain
- MTCMB: multi-task TCM benchmark — 12 subsets, ~7,100 samples covering knowledge, reasoning, formulas, and safety
- TCM-Eval: evaluation dataset for TCM LLMs
- TCMChat-dataset-600k: 600k herbal-knowledge dialogue samples for fine-tuning
- ShenNong_TCM_Dataset: instruction data released with the ShenNong model
- Survey of key technologies for TCM LLMs (IJPRAI): systematic review of knowledge organization, assisted diagnosis, and clinical decision support
Browse by topic
- Open TCM LLMs: public weights and representative systems
- TCM LLM datasets: instruction data, cases, benchmarks
- TCM LLM surveys: diagnosis, multimodal, agents, evaluation
- TCM LLM patents: inquiry, formulas, knowledge graphs, and RAG
What you will find
- Selected news on products, policy, and open releases
- Open-weight and multimodal TCM LLM systems
- Benchmarks, evaluation suites, and surveys
- Pretraining / SFT datasets and related resources
- Invention patents on TCM LLMs, knowledge graphs, and intelligent inquiry
Catalog
Updated: 2026-09-13 · 432 entries (12 open models · 85 datasets · 47 surveys · 22 patents)
All resources
TCM LLM news, papers, surveys, datasets, open weights, and patents — filter by type and year.
Beijing University of Chinese Medicine debuts a TCM constitution-identification system and an embodied tuina robot at the CIFTIS 2026 TCM pavilion.
Guang'anmen Hospital shows Guangyi Qizhi 2.0 at CIFTIS 2026, with AI doctor An'an covering six hospital scenes from patient service to ward management.
TCM dialogue generation with a large language model and long-term memory
Generates personalized TCM dialogue from a long-term memory bank and user portrait (published, not granted).
从大语言模型到智能体(兰州大学学报医学版综述)
Chinese-language systematic review organized around the LLM-to-agent transition for TCM clinical assisted diagnosis and treatment (J. Lanzhou Univ. Med. Sci. 2026;52(4):49-57).
Digital inheritance and analysis of Zhang Xichun's theory, method, formula, and herbs
Rule engine plus three-way retrieval and an LLM for Zhang Xichun school analysis (published, not granted).
Beyond the Poetic Bard(中医AI翻译评论)
Beyond the Poetic Bard: a perspective on accuracy, epistemology, and medical-context limits of generative-AI translation of TCM texts (Translation Review).
Hybrid Retrieval + Re-ranking TCM Prescription Generation
Hybrid retrieval with re-ranking to enhance LLM-based TCM prescription generation (Springer CCIS conference paper).
TCM-RobustSDT
TCM-RobustSDT: a robustness benchmark dataset for LLM clinical reasoning in TCM (Figshare).
Agentic and Knowledge-Grounded LLMs in TCM(预注册)
OSF preregistration (not a completed review) of a systematic review on agentic and knowledge-grounded LLMs in TCM: evidence mapping, text mining, and translation readiness.
Patient-Conditioned Dual Hypergraph Reasoning
Patient-conditioned dual-hypergraph reasoning for auditable TCM prescription support, organizing symptom/tongue/pulse evidence around patterns and treatment principles (Tianjin University).
Insightful Eye presents Bianshi Cloud TCM at WAIC 2026, built on a registered Bianshi multimodal model with four-diagnosis devices and assisted-care systems.
Andun Health debuts a seven-diagnosis TCM robot at WAIC 2026, integrating face/IR face/tongue/ear/auscultation/inquiry/pulse sensing with TianHui pulse algorithms and a TCM clinical LLM.
Guizhou Medical University and partners launch HuaZu BenCao, a national ethnic-medicine AI platform built on ShuZhi QiHuang + Qwen integrating multi-ethnic materia medica classics.
Shanghai Seventh People's Hospital (SHUTCM) releases the Qiyuan TCM LLM with pretraining, domain fine-tuning, and expert RL; supports generative medical records and master-physician Agent digital twins; showcased at WAIC.
中医大模型关键技术综述(IJPRAI)
Survey of key technologies for TCM LLMs: knowledge organization, aided diagnosis, and clinical decision support (formally published in IJPRAI, World Scientific).
DeepTCM1.0
DeepTCM1.0: 11-expert multi-agent system on DeepSeek V3.2 for interpreting TCM formula mechanisms (Guizhi Decoction case); now on arXiv after the Research Square preprint.
DeepRoot
Multi-agent pipeline that turns the Shen Nong Ben Cao Jing into a verified Neo4j graph for therapeutic reasoning; code and evaluation scripts are public.
Evidence-Based TCM Visualization Diagnosis System
Evidence-based TCM visualization diagnosis system: Neo4j knowledge graph (241 patterns, 1,263 symptoms) with four-stage symptom matching (LLM-verified) and information-gain-driven active inquiry.
Zhongke Wenge passes HKEX listing hearing; reports note the DaYi JinKui TCM LLM (with CACMS) has obtained top-tier CAICT Trusted AI certification.
Causal contrastive learning and LLM graph attention for TCM prescription recommendation
Heterogeneous and homogeneous symptom-herb graphs plus an LLM-enhanced causal mechanism (published, not granted).
Multi-task joint optimization for TCM prescription generation
A pretrained LLM predicts herb sequences and herb sets jointly, with extra weight on rare herbs (published, not granted).
LLM-based intelligent acupuncture diagnosis agent
Acupuncture knowledge bases, graph, vector search and reranking, with an LLM agent that classifies questions and asks follow-ups (published, not granted).
TCM diagnosis-and-treatment platform combining an LLM with dual-channel retrieval
LLM plus TCM knowledge-graph dual-channel retrieval for inheriting master-physician experience and generating formulas (published, not granted).
RAG+LoRA 中医执照考试推理架构
RAG+LoRA generative architecture with an 11,476-item Taiwan TCM licensing-exam dataset (2005-2025), raising accuracy from 61.0% to 89.0%+ (Future Internet, MDPI).
AI驱动中医诊断智能化综述(JTCMS)
Survey on multimodal fusion and LLMs for intelligent four-diagnosis in TCM: applications, challenges and outlook (JTCMS).
Pediatric influenza Chinese patent-medicine recommender (KG + LLM)
Knowledge graph of Chinese patent medicines for pediatric influenza built from authoritative guidelines and integrated with an LLM (JMIR Preprints).
HSQ-TD(健身气功指令微调数据集)
HSQ-TD: the first instruction-tuning dataset for health Qigong/wellness, with 57,843 instructions distilled from official textbooks and professional literature (ScienceDB).
Pathogenesis-reasoning CoT supervision for spleen-stomach disorders
Pathogenesis-reasoning chain-of-thought supervision replacing fixed-label classification for spleen-stomach disease syndrome recognition and multi-dimensional evaluation (Prog. Biochem. Biophys.).
GAT+LLM TCM Prescription Generation
Intelligent TCM prescription generation combining graph attention networks with LLMs (formally published in KSII TIIS).
TCMIIES
TCMIIES: a browser-based, zero-installation LLM system for structured information extraction from academic literature, aimed at TCM and other specialty researchers.
Knowledge-graph and LLM system for TCM syndrome-differentiation QA
Chunks classics and cases, extracts a graph, then retrieves local and global keywords for syndrome-differentiation answers (published, not granted).
Affiliated Hospital of Shandong University of TCM launches the provincial AI+ TCM scenario project ZhiHui QiHuang.
Beijing University of Chinese Medicine's XinHuo ZhongGuoYao TCM education LLM completes national generative-AI service filing—the first publicly approved TCM-domain model of its kind.
Structured evidence-subgraph RAG for TCM herb recommendation
Multi-hop evidence subgraphs plus Hopfield associative retrieval for herb recommendation, with graph constraints against hallucination (published, not granted).
Knowledge-enhanced generation of TCM case commentaries
Retrieves knowledge-graph subgraphs, then writes case notes along etiology, treatment principles, formulas, and prognosis (published, not granted).
Knowledge-distillation and RL for oncology TCM prescription recommendation
Distills GPT-4o into a Qwen student, then DPO, for interpretable oncology TCM prescriptions (published, not granted).
Eight ministries issue the TCM Industry High-Quality Development Plan (2026–2030), calling for AI and knowledge graphs to empower classical formulas and master physicians' prescriptions.
Multimodal knowledge-graph construction from TCM texts and records
An LLM extracts triples from TCM texts, then fuses a real-world case layer into a multimodal graph (published, not granted).
人工智能驱动下的中医智能诊疗研究进展与挑战
Chinese review structured on the six-step TCM diagnosis-treatment chain (four diagnoses, pattern differentiation, prescription, outcome prediction), contrasting supervised/unsupervised/RL/deep-learning paradigms (Shanghai J. TCM 2026;60(1)).
Knowledge-graph recommendation of classic TCM formulas
Maps colloquial symptoms to terms, filters by constitution, and recommends classic formulas (granted).
ZMT-M1 scores 96.26 on a simulated national TCM practitioner exam and is piloted in 100+ clinics.
Gushengtang launches a “TCM Brain” product and AI digital twin of National TCM Master Shi Qi; 14 expert twins cover eight core specialties.
Jilin releases ZhongXing·Changbai Qihuang 1.0, an AI-native multimodal TCM LLM.
Five national health ministries issue AI+ healthcare guidelines that explicitly support building TCM diagnostic LLMs.
YiYin classical TCM LLM goes live in Song County, Henan, combining TCM education with AI-assisted care for county clinics.
ZhiFu Qihuang Tiangong TCM AI model completes national deep-synthesis algorithm filing for four-diagnosis devices and constitution assessment.
Transn's RenDu·SuWen passes CAICT Trusted AI TCM LLM Level 4+ evaluation.
NSCC-TJ and Tianjin University of TCM release TianHe·LingShu 2.0 (expanded beyond acupuncture to 20+ specialties) and launch a TCM intelligent-model evaluation system.
Gushengtang releases ten “National Master AI twins” trained on master clinicians' experience, reporting >86% pattern-differentiation accuracy.
Guang'anmen Hospital forms the GuangYi·QiZhi LLM agent alliance with medical consortium partners for cross-institution intelligent care.
Knowledge-graph and case-enhanced RAG for intelligent TCM inquiry
Leiden subgraph search plus hybrid global/local recall, then a single LLM response (published, not granted).
AI-LLM method and system for intelligent TCM inquiry
Multimodal inquiry fusing face/body/tongue images, speech prosody, and a disease library (granted).
Multimodal knowledge-graph and LLM system for TCM rehabilitation diagnosis
Aligns tongue, pulse, and inquiry features with knowledge-graph embeddings before a rehabilitation expert LLM (published, not granted).
TCM Hengqin vertical LLM is officially released.
China Academy of Chinese Medical Sciences releases evaluation standards for TCM large models.
LLM-based Chinese-medicine prescription recommendation with soft prompts
Soft prompts and cross-attention feed patient text into an LLM for controllable herb recommendations (published, not granted).
Transn releases RenDu·SuWen TCM LLM (mixture-of-entropy architecture) for inquiry, pattern differentiation, and formula recommendation.
Guang'anmen Hospital releases GuangYi·QiZhi, described as the first TCM hospital with integrated local compute + model + application deployment.
LLM method for learning famous-physician TCM formulas
Aligns classical-Chinese case terms with vernacular text using models such as Qwen (published, not granted).
LLM and knowledge-graph method for TCM formula compatibility
Pairs a TCM LLM with a knowledge graph for formula compatibility (published, not granted).
LLM-based Chinese-materia-medica question answering
Chinese-materia-medica QA fine-tuned from Baichuan2-7B-Chat (granted February 2025).
China UnionPay Consumer Finance with Sun Yat-sen University and GZUCMS Shenzhen Hospital release vertical TCM LLM ZhongSi for community clinic inquiry.
Building a TCM QA system from a large language model and a knowledge graph
Closed loop of LLM generation and knowledge-graph completion for TCM QA (published, not granted).
Zhongke Wenge releases DaYi JinKui TCM LLM and health platform trained on 1,500+ TCM classics.
AI-large-model system for intelligent TCM prescription recommendation
Trains an AI large model on medical records to recommend personalized TCM prescriptions, wired into a hospital IS (published, not granted).
Tasly and Huawei Cloud release ShuZhi BenCao (Pangu language + molecular models) covering TCM R&D; later earns CAICT TCM LLM Level 4+.
Long-document retrieval-augmented generation for TCM question answering
Long-document RAG (expansion, recall, reranking, source citations) for TCM QA (granted July 2024).
ECNU, SHUTCM, ECUST, NMMU, Lingang Lab, and CR Jiangzhong jointly develop the ShuZhi QiHuang TCM LLM. The linked ECNU page is a later write-up of ShuZhi QiHuang 2.0, not the original 2024.03 release announcement.
Nanjing Dajing TCM releases QiHuang Wendao LLM with large knowledge graphs and clinical data for institutional beta use.
Cong H et al. TCM×LLM 综述(OSF 预印本)
OSF-preprint review of LLMs in TCM (not peer-reviewed; archival).
Cai R et al. TCM×LLM scoping review(OSF 预印本)
OSF-preprint scoping review of LLMs in TCM (not peer-reviewed; archival).
AI in TCM: multimodal data to pharmacology and clinical decision(PRMCM 综述)
Broad AI-in-TCM review from multimodal data integration to pharmacological research and clinical decision support (Pharmacol. Res. Mod. Chin. Med. 2026; found in the third-round scan).
TCMBank
Historical anchor: TCMBank.
ETCM v2.0
Historical anchor: ETCM v2.0.
ETCM
Historical anchor: ETCM.
TCMSP
Historical anchor: TCMSP.
TCMID
Historical anchor: TCMID.
Data-driven based four examinations in TCM: a survey
Historical anchor: Data-driven based four examinations in TCM: a survey.
Research and application of tongue and face diagnosis based on deep learning
Historical anchor: Research and application of tongue and face diagnosis based on deep learning.
Deep Learning Multi-label Tongue Image Analysis and Its Application in a Population Underg
Historical anchor: Deep Learning Multi-label Tongue Image Analysis and Its Application in a Population Underg.
Automatic Construction of Chinese Herbal Prescriptions From Tongue Images Using CNNs and A
Historical anchor: Automatic Construction of Chinese Herbal Prescriptions From Tongue Images Using CNNs and A.
Artificial intelligence in tongue diagnosis: Using deep convolutional neural network for r
Historical anchor: Artificial intelligence in tongue diagnosis: Using deep convolutional neural network for r.
Constitution Identification of Tongue Image Based on CNN
Historical anchor: Constitution Identification of Tongue Image Based on CNN.
Tooth-Marked Tongue Recognition Using Multiple Instance Learning and CNN Features
Historical anchor: Tooth-Marked Tongue Recognition Using Multiple Instance Learning and CNN Features.
Ensemble Learning-Based Pulse Signal Recognition: Classification Model Development Study
Historical anchor: Ensemble Learning-Based Pulse Signal Recognition: Classification Model Development Study.
Diagnostic Method of Diabetes Based on Support Vector Machine and Tongue Images
Historical anchor: Diagnostic Method of Diabetes Based on Support Vector Machine and Tongue Images.
Pulse Waveform Classification Using Support Vector Machine with Gaussian Time Warp Edit Di
Historical anchor: Pulse Waveform Classification Using Support Vector Machine with Gaussian Time Warp Edit Di.
Automated Tongue Feature Extraction for ZHENG Classification in Traditional Chinese Medici
Historical anchor: Automated Tongue Feature Extraction for ZHENG Classification in Traditional Chinese Medici.
Classification of Pulse Waveforms Using Edit Distance with Real Penalty
Historical anchor: Classification of Pulse Waveforms Using Edit Distance with Real Penalty.
Feature extraction and recognition of traditional Chinese medicine pulse based on hemodyna
Historical anchor: Feature extraction and recognition of traditional Chinese medicine pulse based on hemodyna.
A Novel Computerized Method Based on Support Vector Machine for Tongue Diagnosis
Historical anchor: A Novel Computerized Method Based on Support Vector Machine for Tongue Diagnosis.
An ontological framework for the formalization, organization and usage of TCM-Knowledge
Historical anchor: An ontological framework for the formalization, organization and usage of TCM-Knowledge.
Text mining for traditional Chinese medical knowledge discovery: a survey
Historical anchor: Text mining for traditional Chinese medical knowledge discovery: a survey.
Development of traditional Chinese medicine clinical data warehouse for medical knowledge
Historical anchor: Development of traditional Chinese medicine clinical data warehouse for medical knowledge .
Building Clinical Data Warehouse for Traditional Chinese Medicine Knowledge Discovery
Historical anchor: Building Clinical Data Warehouse for Traditional Chinese Medicine Knowledge Discovery.
Information retrieval and knowledge discovery on the semantic web of traditional Chinese m
Historical anchor: Information retrieval and knowledge discovery on the semantic web of traditional Chinese m.
Knowledge discovery in traditional Chinese medicine: State of the art and perspectives
Historical anchor: Knowledge discovery in traditional Chinese medicine: State of the art and perspectives.
Ontology development for unified traditional Chinese medical language system
Historical anchor: Ontology development for unified traditional Chinese medical language system.
Historical Analysis of Medical Artificial Intelligence Development in China: Research Cent
Historical anchor: Historical Analysis of Medical Artificial Intelligence Development in China: Research Cent.
Syndrome Differentiation in Intelligent TCM Diagnosis System
Historical anchor: Syndrome Differentiation in Intelligent TCM Diagnosis System.
Traditional Chinese medical diagnosis based on fuzzy and certainty reasoning
Historical anchor: Traditional Chinese medical diagnosis based on fuzzy and certainty reasoning.
Establishment of a fuzzy mathematical model for syndrome differentiation of gastric cancer
Historical anchor: Establishment of a fuzzy mathematical model for syndrome differentiation of gastric cancer.
Fuzzy match and floating threshold strategy for expert system in traditional Chinese medicine
Historical anchor: Fuzzy match and floating threshold strategy for expert systems in traditional Chinese medicine.
Guan Youbo liver-disease diagnosis and treatment program
Historical anchor: an early computer-based expert system encoding renowned TCM physician Guan Youbo's approach to liver disease.
An artificial intelligence program to advise physicians regarding antimicrobial therapy
Historical anchor: An artificial intelligence program to advise physicians regarding antimicrobial therapy.
Mathematical modeling of Chinese medicine by complex-valued five-agent network
Historical anchor: Mathematical modeling of Chinese medicine by complex-valued five-agent network.
Discovering golden ratio in the world’s first five-agent network in ancient China
Historical anchor: Discovering golden ratio in the world’s first five-agent network in ancient China.
A disturbance rejection framework for the study of traditional Chinese medicine
Historical anchor: A disturbance rejection framework for the study of traditional Chinese medicine.
Equilibrium and nonequilibrium modeling of YinYang WuXing for diagnostic decision support
Historical anchor: Equilibrium and nonequilibrium modeling of YinYang WuXing for diagnostic decision support .
YinYang bipolar logic and bipolar fuzzy logic
Historical anchor: YinYang bipolar logic and bipolar fuzzy logic.
A computer model of the “five elements” theory of traditional Chinese medicine
Historical anchor: A computer model of the “five elements” theory of traditional Chinese medicine.
Functional structure model of human body and Yinyang-Wuxing equations
Historical anchor: Functional structure model of human body and Yinyang-Wuxing equations.
GPT vs ERNIE 中医文化背景对比研究
A culture-framed comparison of GPT versus ERNIE on TCM tasks (J. Integr. Complement. Med. 2024).
Tree-organized self-reflective retrieval for TCM QA
Tree-organized self-reflective retrieval for TCM question answering (Frontiers in Medicine 2026).
RACE-Align
RACE-Align: retrieval-augmented, CoT-style DPO alignment of a compact Qwen3-1.7B for TCM reasoning. arXiv PDF authors at ShanghaiTech, Henan University, and Liaoning University of TCM.
仲景(CMtMedQA 线,Yang et al.)
ZhongJing (CMtMedQA line, Yang et al.): a TCM LLM distinct from the Kang-line ZhongJingGPT—full CPT+SFT+RLHF pipeline on Ziya-LLaMA-13B over ~70K real multi-turn doctor-patient dialogues (AAAI 2024).
LLM-Based Multi-Agent Systems for Clinical Workflows(ACL 2026,邻近)
Adjacent ACL 2026 survey of workflow-level multi-agent clinical systems with a four-layer evaluation stack; no TCM coverage but methodologically isomorphic process-evaluation claims.
医学大语言模型的研发与应用系统综述(智能系统学报,邻近)
Adjacent systematic review of 129 medical-domain LLMs (to 2024-06) and four clinical application categories; methodologically comparable search protocol.
医疗领域的大型语言模型综述(智能系统学报,邻近)
Adjacent Chinese general survey of medical LLMs (training pipeline, strategies, scenarios, challenges), a superset-context reference for TCM LLM surveys.
Intelligent Question-Answering Systems in Healthcare(Healthcare,邻近)
Adjacent review (not TCM-specific): 2018-2025 healthcare QA survey with CiteSpace bibliometrics, explicitly covering TCM formula-development scenarios (Healthcare 2025).
人工智能实现中医四诊的发展现状、问题及解决路径(中华中医药学刊)
Short Chinese review of AI-based four-diagnosis objectification: face/tongue acquisition, electronic nose, pulse sensing, and low fusion of multi-diagnosis data (bibliographic record only).
人工智能赋能中医数字化诊断:现状与挑战(中华中医药学刊)
Short Chinese review of AI-empowered digital TCM diagnosis: applications, data-quality, interpretability, and theory-integration challenges (bibliographic record only).
AI for Spleen-Stomach Disorders in TCM(Curr Med Sci)
Single-disease-area (spleen-stomach) review of KG plus intelligent diagnosis/treatment with a symptom-syndrome-disease-formula framework (Curr. Med. Sci. 2025;45(6)).
AI and Big Data in TCM Standardization and Internationalization(Chin Med Cult)
Perspective on AI and big data for TCM standardization and internationalization (Chin. Med. Cult. 2026, ahead of print).
AI empowers the innovation of TCM(J Integr Med 评论)
Single-author perspective on AI for TCM innovation: classics mining, diagnosis standardization, drug R&D cycles (J. Integr. Med. 2026).
古籍知识图谱×多智能体融合综述(Chin Med)
Challenges-and-prospects review of knowledge-graph construction over ancient TCM classics, first to frame multi-agent convergence in this area (Chin. Med. 2025;20:168).
The integration of machine learning into TCM(J Pharm Anal)
Review of machine-learning integration into TCM along diagnostic objectification and mechanism-elucidation lines (J. Pharm. Anal. 2025;15(8):101157).
Deep learning in TCM(J Integr Med)
Single-technology review of deep learning in TCM: medical imaging, herbal material research, data mining (J. Integr. Med. 2026;24(4):471-480).
AI in TCM: Unraveling Herbal Medicine's Mechanisms(Research)
Broad AI-in-TCM review arguing AI should move beyond correlational analysis toward reconstructing the biological logic of syndrome differentiation and formula compatibility (Research 2026;9:1224).
多模态大模型驱动舌脉面诊智能化综述(Springer 书章)
The only review text dedicated to multimodal-LLM-driven tongue, pulse, and facial diagnosis in TCM (Springer CCIS book chapter; weaker peer review than journals).
OASIS (KIOM)
KIOM traditional-medicine literature portal for Korean-medicine papers and herbal resources.
Korean Medicine Embedding Dataset
Query–positive–negatives (~113k pairs) built from Korean-medicine terms and an ontology, for embedding fine-tunes such as BGE-M3.
KNApSAcK KAMPO
NAIST Kampo public database (~1,581 formulas, 278 crude drugs), downloadable from the NBDC life-science archive.
webMedQA
Early Chinese non-factoid medical QA from health-consult sites (~63k questions, one positive and four negative answers each).
IMCS-21
About 4,116 pediatric online consults annotated for entities, intents, symptoms, and reports; later wired into four CBLUE dialogue tasks.
CBLUE
Chinese biomedical NLU benchmark (NER, relations, diagnosis normalization, classification); the source-task suite behind PromptCBLUE, with a Tianchi submission portal.
CMeKG
Chinese medical knowledge graph of diseases, drugs, and symptoms; main source for BenCao/HuaTuo and ChatGLM-Med instruction data. Official portal is unstable; verify via the tools repo.
MedDialog
Large doctor–patient dialogue corpus (about 1.1M Chinese encounters) used for multi-turn Chinese medical fine-tuning.
cMedQA2
Chinese community medical QA (~108k questions / 200k answers), a common source for BianQue-style SFT mixtures.
cMedQA
Chinese community medical QA-matching set (repo table ~54k questions / 102k answers; non-commercial research). Paper DOI matches the README; see cMedQA2 for the later release.
ChiMed (Qilin)
Qilin-Med's ~3GB Chinese medical corpus (CPT/SFT/DPO); not the same resource as the ChiMed 2.0 pretraining set.
DISC-Med-SFT
Fudan DISC medical-dialogue SFT set (~470k examples from KG triples and reconstructed consults; no preference data).
PromptCBLUE
CBLUE's 16 Chinese medical NLP tasks rewritten as generative instructions; an early unified Chinese medical LLM leaderboard (CCKS 2023).
Huatuo-26M
Largest open Chinese medical QA resource (~26M pairs from encyclopedias, KGs, and consults); Huatuo-Lite is the usual SFT/RAG subset.
CMExam
Chinese medical licensing-exam set (~68k annotated items) used as a knowledge-recall baseline by Chinese medical and TCM LLMs.
CMB
FreedomIntelligence comprehensive Chinese medical benchmark (CMB-Exam ~280k items plus CMB-Clin cases); the most common non-TCM comparison board in TCM LLM papers.
TM-MC
KIOM literature-derived Northeast Asian medicinal-material–compound database; the 2015 release covers ~536 materials, and the 2024 2.0 paper expands to ~34k compounds. Official site currently times out.
CVDHD
Cardiovascular herbal database of 3D compound structures, targets, and pathways for virtual screening and network pharmacology. Original Peking University site currently times out.
TCMGeneDIT
Text-mined associations among TCM, genes, diseases, effects, and ingredients, with pathway and PPI links. Official NTU site no longer resolves.
TCM-Mesh
Herb–compound–gene–disease network with toxicity/side-effect records (~6,235 herbs). Official portal currently returns 403; verify via the open paper.
TCMAnalyzer
RCDD chemo-/bioinformatics service for formula/herb/ingredient networks and scaffold search (~1,493 formulas, 618 herbs). Official rcdd.org.cn currently times out.
SuperTCM
Charité biocultural TCM resource linking drugs, botanical species, ingredients, targets, KEGG pathways, and diseases (~6,516 drugs). Official tcm.charite.de no longer resolves.
CMAUP
BIDD landscape of multi-target activities, pathways, and diseases for useful plants including TCM herbs; 2024 update, downloadable from the official site.
DCABM-TCM
Literature-mined blood constituents and metabolites of TCM prescriptions and herbs, with experimental detection conditions (~1,816 structured absorbed constituents).
ITCM
Integrated formula/herb/ingredient/target platform plus 1,488 pharmacotranscriptomic profiles for 496 TCM ingredients (expression data also on Synapse).
TCMBank
Large downloadable herb–ingredient–target–disease resource with literature-mining updates after manual checks.
YaTCM
About 1,813 prescriptions, 6,220 herbs, and 47k natural products with target/pathway tools. Nankai site currently returns 403; verify via the open-access paper.
CEMTDD
Ethnic-minority herbal–compound–target–disease resource (~621 herbs, mainly Uygur/Kazakh). Original cemtdd.com now hosts something else; verify via the PMC paper.
TCMIO
Immuno-oncology TCM database of prescriptions, herbs, ingredients, targets, and pathways, with downloads and a REST API.
LTM-TCM
Symptom–prescription–plant–ingredient–target platform linking 14 source databases plus clinical and classical records (~48k formulas). Official Tasly cloud no longer resolves; verify via the paper DOI.
HIT 2.0
Manually curated herbal-ingredient–target activity pairs (~1,237 ingredients / 2,208 targets, 2000–2020 literature). Portal opens; the analysis backend port is currently down (marked site issue).
TCM Database@Taiwan
About 20k isolated-compound 2D/3D structures from 453 TCM materials for virtual screening. Original tcm.cmu.edu.tw is unreachable; verify via the PLOS paper.
BATMAN-TCM 2.0
Known and predicted TCM ingredient–target protein interactions, with greatly expanded TTI coverage and target-to-ingredient search.
SymMap 2.0
Herb–TCM symptom–modern symptom–ingredient–target–disease maps, expanded with newer pharmacopoeia records and downloadable relationship tables.
HERB 2.0
Evidence-centered TCM resource integrating clinical trials, meta-analyses, high-throughput experiments, literature, and a knowledge graph.
ETCM 2.0
Encyclopedia of TCM formulas, patent drugs, materia medica, and ingredients with target prediction and multi-scale networks; v1 site remains online.
TCM-ID
NUS BIDD formula–herb–ingredient–target resource covering pharmacopoeia, classical, and CFDA-approved prescriptions; not the same database as TCMID 2.0.
TCMID 2.0
Integrative formula–herb–ingredient–target database (distinct from NUS TCM-ID). Original megabionet site is down; verify via Zenodo extract and the NAR paper.
TCMSP
Herb–ingredient–target–disease networks with ADME parameters; public site is TCMSP 2.3 with downloadable relationship tables.
知方丹台
ZhiFangDanTai formula-generation model weights.
白泽 (Baize)
Baize TCM LLM weights.
medchatzh
MedChatZH weights.
ZhongJing
ZhongJing GPT weights.
TCMChat
TCMChat weights.
ShizhenGPT
ShizhenGPT multimodal weight series.
ShenNong-TCM-LLM
ShenNong-TCM-LLM weights.
Lingdan
Lingdan / TCMLLM weights.
ChatTCM-7B-SFT
ChatTCM full-parameter SFT checkpoint.
ChatTCM
ChatTCM pretrained weights.
BianCang
BianCang open-weight series.
杏核 (Xinghe)
Xinghe Neijing reasoning model weights.
TCM-QG
About 5,000 TCM documents and 13,000 question-answer pairs from CHIP2020, for knowledge-base expansion and question generation (CC BY-SA 4.0).
TCM-PD
Yao et al. TKDE 2018 prescription topic-model set (98,334 raw / 33,765 processed symptom–herb ID pairs). CKCEST copyright, research use only. PresRecST's prescript_1195.csv is a reproduction table.
TCM-Lung
Pulmonary-disease cases from FAH-HUCM (14,948 processed; 4,484 encoded public rows of symptom/syndrome/method/prescription IDs). Full names on request. Not the same resource as TCMNSCLC.
TCM-SD
First large public TCM syndrome-differentiation text benchmark (54,152 real records, 148 syndromes, CC BY-NC-SA 4.0). Full set is in the repo folder TCM_SD_with_knowledge; Tianchi id 139034.
TCM-NER
1,997 Chinese-medicine package inserts with 59,803 entities in 13 types for building a medication knowledge graph (OpenKG / CHIP, CC BY-SA 4.0).
TCM_KG
ChatMed knowledge graph.
TCM-MKG
TCM multi-dimensional knowledge graph.
LingShu(灵枢知识图谱)
Symptom-centric contextual KG bridging TCM and biomedicine (~17.33M entities, ~39.47M relations including triples and contextual quadruples), with a portal for visualization, reasoning, and evidence-grounded QA.
OpenTCM-KG
OpenTCM gynecology classics KG (~48k entities / ~152k relations).
TCMNSCLC
Real-world NSCLC TCM reasoning dataset with fully annotated cases (pattern differentiation / treatment method / decoction / patent medicine).
ChP-TCM
KnowledgeQA and PrescriptionWriting instructions built from Chinese Pharmacopoeia Vol. I.
Traditional-Chinese-Medicine-Dataset-SFT
High-quality TCM supervised fine-tuning dataset.
TCMChat-dataset-600k
TCMChat herbal QA and recommendation instruction data (~600k).
TCM-Instruction-Tuning-ShizhenGPT
ShizhenGPT multimodal SFT data (text/vision/speech/ECG etc.; ~311k items total per paper Table 3).
ShenNong_TCM_Dataset
ShenNong TCM instruction dataset.
MedChatZH
MedChatZH TCM consultation dataset.
ChatMed_Consult_Dataset
Chinese online medical consult dataset (500k+ consults with ChatGPT replies).
CMtMedQA
ZhongJing real multi-turn doctor–patient dialogues (~70k).
Baize-TCM-Corpus-V3
~157k TCM QA items covering theory, herbs, formulas, diagnosis, acupuncture, and clinic.
neijing-sft-v1.2
~2,009 Neijing-related instruction samples for Xinghe, with thinking/output fields.
TCM-Text-Exams
Recent TCM licensure / graduate-exam text benchmark.
Medical-LLMs-Chinese-Exam
Chinese medical exam evaluation for medical LLMs.
ZhongJing-OMNI
ZhongJing-OMNI multimodal TCM eval (including tongue).
TCMEval-SDT
TCMEval-SDT: a benchmark of 300 syndrome-diagnosis cases (web, classical texts, hospital records) for evaluating TCM syndrome-differentiation reasoning, with FAIR metadata (Sci. Data 2025).
TCMBench
TCMBench: a comprehensive benchmark for evaluating LLMs in traditional Chinese medicine (arXiv 2024).
TCM-Vision-Benchmark
TCM vision benchmark (herb recognition / inspection, ~7k items).
TCM-Tongue
6,719 standardized tongue images with 20-class multi-label pathology annotations and detection baselines.
TCM-Ladder
TCM-Ladder: a multimodal QA benchmark for comprehensively evaluating TCM multimodal LLMs on real-world tasks (arXiv 2025).
TCM-Eval
Dynamic, extensible TCM evaluation platform.
TCM-BEST4SDT
Case benchmark for syndrome differentiation and treatment.
TCM-5CEval
Five-dimension deep TCM evaluation suite.
TCM-3CEval
Three-axis eval: core knowledge, classics, clinical decisions.
MTCMB
MTCMB dataset: a multi-task TCM benchmark covering knowledge, reasoning and safety, 12 subsets with ~7,100 samples (arXiv 2025).
HWTCMBench
HWTCMBench TCM capability evaluation set.
TCMEval-PA
328 multiple-choice items on prescription normative quality and safety auditing.
TCM-AQA61
Dual-view acupuncture and Tuina action-quality videos from 61 subjects each (first- and third-person), with expert categorical and continuous ratings; paired with the CME-AQA cross-view multimodal assessment framework.
LingLan
LingLan large multi-task TCM evaluation benchmark (2026).
ChiMed 2.0
Upgraded Chinese medical pretraining dataset covering TCM corpora for LLM pretraining.
classical-tcm-canon
Full-text digitizations of the TCM canon: Neijing, Nanjing, Shanghan Lun, Jingui Yaolue and warm-disease classics.
Traditional-Chinese-Medicine-Dataset-Pretrain
High-quality TCM pretraining dataset from non-Internet sources (~1GB; clinical cases, classics, encyclopedia), 99% simplified Chinese.
TCM-Pretrain-Data-ShizhenGPT
ShizhenGPT pretraining corpus (15B+ tokens reported in the paper — Stage-1 text 11.92B incl. 6.3B TCM, plus Stage-2 multimodal ~3.6B).
TCM-Ancient-Books
A corpus of nearly 700 TCM ancient-book texts.
awesome_Chinese_medical_NLP
Curated list of Chinese medical NLP resources: terminologies, corpora, word vectors, pretrained models, KGs, NER and QA (incl. CBLUE).
CPM中成药数据集
Living large-scale public Chinese patent medicine data accompanying RAG-CPMF.
Lukman et al. 2007: 中医计算方法综述
Foundational survey of computational methods for TCM (expert systems, ML, data mining).
Gu & Chen 2013: 生物信息学遇见中医
Historical review of bioinformatics meeting TCM (omics and text mining).
Zhao et al. 2015: 中医患者分类进展(ML 视角)
Review of ML-driven advances in patient classification for TCM.
Chu et al. 2020: 中医定量知识表示模型综述
Review of quantitative knowledge representation models of TCM (ontologies, rules, statistics).
Zhang et al. 2021: 计算中医诊断文献综述
Literature survey of computational TCM diagnosis — symptom acquisition, pattern modeling, and systems.
Tian et al. 2024: 四诊机器学习综述
Review of machine learning for TCM four diagnoses — inspection, auscultation-olfaction, inquiry, and palpation.
Qu et al. 2024: 中医知识图谱综述
Review of knowledge graphs in TCM — analysis, construction, applications, and prospects.
Song et al. 2024: AI 辅助中医辨证关键问题与技术挑战
Strategic-study review of key issues in AI-assisted TCM syndrome differentiation — multimodal fusion, symptom association, pattern quantification and reasoning, and TCM LLMs.
Su et al. 2024 — Review of AI in TCM diagnosis and treatment (Chinese)
Chinese-language review of three AI stages in TCM care — expert systems, ML, and deep learning — with challenges.
Li et al. 2024 — Research progress and prospects of LLMs in TCM (Chinese)
Chinese-language review of TCM LLM pipelines, frontier techniques (prompting/RAG/RLHF), and application prospects.
Yip et al. 2025: 中西医结合 LLM 进展与挑战
Review of LLMs in integrative medicine — progress, challenges, and opportunities.
Wang et al. 2025: AI 驱动中医诊断模型进展
Systematic review of AI-driven TCM diagnostic models (four-diagnosis objectification, pattern differentiation).
Meng et al. 2025: 大模型+虚拟细胞助力中医变革
Review of large models and virtual cells aiding modern analysis of stroke treatment with TCM formulas.
Zhang et al. 2025: 中医 LLM 短综述与展望
Short survey and outlook on TCM LLM models and tasks.
Shataer et al. 2025: LLM 在中医应用(State-of-the-Art Review)
State-of-the-art review scanning TCM LLM application scenarios (care, education, translation, research).
Guo et al. 2025: GPT 能否加速中医智能诊疗(综述+实证)
Survey plus empirical analysis of whether GPTs can accelerate intelligent TCM diagnosis and treatment.
Chen et al. 2025: 中医大语言模型系统综述
Systematic review of 10 studies (to mid-2024) on LLMs in TCM generative tasks.
Ren et al. 2025: 中医大语言模型(Scoping Review)
Arksey-O'Malley scoping review (29 studies to 2024-04) covering knowledge management, assisted care, and exam accuracy.
Lu et al. 2026: 深度学习中医诊断方法学质量审计
Systematic review and validation-gap analysis of deep learning for TCM disease diagnosis.
Wu et al. 2026: AI 在中药材中的应用综述
Full-stack survey of AI in TCM herbs — compounds, targets, quality control, with an LLM section.
Guo et al. 2026: AI 与多模态数据融合推动中医现代化
Panoramic AI review (ML/DL/KG/NLP/LLM) for TCM modernization with multimodal data integration.
Chen et al. 2026: LLM 在中医的下一步(叙述性综述)
Narrative review on the next step of LLMs in TCM — multimodality, agents, and clinical translation.
Xu et al. 2026: 基于 LLM 的中医智能问答系统综述
Review of intelligent TCM question-answering systems based on LLMs (KG-QA to LLM-QA and RAG).
Yao et al. 2026: LLM 与循证中医整合(Scoping Review)
PRISMA scoping review (12 studies, 2022-11 to 2026-01) on integrating LLMs with evidence-based Chinese medicine.
Han et al. 2026: LLM 在中医中的调优与临床应用(Scoping Review)
PRISMA-ScR scoping review (27 studies to 2025-05) on tuning (LoRA/CPT/RAG) and clinical application of TCM LLMs.
中医症状名识别
Supervised methods for symptom name recognition in free-text TCM clinical records.
中医临床细粒度 NER 语料
Fine-grained entity-recognition corpus built from TCM clinical records.
TCMPR 子网术语映射处方推荐
Herb-symptom knowledge graph (~18k entities / ~100k relations) plus subnetwork term mapping and a CNN for prescription recommendation.
中医期刊关系抽取 (Wang & Poon 2016)
Relation extraction from TCM journal full text to support later knowledge-graph construction.
语义中医方剂知识图谱 (Miao et al. 2018)
Semantic TCM prescription knowledge graph built top-down from formula texts.
中医养生知识图谱 (Yu et al. 2017)
Large TCM health-preservation knowledge graph integrating terms, literature, and databases, with retrieval, visualization, and recommendation.
TCMKG
Deep-learning-based TCM knowledge graph platform.
乙肝中医 KG 问答系统
Knowledge-graph-based QA system for TCM diagnosis and treatment of viral hepatitis B.
中医新冠文献 LLM 命名实体识别
Comparative study of LLMs for named entity recognition in TCM COVID-19 literature (preprint).
PreGenerator
TCM prescription recommendation model combining retrieval and generation.
草药智能配送聊天机器人
Smarter herbal medication delivery system employing an AI-powered chatbot.
LLM+GNN 中医处方推荐
TCM prescription recommendation combining large language models with graph neural networks.
中医方剂 LLM 分类
Fine-tuned LLMs with refined prompt templates for TCM formula classification, using data sources such as the national medical-insurance catalog of proprietary Chinese medicines (IEEE BIBM 2023).
中医疫病防治问答模型
LLM-based QA model for TCM epidemic prevention and treatment.
ChatGPT 针灸教育研究
Comparative study of ChatGPT as a learning tool in acupuncture education.
大模型融合知识图谱问答系统
Vertical-domain QA system deeply integrating LLMs with knowledge graphs for TCM formulas.
黄帝 (HuangDi)
HuangDi: a TCM classics QA LLM built on Ziya-LLaMA-13B, pretrained on 22 TCM textbooks plus TCM web corpora and SFT-tuned with ancient-book instruction data (Library Tribune 2024).
女娲 (Nüwa / TCM-Nvwa)
Nüwa TCM LLM training stack (continual pretraining, SFT, reward modeling, RLAIF) on Ziya-LLaMA-13B. Repo ships only partial pretrain/TCM-QR/reward data and no standalone weights; GitHub created April 2025, distinct from arXiv 2411.00897.
神农大模型 (ShenNong-TCM-LLM)
ShenNong-TCM-LLM, the first TCM large language model, released with the ShenNong_TCM_Dataset and open weights.
KM-Agent
Tool-augmented agent for Korean / East Asian traditional medicine over 4,780 herb–syndrome–acupoint metadata records, evaluated on TCMBench-style sets. MDPI/OpenAlex authors at Wonkwang University and Pusan National University et al.
KAMPO LLM
Official site still live. Closed Kampo API from VARYTEX with the Japan Society for Oriental Medicine; page reports 97.4% (Pro) and 92.1% (Flash) on 471 specialist-training items; no public weights.
Ant Afu
Closed Ant multimodal medical LLM, first shipped as Alipay AQ and later rebranded Afu for consults, report and pill-box reading; announced at WAIC 2024, no public weights.
PanGu Drug Model
Closed Huawei Cloud / SIMM molecule foundation model (~1.7B small molecules) for property prediction, generation, and optimization; later used under Shuzhi Bencao.
Tencent Hunyuan Medical
Closed Tencent Health medical LLM on Hunyuan for QA, triage, records, and imaging; announced by Tencent Jarvis Lab, no public weights.
iFlytek Spark Medical
Closed iFlytek medical LLM (Spark Medical X1/X2) behind Zhiyi Assistant and Xiaoyi; no public weights, listed as an industry baseline.
ClinicalGPT
BUPT clinical Chinese medical model (BLOOM-7B) fine-tuned on records, knowledge, exams, and multi-turn consults; a common CMB-era baseline. HF hosts a medicalai snapshot.
ChatGLM-Med
HIT-SCIR ChatGLM-6B instruction-tuned on Chinese medical KGs, sharing data lineage with BenCao/HuaTuo; a frequent CMB baseline.
IvyGPT
LLaMA-based Chinese medical QA model fine-tuned with curated clinical QA and RLHF; listed as an open baseline in the CMB paper. arXiv PDF authors at Macao Polytechnic University.
BianQue-2
Second BianQue open consultation model with stronger multi-turn questioning; a common CMB baseline.
CareGPT
Open Chinese medical LLM training stack (pretrain through DPO) with accompanying weights, often used to reproduce Chinese medical fine-tunes. README attributes the work to Macao Polytechnic University.
SoulChat
SCUT mental-health dialogue LLM from the same lab as BianQue; a common Chinese health-conversation baseline.
MedicalGPT
Open training stack for Chinese medical LLMs (pretrain/SFT/RLHF/DPO), often reused as a baseline pipeline in TCM fine-tuning.
DoctorGLM
Early open Chinese consultation model on ChatGLM-6B with LoRA/P-Tuning; a frequent 2023 Chinese medical baseline. arXiv HTML authors at ShanghaiTech and United Imaging Intelligence et al.
Apollo
FreedomIntelligence multilingual medical LLM (including Chinese) with ApolloCorpus and XMedBench.
ChiMed-GPT
USTC Chinese medical LLM continued from Ziya-v2 with pretraining, SFT, and RLHF for extraction, QA, and multi-turn dialogue.
Qilin-Med
Multi-stage Chinese medical LLM (CPT+SFT+DPO on Baichuan-7B) releasing the ~3GB ChiMed corpus, optionally with RAG. arXiv PDF authors at Peking University and HKUST (Guangzhou) et al.
Taiyi-2
Second Taiyi open biomedical model, moving from Qwen-7B to GLM4-9B with tighter data filters and task instructions; the official replacement for Taiyi-1.
Taiyi
DUTIR bilingual biomedical LLM on Qwen-7B for QA, doctor–patient dialogue, report generation, and information extraction (JAMIA 2024).
WiNGPT3
Winning Health's third medical reasoning model (32B on Qwen2.5) with multi-stage SFT+RL and WiNEX hospital integration; tech report and code are public, weights are not verified as downloadable.
WiNGPT2
Winning Health open medical LLM on Qwen for medical QA, record understanding, and multi-turn consults; 7B/14B weights are public.
PULSE
OpenMEDLab Chinese medical LLM (Bloom 7B/14B) for exams, report reading, record structuring, and simulated diagnosis.
DISC-MedLLM
Fudan DISC conversational medical LLM on Baichuan-13B, with DISC-Med-SFT built from knowledge graphs and reconstructed consultations.
HuatuoGPT-Vision
HuatuoGPT multimodal medical model that injects visual knowledge at scale; a common Chinese medical vision baseline.
HuatuoGPT-o1
HuatuoGPT medical complex-reasoning model trained with verifiable problems and a medical verifier; 7B/72B cover Chinese and English.
HuatuoGPT-II
Second HuatuoGPT generation with one-stage medical adaptation; 7B/13B use Baichuan2 backbones and remain a standard Chinese medical baseline.
Baichuan-M3
Baichuan's third open medical reasoning model (235B on Qwen3) with SPAR staged RL and fact-aware RL for active inquiry and hallucination control; strong HealthBench and SCAN-bench results.
Baichuan-M2
Baichuan's second open medical reasoning model (32B on Qwen2.5-32B) with a large verifier system and multi-stage RL; strong open-source HealthBench results.
Baichuan-M1
Baichuan's from-scratch open medical LLM (14B), trained on ~20T medical+general tokens across 20+ specialties; a common base for later TCM fine-tunes.
BianQue
Chinese proactive health LLM for everyday living spaces (BianQue).
孙思邈 (Sunsimiao)
Sunsimiao Chinese medical LLM; Sunsimiao-7B fine-tuned from Qwen2-7B on curated medical data, reaching 30B-level SOTA on CMB-Exam.
QiZhenGPT
Chinese clinical QA model for drugs, diseases, procedures, and labs (QiZhenGPT).
HuaTuoGPT
Large language model trained on Chinese medical corpora (HuaTuoGPT).
XrayGLM
Chinese multimodal medical LLM for chest X-ray interpretation.
ChatMed
ChatMed series of Chinese medical LLMs, including ChatMed-Consult trained on 500k+ online consultation dialogues. GitHub README cites Wei Zhu and Xiaoling Wang; no standalone journal paper.
RAG 增强中医问答置信度
Implementing retrieval-augmented generation to build LLM confidence in TCM (preprint).
GPT-4 中医研究生考试评估
GPT-4 vs mainstream Chinese LLMs on a TCM postgraduate examination dataset (preprint).
中医药问答大语言模型
TCM QA LLM combining RAG with P-Tuning v2 fine-tuning on ChatGLM2-6B.
中医标准化评估基准
Standardized TCM evaluation benchmark of 29,506 questions across 13 subjects; tests 3 general and 5 Chinese medical LLMs.
中医药大模型知识增强方法
Knowledge augmentation for TCM LLMs — a graph built from ~100k classical formulas preserving prescription structure.
ACUBERT
ACUBERT for meridian entity recognition and classification in acupuncture indication knowledge bases.
BSG 中医智能问答
Intelligent QA system for TCM based on a BSG deep-learning model (prescription and materia medica cases).
TCMD
TCMD, a TCM licensing-exam multiple-choice set for LLM evaluation (paper reports ~2,851 train / 600 test). Independent check found no official GitHub or Hugging Face download.
PresRecST
Progressive herb-prescription recommendation following syndrome differentiation then treatment planning (JAMIA 2024); ships a public encoded TCM-Lung subset and a TCM-PD reproduction table.
RLAIF 中医对齐
Enhancing LLMs' TCM capabilities through reinforcement learning from AI feedback.
中医提示工程框架
Prompt-engineering framework for LLM intelligent understanding in TCM.
ChatGPT 中医知识理解探究
Evaluating ChatGPT's comprehension of Traditional Chinese Medicine knowledge.
TCMSF
TCMSF: a construction framework for a TCM syndrome ancient-book knowledge graph that organizes syndrome knowledge from classical texts in a structured, semantically oriented way (Methods Inf. Med. 2024).
TCM MLKG-RAG
TCM intelligent diagnosis based on multi-layer knowledge graph retrieval-augmented generation.
TCM-BERT
BERT further pretrained on TCM clinical text for five-way disease classification (JAMIA 2019). CKCEST holds copyright; the full 46,205 records are not released, only splits plus drive-hosted fine-tuned weights.
ZY-BERT
Domain TCM encoder from the TCM-SD paper (~0.4B-token corpus); weights are on cloud drive, with syndrome-differentiation fine-tune code in the repo. Not the same work as arXiv 2411.00897.
Evi-BERT
Automated information-extraction model (Evi-BERT) enhancing RCT evidence extraction for TCM.
ChatGPT 中医交互可行性研究
Feasibility and challenges of interactive AI for TCM, using ChatGPT as an example.
中医领域知识图谱补全
Domain knowledge graph completion and quality evaluation for Traditional Chinese Medicine.
LLM 构建中医知识图谱
Constructing Traditional Chinese Medicine knowledge graphs based on large language models.
LLM 中医语言文化偏差研究
Comparing LLMs developed in different countries on TCM; highlights language/cultural bias and the need for localized models.
LLM 腧穴定位关系抽取
Relation extraction with LLMs — a case study on acupuncture point locations.
TCM-FTP
Fine-tuning LLMs for herbal prescription prediction.
CPMI-ChatGLM
Parameter-efficient fine-tuning of ChatGLM with Chinese patent medicine instructions. Authors at Anhui University of Chinese Medicine and CACMS Anhui Computer Application Research Institute.
TCM-GPT
Efficient pre-training of LLMs for domain adaptation in Traditional Chinese Medicine. Journal metadata lists BUPT and UCL.
BenCao (formerly HuaTuo)
Instruction-tuned Chinese medical LLM (BenCao / formerly HuaTuo).
明医 (MING)
MING: a Chinese medical consultation LLM using a sparse mixture of low-rank adapter experts (MING-MoE) for medical multi-task learning (arXiv 2024).
大数中医 (BigDataTCM)
BigDataTCM (34B): a vertical TCM LLM co-developed by HAUT's Complexity Science institute and Apus, offering medical QA, diagnostic support and TCM knowledge services.
DFGLM-TCM
Dongfang Hospital / Zhipu TCM clinical system that models textbook knowledge and practitioner experience in separate modules; paper is out, weights are not.
Lingdan-V2
BJTU's second Lingdan TCM reasoning family (Qwen3 4B/8B/14B with CPT, SFT, and prescription GRPO). ModelScope checkpoints exist but require access requests, so they are not marked freely downloadable.
TCMLLM / Lingdan
TCMLLM / Lingdan for TCM modeling and prescription recommendation.
MedChatZH
MedChatZH: a fine-tuned LLM for TCM consultation dialogues, released with open dataset and weights (Comput. Biol. Med. 2024).
Chinese-LLaVA-Med
Chinese medical multimodal LLM based on the LLaVA architecture, with the llava-med-zh-eval benchmark and open 7B weights.
GPT 台湾中医执业考试评估
GPT-3.5/GPT-4/GPT-4o performance on the Taiwan TCM licensing examination with reliability analysis (preprint).
双通道知识注意力辨证模型
Dual-channel knowledge-attention NLP model for TCM syndrome differentiation, addressing rare characters and terminology extraction.
中医药标准知识问答系统
Retrieval-augmented QA system for TCM standards knowledge, built and evaluated in practice.
Hengqin-RA-v1
LLM and companion dataset for TCM diagnosis and treatment of rheumatoid arthritis. arXiv/DOI page lists Chinese Medicine Guangdong Laboratory and Southern University of Science and Technology.
Gen-SynDi
Knowledge-guided generative-AI framework for dual education of syndrome differentiation and disease diagnosis.
辨证思维评测 (Syndrome Differentiation Thinking)
Method-development study evaluating and improving LLMs' TCM syndrome-differentiation thinking ability.
针灸大模型驯化与生成评估 (Taming LLMs for Acupuncture)
Taming LLMs for acupuncture & moxibustion diagnosis, with generation quality evaluated at the semantic-similarity level.
TCM-Sage
Evidence-synthesis RAG assistant for TCM practitioners (hybrid vector + knowledge graph).
From Metaphor to Mechanism
LLMs decode TCM metaphor / imagistic-thinking language and map it to modern medical concepts.
New Snow Tablets
Reveals systematic flaws of general and TCM-specific LLMs that guess formula ingredients from drug names.
TCDiff
Triplet cascaded diffusion model generating high-fidelity multimodal TCM EHRs, with the TCM-SZ1 benchmark dataset.
Ladder-base (GRPO-TCM)
First GRPO reinforcement-learning-aligned TCM LLM, from the TCM-Ladder team. arXiv PDF authors at University of Missouri and Shanghai University of TCM et al.
MRD-RAG
Multi-round diagnostic RAG simulating clinical reasoning; builds DiagnosGraph spanning TCM and Western medicine (876 diseases / 7,997 nodes / 37,201 triples).
TCM compound retrieval agent
AI agent-based system for retrieving TCM compound information.
LM extraction for complementary medicine
Language models for data extraction and risk-of-bias assessment in complementary medicine literature.
LLM-driven TCM KG construction
LLM-driven construction and application of a TCM knowledge graph.
中医医案问答系统
A TCM case-based QA system integrating LLMs and knowledge graphs for efficient case retrieval and analysis (Front. Med. 2025).
LLM + RAG TCM inference
Combining LLMs with RAG for TCM inference.
Weighted-voting TCM formula classification
Weighted-voting LLM approach for TCM formula classification.
TCM guideline adherence evaluation
Content-analysis evaluation of LLM adherence to clinical practice guidelines in Chinese medicine.
5-LLM TCM clinical decision comparison
Comparative study of 5 LLMs for TCM clinical decision-making.
TCM stroke LLM benchmark
Quantitative benchmark study of LLMs in the TCM stroke domain.
LLM herb–drug interaction prediction
LLM-enhanced herbal medicine–drug interaction prediction.
TCMLCM
KG2T-based intelligent QA model for TCM lung cancer.
Chinese patent medicine knowledge system
Constructing a knowledge system for traditional Chinese patent medicine using LLMs and KGs.
TCMRD-KG
Innovative design of a rheumatology TCM knowledge graph from ancient literature.
TCM-Eval (WISE 2025)
Multi-dimensional TCM evaluation framework (WISE 2025); a different work from the ZMT-M1 TCM-Eval (arXiv 2511.07148) despite the identical name.
Few-shot tongue-diagnosis in-context multitask learning
Few-shot in-context multitask fine-tuning of LLMs mapping tongue images directly to constitutions.
TCM misinformation detection evaluation
Safety evaluation framework with 3,000+ TCM exam items × 4 paradigms, covering wrong-option, misleading, and fabrication detection.
ChatGLM-FGIDs-TCM
Knowledge-fused ChatGLM clinical decision-support model for functional gastrointestinal disorders (FGIDs).
TCM-VisResolve (TCM-VR)
Qwen2.5-VL-based TCM multimodal LLM — 163-class dried-herb recognition over 220k images plus clinical MCQs with 880k candidate answers.
TCM-DS
Domain LLM for medicine–food homology dietary-therapy recommendation.
XuanHuGPT
TCM domain LLM built with parameter-efficient fine-tuning (PEFT).
Jingfang
LLM-based multi-agent TCM diagnosis/treatment system reporting large relative SDT gains under the authors' protocol.
神农Alpha
ShennongAlpha (Westlake University): an AI-driven sharing and collaboration platform for intelligent curation, acquisition and translation of natural-medicinal-material knowledge (Cell Discov. 2025).
ZhiFangDanTai
GraphRAG + LLM fine-tuning for interpretable formula generation (sovereign–minister–assistant–courier, efficacy, contraindications) with open weights. Authors at Capital Normal University and the University of Queensland.
Baize-TCM-LLM
ICMM Baize TCM QA models on Qwen3 (0.6B/8B) with ~157k LoRA-tuning examples.
ZMT-M1
ZMT-M1 TCM LLM and the dynamic, extensible TCM-Eval benchmark platform.
BianCang
BianCang TCM LLM series (IEEE JBHI); 14B open-weight release in Dec 2025.
Qibo
TCM LLM and Qibo Benchmark from Tianjin University et al.; CPT + SFT for SDT and QA.
TianHui
Domain LLM for 12 TCM scenarios (DeepSeek-R1-Distill-Qwen-14B + PT/SFT) with open code and eval scripts. arXiv PDF authors at Chengdu University of TCM.
Tianyi
~7B TCM LLM from NJUCM et al. with reading–clinic–apprenticeship training stages, TCMEval, and real-world validation.
仲景 (ZhongJing)
ZhongJingGPT, an expert-knowledge-guided TCM LLM combining vertical-domain fine-tuning with cognitive-psychology insights and multi-scenario TCM knowledge instructions (Tsinghua Sci. Technol. 2025).
RenShu-AI
FastAPI + LangGraph multi-agent TCM consultation system combining GraphRAG and DeepSeek-TCM.
Yaoshi-RAG
Uncertain-KG RAG for medicine–food homology dietary recommendation with personalization and explainability.
ViTCM-LLM
Qwen2.5-VL + RAG tongue multimodal clinical framework; MedTCM dataset and TDEU metric (precursor to MMIR-TCM).
TCMChat
Generative TCM LLM built via pre-training and supervised fine-tuning, released with the 600k-sample TCMChat-600k dialogue dataset (Pharmacol. Res. 2024).
TCM-R1
TCM LLM with GRPO-enhanced reasoning.
TCM-Ladder
First large multimodal TCM QA benchmark with 52,000+ items (NeurIPS 2025).
TCM-KLLaMA
KG-fused LLM for intelligent TCM formula generation. PubMed 40056842 lists Zhejiang Gongshang University and Hangzhou First People's Hospital (Westlake University School of Medicine).
TCM-BEST4SDT
Case benchmark for syndrome differentiation and treatment (knowledge / ethics / safety / SDT).
TCM-5CEval
Five-dimension deep evaluation extending TCM-3CEval with materia medica and non-drug therapies.
TCM-3CEval
Three-axis TCM LLM evaluation: core knowledge, classics comprehension, and clinical decision-making.
TCM LLM acupuncture clinical evaluation
Real-case evaluation of 7 general LLMs vs licensed acupuncturists on SDT, point selection, needling, and herbs (*npj Digital Medicine*).
ShizhenGPT
Multimodal TCM LLM supporting the four diagnoses (inspection, auscultation-olfaction, inquiry, palpation).
RAG-CPMF
Multi-LLM verification + RAG for Chinese patent medicine recommendation, with a living public CPM dataset.
OpenTCM
GraphRAG TCM retrieval and diagnosis system with a gynecology classics knowledge graph.
MTCMB
Multi-task TCM benchmark (~12 subsets, ~7.1k samples) covering knowledge, reasoning, formulas, and safety.
MCM
Multi-agent collaborative multimodal TCM diagnosis framework (IEEE ICIP 2025).
DoPI
Doctor-like proactive inquiry TCM LLM (guide + expert models); reported inquiry accuracy 84.68%. arXiv HTML authors at Tianjin University and CUHK et al.
DiagX-DT
Exclusionary syndrome-differentiation reasoning with CoT and an external TCM knowledge base.
ChatTCM
SylvanL open TCM LLM (Qwen2-7B continual pretrain then full SFT) with public data and weights. Contact email sl18n19@soton.ac.uk only; no stated affiliation or paper.
BenCao
Instruction-aligned multimodal TCM assistant (ChatGPT/GPTs Store) with tongue APIs and knowledge bases (distinct from HuaTuo/BenCao). arXiv HTML authors at University of Missouri and Shanghai University of TCM et al.
Medicinal-plant MLLM benchmark
Benchmarking multimodal LLMs for medicinal plant identification.
Large vs lightweight LLMs on TCM exams
Systematic comparison of large-scale vs lightweight LLMs on TCM exam questions.
TCM AI-tutor evaluation
Multimodal LLM evaluation for TCM education across cognitive levels.
Three-LLM TCM licensing exam evaluation
Systematic evaluation of 3 LLMs (incl. Gemini) on the national TCM medical licensing examination.
TCM intelligent pre-consultation clinical evaluation
Tertiary-hospital clinical evaluation of an LLM intelligent pre-consultation system using a physician–AI–patient triad model.
TCM Data Hub (YiYuan)
YiYuan LLM-driven TCM data platform.
TCMNet
LLM-assisted disease knowledge mining with PPI networks and binding prediction for formula optimization.
CMM-EmbedCluster
LLM + medicinal-property-theory clustering framework for Chinese materia medica, with a 567-herb property knowledge base.
KDC-NER
Knowledge-guided data augmentation + LLM fine-tuning framework for nested NER in TCM.
Jin San Zhen KG-QA
Knowledge graph + LLM QA tool for the Jin San Zhen acupuncture school.
TCMI-F-6D
Six-dimensional benchmark of interdisciplinary foundational competence in TCM informatics.
Tongue–face multimodal fusion diagnosis
Tongue–face multimodal feature fusion with LLM-driven intelligent TCM diagnosis.
TCMBenchEval
Benchmark evaluating LLMs on real clinical TCM case records (ICIC 2026).
Med-Bench-Arena
Open evaluation platform for medical and TCM LLMs/Agents (HF/vLLM/LiteLLM, multimodal, TCM-specific metrics), from the ZhongJing team.
TongueDx2
Systematic ablation of the tongue-diagnosis DL design space (20+ model variants); TongueDx2 includes 5,109 images / 976 expert annotations.
MACAT
Multi-agent culture-aware translation framework, evaluated on culture-loaded terms from TCM classics and the Analects.
DongYuan
Integrative spleen–stomach disease diagnosis LLM framework combining TCM pattern differentiation with Western diagnostic reasoning. arXiv PDF authors at Hebei Provincial Hospital of TCM and CASIA et al.
Med-Shicheng
Lightweight master-physician experience-inheritance framework built on Tianyi; a single model internalizes 5 national masters' knowledge systems across 7 task types.
HerbWise
Domain LLM for traditional herbal medicine (THM), serving herbal modernization and standardization.
QingNangTCM
Parameter-efficient fine-tuned TCM QA and clinical reasoning model; builds the 100k-item QnTCM_Dataset.
End-to-end TCM clinical support benchmark
Benchmark for end-to-end TCM clinical support across the full LLM care pipeline.
ATCMD-Bench
First agentic TCM diagnosis benchmark, evaluating LLMs through multi-agent simulated consultations.
LingLan
Large multi-task TCM benchmark: 5 domains, 13 subtasks, 25,624 instances.
Xinghe
Xinghe-TCM Neijing reasoning model on Qwen3.5-9B (QLoRA, ~2009 instructions, no-prescription safety). HF page lists no paper and no university affiliation.
TongueVLM
Multimodal VLM for TCM tongue diagnosis, description generation, and constitution reasoning. JMIR page blocked; OpenAlex lists HFUT and Anhui University of Chinese Medicine et al.
TCM-DiffRAG
Syndrome-differentiation RAG with a general KG, a personalized KG, and chain-of-thought.
TCM-Agent
LLM multi-agent system for network pharmacology and herbal discovery.
Qwen-TCM-Dia
Specialty fine-tuned model for TCM diarrhea care (CPT + CoT SFT) covering symptom→pathomechanism→method→formula chains. Digital Chinese Medicine authors at Beijing Hospital of TCM (CMU) and BUCM et al.
MMIR-TCM
Memory-augmented multimodal tongue diagnosis and clinical decision framework; proposes MedTCM dataset and TDEU metric.
GastroTCM
TCM gastroenterology LLM fine-tuned from Llama3-8B with RAG and agent scaffolding. Chinese Medicine paper from Tsinghua TCM-X and China-Japan Friendship Hospital et al.
DERM-3R
Resource-constrained multimodal multi-agent framework for TCM dermatology (recognition / representation / SDT agents).
CORE-Acu
Acupuncture clinical decision support with structured reasoning traces and a knowledge-graph safety veto loop.
No matching resources.
Full index (432 entries)
- 北京中医药大学在服贸会首发展示中医体质辨识体系与具身智能推拿机器人
- 广医·岐智大模型2.0亮相2026服贸会,落地AI医生安安与六大医院场景
- TCM dialogue generation with a large language model and long-term memory
- 从大语言模型到智能体(兰州大学学报医学版综述)
- Digital inheritance and analysis of Zhang Xichun's theory, method, formula, and herbs
- Beyond the Poetic Bard(中医AI翻译评论)
- Hybrid Retrieval + Re-ranking TCM Prescription Generation
- TCM-RobustSDT
- Agentic and Knowledge-Grounded LLMs in TCM(预注册)
- Patient-Conditioned Dual Hypergraph Reasoning
- 智慧眼砭石多模态中医大模型亮相WAIC 2026
- 安顿七诊合参中医机器人与天回大模型亮相WAIC 2026
- 贵州医大等发布华族本草民族药AI大模型
- 上海市第七人民医院发布岐元中医大模型,支持名老中医Agent数字孪生
- 中医大模型关键技术综述(IJPRAI)
- DeepTCM1.0
- DeepRoot
- Evidence-Based TCM Visualization Diagnosis System
- 中科闻歌通过港交所上市聆讯;报道提及与中国中医科学院合作的大医金匮中医大模型已通
- Causal contrastive learning and LLM graph attention for TCM prescription recommendation
- Multi-task joint optimization for TCM prescription generation
- LLM-based intelligent acupuncture diagnosis agent
- TCM diagnosis-and-treatment platform combining an LLM with dual-channel retrieval
- RAG+LoRA 中医执照考试推理架构
- AI驱动中医诊断智能化综述(JTCMS)
- Pediatric influenza Chinese patent-medicine recommender (KG + LLM)
- HSQ-TD(健身气功指令微调数据集)
- Pathogenesis-reasoning CoT supervision for spleen-stomach disorders
- GAT+LLM TCM Prescription Generation
- TCMIIES
- Knowledge-graph and LLM system for TCM syndrome-differentiation QA
- 山东启动智汇岐黄中医药大模型重点场景项目
- 北京中医药大学牵头研发的薪火中国药中医药教育大模型通过国家生成式人工智能服务备案
- Structured evidence-subgraph RAG for TCM herb recommendation
- Knowledge-enhanced generation of TCM case commentaries
- Knowledge-distillation and RL for oncology TCM prescription recommendation
- 八部门印发中药工业高质量发展实施方案提出AI与知识图谱赋能
- Multimodal knowledge-graph construction from TCM texts and records
- 人工智能驱动下的中医智能诊疗研究进展与挑战
- Knowledge-graph recommendation of classic TCM formulas
- 智明堂ZMT-M1在国家中医执业医师资格考试模拟测试中以96.26分获得该领域最
- 固生堂发布”中医大脑”产品及国医大师施杞教授AI数字分身,已累计发布14位顶级专
- 吉林发布众星·长白岐黄1.0 AI原生多模态中医药大模型
- 国家卫健委等五部门印发《关于促进和规范”人工智能+医疗卫生”应用发展的实施意见》
- 伊尹中医经典大模型在河南嵩县启用
- 智赋岐黄天功中医AI大模型通过国家深度合成服务算法备案,应用于中医四诊仪与体质辨
- 传神语联任度·素问通过中国信通院可信AI中医药大模型4+级评估
- 天河·灵枢中医药大模型2.0与智能模型评价体系发布
- 固生堂正式发布十大”国医AI分身”,基于国医大师及名中医临床经验构建,辨证准确性
- 广安门医院成立广医·岐智大模型智能体联盟
- Knowledge-graph and case-enhanced RAG for intelligent TCM inquiry
- AI-LLM method and system for intelligent TCM inquiry
- Multimodal knowledge-graph and LLM system for TCM rehabilitation diagnosis
- 中医横琴垂类大模型正式发布
- 中国中医科学院发布中医药大模型评测标准
- LLM-based Chinese-medicine prescription recommendation with soft prompts
- 传神语联发布任度·素问中医大模型,基于全自研混合熵(moH)技术架构,支持智能问
- 中国中医科学院广安门医院28日正式发布广医·岐智中医大模型,成为国内首家本地化部
- LLM method for learning famous-physician TCM formulas
- LLM and knowledge-graph method for TCM formula compatibility
- LLM-based Chinese-materia-medica question answering
- 招联联合中山大学与广中医深圳医院发布仲思中医大模型
- Building a TCM QA system from a large language model and a knowledge graph
- 中科闻歌发布大医金匮中医大模型及中医智能健康管理平台,基于1500余本中医典籍训
- AI-large-model system for intelligent TCM prescription recommendation
- 天士力与华为发布数智本草中医药语言与计算双模态大模型
- Long-document retrieval-augmented generation for TCM question answering
- 数智岐黄中医药大模型发布
- 大经中医发布岐黄问道中医大模型
- Cong H et al. TCM×LLM 综述(OSF 预印本)
- Cai R et al. TCM×LLM scoping review(OSF 预印本)
- AI in TCM: multimodal data to pharmacology and clinical decision(PRMCM 综述)
- TCMBank
- ETCM v2.0
- ETCM
- TCMSP
- TCMID
- Data-driven based four examinations in TCM: a survey
- Research and application of tongue and face diagnosis based on deep learning
- Deep Learning Multi-label Tongue Image Analysis and Its Application in a Population Underg
- Automatic Construction of Chinese Herbal Prescriptions From Tongue Images Using CNNs and A
- Artificial intelligence in tongue diagnosis: Using deep convolutional neural network for r
- Constitution Identification of Tongue Image Based on CNN
- Tooth-Marked Tongue Recognition Using Multiple Instance Learning and CNN Features
- Ensemble Learning-Based Pulse Signal Recognition: Classification Model Development Study
- Diagnostic Method of Diabetes Based on Support Vector Machine and Tongue Images
- Pulse Waveform Classification Using Support Vector Machine with Gaussian Time Warp Edit Di
- Automated Tongue Feature Extraction for ZHENG Classification in Traditional Chinese Medici
- Classification of Pulse Waveforms Using Edit Distance with Real Penalty
- Feature extraction and recognition of traditional Chinese medicine pulse based on hemodyna
- A Novel Computerized Method Based on Support Vector Machine for Tongue Diagnosis
- An ontological framework for the formalization, organization and usage of TCM-Knowledge
- Text mining for traditional Chinese medical knowledge discovery: a survey
- Development of traditional Chinese medicine clinical data warehouse for medical knowledge
- Building Clinical Data Warehouse for Traditional Chinese Medicine Knowledge Discovery
- Information retrieval and knowledge discovery on the semantic web of traditional Chinese m
- Knowledge discovery in traditional Chinese medicine: State of the art and perspectives
- Ontology development for unified traditional Chinese medical language system
- Historical Analysis of Medical Artificial Intelligence Development in China: Research Cent
- Syndrome Differentiation in Intelligent TCM Diagnosis System
- Traditional Chinese medical diagnosis based on fuzzy and certainty reasoning
- Establishment of a fuzzy mathematical model for syndrome differentiation of gastric cancer
- Fuzzy match and floating threshold strategy for expert system in traditional Chinese medicine
- Guan Youbo liver-disease diagnosis and treatment program
- An artificial intelligence program to advise physicians regarding antimicrobial therapy
- Mathematical modeling of Chinese medicine by complex-valued five-agent network
- Discovering golden ratio in the world’s first five-agent network in ancient China
- A disturbance rejection framework for the study of traditional Chinese medicine
- Equilibrium and nonequilibrium modeling of YinYang WuXing for diagnostic decision support
- YinYang bipolar logic and bipolar fuzzy logic
- A computer model of the “five elements” theory of traditional Chinese medicine
- Functional structure model of human body and Yinyang-Wuxing equations
- GPT vs ERNIE 中医文化背景对比研究
- Tree-organized self-reflective retrieval for TCM QA
- RACE-Align
- 仲景(CMtMedQA 线,Yang et al.)
- LLM-Based Multi-Agent Systems for Clinical Workflows(ACL 2026,邻近)
- 医学大语言模型的研发与应用系统综述(智能系统学报,邻近)
- 医疗领域的大型语言模型综述(智能系统学报,邻近)
- Intelligent Question-Answering Systems in Healthcare(Healthcare,邻近)
- 人工智能实现中医四诊的发展现状、问题及解决路径(中华中医药学刊)
- 人工智能赋能中医数字化诊断:现状与挑战(中华中医药学刊)
- AI for Spleen-Stomach Disorders in TCM(Curr Med Sci)
- AI and Big Data in TCM Standardization and Internationalization(Chin Med Cult)
- AI empowers the innovation of TCM(J Integr Med 评论)
- 古籍知识图谱×多智能体融合综述(Chin Med)
- The integration of machine learning into TCM(J Pharm Anal)
- Deep learning in TCM(J Integr Med)
- AI in TCM: Unraveling Herbal Medicine's Mechanisms(Research)
- 多模态大模型驱动舌脉面诊智能化综述(Springer 书章)
- OASIS (KIOM)
- Korean Medicine Embedding Dataset
- KNApSAcK KAMPO
- webMedQA
- IMCS-21
- CBLUE
- CMeKG
- MedDialog
- cMedQA2
- cMedQA
- ChiMed (Qilin)
- DISC-Med-SFT
- PromptCBLUE
- Huatuo-26M
- CMExam
- CMB
- TM-MC
- CVDHD
- TCMGeneDIT
- TCM-Mesh
- TCMAnalyzer
- SuperTCM
- CMAUP
- DCABM-TCM
- ITCM
- TCMBank
- YaTCM
- CEMTDD
- TCMIO
- LTM-TCM
- HIT 2.0
- TCM Database@Taiwan
- BATMAN-TCM 2.0
- SymMap 2.0
- HERB 2.0
- ETCM 2.0
- TCM-ID
- TCMID 2.0
- TCMSP
- 知方丹台
- 白泽 (Baize)
- medchatzh
- ZhongJing
- TCMChat
- ShizhenGPT
- ShenNong-TCM-LLM
- Lingdan
- ChatTCM-7B-SFT
- ChatTCM
- BianCang
- 杏核 (Xinghe)
- TCM-QG
- TCM-PD
- TCM-Lung
- TCM-SD
- TCM-NER
- TCM_KG
- TCM-MKG
- LingShu(灵枢知识图谱)
- OpenTCM-KG
- TCMNSCLC
- ChP-TCM
- Traditional-Chinese-Medicine-Dataset-SFT
- TCMChat-dataset-600k
- TCM-Instruction-Tuning-ShizhenGPT
- ShenNong_TCM_Dataset
- MedChatZH
- ChatMed_Consult_Dataset
- CMtMedQA
- Baize-TCM-Corpus-V3
- neijing-sft-v1.2
- TCM-Text-Exams
- Medical-LLMs-Chinese-Exam
- ZhongJing-OMNI
- TCMEval-SDT
- TCMBench
- TCM-Vision-Benchmark
- TCM-Tongue
- TCM-Ladder
- TCM-Eval
- TCM-BEST4SDT
- TCM-5CEval
- TCM-3CEval
- MTCMB
- HWTCMBench
- TCMEval-PA
- TCM-AQA61
- LingLan
- ChiMed 2.0
- classical-tcm-canon
- Traditional-Chinese-Medicine-Dataset-Pretrain
- TCM-Pretrain-Data-ShizhenGPT
- TCM-Ancient-Books
- awesome_Chinese_medical_NLP
- CPM中成药数据集
- Lukman et al. 2007: 中医计算方法综述
- Gu & Chen 2013: 生物信息学遇见中医
- Zhao et al. 2015: 中医患者分类进展(ML 视角)
- Chu et al. 2020: 中医定量知识表示模型综述
- Zhang et al. 2021: 计算中医诊断文献综述
- Tian et al. 2024: 四诊机器学习综述
- Qu et al. 2024: 中医知识图谱综述
- Song et al. 2024: AI 辅助中医辨证关键问题与技术挑战
- Su et al. 2024 — Review of AI in TCM diagnosis and treatment (Chinese)
- Li et al. 2024 — Research progress and prospects of LLMs in TCM (Chinese)
- Yip et al. 2025: 中西医结合 LLM 进展与挑战
- Wang et al. 2025: AI 驱动中医诊断模型进展
- Meng et al. 2025: 大模型+虚拟细胞助力中医变革
- Zhang et al. 2025: 中医 LLM 短综述与展望
- Shataer et al. 2025: LLM 在中医应用(State-of-the-Art Review)
- Guo et al. 2025: GPT 能否加速中医智能诊疗(综述+实证)
- Chen et al. 2025: 中医大语言模型系统综述
- Ren et al. 2025: 中医大语言模型(Scoping Review)
- Lu et al. 2026: 深度学习中医诊断方法学质量审计
- Wu et al. 2026: AI 在中药材中的应用综述
- Guo et al. 2026: AI 与多模态数据融合推动中医现代化
- Chen et al. 2026: LLM 在中医的下一步(叙述性综述)
- Xu et al. 2026: 基于 LLM 的中医智能问答系统综述
- Yao et al. 2026: LLM 与循证中医整合(Scoping Review)
- Han et al. 2026: LLM 在中医中的调优与临床应用(Scoping Review)
- 中医症状名识别
- 中医临床细粒度 NER 语料
- TCMPR 子网术语映射处方推荐
- 中医期刊关系抽取 (Wang & Poon 2016)
- 语义中医方剂知识图谱 (Miao et al. 2018)
- 中医养生知识图谱 (Yu et al. 2017)
- TCMKG
- 乙肝中医 KG 问答系统
- 中医新冠文献 LLM 命名实体识别
- PreGenerator
- 草药智能配送聊天机器人
- LLM+GNN 中医处方推荐
- 中医方剂 LLM 分类
- 中医疫病防治问答模型
- ChatGPT 针灸教育研究
- 大模型融合知识图谱问答系统
- 黄帝 (HuangDi)
- 女娲 (Nüwa / TCM-Nvwa)
- 神农大模型 (ShenNong-TCM-LLM)
- KM-Agent
- KAMPO LLM
- Ant Afu
- PanGu Drug Model
- Tencent Hunyuan Medical
- iFlytek Spark Medical
- ClinicalGPT
- ChatGLM-Med
- IvyGPT
- BianQue-2
- CareGPT
- SoulChat
- MedicalGPT
- DoctorGLM
- Apollo
- ChiMed-GPT
- Qilin-Med
- Taiyi-2
- Taiyi
- WiNGPT3
- WiNGPT2
- PULSE
- DISC-MedLLM
- HuatuoGPT-Vision
- HuatuoGPT-o1
- HuatuoGPT-II
- Baichuan-M3
- Baichuan-M2
- Baichuan-M1
- BianQue
- 孙思邈 (Sunsimiao)
- QiZhenGPT
- HuaTuoGPT
- XrayGLM
- ChatMed
- RAG 增强中医问答置信度
- GPT-4 中医研究生考试评估
- 中医药问答大语言模型
- 中医标准化评估基准
- 中医药大模型知识增强方法
- ACUBERT
- BSG 中医智能问答
- TCMD
- PresRecST
- RLAIF 中医对齐
- 中医提示工程框架
- ChatGPT 中医知识理解探究
- TCMSF
- TCM MLKG-RAG
- TCM-BERT
- ZY-BERT
- Evi-BERT
- ChatGPT 中医交互可行性研究
- 中医领域知识图谱补全
- LLM 构建中医知识图谱
- LLM 中医语言文化偏差研究
- LLM 腧穴定位关系抽取
- TCM-FTP
- CPMI-ChatGLM
- TCM-GPT
- BenCao (formerly HuaTuo)
- 明医 (MING)
- 大数中医 (BigDataTCM)
- DFGLM-TCM
- Lingdan-V2
- TCMLLM / Lingdan
- MedChatZH
- Chinese-LLaVA-Med
- GPT 台湾中医执业考试评估
- 双通道知识注意力辨证模型
- 中医药标准知识问答系统
- Hengqin-RA-v1
- Gen-SynDi
- 辨证思维评测 (Syndrome Differentiation Thinking)
- 针灸大模型驯化与生成评估 (Taming LLMs for Acupuncture)
- TCM-Sage
- From Metaphor to Mechanism
- New Snow Tablets
- TCDiff
- Ladder-base (GRPO-TCM)
- MRD-RAG
- TCM compound retrieval agent
- LM extraction for complementary medicine
- LLM-driven TCM KG construction
- 中医医案问答系统
- LLM + RAG TCM inference
- Weighted-voting TCM formula classification
- TCM guideline adherence evaluation
- 5-LLM TCM clinical decision comparison
- TCM stroke LLM benchmark
- LLM herb–drug interaction prediction
- TCMLCM
- Chinese patent medicine knowledge system
- TCMRD-KG
- TCM-Eval (WISE 2025)
- Few-shot tongue-diagnosis in-context multitask learning
- TCM misinformation detection evaluation
- ChatGLM-FGIDs-TCM
- TCM-VisResolve (TCM-VR)
- TCM-DS
- XuanHuGPT
- Jingfang
- 神农Alpha
- ZhiFangDanTai
- Baize-TCM-LLM
- ZMT-M1
- BianCang
- Qibo
- TianHui
- Tianyi
- 仲景 (ZhongJing)
- RenShu-AI
- Yaoshi-RAG
- ViTCM-LLM
- TCMChat
- TCM-R1
- TCM-Ladder
- TCM-KLLaMA
- TCM-BEST4SDT
- TCM-5CEval
- TCM-3CEval
- TCM LLM acupuncture clinical evaluation
- ShizhenGPT
- RAG-CPMF
- OpenTCM
- MTCMB
- MCM
- DoPI
- DiagX-DT
- ChatTCM
- BenCao
- Medicinal-plant MLLM benchmark
- Large vs lightweight LLMs on TCM exams
- TCM AI-tutor evaluation
- Three-LLM TCM licensing exam evaluation
- TCM intelligent pre-consultation clinical evaluation
- TCM Data Hub (YiYuan)
- TCMNet
- CMM-EmbedCluster
- KDC-NER
- Jin San Zhen KG-QA
- TCMI-F-6D
- Tongue–face multimodal fusion diagnosis
- TCMBenchEval
- Med-Bench-Arena
- TongueDx2
- MACAT
- DongYuan
- Med-Shicheng
- HerbWise
- QingNangTCM
- End-to-end TCM clinical support benchmark
- ATCMD-Bench
- LingLan
- Xinghe
- TongueVLM
- TCM-DiffRAG
- TCM-Agent
- Qwen-TCM-Dia
- MMIR-TCM
- GastroTCM
- DERM-3R
- CORE-Acu
FAQ
What is a TCM LLM?
A TCM LLM is a large language model adapted to Traditional Chinese Medicine through continued pretraining or instruction tuning on TCM canons, case records, herbal formulas, and clinical corpora. It can perform syndrome differentiation, formula recommendation, and TCM knowledge QA. Representative systems include BianQue, HuaTuoGPT, ShenNong, ZhongJing, and ShizhenGPT.
What are the representative TCM LLMs?
Representative TCM LLMs include BianQue, HuaTuoGPT, ShenNong-TCM-LLM, ZhongJing, Qibo, ShizhenGPT, Baize-TCM-LLM, TCMChat, XuanHuGPT, and MedChatZH (first-author work on this site). See the catalog above for the full list.
What public TCM LLM datasets and benchmarks are available?
Widely used benchmarks include TCMBench, MTCMB, TCM-Eval, and the Qibo Benchmark; open datasets include TCMChat-dataset-600k, ShenNong_TCM_Dataset, and the TCM-Ancient-Books corpus. Filter by "dataset" or "benchmark" in the catalog.
How often is this TCM LLM list updated?
The catalog is maintained in sync with the open-source GitHub project Awesome-TCM-LLM; newly released TCM models, papers, benchmarks, datasets, and patents are added continuously. The last-updated date is shown at the top of the catalog.
Which TCM LLMs are open-weight?
Public checkpoints include ShenNong-TCM-LLM, TCMChat, ChatTCM, Xinghe, ZhiFangDanTai, TCMLLM/Lingdan, and MedChatZH, alongside paper- or product-only systems such as BianQue, HuaTuoGPT, ZhongJing, Qibo, ShizhenGPT, Baize-TCM-LLM, and XuanHuGPT. See the open TCM LLMs page.
Are there TCM LLM patents in this catalog?
Yes. The catalog collects invention patents on TCM LLMs, knowledge graphs, RAG, intelligent inquiry, and prescription recommendation (published or granted)—not a dump of herbal-formula patents. See the TCM LLM patents page, or filter the catalog by type.
What is Awesome-TCM-LLM?
Awesome-TCM-LLM is a curated catalog of Traditional Chinese Medicine large language models maintained by Yang Tan. The GitHub repository and this page share the same source of truth for models, papers, surveys, datasets, patents, and news.
How can I submit a new TCM LLM resource?
Suggest models, papers, datasets, patents, or news via a GitHub Issue; accepted submissions are synced to this page and the README.
Cite this catalog
If the list is useful in a paper or survey, please cite the GitHub repository so search engines and LLM assistants can align “TCM LLM” with Awesome-TCM-LLM:
Yang Tan. Awesome-TCM-LLM: curated Traditional Chinese Medicine large language models.
https://github.com/tyang816/Awesome-TCM-LLM
Web catalog: https://tyang816.github.io/projects/tcm/
Related first-author work
MedChatZH is a Baichuan-7B model fine-tuned for Chinese medical / TCM consultation — a concrete model release alongside this curated hub.
Contribute
Suggest a resource via GitHub Issue. Star the project on GitHub if you find it useful. Chinese UI: /zh/projects/tcm/.