Text2MDT: Extracting Medical Decision Trees from Medical Texts

Zhu, Wei; Li, Wenfeng; Tian, Xing; Wang, Pengfei; Wang, Xiaoling; Chen, Jin; Wu, Yuanbin; Ni, Yuan; Xie, Guotong

Abstract:Knowledge of the medical decision process, which can be modeled as medical decision trees (MDTs), is critical to build clinical decision support systems. However, the current MDT construction methods rely heavily on time-consuming and laborious manual annotation. In this work, we propose a novel task, Text2MDT, to explore the automatic extraction of MDTs from medical texts such as medical guidelines and textbooks. We normalize the form of the MDT and create an annotated Text-to-MDT dataset in Chinese with the participation of medical experts. We investigate two different methods for the Text2MDT tasks: (a) an end-to-end framework which only relies on a GPT style large language models (LLM) instruction tuning to generate all the node information and tree structures. (b) The pipeline framework which decomposes the Text2MDT task to three subtasks. Experiments on our Text2MDT dataset demonstrate that: (a) the end-to-end method basd on LLMs (7B parameters or larger) show promising results, and successfully outperform the pipeline methods. (b) The chain-of-thought (COT) prompting method \cite{Wei2022ChainOT} can improve the performance of the fine-tuned LLMs on the Text2MDT test set. (c) the lightweight pipelined method based on encoder-based pretrained models can perform comparably with LLMs with model complexity two magnititudes smaller. Our Text2MDT dataset is open-sourced at \url{this https URL}, and the source codes are open-sourced at \url{this https URL}.

Subjects:	Computation and Language (cs.CL)
Cite as:	arXiv:2401.02034 [cs.CL]
	(or arXiv:2401.02034v1 [cs.CL] for this version)
	https://doi.org/10.48550/arXiv.2401.02034

Computer Science > Computation and Language

Title:Text2MDT: Extracting Medical Decision Trees from Medical Texts

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators