Back to Search Start Over

CMedTEX: A Rule-based Temporal Expression Extraction and Normalization System for Chinese Clinical Notes

Authors :
Zengjian, Liu
Buzhou, Tang
Xiaolong, Wang
Qingcai, Chen
Haodi, Li
Junzhao, Bu
Jingzhi, Jiang
Qiwen, Deng
Suisong, Zhu
Source :
AMIA ... Annual Symposium proceedings. AMIA Symposium. 2016
Publication Year :
2017

Abstract

Time is an important aspect of information and is very useful for information utilization. The goal of this study was to analyze the challenges of temporal expression (TE) extraction and normalization in Chinese clinical notes by assessing the performance of a rule-based system developed by us on a manually annotated corpus (including 1,778 clinical notes of 281 hospitalized patients). In order to develop system conveniently, we divided TEs into three categories: direct, indirect and uncertain TEs, and designed different rules for each category of them. Evaluation on the independent test set shows that our system achieves an F-score of93.40% on TE extraction, and an accuracy of 92.58% on TE normalization under “exact-match” criterion. Compared with HeidelTime for Chinese newswire text, our system is much better, indicating that it is necessary to develop a specific TE extraction and normalization system for Chinese clinical notes because of domain difference.

Details

ISSN :
1942597X
Volume :
2016
Database :
OpenAIRE
Journal :
AMIA ... Annual Symposium proceedings. AMIA Symposium
Accession number :
edsair.pmid..........a5ca0314fb927c86f46d14c8b0eaac4a