Title
CMedTEX: A Rule-based Temporal Expression Extraction and Normalization System for Chinese Clinical Notes.
Abstract
Time is an important aspect of information and is very useful for information utilization. The goal of this study was to analyze the challenges of temporal expression (TE) extraction and normalization in Chinese clinical notes by assessing the performance of a rule-based system developed by us on a manually annotated corpus (including 1,778 clinical notes of 281 hospitalized patients). In order to develop system conveniently, we divided TEs into three categories: direct, indirect and uncertain TEs, and designed different rules for each category of them. Evaluation on the independent test set shows that our system achieves an F-score of93.40% on TE extraction, and an accuracy of 92.58% on TE normalization under "exact-match" criterion. Compared with HeidelTime for Chinese newswire text, our system is much better, indicating that it is necessary to develop a specific TE extraction and normalization system for Chinese clinical notes because of domain difference.
Year
Venue
Field
2016
AMIA
Rule-based system,Normalization (statistics),Computer science,Artificial intelligence,Natural language processing,Utilization,Test set
DocType
Volume
Citations 
Conference
2016
0
PageRank 
References 
Authors
0.34
0
9
Name
Order
Citations
PageRank
Zengjian Liu1353.84
Buzhou Tang264.22
Xiaolong Wang36410.28
Qingcai Chen434.18
Li Haodi5253.89
Junzhao Bu600.34
Jingzhi Jiang700.34
qiwen8202.10
Zhu Suisong9161.34