Unknown

Dataset Information

0

Automatic ICD-10 coding algorithm using an improved longest common subsequence based on semantic similarity.


ABSTRACT: ICD-10(International Classification of Diseases 10th revision) is a classification of a disease, symptom, procedure, or injury. Diseases are often described in patients' medical records with free texts, such as terms, phrases and paraphrases, which differ significantly from those used in ICD-10 classification. This paper presents an improved approach based on the Longest Common Subsequence (LCS) and semantic similarity for automatic Chinese diagnoses, mapping from the disease names given by clinician to the disease names in ICD-10. LCS refers to the longest string that is a subsequence of every member of a given set of strings. The proposed method of improved LCS in this paper can increase the accuracy of processing in Chinese disease mapping.

SUBMITTER: Chen Y 

PROVIDER: S-EPMC5356997 | biostudies-literature |

REPOSITORIES: biostudies-literature

Similar Datasets

| S-EPMC7157985 | biostudies-literature
| S-EPMC6180005 | biostudies-literature
| S-EPMC2944781 | biostudies-literature
| S-EPMC8432850 | biostudies-literature
| S-EPMC7079900 | biostudies-literature
| S-EPMC8783617 | biostudies-literature
| S-EPMC9460974 | biostudies-literature
| S-EPMC6320017 | biostudies-literature
| S-EPMC8205565 | biostudies-literature
| S-EPMC7830822 | biostudies-literature