Automatic ICD-10 coding algorithm using an improved longest common subsequence based on semantic similarity

ICD-10(International Classification of Diseases 10th revision) is a classification of a disease, symptom, procedure, or injury. Diseases are often described in patients' medical records with free texts, such as terms, phrases and paraphrases, which differ significantly from those used in ICD-10 classification. This paper presents an improved approach based on the Longest Common Subsequence (LCS) and semantic similarity for automatic Chinese diagnoses, mapping from the disease names given by clinician to the disease names in ICD-10. LCS refers to the longest string that is a subsequence of every member of a given set of strings. The proposed method of improved LCS in this paper can increase the accuracy of processing in Chinese disease mapping.

Medienart:

E-Artikel

Erscheinungsjahr:

2017

Erschienen:

2017

Enthalten in:

Zur Gesamtaufnahme - volume:12

Enthalten in:

PloS one - 12(2017), 3 vom: 10., Seite e0173410

Sprache:

Englisch

Beteiligte Personen:

Chen, YunZhi [VerfasserIn]
Lu, HuiJuan [VerfasserIn]
Li, LanJuan [VerfasserIn]

Links:

Volltext

Themen:

Journal Article

Anmerkungen:

Date Completed 07.09.2017

Date Revised 08.02.2019

published: Electronic-eCollection

Citation Status MEDLINE

doi:

10.1371/journal.pone.0173410

funding:

Förderinstitution / Projekttitel:

PPN (Katalog-ID):

NLM269965203