Back to Search Start Over

Unsupervised information extraction from italian clinical records.

Authors :
Alicante A
Corazza A
Isgrò F
Silvestri S
Source :
Studies in health technology and informatics [Stud Health Technol Inform] 2014; Vol. 207, pp. 340-9.
Publication Year :
2014

Abstract

This paper discusses the application of an unsupervised text mining technique for the extraction of information from clinical records in Italian. The approach includes two steps. First of all, a metathesaurus is exploited together with natural language processing tools to extract the domain entities. Then, clustering is applied to explore relations between entity pairs. The results of a preliminary experiment, performed on the text extracted from 57 medical records containing more than 20,000 potential relations, show how the clustering should be based on the cosine similarity distance rather than the City Block or Hamming ones.

Details

Language :
English
ISSN :
1879-8365
Volume :
207
Database :
MEDLINE
Journal :
Studies in health technology and informatics
Publication Type :
Academic Journal
Accession number :
25488240