An efficient representation of chronological events in medical texts

Kormilitzin, Andrey; Vaci, Nemanja; Liu, Qiang; Ni, Hao; Nenadic, Goran; Nevado-Holgado, Alejo

Computer Science > Computation and Language

arXiv:2010.08433 (cs)

[Submitted on 16 Oct 2020 (v1), last revised 24 Oct 2020 (this version, v2)]

Title:An efficient representation of chronological events in medical texts

Authors:Andrey Kormilitzin, Nemanja Vaci, Qiang Liu, Hao Ni, Goran Nenadic, Alejo Nevado-Holgado

View PDF

Abstract:In this work we addressed the problem of capturing sequential information contained in longitudinal electronic health records (EHRs). Clinical notes, which is a particular type of EHR data, are a rich source of information and practitioners often develop clever solutions how to maximise the sequential information contained in free-texts. We proposed a systematic methodology for learning from chronological events available in clinical notes. The proposed methodological {\it path signature} framework creates a non-parametric hierarchical representation of sequential events of any type and can be used as features for downstream statistical learning tasks. The methodology was developed and externally validated using the largest in the UK secondary care mental health EHR data on a specific task of predicting survival risk of patients diagnosed with Alzheimer's disease. The signature-based model was compared to a common survival random forest model. Our results showed a 15.4$\%$ increase of risk prediction AUC at the time point of 20 months after the first admission to a specialist memory clinic and the signature method outperformed the baseline mixed-effects model by 13.2 $\%$.

Comments:	4 pages, 2 figures, 7 tables
Subjects:	Computation and Language (cs.CL); Information Retrieval (cs.IR)
Cite as:	arXiv:2010.08433 [cs.CL]
	(or arXiv:2010.08433v2 [cs.CL] for this version)
	https://doi.org/10.48550/arXiv.2010.08433

Submission history

From: Andrey Kormilitzin [view email]
[v1] Fri, 16 Oct 2020 14:54:29 UTC (96 KB)
[v2] Sat, 24 Oct 2020 21:52:03 UTC (96 KB)

Computer Science > Computation and Language

Title:An efficient representation of chronological events in medical texts

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computation and Language

Title:An efficient representation of chronological events in medical texts

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators