Genomics Inform.  2021 Sep;19(3):e24. 10.5808/gi.21008.

COVID-19 recommender system based on an annotated multilingual corpus

Affiliations
  • 1LASIGE, Faculdade de Ciências, Universidade de Lisboa, Portugal
  • 2CENTRA, Faculdade de Ciências, Universidade de Lisboa, Portugal
  • 3Shifa College of Medicine, STMU, Islamabad, Pakistan
  • 4Working Group 3, COST Action EVidence-Based RESearch (EVBRES)

Abstract

Tracking the most recent advances in Coronavirus disease 2019 (COVID-19)‒related research is essential, given the disease's novelty and its impact on society. However, with the publication pace speeding up, researchers and clinicians require automatic approaches to keep up with the incoming information regarding this disease. A solution to this problem requires the development of text mining pipelines; the efficiency of which strongly depends on the availability of curated corpora. However, there is a lack of COVID-19‒related corpora, even more, if considering other languages besides English. This project's main contribution was the annotation of a multilingual parallel corpus and the generation of a recommendation dataset (EN-PT and EN-ES) regarding relevant entities, their relations, and recommendation, providing this resource to the community to improve the text mining research on COVID-19‒related literature. This work was developed during the 7th Biomedical Linked Annotation Hackathon (BLAH7).

Keyword

COVID-19; entity extraction; recommendation; relation extraction; text mining
Full Text Links
  • GNI
Actions
Cited
CITED
export Copy
Close
Share
  • Twitter
  • Facebook
Similar articles
Copyright © 2024 by Korean Association of Medical Journal Editors. All rights reserved.     E-mail: koreamed@kamje.or.kr