Toward a Coronavirus Knowledge Graph

Zhang, Peng; Bu, Yi; Jiang, Peng; Shi, Xiaowen; Lun, Bing; Chen, Chongyan; Syafiandini, Arida Ferti; Ding, Ying; Song, Min

doi:10.3390/genes12070998

Citation Details

Toward a Coronavirus Knowledge Graph

This study builds a coronavirus knowledge graph (KG) by merging two information sources. The first source is Analytical Graph (AG), which integrates more than 20 different public datasets related to drug discovery. The second source is CORD-19, a collection of published scientific articles related to COVID-19. We combined both chemo genomic entities in AG with entities extracted from CORD-19 to expand knowledge in the COVID-19 domain. Before populating KG with those entities, we perform entity disambiguation on CORD-19 collections using Wikidata. Our newly built KG contains at least 21,700 genes, 2500 diseases, 94,000 phenotypes, and other biological entities (e.g., compound, species, and cell lines). We define 27 relationship types and use them to label each edge in our KG. This research presents two cases to evaluate the KG’s usability: analyzing a subgraph (ego-centered network) from the angiotensin-converting enzyme (ACE) and revealing paths between biological entities (hydroxychloroquine and IL-6 receptor; chloroquine and STAT1). The ego-centered network captured information related to COVID-19. We also found significant COVID-19-related information in top-ranked paths with a depth of three based on our path evaluation. more »

Award ID(s):: 2028717

PAR ID:: 10312660

Author(s) / Creator(s):: Zhang, Peng; Bu, Yi; Jiang, Peng; Shi, Xiaowen; Lun, Bing; Chen, Chongyan; Syafiandini, Arida Ferti; Ding, Ying; Song, Min

Date Published:: 2021-07-01

Journal Name:: Genes

Volume:: 12

Issue:: 7

ISSN:: 2073-4425

Format(s):: Medium: X

Sponsoring Org:: National Science Foundation

Free Publicly Accessible Full Text
Accepted Manuscript1.0
Journal Article:
https://doi.org/10.3390/genes12070998

More Like this