Linking chemical and disease entities to ontologies by integrating PageRank with extracted relations from literature | Journal of Cheminformatics | Full Text
Background Named Entity Linking systems are a powerful aid to the manual curation of digital libraries, which is getting increasingly costly and inefficient due to the information overload. Models based on the Personalized PageRank (PPR) algorithm are one of the state-of-the-art approaches, but these have low performance when the disambiguation graphs are sparse. Findings This work proposes a Named Entity Linking framework designated by Relation Extraction for Entity Linking (REEL) that uses automatically extracted relations to overcome this limitation. Our method builds a disambiguation graph, where the nodes are the ontology candidates for the entities and the edges are added according to the relations established in the text, which the method extracts automatically. The PPR algorithm and the information content of each ontology are then applied to choose the candidate for each entity that maximises the coherence of the disambiguation graph. We evaluated the method on three gold standards: the subset of the CRAFT corpus with ChEBI annotations (CRAFT-ChEBI), the subset of the BC5CDR corpus with disease annotations from the MEDIC vocabulary (BC5CDR-Diseases) and the subset with chemical annotations from the CTD-Chemical vocabulary (BC5CDR-Chemicals). The F1-Score achieved by REEL was 85.8%, 80.9% and 90.3% in these gold standards, respectively, outperforming baseline approaches. Conclusions We demonstrated that RE tools can improve Named Entity Linking by capturing semantic information expressed in text missing in Knowledge Bases and use it to improve the disambiguation graph of Named Entity Linking models. REEL can be adapted to any text mining pipeline and potentially to any domain, as long as there is an ontology or other knowledge Base available.
Linking chemical and disease entities to ontologies by integrating PageRank with extracted relations from literature Research article Open access Published: 21 September 2020 Volume 12 , article number 57 ( 2020 ) Cite this article You have full access to this open access article Download PDF Save article View saved research Journal of Cheminformatics Aims and scope Submit manuscript Linking chemical and disease entities to ontologies by integrating PageRank with extracted relations from literature Download PDF Abstract Background Named Entity Linking systems are a powerful aid to the manual c
Explore this link on the map →related reading
- Elicit: AI for scientific researchelicit.com
- Elicit: AI for scientific researchelicit.org
- Building a PubMed knowledge graph | Scientific Datanature.com
- B-LBConA: a medical entity disambiguation model based on Bio-LinkBERT and context-aware mechanism | BMC Bioinformatics | Springer Nature Linkbmcbioinformatics.biomedcentral.com
- [PDF] Pynsett: A programmable relation extractor | Semantic Scholarsemanticscholar.org
- Request for new ontology pbpko · Issue #2563 · OBOFoundry/OBOFoundry.github.io · GitHubgithub.com
- Practical Cheminformatics Index - Practical Cheminformaticspatwalters.github.io
- Connecting the dots in early drug discovery at Novartisneo4j.com
- NLPContributionGraph -- Structuring Scholarly NLP Contributions in the Open Research Knowledge Graphncg-task.github.io
- Systematic integration of biomedical knowledge prioritizes drugs for repurposinggit.dhimmel.com
- GitHub - leonweber/pedl: Search the biomedical literature for protein interactions and protein associations · GitHubgithub.com
- biochem4j: Integrated and extensible biochemical knowledge through graph databases | PLOS Onejournals.plos.org