BELB: a biomedical entity linking benchmark.

Bioinformatics

Computer Science Department, Humboldt-Universität zu Berlin, Berlin 10099, Germany.

Published: November 2023

Motivation: Biomedical entity linking (BEL) is the task of grounding entity mentions to a knowledge base (KB). It plays a vital role in information extraction pipelines for the life sciences literature. We review recent work in the field and find that, as the task is absent from existing benchmarks for biomedical text mining, different studies adopt different experimental setups making comparisons based on published numbers problematic. Furthermore, neural systems are tested primarily on instances linked to the broad coverage KB UMLS, leaving their performance to more specialized ones, e.g. genes or variants, understudied.

Results: We therefore developed BELB, a biomedical entity linking benchmark, providing access in a unified format to 11 corpora linked to 7 KBs and spanning six entity types: gene, disease, chemical, species, cell line, and variant. BELB greatly reduces preprocessing overhead in testing BEL systems on multiple corpora offering a standardized testbed for reproducible experiments. Using BELB, we perform an extensive evaluation of six rule-based entity-specific systems and three recent neural approaches leveraging pre-trained language models. Our results reveal a mixed picture showing that neural approaches fail to perform consistently across entity types, highlighting the need of further studies towards entity-agnostic models.

Availability And Implementation: The source code of BELB is available at: https://github.com/sg-wbi/belb. The code to reproduce our experiments can be found at: https://github.com/sg-wbi/belb-exp.

Download full-text PDF

Source
http://www.ncbi.nlm.nih.gov/pmc/articles/PMC10681865PMC
http://dx.doi.org/10.1093/bioinformatics/btad698DOI Listing

Publication Analysis

Top Keywords

biomedical entity
12
entity linking
12
belb biomedical
8
linking benchmark
8
entity types
8
neural approaches
8
entity
6
belb
5
benchmark motivation
4
motivation biomedical
4

Similar Publications

Objective: Extracting PICO elements-Participants, Intervention, Comparison, and Outcomes-from clinical trial literature is essential for clinical evidence retrieval, appraisal, and synthesis. Existing approaches do not distinguish the attributes of PICO entities. This study aims to develop a named entity recognition (NER) model to extract PICO entities with fine granularities.

View Article and Find Full Text PDF

Disruption of developmental processes affecting the fetal lung leads to pulmonary hypoplasia. Pulmonary hypoplasia results from several conditions including congenital diaphragmatic hernia (CDH) and oligohydramnios. Both entities have high morbidity and mortality, and no effective therapy that fully restores normal lung development.

View Article and Find Full Text PDF

While the role of cancer stem cells (CSCs) in tumorigenesis, chemoresistance, metastasis, and relapse has been extensively studied in solid tumors, such as adenocarcinomas or sarcomas, the same cannot be said for neuroendocrine neoplasms (NENs). While lagging, CSCs have been described in numerous NENs, including gastrointestinal and pancreatic NENs (PanNENs), and they have been found to play critical roles in tumor initiation, progression, and treatment resistance. However, it seems that there is still skepticism regarding the role of CSCs in NENs, even in light of studies that support the CSC model in these tumors and the therapeutic benefits of targeting them.

View Article and Find Full Text PDF

The sacrum can harbor a diverse group of both benign and malignant tumors, including metastases. Primary tumors of the sacrum can arise from bone, cartilage, marrow, notochordal remnants, or surrounding nerves and vessels. Among a variety of primary tumors of the spine, chordoma, germ cell tumors and Ewing's sarcoma are recognized for their propensity to occur in the sacrum.

View Article and Find Full Text PDF

Background: Central sensitisation (CS) increases musculoskeletal pain. Quantitative sensory testing (QST) or self-report questionnaires might indicate CS. Indices of CS might be suppressed by exercise, although the optimal exercise regimen remains unclear.

View Article and Find Full Text PDF

Want AI Summaries of new PubMed Abstracts delivered to your In-box?

Enter search terms and have AI summaries delivered each week - change queries or unsubscribe any time!