Unknown

Dataset Information

0

Bio-AnswerFinder: a system to find answers to questions from biomedical texts.


ABSTRACT: The ever accelerating pace of biomedical research results in corresponding acceleration in the volume of biomedical literature created. Since new research builds upon existing knowledge, the rate of increase in the available knowledge encoded in biomedical literature makes the easy access to that implicit knowledge more vital over time. Toward the goal of making implicit knowledge in the biomedical literature easily accessible to biomedical researchers, we introduce a question answering system called Bio-AnswerFinder. Bio-AnswerFinder uses a weighted-relaxed word mover's distance based similarity on word/phrase embeddings learned from PubMed abstracts to rank answers after question focus entity type filtering. Our approach retrieves relevant documents iteratively via enhanced keyword queries from a traditional search engine. To improve document retrieval performance, we introduced a supervised long short term memory neural network to select keywords from the question to facilitate iterative keyword search. Our unsupervised baseline system achieves a mean reciprocal rank score of 0.46 and Precision@1 of 0.32 on 936 questions from BioASQ. The answer sentences are further ranked by a fine-tuned bidirectional encoder representation from transformers (BERT) classifier trained using 100 answer candidate sentences per question for 492 BioASQ questions. To test ranking performance, we report a blind test on 100 questions that three independent annotators scored. These experts preferred BERT based reranking with 7% improvement on MRR and 13% improvement on Precision@1 scores on average.

SUBMITTER: Ozyurt IB 

PROVIDER: S-EPMC7053013 | biostudies-literature | 2020 Jan

REPOSITORIES: biostudies-literature

altmetric image

Publications

Bio-AnswerFinder: a system to find answers to questions from biomedical texts.

Ozyurt Ibrahim Burak IB   Bandrowski Anita A   Grethe Jeffrey S JS  

Database : the journal of biological databases and curation 20200101


The ever accelerating pace of biomedical research results in corresponding acceleration in the volume of biomedical literature created. Since new research builds upon existing knowledge, the rate of increase in the available knowledge encoded in biomedical literature makes the easy access to that implicit knowledge more vital over time. Toward the goal of making implicit knowledge in the biomedical literature easily accessible to biomedical researchers, we introduce a question answering system c  ...[more]

Similar Datasets

| S-EPMC4015759 | biostudies-literature
| S-EPMC8477173 | biostudies-literature
| S-EPMC8388038 | biostudies-literature
| S-EPMC7756489 | biostudies-literature
| S-EPMC7771841 | biostudies-literature
| S-EPMC3498727 | biostudies-literature
| S-EPMC7532186 | biostudies-literature
| S-EPMC3278165 | biostudies-literature
| S-EPMC3821470 | biostudies-literature