Title
A SNPshot of PubMed to associate genetic variants with drugs, diseases, and adverse reactions.
Abstract
Genetic factors determine differences in pharmacokinetics, drug efficacy, and drug responses between individuals and sub-populations. Wrong dosages of drugs can lead to severe adverse drug reactions in individuals whose drug metabolism drastically differs from the "assumed average". Databases such as PharmGKB are excellent sources of pharmacogenetic information on enzymes, genetic variants, and drug response affected by changes in enzymatic activity. Here, we seek to aid researchers, database curators, and clinicians in their search for relevant information by automatically extracting these data from literature.We automatically populate a repository of information on genetic variants, relations to drugs, occurrence in sub-populations, and associations with disease. We mine textual data from PubMed abstracts to discover such genotype-phenotype associations, focusing on SNPs that can be associated with variations in drug response. The overall repository covers relations found between genes, variants, alleles, drugs, diseases, adverse drug reactions, populations, and allele frequencies. We cross-reference these data to EntrezGene, PharmGKB, PubChem, and others.The performance regarding entity recognition and relation extraction yields a precision of 90-92% for the major entity types (gene, drug, disease), and 76-84% for relations involving these types. Comparison of our repository to PharmGKB reveals a coverage of 93% of gene-drug associations in PharmGKB and 97% of the gene-variant mappings based on 180,000 PubMed abstracts.http://bioai4core.fulton.asu.edu/snpshot.
Year
DOI
Venue
2012
10.1016/j.jbi.2012.04.006
Journal of Biomedical Informatics
Keywords
Field
DocType
adverse drug reaction,associate genetic variant,adverse reaction,pharmacogenetic information,pubmed abstract,overall repository,drug metabolism,drug efficacy,severe adverse drug reaction,genetic variant,genetic factor,drug response,databases,information extraction,text mining,pharmacogenomics
Pharmacogenetics,Data mining,Disease,Allele frequency,PubChem,PharmGKB,Single-nucleotide polymorphism,Bioinformatics,Drug,Medicine,Pharmacogenomics
Journal
Volume
Issue
ISSN
45
5
1532-0480
Citations 
PageRank 
References 
18
0.80
9
Authors
9
Name
Order
Citations
PageRank
Jörg Hakenberg147223.88
Dmitry Voronov2180.80
Nguyen Ha Vo3454.96
Shanshan Liang4724.05
Saadat Anwar5774.98
Barry Lumpkin6221.61
Robert Leaman791439.98
Luis Tari817813.56
Chitta Baral92353269.58