• Pain Med · Nov 2020

    Forty-two Million Ways to Describe Pain: Topic Modeling of 200,000 PubMed Pain-Related Abstracts Using Natural Language Processing and Deep Learning-Based Text Generation.

    • Patrick J Tighe, Bharadwaj Sannapaneni, Roger B Fillingim, Charlie Doyle, Michael Kent, Ben Shickel, and Parisa Rashidi.
    • Department of Anesthesiology, University of Florida College of Medicine, Gainesville, Florida.
    • Pain Med. 2020 Nov 1; 21 (11): 3133-3160.

    ObjectiveRecent efforts to update the definitions and taxonomic structure of concepts related to pain have revealed opportunities to better quantify topics of existing pain research subject areas.MethodsHere, we apply basic natural language processing (NLP) analyses on a corpus of >200,000 abstracts published on PubMed under the medical subject heading (MeSH) of "pain" to quantify the topics, content, and themes on pain-related research dating back to the 1940s.ResultsThe most common stemmed terms included "pain" (601,122 occurrences), "patient" (508,064 occurrences), and "studi-" (208,839 occurrences). Contrarily, terms with the highest term frequency-inverse document frequency included "tmd" (6.21), "qol" (6.01), and "endometriosis" (5.94). Using the vector-embedded model of term definitions available via the "word2vec" technique, the most similar terms to "pain" included "discomfort," "symptom," and "pain-related." For the term "acute," the most similar terms in the word2vec vector space included "nonspecific," "vaso-occlusive," and "subacute"; for the term "chronic," the most similar terms included "persistent," "longstanding," and "long-standing." Topic modeling via Latent Dirichlet analysis identified peak coherence (0.49) at 40 topics. Network analysis of these topic models identified three topics that were outliers from the core cluster, two of which pertained to women's health and obstetrics and were closely connected to one another, yet considered distant from the third outlier pertaining to age. A deep learning-based gated recurrent units abstract generation model successfully synthesized several unique abstracts with varying levels of believability, with special attention and some confusion at lower temperatures to the roles of placebo in randomized controlled trials.ConclusionsQuantitative NLP models of published abstracts pertaining to pain may point to trends and gaps within pain research communities.© The Author(s) 2020. Published by Oxford University Press on behalf of the American Academy of Pain Medicine. All rights reserved. For permissions, please e-mail: journals.permissions@oup.com.

      Pubmed     Full text   Copy Citation     Plaintext  

      Add institutional full text...

    Notes

     
    Knowledge, pearl, summary or comment to share?
    300 characters remaining
    help        
    You can also include formatting, links, images and footnotes in your notes
    • Simple formatting can be added to notes, such as *italics*, _underline_ or **bold**.
    • Superscript can be denoted by <sup>text</sup> and subscript <sub>text</sub>.
    • Numbered or bulleted lists can be created using either numbered lines 1. 2. 3., hyphens - or asterisks *.
    • Links can be included with: [my link to pubmed](http://pubmed.com)
    • Images can be included with: ![alt text](https://bestmedicaljournal.com/study_graph.jpg "Image Title Text")
    • For footnotes use [^1](This is a footnote.) inline.
    • Or use an inline reference [^1] to refer to a longer footnote elseweher in the document [^1]: This is a long footnote..

    hide…

What will the 'Medical Journal of You' look like?

Start your free 21 day trial now.

We guarantee your privacy. Your email address will not be shared.