Menopause and Big Data: Word Adjacency Graph Modeling of Menopause-Related ChaCha® Data

We used Word Adjacency Graph (WAG) modeling to detect clusters and visualize the range of menopause-related topics and their mutual proximity. The subset of relevant queries was fully modeled. We split each query into token words (ie, meaningful words and phrases) and removed stopwords (ie, not meaningful functional words). The remaining words were considered in sequence to build summary tables of words and two and three-word phrases. Phrases occurring at least 10 times were used to build a network graph model that was iteratively refined by observing and removing clusters of unrelated content. RESULTS:

We identified two menopause-related subsets of queries by searching for questions containing menopause and menopause-related terms (eg, climacteric, hot flashes, night sweats, hormone replacement). The first contained 263,363 queries from individuals aged 13 and older and the second contained 5,892 queries from women aged 40 to 62 years. In the first set, we identified 12 topic clusters: 6 relevant to menopause and 6 less relevant. In the second set, we identified 15 topic clusters: 11 relevant to menopause and 4 less relevant. Queries about hormones were pervasive within both WAG models. Many of the queries reflected low literacy levels and/or feelings of embarrassment. CONCLUSIONS:

We modeled menopause-related queries posed by ChaCha users between 2009 and 2012. ChaCha data may be used on its own or in combination with other Big Data sources to identify patient-driven educational needs and create patient-centered interventions.

Keywords

Adolescent, Adult, Climacteric, Estrogen Replacement Therapy, Hot Flashes, Information Storage and Retrieval, Menopause, Middle Aged, Models, Theoretical, Terminology as Topic

Cite As

Carpenter, J. S., Groves, D., Chen, C. X., Otte, J. L., & Miller, W. R. (2017). Menopause and big data: Word Adjacency Graph modeling of menopause-related ChaCha data. Menopause (New York, N.Y.), 24(7), 783–788. doi:10.1097/GME.0000000000000833

Journal

Menopause

Rights

Publisher Policy

Source

PMC

Type

Article

Permanent Link

https://hdl.handle.net/1805/19295

DOI

https://doi.org/10.1097/GME.0000000000000833

Version

Author's manuscript

Collections

Open Access Policy Articles
IU School of Nursing Works
Wendy Miller

Full item page