Skip to main content
medRxiv
  • Home
  • About
  • Submit
  • ALERTS / RSS
Advanced Search

Subtyping of common complex diseases and disorders by integrating heterogeneous data. Identifying clusters among women with lower urinary tract symptoms in the LURN study

Victor P. Andreev, Margaret E. Helmuth, Gang Liu, Abigail R. Smith, Robert M. Merion, Claire C. Yang, Anne P. Cameron, J. Eric Jelovsek, Cindy L. Amundsen, Brian T. Helfand, Catherine S. Bradley, John O. L. DeLancey, James W. Griffith, Alexander P. Glaser, Brenda W. Gillespie, J. Quentin Clemens, H. Henry Lai, the LURN Study Group
doi: https://doi.org/10.1101/2021.09.17.21263124
Victor P. Andreev
1Arbor Research Collaborative for Health, Ann Arbor, Michigan, United States of America
  • Find this author on Google Scholar
  • Find this author on PubMed
  • Search for this author on this site
  • For correspondence: victor.andreev{at}arborresearch.org
Margaret E. Helmuth
1Arbor Research Collaborative for Health, Ann Arbor, Michigan, United States of America
  • Find this author on Google Scholar
  • Find this author on PubMed
  • Search for this author on this site
Gang Liu
2Department of Computational Medicine and Bioinformatics, University of Michigan, Ann Arbor, Michigan, United States of America
  • Find this author on Google Scholar
  • Find this author on PubMed
  • Search for this author on this site
Abigail R. Smith
1Arbor Research Collaborative for Health, Ann Arbor, Michigan, United States of America
  • Find this author on Google Scholar
  • Find this author on PubMed
  • Search for this author on this site
Robert M. Merion
1Arbor Research Collaborative for Health, Ann Arbor, Michigan, United States of America
  • Find this author on Google Scholar
  • Find this author on PubMed
  • Search for this author on this site
Claire C. Yang
3Department of Urology, University of Washington, Seattle, Washington, United States of America
  • Find this author on Google Scholar
  • Find this author on PubMed
  • Search for this author on this site
Anne P. Cameron
4Department of Urology, University of Michigan, Ann Arbor, Michigan, United States of America
  • Find this author on Google Scholar
  • Find this author on PubMed
  • Search for this author on this site
J. Eric Jelovsek
5Department of Obstetrics and Gynecology, Duke University, Raleigh, North Carolina, United States of America
  • Find this author on Google Scholar
  • Find this author on PubMed
  • Search for this author on this site
Cindy L. Amundsen
5Department of Obstetrics and Gynecology, Duke University, Raleigh, North Carolina, United States of America
  • Find this author on Google Scholar
  • Find this author on PubMed
  • Search for this author on this site
Brian T. Helfand
6Department of Urology, North Shore University, Evanston, Illinois, United States of America
  • Find this author on Google Scholar
  • Find this author on PubMed
  • Search for this author on this site
Catherine S. Bradley
7Department of Obstetrics and Gynecology, University of Iowa, Iowa City, Iowa, United States of America
  • Find this author on Google Scholar
  • Find this author on PubMed
  • Search for this author on this site
John O. L. DeLancey
4Department of Urology, University of Michigan, Ann Arbor, Michigan, United States of America
  • Find this author on Google Scholar
  • Find this author on PubMed
  • Search for this author on this site
James W. Griffith
8Department of Medical Social Sciences, Northwestern University, Chicago, Illinois, United States of America
  • Find this author on Google Scholar
  • Find this author on PubMed
  • Search for this author on this site
Alexander P. Glaser
6Department of Urology, North Shore University, Evanston, Illinois, United States of America
  • Find this author on Google Scholar
  • Find this author on PubMed
  • Search for this author on this site
Brenda W. Gillespie
9Department of Biostatistics, University of Michigan, Ann Arbor, Michigan, United States of America
  • Find this author on Google Scholar
  • Find this author on PubMed
  • Search for this author on this site
J. Quentin Clemens
4Department of Urology, University of Michigan, Ann Arbor, Michigan, United States of America
  • Find this author on Google Scholar
  • Find this author on PubMed
  • Search for this author on this site
H. Henry Lai
10Department of Surgery, Washington University, St Louis, Missouri, United States of America
  • Find this author on Google Scholar
  • Find this author on PubMed
  • Search for this author on this site
  • Abstract
  • Full Text
  • Info/History
  • Metrics
  • Supplementary material
  • Data/Code
  • Preview PDF
Loading

ABSTRACT

We present a novel methodology for subtyping of persons with a common clinical symptom complex by integrating heterogeneous continuous and categorical data. We illustrate it by clustering women with lower urinary tract symptoms (LUTS), who represent a heterogeneous cohort with overlapping symptoms and multifactorial etiology. Identifying subtypes within this group would potentially lead to better diagnosis and treatment decision-making. Data collected in the Symptoms of Lower Urinary Tract Dysfunction Research Network (LURN), a multi-center prospective observational cohort study, included self-reported urinary and non-urinary symptoms, bladder diaries, and physical examination data for 545 women. Heterogeneity in these multidimensional data required thorough and non-trivial preprocessing, including scaling by controls and weighting to mitigate data redundancy, while the various data types (continuous and categorical) required novel methodology using a weighted Tanimoto indices approach. Data domains only available on a subset of the cohort were integrated using a semi-supervised clustering approach. Novel contrast criterion for determination of the optimal number of clusters in consensus clustering was introduced and compared with existing criteria. Distinctiveness of the clusters was confirmed by using multiple criteria for cluster quality, and by testing for significantly different variables in pairwise comparisons of the clusters. Cluster dynamics were explored by analyzing longitudinal data at 3- and 12-month follow-up. Five distinct clusters of women with LUTS were identified using the developed methodology. The clinical relevance of the identified clusters is discussed and compared with the current conventional approaches to the evaluation of LUTS patients. Rationale and thought process are described for selection of procedures for data preprocessing, clustering, and cluster evaluation. Suggestions are provided for minimum reporting requirements in publications utilizing clustering methodology with multiple heterogeneous data domains.

Competing Interest Statement

The authors have declared no competing interest.

Clinical Trial

NCT02485808

Funding Statement

This study is supported by the National Institute of Diabetes & Digestive & Kidney Diseases through cooperative agreements. Grant Numbers: DK097780 DK097772 DK097779 DK099932 DK100011 DK100017 DK099879. Dr. Andreev Biomarker Ancillary LURN R01 Grant Number: 5R01DK125251. Research reported in this publication was supported at Northwestern University, in part, by the National Institutes of Health National Center for Advancing Translational Sciences. Grant Number: UL1TR001422. The content is solely the responsibility of the authors and does not necessarily represent the official views of the National Institutes of Health.

Author Declarations

I confirm all relevant ethical guidelines have been followed, and any necessary IRB and/or ethics committee approvals have been obtained.

Yes

The details of the IRB/oversight body that provided approval or exemption for the research described are given below:

The authors confirm all relevant ethical guidelines have been followed, and all research has been conducted according to the principles expressed in the Declaration of Helsinki. Informed consent has been obtained from participants. Institutional Review Board (IRB) approval has been obtained from: Ethical and Independent Review Services (E&I) IRB, an Association for the Accreditation of Human Research Protection Programs (AAHRPP) Accredited Board, Registration #IRB 00007807.

All necessary patient/participant consent has been obtained and the appropriate institutional forms have been archived.

Yes

I understand that all clinical trials and any other prospective interventional studies must be registered with an ICMJE-approved registry, such as ClinicalTrials.gov. I confirm that any such study reported in the manuscript has been registered and the trial registration ID is provided (note: if posting a prospective study registered retrospectively, please provide a statement in the trial ID field explaining why the study was not registered in advance).

Yes

I have followed all appropriate research reporting guidelines and uploaded the relevant EQUATOR Network research reporting checklist(s) and other pertinent material as supplementary files, if applicable.

Yes

Data Availability

The data that support the findings of this study are openly available in the NIDDK Central Repository at https://repository.niddk.nih.gov/; please reference the acronym LURN.

https://repository.niddk.nih.gov/

Copyright 
The copyright holder for this preprint is the author/funder, who has granted medRxiv a license to display the preprint in perpetuity. All rights reserved. No reuse allowed without permission.
Back to top
PreviousNext
Posted September 22, 2021.
Download PDF

Supplementary Material

Data/Code
Email

Thank you for your interest in spreading the word about medRxiv.

NOTE: Your email address is requested solely to identify you as the sender of this article.

Enter multiple addresses on separate lines or separate them with commas.
Subtyping of common complex diseases and disorders by integrating heterogeneous data. Identifying clusters among women with lower urinary tract symptoms in the LURN study
(Your Name) has forwarded a page to you from medRxiv
(Your Name) thought you would like to see this page from the medRxiv website.
CAPTCHA
This question is for testing whether or not you are a human visitor and to prevent automated spam submissions.
Share
Subtyping of common complex diseases and disorders by integrating heterogeneous data. Identifying clusters among women with lower urinary tract symptoms in the LURN study
Victor P. Andreev, Margaret E. Helmuth, Gang Liu, Abigail R. Smith, Robert M. Merion, Claire C. Yang, Anne P. Cameron, J. Eric Jelovsek, Cindy L. Amundsen, Brian T. Helfand, Catherine S. Bradley, John O. L. DeLancey, James W. Griffith, Alexander P. Glaser, Brenda W. Gillespie, J. Quentin Clemens, H. Henry Lai, the LURN Study Group
medRxiv 2021.09.17.21263124; doi: https://doi.org/10.1101/2021.09.17.21263124
Twitter logo Facebook logo LinkedIn logo Mendeley logo
Citation Tools
Subtyping of common complex diseases and disorders by integrating heterogeneous data. Identifying clusters among women with lower urinary tract symptoms in the LURN study
Victor P. Andreev, Margaret E. Helmuth, Gang Liu, Abigail R. Smith, Robert M. Merion, Claire C. Yang, Anne P. Cameron, J. Eric Jelovsek, Cindy L. Amundsen, Brian T. Helfand, Catherine S. Bradley, John O. L. DeLancey, James W. Griffith, Alexander P. Glaser, Brenda W. Gillespie, J. Quentin Clemens, H. Henry Lai, the LURN Study Group
medRxiv 2021.09.17.21263124; doi: https://doi.org/10.1101/2021.09.17.21263124

Citation Manager Formats

  • BibTeX
  • Bookends
  • EasyBib
  • EndNote (tagged)
  • EndNote 8 (xml)
  • Medlars
  • Mendeley
  • Papers
  • RefWorks Tagged
  • Ref Manager
  • RIS
  • Zotero
  • Tweet Widget
  • Facebook Like
  • Google Plus One

Subject Area

  • Urology
Subject Areas
All Articles
  • Addiction Medicine (349)
  • Allergy and Immunology (668)
  • Allergy and Immunology (668)
  • Anesthesia (181)
  • Cardiovascular Medicine (2648)
  • Dentistry and Oral Medicine (316)
  • Dermatology (223)
  • Emergency Medicine (399)
  • Endocrinology (including Diabetes Mellitus and Metabolic Disease) (942)
  • Epidemiology (12228)
  • Forensic Medicine (10)
  • Gastroenterology (759)
  • Genetic and Genomic Medicine (4103)
  • Geriatric Medicine (387)
  • Health Economics (680)
  • Health Informatics (2657)
  • Health Policy (1005)
  • Health Systems and Quality Improvement (985)
  • Hematology (363)
  • HIV/AIDS (851)
  • Infectious Diseases (except HIV/AIDS) (13695)
  • Intensive Care and Critical Care Medicine (797)
  • Medical Education (399)
  • Medical Ethics (109)
  • Nephrology (436)
  • Neurology (3882)
  • Nursing (209)
  • Nutrition (577)
  • Obstetrics and Gynecology (739)
  • Occupational and Environmental Health (695)
  • Oncology (2030)
  • Ophthalmology (585)
  • Orthopedics (240)
  • Otolaryngology (306)
  • Pain Medicine (250)
  • Palliative Medicine (75)
  • Pathology (473)
  • Pediatrics (1115)
  • Pharmacology and Therapeutics (466)
  • Primary Care Research (452)
  • Psychiatry and Clinical Psychology (3432)
  • Public and Global Health (6527)
  • Radiology and Imaging (1403)
  • Rehabilitation Medicine and Physical Therapy (814)
  • Respiratory Medicine (871)
  • Rheumatology (409)
  • Sexual and Reproductive Health (410)
  • Sports Medicine (342)
  • Surgery (448)
  • Toxicology (53)
  • Transplantation (185)
  • Urology (165)