Asset Details
MbrlCatalogueTitleDetail
Do you wish to reserve the book?
Generating Realistic Synthetic Patient Cohorts: Enforcing Statistical Distributions, Correlations, and Logical Constraints
by
ElFass, Kareem
, Fasseeh, Ahmad Nader
, Vokó, Zoltán
, Imre, Attila
, Nagy, Balázs
, Nagy, Dávid
, Ashmawy, Rasha
, Hren, Rok
, Németh, Bertalan
in
Analysis
/ Cholesky decomposition
/ Clinical trials
/ Correlation
/ correlation enforcement
/ Datasets
/ Decomposition
/ Distribution (Probability theory)
/ economic modeling
/ Electronic health records
/ Evidence-based medicine
/ individual patient simulation (IPS)
/ Medical research
/ Medicine, Experimental
/ Methods
/ patient cohort generation
/ Pediatrics
/ Simulation
/ Simulation methods
/ Statistical analysis
/ Statistical distributions
/ Synthetic data
/ synthetic patient data
/ Variables
2025
Hey, we have placed the reservation for you!
By the way, why not check out events that you can attend while you pick your title.
You are currently in the queue to collect this book. You will be notified once it is your turn to collect the book.
Oops! Something went wrong.
Looks like we were not able to place the reservation. Kindly try again later.
Are you sure you want to remove the book from the shelf?
Generating Realistic Synthetic Patient Cohorts: Enforcing Statistical Distributions, Correlations, and Logical Constraints
by
ElFass, Kareem
, Fasseeh, Ahmad Nader
, Vokó, Zoltán
, Imre, Attila
, Nagy, Balázs
, Nagy, Dávid
, Ashmawy, Rasha
, Hren, Rok
, Németh, Bertalan
in
Analysis
/ Cholesky decomposition
/ Clinical trials
/ Correlation
/ correlation enforcement
/ Datasets
/ Decomposition
/ Distribution (Probability theory)
/ economic modeling
/ Electronic health records
/ Evidence-based medicine
/ individual patient simulation (IPS)
/ Medical research
/ Medicine, Experimental
/ Methods
/ patient cohort generation
/ Pediatrics
/ Simulation
/ Simulation methods
/ Statistical analysis
/ Statistical distributions
/ Synthetic data
/ synthetic patient data
/ Variables
2025
Oops! Something went wrong.
While trying to remove the title from your shelf something went wrong :( Kindly try again later!
Do you wish to request the book?
Generating Realistic Synthetic Patient Cohorts: Enforcing Statistical Distributions, Correlations, and Logical Constraints
by
ElFass, Kareem
, Fasseeh, Ahmad Nader
, Vokó, Zoltán
, Imre, Attila
, Nagy, Balázs
, Nagy, Dávid
, Ashmawy, Rasha
, Hren, Rok
, Németh, Bertalan
in
Analysis
/ Cholesky decomposition
/ Clinical trials
/ Correlation
/ correlation enforcement
/ Datasets
/ Decomposition
/ Distribution (Probability theory)
/ economic modeling
/ Electronic health records
/ Evidence-based medicine
/ individual patient simulation (IPS)
/ Medical research
/ Medicine, Experimental
/ Methods
/ patient cohort generation
/ Pediatrics
/ Simulation
/ Simulation methods
/ Statistical analysis
/ Statistical distributions
/ Synthetic data
/ synthetic patient data
/ Variables
2025
Please be aware that the book you have requested cannot be checked out. If you would like to checkout this book, you can reserve another copy
We have requested the book for you!
Your request is successful and it will be processed during the Library working hours. Please check the status of your request in My Requests.
Oops! Something went wrong.
Looks like we were not able to place your request. Kindly try again later.
Generating Realistic Synthetic Patient Cohorts: Enforcing Statistical Distributions, Correlations, and Logical Constraints
Journal Article
Generating Realistic Synthetic Patient Cohorts: Enforcing Statistical Distributions, Correlations, and Logical Constraints
2025
Request Book From Autostore
and Choose the Collection Method
Overview
Large, high-quality patient datasets are essential for applications like economic modeling and patient simulation. However, real-world data is often inaccessible or incomplete. Synthetic patient data offers an alternative, and current methods often fail to preserve clinical plausibility, real-world correlations, and logical consistency. This study presents a patient cohort generator designed to produce realistic, statistically valid synthetic datasets. The generator uses predefined probability distributions and Cholesky decomposition to reflect real-world correlations. A dependency matrix handles variable relationships in the right order. Hard limits block unrealistic values, and binary variables are set using percentiles to match expected rates. Validation used two datasets, NHANES (2021–2023) and the Framingham Heart Study, evaluating cohort diversity (general, cardiac, low-dimensional), data sparsity (five correlation scenarios), and model performance (MSE, RMSE, R2, SSE, correlation plots). Results demonstrated strong alignment with real-world data in central tendency, dispersion, and correlation structures. Scenario A (empirical correlations) performed best (R2 = 86.8–99.6%, lowest SSE and MAE). Scenario B (physician-estimated correlations) also performed well, especially in a low-dimensions population (R2 = 80.7%). Scenario E (no correlation) performed worst. Overall, the proposed model provides a scalable, customizable solution for generating synthetic patient cohorts, supporting reliable simulations and research when real-world data is limited. While deep learning approaches have been proposed for this task, they require access to large-scale real datasets and offer limited control over statistical dependencies or clinical logic. Our approach addresses this gap.
Publisher
MDPI AG
This website uses cookies to ensure you get the best experience on our website.