Resources


Database Credentialed Access

NCH Sleep DataBank: A Large Collection of Real-world Pediatric Sleep Studies with Longitudinal Clinical Data

Harlin Lee, Boyue Li, Yungui Huang, Yuejie Chi, Simon Lin

The NCH Sleep DataBank includes 3,984 pediatric sleep studies on 3,673 unique patients conducted at Nationwide Children's Hospital between 2017 and 2019. It contains polysomnography (PSG), clinical annotations, and longitudinal clinical data.

eeg ehr polysomnography pediatrics clinical decision support sleep disorders sleep study electronic health records ecg

Published: Oct. 27, 2021. Version: 3.1.0


Database Open Access

Wearable-based signals during physical exercises from patients with frailty after open-heart surgery

Daivaras Sokas, Monika Butkuvienė, Egle Tamulevičiūtė-Prascienė, Aurelija Beigienė, Raimondas Kubilius, Andrius Petrėnas, Birutė Paliakaitė

A data collection contains a wearable-based electrocardiogram and triaxial acceleration signals of 80 elderly patients with frailty after an open-heart surgery. The signals were collected while the patients were performing a series of exercise tests.

aging heart rate response frailty posture veloergometry timed up and go stair climbing physical exercise wearable heart rate reserve rehabilitation heart rate accelerometer gait electrocardiogram walk test balance

Published: March 31, 2022. Version: 1.0.0

Visualize waveforms

Database Credentialed Access

MIMIC-IV-Note: Deidentified free-text clinical notes

Alistair Johnson, Tom Pollard, Steven Horng, Leo Anthony Celi, Roger Mark

Deidentified free-text clinical notes for patients in the MIMIC-IV Clinical Database.

mimic deidentification critical care electronic health record clinical notes natural language processing

Published: Jan. 6, 2023. Version: 2.2


Database Open Access

NInFEA: Non-Invasive Multimodal Foetal ECG-Doppler Dataset for Antenatal Cardiology Research

Danilo Pani, Eleonora Sulas, Monica Urru, Reza Sameni, Luigi Raffo, Roberto Tumbarello

Open dataset featuring non-invasive electrophysiological recordings, fetal pulsed-wave Doppler and maternal respiration signals. It provides a ground truth on the fetal heart activity when an invasive scalp lead is unavailable.

foetus pwd doppler foetal ecg maternal ecg pwd envelope non-invasive cardiology early pregnancy antenatal fecg ecg

Published: Nov. 12, 2020. Version: 1.0.0

Visualize waveforms

Database Contributor Review

BRATECA (Brazilian Tertiary Care Dataset): a Clinical Information Dataset for the Portuguese Language

Henrique Dias, Ana Helena Dias Pereira dos Ulbrich

Brazilian clinical dataset containing over 70,000 admissions from 10 hospitals in two Brazilian states.

prescriptions exams tertiary care clinical notes natural language processing

Published: July 14, 2022. Version: 1.1


Database Credentialed Access

MIMIC-III Clinical Database

Alistair Johnson, Tom Pollard, Roger Mark

MIMIC-III is a large, freely-available database comprising deidentified health-related data associated with over forty thousand patients who stayed in critical care units of the Beth Israel Deaconess Medical Center between 2001 and 2012. The databas…

clinical intensive care machine learning critical care natural language processing

Published: Sept. 4, 2016. Version: 1.4


Database Credentialed Access

RadGraph: Extracting Clinical Entities and Relations from Radiology Reports

Saahil Jain, Ashwin Agrawal, Adriel Saporta, Steven QH Truong, Du Nguyen Duong, Tan Bui, Pierre Chambon, Matthew Lungren, Andrew Ng, Curtis Langlotz, Pranav Rajpurkar

RadGraph is a dataset of entities and relations in full-text chest X-ray radiology reports, which are obtained using a novel information extraction (IE) schema to capture clinically relevant information in a radiology report.

entity and relation extraction graph multi-modal radiology natural language processing

Published: June 3, 2021. Version: 1.0.0


Database Credentialed Access

AMR-UTI: Antimicrobial Resistance in Urinary Tract Infections

Michael Oberst, Soorajnath Boominathan, Helen Zhou, Sanjat Kanjilal, David Sontag

AMR-UTI is a freely accessible dataset, derived from electronic health record (EHR) information on over 100,000 urinary tract infections (UTI) treated at Massachusetts General Hospital and Brigham & Women's Hospital in Boston, MA, USA.

antibiotic resistance causal inference policy learning antimicrobial resistance urinary tract infection clinical decision support machine learning

Published: Nov. 4, 2020. Version: 1.0.0


Database Restricted Access

Upper body thermal images and associated clinical data from a pilot cohort study of COVID-19

Jose Tamez-Peña, Adam Yala, Servando Cardona, Rocio Ortiz-Lopez, Victor Trevino

Thermal videos of people with positive and negative COVID-19 tests.

covid-19 thermal videos sars-cov-2 clinical symptoms

Published: Aug. 16, 2021. Version: 1.1


Database Credentialed Access

Annotated Question-Answer Pairs for Clinical Notes in the MIMIC-III Database

Xiang Yue, Xinliang Frederick Zhang, Huan Sun

Annotated Question Answering Pairs for Clinical Notes in the MIMIC-III Database

clinical question answering clinical nlp clinical reading comprehension

Published: Jan. 15, 2021. Version: 1.0.0