Resources


Database Open Access

PTB-XL, a large publicly available electrocardiography dataset

Patrick Wagner, Nils Strodthoff, Ralf-Dieter Bousseljot, et al.

The PTB-XL ECG dataset is a large dataset of 21801 clinical 12-lead ECGs from 18869 patients of 10 second length. The raw signal data has been annotated by up to two cardiologists with 71 different ECG statements and is supplemented by rich metadata.

ptb-xl ptb ecg electrocardiography

Published: Nov. 9, 2022. Version: 1.0.3

Visualize waveforms

Database Credentialed Access

BRAX, a Brazilian labeled chest X-ray dataset

Eduardo Pontes Reis, Joselisa Paiva, Maria Carolina Bueno da Silva, et al.

BRAX contains 24,959 chest radiography exams and 40,967 images acquired in a large general Brazilian hospital. All images have been read by trained radiologists and 14 labels were derived from Brazilian Portuguese reports using NLP.

chest x-ray dataset artificial intelligence

Published: June 17, 2022. Version: 1.1.0


Database Open Access

Continuous Cuffless Monitoring of Arterial Blood Pressure via Graphene Bioimpedance Tattoos

Bassem Ibrahim, Dmitry Kireev, Kaan Sel, et al.

Cuffless blood pressure data repository that includes raw time data for 4-channel Bioimpedance signals using Graphene Tattoos from the wrist with synchronized continuous blood pressure and PPG signals from 7 subjects

blood pressure graphene bioimpedance ppg

Published: June 4, 2022. Version: 1.0.0


Database Credentialed Access

Paediatric Intensive Care database

Haomin Li, Xian Zeng, Gang Yu

PIC (Paediatric Intensive Care) is a large paediatric-specific, single-centre, bilingual database comprising information relating to children admitted to critical care units at a large children’s hospital in China.

intensive care pediatrics critical care natural language processing

Published: Nov. 12, 2020. Version: 1.1.0


Database Credentialed Access

MIMIC-III - SequenceExamples for TensorFlow modeling

Jonas Kemp, Kun Zhang, Andrew Dai

MIMIC-III data converted into TensorFlow SequenceExample format, for use in modeling pipelines.

tensorflow sequence modeling machine learning deep learning

Published: Sept. 29, 2020. Version: 1.0.0


Software Open Access

Random Search Toolbox

A major issue with many signal processing and machine learning algorithms is the lack of optimisation methods for determining the numerous hyper-parameters associated with the model as well as the knowledge of which hyper-parameters are relevant. Th…

Published: Nov. 4, 2014. Version: 1.0.0


Database Contributor Review

Salzburg Intensive Care database (SICdb), a freely accessible intensive care database

Niklas Rodemund, Andreas Kokoefer, Bernhard Wernly, et al.

The SICdb dataset, version 1.0.8 contains 27350 admissions to an ICU in an Austrian tertiary care institution.

clinical intensive care critical care machine learning open data

Published: Sept. 10, 2024. Version: 1.0.8


Database Open Access

A multi-camera and multimodal dataset for posture and gait analysis

Manuel Palermo, João Mendes Lopes, João André, et al.

Multimodal dataset with 166k samples for vision-based applications with a smart walker used in gait and posture rehabilitation. It is equipped with a pair of Depth cameras with data synchronized with an inertial MoCap system worn by the participant.

computer vision inertial motion capture smart walker gait and posture analysis depth rehabilitation human pose estimation deep learning

Published: Nov. 1, 2021. Version: 1.0.0


Database Open Access

Wide-field calcium imaging sleep state database

Eric Landsness, Xiaohui Zhang, Wei Chen, et al.

Wide-field calcium imaging database that consists of annotated sleep recording collected from transgenic mice at Washington University of St Louis School of Medicine.

sleep wide-field calcium imaging sleep state classification sleep staging machine learning

Published: March 17, 2022. Version: 1.0.1


Database Credentialed Access

Predictors of Hospital Onset Infection: A Matched Retrospective Cohort Dataset

Ziming Wei, Luke Sagers, Caroline McKenna, et al.

NPA-CP is a freely accessible dataset derived from electronic health record (EHR) information at MGB between 2015 and 2024. The dataset includes 11 different pathogens and can be used to predict hospital-onset infections for these pathogens.

electronic health records infection control clinical machine learning infectious diseases hospital onset infection colonization pressure

Published: Nov. 4, 2025. Version: 1.0.0