Resources


Database Credentialed Access

MIMIC-Ext-MIMIC-CXR-VQA: A Complex, Diverse, And Large-Scale Visual Question Answering Dataset for Chest X-ray Images

Seongsu Bae, Daeun Kyung, Jaehee Ryu, Eunbyeol Cho, Gyubok Lee, Sunjun Kweon, Jungwoo Oh, Lei JI, Eric Chang, Tackeun Kim, Edward Choi

We introduce MIMIC-Ext-MIMIC-CXR-VQA, a complex, diverse, and large-scale dataset designed for Visual Question Answering (VQA) tasks within the medical domain, focusing primarily on chest radiographs.

question answering machine learning evaluation chest x-ray radiology benchmark electronic health records multimodal deep learning visual question answering

Published: July 19, 2024. Version: 1.0.0


Database Credentialed Access

DrugEHRQA: A Question Answering Dataset on Structured and Unstructured Electronic Health Records For Medicine Related Queries

Jayetri Bardhan, Anthony Colas, Kirk Roberts, Daisy Zhe Wang

DrugEHRQA is a QA dataset containing question-answers from MIMIC-III tables and discharge summaries.

question-answer qa

Published: April 12, 2022. Version: 1.0.0


Database Open Access

bigP3BCI: An Open, Diverse and Machine Learning Ready P300-based Brain-Computer Interface Dataset

Boyla Mainsah, Chance Fleeting, Thomas Balmat, Eric Sellers, Leslie Collins

A collection of data from P300-based brain-computer interface studies.

brain-computer interface electroencephalography ieee p2731 working group standard amyotrophic lateral sclerosis p300 speller p300 event related potential oddball paradigm error-related potential

Published: May 19, 2025. Version: 1.0.0


Software Open Access

PhysioTag: An Open-Source Platform for Collaborative Annotation of Physiological Waveforms

Lucas McCullum, Benjamin Moody, Hasan Saeed, Tom Pollard, Xavier Borrat Frigola, Li-wei Lehman, Roger Mark

Platform for collaborative and interactive annotation of physiological waveform data.

annotation

Published: April 25, 2023. Version: 1.0.0


Database Open Access

PTB-XL+, a comprehensive electrocardiographic feature dataset

Nils Strodthoff, Temesgen Mehari, Claudia Nagel, Philip Aston, Ashish Sundar, Claus Graff, Jørgen Kanters, Wilhelm Haverkamp, Olaf Doessel, Axel Loewe, Markus Bär, Tobias Schaeffter

ECG feature dataset accompanying the PTB-XL ECG dataset

ptb-xl ptb ecg electrocardiography

Published: June 27, 2023. Version: 1.0.1

Visualize waveforms

Database Open Access

MIMIC-IV Clinical Database Demo

Alistair Johnson, Lucas Bulgarelli, Tom Pollard, Steven Horng, Leo Anthony Celi, Roger Mark

An openly available subset of patients in the MIMIC-IV database.

critical care electronic health record mimic

Published: Jan. 31, 2023. Version: 2.2


Database Credentialed Access

Northwestern ICU (NWICU) database

Dana Moukheiber, William Temps, Bhadrappa Molgi, Yikuan Li, Alice Lu, Prasanth Nannapaneni, Abdulrahman Chahin, Sicheng Hao, Felipe Torres Fabregas, Leo Anthony Celi, Adrian Wong, Maxwell Lloyd, Xavier Borrat Frigola, Hyung-Chul Lee, Daniel Schneider, Tom Pollard, Yuan Luo, Abel Kho, Roger Mark

A freely available COVID-rich ICU database comprising de-identified health-related data from Northwestern Memorial Health Center (NHMC).

Published: Nov. 19, 2024. Version: 0.1.0


Model Credentialed Access

Me-LLaMA: Foundation Large Language Models for Medical Applications

Qianqian Xie, Qingyu Chen, Aokun Chen, Cheng Peng, Yan Hu, Fongci Lin, Xueqing Peng, Jimin Huang, Jeffrey Zhang, Vipina Keloth, Xinyu Zhou, Huan He, Lucila Ohno-Machado, Yonghui Wu, Hua Xu, Jiang Bian

Me-LLaMA is a family of large language models for medical applications trained using clinical text with LLaMA2 models as the base. We release model weights for the foundation models as well as the chat-enhanced models.

large language models

Published: June 5, 2024. Version: 1.0.0


Challenge Open Access

Predicting Acute Hypotensive Episodes: The PhysioNet/Computing in Cardiology Challenge 2009

This year's challenge is the tenth in the annual series of open challenges hosted by PhysioNet in cooperation with Computers in Cardiology. The goal of the challenge is to predict which patients in the challenge dataset will experience an acute …

hypertension multiparameter challenge ehr mimic

Published: Jan. 27, 2009. Version: 1.0.0


Database Open Access

HeartCycle: A comprehensive dataset of synchronized impedance cardiography and echocardiography for accurate hemodynamic predictions

Eduardo Illueca Fernandez, Ricardo Couceiro, Farhad Abtahi, Jorge Henriques, Rui Pedro Paiva, Lino Goncalves, Jose Millet, Fernando Seoane, Jens Muehlsteff, Paulo Carvalho

Impedance cardiography dataset (ICG) which combines the ICG signals and other methodologies with the golden standard echocardiographys (ECG). Researchers can use this dataset to compare the ICG points with the real hemodynamic events.

machine learning cardiovascular physiology electrophysiological study echocardiography impedance cardiography

Published: Nov. 2, 2025. Version: 1.0.0