Resources


Database Credentialed Access

Chest ImaGenome Dataset

Joy Wu, Nkechinyere Agu, Ismini Lourentzou, Arjun Sharma, Joseph Paguio, Jasper Seth Yao, Edward Christopher Dee, William Mitchell, Satyananda Kashyap, Andrea Giovannini, Leo Anthony Celi, Tanveer Syeda-Mahmood, Mehdi Moradi

The Chest ImaGenome dataset is a scene graph dataset with additional chronological comparison relations for chest X-rays. It is automatically derived from the MIMIC-CXR dataset. A manually annotated gold standard is also available for 500 patients.

multimodal chest x-ray radiology machine learning scene graph visual question answering visual dialogue object detection disease progression semantic reasoning bounding box relation extraction knowledge graph cxr chest explainability reasoning deep learning

Published: July 13, 2021. Version: 1.0.0


Database Credentialed Access

Eye Gaze Data for Chest X-rays

Alexandros Karargyris, Satyananda Kashyap, Ismini Lourentzou, Joy Wu, Matthew Tong, Arjun Sharma, Shafiq Abedin, David Beymer, Vandana Mukherjee, Elizabeth Krupinski, Mehdi Moradi

This dataset was a collected using an eye tracking system while a radiologist interpreted and read 1,083 public CXR images. The dataset contains the following aligned modalities: image, transcribed report text, dictation audio and eye gaze data.

audio convolutional network heatmap multimodal chest x-ray radiology machine learning eye tracking cxr chest explainability deep learning

Published: Sept. 12, 2020. Version: 1.0.0


Database Credentialed Access

Deidentified Medical Text

Margaret Douglass, Bill Long, George Moody, Peter Szolovits, Li-wei Lehman, Roger Mark, Gari Clifford

Gold standard corpus of 2,434 deidentified nursing notes

medical text nursing notes de-identification hipaa

Published: Dec. 18, 2007. Version: 1.0


Database Credentialed Access

MIMIC-CXR Database

Alistair Johnson, Tom Pollard, Roger Mark, Seth Berkowitz, Steven Horng

Chest radiographs in DICOM format with associated free-text reports.

mimic natural language processing computer vision chest x-rays radiology machine learning

Published: Sept. 19, 2019. Version: 2.0.0


Database Credentialed Access

RadGraph: Extracting Clinical Entities and Relations from Radiology Reports

Saahil Jain, Ashwin Agrawal, Adriel Saporta, Steven QH Truong, Du Nguyen Duong, Tan Bui, Pierre Chambon, Matthew Lungren, Andrew Ng, Curtis Langlotz, Pranav Rajpurkar

RadGraph is a dataset of entities and relations in full-text chest X-ray radiology reports, which are obtained using a novel information extraction (IE) schema to capture clinically relevant information in a radiology report.

natural language processing radiology entity and relation extraction graph multi-modal

Published: June 3, 2021. Version: 1.0.0


Database Credentialed Access

MIMIC-CXR-JPG - chest radiographs with structured labels

Alistair Johnson, Matt Lungren, Yifan Peng, Zhiyong Lu, Roger Mark, Seth Berkowitz, Steven Horng

Chest x-rays in JPG format with structured labels derived from the associated radiology report.

mimic computer vision chest x-ray radiology deep learning

Published: Nov. 14, 2019. Version: 2.0.0


Database Credentialed Access

FFA-IR: Towards an Explainable and Reliable Medical Report Generation Benchmark

Mingjie Li, Wenjia Cai, Rui Liu, Yuetian Weng, Xiaoyun Zhao, Cong Wang, Xin Chen, Zhong Liu, Caineng Pan, Mengke Li, Yingfeng Zheng, Yizhi Liu, Flora Salim, Karin Verspoor, Xiaodan Liang, Xiaojun Chang

Benchmark dataset for report generation based on fundus fluorescein angiography images and reports.

fundus fluorescein angiography explainable and reliable evaluation vision and language medical report generation

Published: Sept. 21, 2021. Version: 1.0.0


Software Open Access

Lightweight 12-lead ECG viewer for MATLAB

Erick Andres Perez Alday, Larisa Tereshchenko

Clinical Viewer of raw digital 12-lead ECG file (ECG file in .txt format).

clinical 12-lead ecg routine clinical ecg electrocardiogram viewer

Published: Aug. 30, 2021. Version: 1.0.0


Database Open Access

PhysioZoo - mammalian NSR databases

Ori Shemla, Joachim Behar

PhysioZoo is a collaborative platform dedicated to the study of the heart rate variability in electrophysiological recordings from mammals

heart rate variabillity electrophysiology mammals ecg

Published: Aug. 27, 2019. Version: 1.0.0

Visualize waveforms

Challenge Credentialed Access

Analysis of Clinical Text: Task 14 of SemEval 2015

Guergana Savova

This is the dataset for SemEval 2014 and 2015, Analysis of Clinical Text

nlp semeval

Published: Dec. 28, 2014. Version: 2.0