PhysioNet Index

Database Restricted Access

LATTE-CXR: Locally Aligned TexT and imagE, Explainable dataset for Chest X-Rays

Elham Ghelichkhan, Tolga Tasdizen

This dataset includes bounding box-statement pairs for chest X-ray images, derived from radiologists’ eye-tracking data (for explainability) and annotations, for local visual-language models.

eye-tracking chest x-ray dataset automatically generated dataset caption-guided object detection image captioning with region-level description grounded radiology report generation phrase grounding xai multi-modal learning local visual-language models localization

Published: Feb. 4, 2025. Version: 1.0.0

Database Restricted Access

REFLACX: Reports and eye-tracking data for localization of abnormalities in chest x-rays

Ricardo Bigolin Lanfredi, Mingyuan Zhang, William Auffermann, et al.

This dataset contains 3032 cases of eye-tracking data collected while five radiologists dictated reports for frontal chest x-rays, synchronized timestamped dictation transcription, and manual labels for validation of localization of abnormalities.

eye tracking radiology report reflacx fixations chest x-rays computer vision gaze radiology deep learning machine learning

Published: Sept. 27, 2021. Version: 1.0.0

Database Credentialed Access

MIMIC-Ext-CXR-QBA: A Structured, Tagged, and Localized Visual Question Answering Dataset with Question-Box-Answer Triplets and Scene Graphs for Chest X-ray Images

Philip Müller, Friederike Jungmann, Georgios Kaissis, et al.

We present a large-scale CXR VQA dataset derived from MIMIC-CXR with 42M QA pairs, featuring hierarchical answers, bounding boxes, and structured tags. We generated QA-pairs using LLM-based extraction from radiology reports and localization models.

chest x-rays vqa localization scene graphs

Published: July 22, 2025. Version: 1.0.0

Database Open Access

CPAP Pressure and Flow Data from a Local Trial of 30 Adults at the University of Canterbury

Ella Guy, Jennifer Knopp, Geoff Chase

A pressure and flow dataset was collected from a trial of 30 adults at the University of Canterbury undergoing CPAP therapy for a variety of instructed breath rates at PEEP levels of 4cmH2O and 7cmH2O.

peep cpap respiratory mechanics pulmonary mechanics respiratory modelling biomedical engineering

Published: March 24, 2022. Version: 1.0.1

Database Credentialed Access

Chest ImaGenome Dataset

Joy Wu, Nkechinyere Agu, Ismini Lourentzou, et al.

The Chest ImaGenome dataset is a scene graph dataset with additional chronological comparison relations for chest X-rays. It is automatically derived from the MIMIC-CXR dataset. A manually annotated gold standard is also available for 500 patients.

scene graph visual dialogue object detection semantic reasoning bounding box knowledge graph explainability reasoning relation extraction chest disease progression cxr chest x-ray radiology multimodal visual question answering deep learning machine learning

Published: July 13, 2021. Version: 1.0.0

Database Credentialed Access

MS-CXR: Making the Most of Text Semantics to Improve Biomedical Vision-Language Processing

Benedikt Boecking, Naoto Usuyama, Shruthi Bannur, et al.

MS-CXR is a new dataset containing 1162 chest X-ray bounding box labels paired with radiology text descriptions, annotated and verified by two board-certified radiologists.

vision-language processing chest x-ray phrase grounding localization

Published: Nov. 15, 2024. Version: 1.1.0

Database Open Access

A Multi-Night Instantaneous Heart Rate and Accelerometry Dataset with EEG Sleep Stage Labels

Tzu-An Song

This dataset contains multi-night recordings of instantaneous heart rate (IHR) and 3-axis accelerometry from the Apple Watch, paired with EEG-based sleep stage labels from the Dreem 2 headband.

accelerometry sleep staging wearable devices apple watch instantaneous heart rate

Published: May 12, 2026. Version: 1.0.0

Database Restricted Access

MIMIC-IV-Ext-Apixaban-Trial-Criteria-Questions

Elizabeth Woo, Michael Craig Burkhart, Emily Alsentzer, et al.

We created 23 questions resembling eligibility criteria from the apixaban clinical trial and evaluated them on a random sample of 100 patient notes from MIMIC-IV. We release the 2300 total question-answer pairs as a dataset here.

clinical q and a evaluation set clinical trial eligibility

Published: April 30, 2025. Version: 1.0.0

Database Restricted Access

MIMIC-III-Ext-Synthetic-Clinical-Trial-Questions

Elizabeth Woo, Michael Craig Burkhart, Emily Alsentzer, et al.

In our recent study, we used Llama-3.1-70B-Instruct to generate synthetic training examples resembling clinical trial eligibility criteria. We manually reviewed 1000 of these examples and release them here.

large language models synthetic data distillation clinical trial eligibility

Published: April 22, 2025. Version: 1.0.0

Database Open Access

Kiel Cardio Database

Erik Engelhardt, Norbert Frey, Gerhard Schmidt

The Kiel Cardio Database (KCD) contains one-minute 8-lead magnetocardiographic (MCG) measurements from seven subjects. Each subject underwent 25 consecutive measurements using a sensor array comprising four QuSpin QZFMs.

opically pumped magnetometer opm mcg magnetocardiography

Published: Dec. 15, 2023. Version: 1.0.0

Visualize waveforms

Search

Resources

LATTE-CXR: Locally Aligned TexT and imagE, Explainable dataset for Chest X-Rays

REFLACX: Reports and eye-tracking data for localization of abnormalities in chest x-rays

MIMIC-Ext-CXR-QBA: A Structured, Tagged, and Localized Visual Question Answering Dataset with Question-Box-Answer Triplets and Scene Graphs for Chest X-ray Images

CPAP Pressure and Flow Data from a Local Trial of 30 Adults at the University of Canterbury

Chest ImaGenome Dataset

MS-CXR: Making the Most of Text Semantics to Improve Biomedical Vision-Language Processing

A Multi-Night Instantaneous Heart Rate and Accelerometry Dataset with EEG Sleep Stage Labels

MIMIC-IV-Ext-Apixaban-Trial-Criteria-Questions

MIMIC-III-Ext-Synthetic-Clinical-Trial-Questions

Kiel Cardio Database