Data labels and scripts for fastMRI.org

Last update: Dec 22, 2022

Related tags

Overview

fastMRI+: Clinical pathology annotations for the fastMRI dataset

The fastMRI dataset is a publicly available MRI raw (k-space) dataset. It has been used widely to train machine learning models for image reconstruction and has been used in reconstruction challenges.

This repo includes clinical pathology annotations for this dataset. The entire knee dataset and approximately 1000 brain datasets have been labeled. The goal of providing these labels is to enable developers of image reconstruction models and algorithms to evaluate the performance of the developed techniques with a focus on the sections or regions that could contain clinical pathology.

Limitations

Each image has labeled by a single radiologist and without the benefit of looking at other views and angles of the same subject, and should therefore be considered in that context. Specifically, the labels should not be considered clinical ground truth or an exhaustive list of all lesions but rather an indicatition of where a pathology could be present.

Obtaining fastMRI raw data and images

The fastMRI raw data and reference images can be obtained from fastmri.org. You will be able to download and use the data for academic purposes after signing the data sharing agreement. If you are looking for automation for downloading the dataset and training fastMRI models, please see the InnerEye Deep Learning Toolkit.

Labeling procedure and generating DICOM images from fastMRI data

In order to label the data, DICOM files were generated from the fastMRI dataset, and we are providing a fastmri_to_dicom.py to document the procedure. This script can be used like this:

python fastmri_to_dicom.py --filename fastmridatafile.h5

Note: In the process of converting the images to DICOM, the pixel arrays were flipped (up/down) to provide a view that was closer to DICOM orientation and assist with labeling. This should be taken into consideration when using the labels.

The labeling was performed by experienced radiologists using MD.ai.

Working with the annotations

The Annotations folder contains a label file for each of the knee (knee.csv and brain (brain.csv datasets. The files contain one line for each annotation (bounding box) that was labeled by the radiologists. Datasets with no findings (no annotations) are not represented in the label files, however, you can see which files were reviewed in the brain_file_list.csv and knee_file_list.csv. If a dataset (a fastMRI file) is listed in the file lists but not in the label files, it means that it has been reviewed, but there were no findings.

The repo contains an example jupyter notebook, which illustrates how to read the labels and overlay them onto the image pixels.

Contributing

This project welcomes contributions and suggestions. Most contributions require you to agree to a Contributor License Agreement (CLA) declaring that you have the right to, and actually do, grant us the rights to use your contribution. For details, visit https://cla.opensource.microsoft.com.

When you submit a pull request, a CLA bot will automatically determine whether you need to provide a CLA and decorate the PR appropriately (e.g., status check, comment). Simply follow the instructions provided by the bot. You will only need to do this once across all repos using our CLA.

This project has adopted the Microsoft Open Source Code of Conduct. For more information see the Code of Conduct FAQ or contact [email protected] with any additional questions or comments.

Trademarks

This project may contain trademarks or logos for projects, products, or services. Authorized use of Microsoft trademarks or logos is subject to and must follow Microsoft's Trademark & Brand Guidelines. Use of Microsoft trademarks or logos in modified versions of this project must not cause confusion or imply Microsoft sponsorship. Any use of third-party trademarks or logos are subject to those third-party's policies.

Data labels and scripts for fastMRI.org

Related tags

Overview

fastMRI+: Clinical pathology annotations for the fastMRI dataset

Limitations

Obtaining fastMRI raw data and images

Labeling procedure and generating DICOM images from fastMRI data

Working with the annotations

Contributing

Trademarks

Owner

Microsoft

'Aligned mixture of latent dynamical systems' (amLDS) for stimulus decoding probabilistic manifold alignment across animals. P. Herrero-Vidal et al. NeurIPS 2021 code.

Multiple types of NN model optimization environments. It is possible to directly access the host PC GUI and the camera to verify the operation. Intel iHD GPU (iGPU) support. NVIDIA GPU (dGPU) support.

Image Segmentation with U-Net Algorithm on Carvana Dataset using AWS Sagemaker

Predict halo masses from simulations via graph neural networks

EmoTag helps you train emotion detection model for Chinese audios

Syllabic Quantity Patterns as Rhythmic Features for Latin Authorship Attribution

CVPR 2021 - Official code repository for the paper: On Self-Contact and Human Pose.

Interpolation-based reduced-order models

Moiré Attack (MA): A New Potential Risk of Screen Photos [NeurIPS 2021]

A framework for the elicitation, specification, formalization and understanding of requirements.

The Malware Open-source Threat Intelligence Family dataset contains 3,095 disarmed PE malware samples from 454 families

SafePicking: Learning Safe Object Extraction via Object-Level Mapping, ICRA 2022

LegoDNN: a block-grained scaling tool for mobile vision systems

Implementation of SSMF: Shifting Seasonal Matrix Factorization

VIsually-Pivoted Audio and(N) Text

The repository offers the official implementation of our BMVC 2021 paper in PyTorch.

Semantic-aware Grad-GAN for Virtual-to-Real Urban Scene Adaption

A modular active learning framework for Python

Implementation of H-Transformer-1D, Hierarchical Attention for Sequence Learning using 🤗 transformers

Face2webtoon - Despite its importance, there are few previous works applying I2I translation to webtoon.