Dataset Information

Deep-learning-assisted diagnosis for knee magnetic resonance imaging: Development and retrospective validation of MRNet.

ABSTRACT: BACKGROUND:Magnetic resonance imaging (MRI) of the knee is the preferred method for diagnosing knee injuries. However, interpretation of knee MRI is time-intensive and subject to diagnostic error and variability. An automated system for interpreting knee MRI could prioritize high-risk patients and assist clinicians in making diagnoses. Deep learning methods, in being able to automatically learn layers of features, are well suited for modeling the complex relationships between medical images and their interpretations. In this study we developed a deep learning model for detecting general abnormalities and specific diagnoses (anterior cruciate ligament [ACL] tears and meniscal tears) on knee MRI exams. We then measured the effect of providing the model's predictions to clinical experts during interpretation. METHODS AND FINDINGS:Our dataset consisted of 1,370 knee MRI exams performed at Stanford University Medical Center between January 1, 2001, and December 31, 2012 (mean age 38.0 years; 569 [41.5%] female patients). The majority vote of 3 musculoskeletal radiologists established reference standard labels on an internal validation set of 120 exams. We developed MRNet, a convolutional neural network for classifying MRI series and combined predictions from 3 series per exam using logistic regression. In detecting abnormalities, ACL tears, and meniscal tears, this model achieved area under the receiver operating characteristic curve (AUC) values of 0.937 (95% CI 0.895, 0.980), 0.965 (95% CI 0.938, 0.993), and 0.847 (95% CI 0.780, 0.914), respectively, on the internal validation set. We also obtained a public dataset of 917 exams with sagittal T1-weighted series and labels for ACL injury from Clinical Hospital Centre Rijeka, Croatia. On the external validation set of 183 exams, the MRNet trained on Stanford sagittal T2-weighted series achieved an AUC of 0.824 (95% CI 0.757, 0.892) in the detection of ACL injuries with no additional training, while an MRNet trained on the rest of the external data achieved an AUC of 0.911 (95% CI 0.864, 0.958). We additionally measured the specificity, sensitivity, and accuracy of 9 clinical experts (7 board-certified general radiologists and 2 orthopedic surgeons) on the internal validation set both with and without model assistance. Using a 2-sided Pearson's chi-squared test with adjustment for multiple comparisons, we found no significant differences between the performance of the model and that of unassisted general radiologists in detecting abnormalities. General radiologists achieved significantly higher sensitivity in detecting ACL tears (p-value = 0.002; q-value = 0.019) and significantly higher specificity in detecting meniscal tears (p-value = 0.003; q-value = 0.019). Using a 1-tailed t test on the change in performance metrics, we found that providing model predictions significantly increased clinical experts' specificity in identifying ACL tears (p-value < 0.001; q-value = 0.006). The primary limitations of our study include lack of surgical ground truth and the small size of the panel of clinical experts. CONCLUSIONS:Our deep learning model can rapidly generate accurate clinical pathology classifications of knee MRI exams from both internal and external datasets. Moreover, our results support the assertion that deep learning models can improve the performance of clinical experts during medical imaging interpretation. Further research is needed to validate the model prospectively and to determine its utility in the clinical setting.

SUBMITTER: Bien N

PROVIDER: S-EPMC6258509 | biostudies-other | 2018 Nov

REPOSITORIES: biostudies-other

ACCESS DATA

Publications

Deep-learning-assisted diagnosis for knee magnetic resonance imaging: Development and retrospective validation of MRNet.

Bien Nicholas N Rajpurkar Pranav P Ball Robyn L RL Irvin Jeremy J Park Allison A Jones Erik E Bereket Michael M Patel Bhavik N BN Yeom Kristen W KW Shpanskaya Katie K Halabi Safwan S Zucker Evan E Fanton Gary G Amanatullah Derek F DF Beaulieu Christopher F CF Riley Geoffrey M GM Stewart Russell J RJ Blankenberg Francis G FG Larson David B DB Jones Ricky H RH Langlotz Curtis P CP Ng Andrew Y AY Lungren Matthew P MP

PLoS medicine 20181127 11

<h4>Background</h4>Magnetic resonance imaging (MRI) of the knee is the preferred method for diagnosing knee injuries. However, interpretation of knee MRI is time-intensive and subject to diagnostic error and variability. An automated system for interpreting knee MRI could prioritize high-risk patients and assist clinicians in making diagnoses. Deep learning methods, in being able to automatically learn layers of features, are well suited for modeling the complex relationships between medical ima ...[more]

PMID: 30481176

Similar Datasets

Project description:ImportanceDeep learning may be able to use patient magnetic resonance imaging (MRI) data to aid in brain tumor classification and diagnosis.ObjectiveTo develop and clinically validate a deep learning system for automated identification and classification of 18 types of brain tumors from patient MRI data.Design, setting, and participantsThis diagnostic study was conducted using MRI data collected between 2000 and 2019 from 37 871 patients. A deep learning system for segmentation and classification of 18 types of intracranial tumors based on T1- and T2-weighted images and T2 contrast MRI sequences was developed and tested. The diagnostic accuracy of the system was tested using 1 internal and 3 external independent data sets. The clinical value of the system was assessed by comparing the tumor diagnostic accuracy of neuroradiologists with vs without assistance of the proposed system using a separate internal test data set. Data were analyzed from March 2019 through February 2020.Main outcomes and measuresChanges in neuroradiologist clinical diagnostic accuracy in brain MRI scans with vs without the deep learning system were evaluated.ResultsA deep learning system was trained among 37 871 patients (mean [SD] age, 41.6 [11.4] years; 18 519 women [48.9%]). It achieved a mean area under the receiver operating characteristic curve of 0.92 (95% CI, 0.84-0.99) on 1339 patients from 4 centers' data sets in diagnosis and classification of 18 types of tumors. Higher outcomes were found compared with neuroradiologists for accuracy and sensitivity and similar outcomes for specificity (for 300 patients in the Tiantan Hospital test data set: accuracy, 73.3% [95% CI, 67.7%-77.7%] vs 60.9% [95% CI, 46.8%-75.1%]; sensitivity, 88.9% [95% CI, 85.3%-92.4%] vs 53.4% [95% CI, 41.8%-64.9%]; and specificity, 96.3% [95% CI, 94.2%-98.4%] vs 97.9%; [95% CI, 97.3%-98.5%]). With the assistance of the deep learning system, the mean accuracy of neuroradiologists among 1166 patients increased by 12.0 percentage points, from 63.5% (95% CI, 60.7%-66.2%) without assistance to 75.5% (95% CI, 73.0%-77.9%) with assistance.Conclusions and relevanceThese findings suggest that deep learning system-based automated diagnosis may be associated with improved classification and diagnosis of intracranial tumors from MRI data among neuroradiologists.

Project description:Background: Early-stage diagnosis and treatment can improve survival rates of liver cancer patients. Dynamic contrast-enhanced MRI provides the most comprehensive information for differential diagnosis of liver tumors. However, MRI diagnosis is affected by subjective experience, so deep learning may supply a new diagnostic strategy. We used convolutional neural networks (CNNs) to develop a deep learning system (DLS) to classify liver tumors based on enhanced MR images, unenhanced MR images, and clinical data including text and laboratory test results. Methods: Using data from 1,210 patients with liver tumors (N = 31,608 images), we trained CNNs to get seven-way classifiers, binary classifiers, and three-way malignancy-classifiers (Model A-Model G). Models were validated in an external independent extended cohort of 201 patients (N = 6,816 images). The area under receiver operating characteristic (ROC) curve (AUC) were compared across different models. We also compared the sensitivity and specificity of models with the performance of three experienced radiologists. Results: Deep learning achieves a performance on par with three experienced radiologists on classifying liver tumors in seven categories. Using only unenhanced images, CNN performs well in distinguishing malignant from benign liver tumors (AUC, 0.946; 95% CI 0.914-0.979 vs. 0.951; 0.919-0.982, P = 0.664). New CNN combining unenhanced images with clinical data greatly improved the performance of classifying malignancies as hepatocellular carcinoma (AUC, 0.985; 95% CI 0.960-1.000), metastatic tumors (0.998; 0.989-1.000), and other primary malignancies (0.963; 0.896-1.000), and the agreement with pathology was 91.9%.These models mined diagnostic information in unenhanced images and clinical data by deep-neural-network, which were different to previous methods that utilized enhanced images. The sensitivity and specificity of almost every category in these models reached the same high level compared to three experienced radiologists. Conclusion: Trained with data in various acquisition conditions, DLS that integrated these models could be used as an accurate and time-saving assisted-diagnostic strategy for liver tumors in clinical settings, even in the absence of contrast agents. DLS therefore has the potential to avoid contrast-related side effects and reduce economic costs associated with current standard MRI inspection practices for liver tumor patients.

Project description:Background: Multiparametric magnetic resonance imaging (mpMRI) plays an important role in the diagnosis of prostate cancer (PCa) in the current clinical setting. However, the performance of mpMRI usually varies based on the experience of the radiologists at different levels; thus, the demand for MRI interpretation warrants further analysis. In this study, we developed a deep learning (DL) model to improve PCa diagnostic ability using mpMRI and whole-mount histopathology data. Methods: A total of 739 patients, including 466 with PCa and 273 without PCa, were enrolled from January 2017 to December 2019. The mpMRI (T2 weighted imaging, diffusion weighted imaging, and apparent diffusion coefficient sequences) data were randomly divided into training (n = 659) and validation datasets (n = 80). According to the whole-mount histopathology, a DL model, including independent segmentation and classification networks, was developed to extract the gland and PCa area for PCa diagnosis. The area under the curve (AUC) were used to evaluate the performance of the prostate classification networks. The proposed DL model was subsequently used in clinical practice (independent test dataset; n = 200), and the PCa detective/diagnostic performance between the DL model and different level radiologists was evaluated based on the sensitivity, specificity, precision, and accuracy. Results: The AUC of the prostate classification network was 0.871 in the validation dataset, and it reached 0.797 using the DL model in the test dataset. Furthermore, the sensitivity, specificity, precision, and accuracy of the DL model for diagnosing PCa in the test dataset were 0.710, 0.690, 0.696, and 0.700, respectively. For the junior radiologist without and with DL model assistance, these values were 0.590, 0.700, 0.663, and 0.645 versus 0.790, 0.720, 0.738, and 0.755, respectively. For the senior radiologist, the values were 0.690, 0.770, 0.750, and 0.730 vs. 0.810, 0.840, 0.835, and 0.825, respectively. The diagnosis made with DL model assistance for radiologists were significantly higher than those without assistance (P < 0.05). Conclusion: The diagnostic performance of DL model is higher than that of junior radiologists and can improve PCa diagnostic accuracy in both junior and senior radiologists.

Dataset Information

Deep-learning-assisted diagnosis for knee magnetic resonance imaging: Development and retrospective validation of MRNet.

Publications

Deep-learning-assisted diagnosis for knee magnetic resonance imaging: Development and retrospective validation of MRNet.

Similar Datasets

OmicsDI is part of the ELIXIR infrastructure

Tweets