Next Article in Journal
Correction: Yasunaga et al. Associations Among Developmental Coordination Disorder Traits, Neurodevelopmental Difficulties and University Personality Inventory Scores in Undergraduate Students at a Japanese National University: A Cross-Sectional Correlational Study. Brain Sci. 2025, 15, 895
Previous Article in Journal
The Role of Cold-Inducible RNA-Binding Protein (CIRP) in Neurological Disorders
 
 
Font Type:
Arial Georgia Verdana
Font Size:
Aa Aa Aa
Line Spacing:
Column Width:
Background:
Article

A Two-Stage Localization and Refinement Neural Network Structure for Data-Efficient Microbleed Detection

by
Lukas Rau
1,†,
Oliver Granert
2,†,
Nils G. Margraf
2,
Stephan Schneider
3,† and
Ulf Jensen-Kondering
4,5,*,†
1
Faculty of Computer Science and Electrical Engineering, HAW Kiel, 24149 Kiel, Germany
2
Department of Neurology, University Hospital Schleswig-Holstein, Campus Kiel, 24105 Kiel, Germany
3
Department of Business Information Systems, Faculty of Business Management, HAW Kiel, 24149 Kiel, Germany
4
Department of Neuroradiology, University Hospital Schleswig-Holstein, Campus Lübeck, 23538 Lübeck, Germany
5
Department of Radiology and Neuroradiology, University Hospital Schleswig-Holstein, Campus Kiel, 24105 Kiel, Germany
*
Author to whom correspondence should be addressed.
These authors contributed equally to this work.
Brain Sci. 2026, 16(2), 207; https://doi.org/10.3390/brainsci16020207
Submission received: 30 December 2025 / Revised: 30 January 2026 / Accepted: 7 February 2026 / Published: 10 February 2026
(This article belongs to the Section Neurotechnology and Neuroimaging)

Abstract

Background/Objectives: In medical diagnostics, (semi-)automatic detection of pathological structures in images is becoming increasingly important. In particular, detecting cerebral microbleeds (CMBs) poses a challenge in clinical practice because the process is time-consuming and prone to error. Methods: Compared to previous methods of (semi-) automatic CMB detection that rely on large training datasets, we propose a method that can be adapted with a small dataset while still performing well. We propose a workflow that uses a two-stage approach to detect cerebral microbleeds that can be trained with a small dataset. The first stage is a 3D U-Net that retrieves potential CMB locations in the SWI image volume. Then, a 3D convolutional neural network (CNN) is used for discrimination to distinguish between real CMB and CMB mimics. Results: Using a dataset of 15 MRI scans with 40 marked CMBs, we are able to achieve a sensitivity of 97.5%. Conclusions: We showed that it is possible to create a workflow with high sensitivity using only a few training samples, enabling smaller radiological facilities to train networks using their own datasets. Even though the workflow performs well on a small dataset, it still requires further testing with other larger datasets.

1. Introduction

Cerebral microbleeds (CMBs) are 2–10 mm small accumulations of blood in the brain parenchyma. They are defined as round or ovoid signal hypointensities on sequences sensitive to blood products, such as susceptibility-weighted images (SWI) or gradient echo (GRE) T2*-weighted images, and are typically not visualized on other sequences or other imaging modalities [1]. They have been associated with cognitive decline [2,3] and vascular diseases [4,5]. While single CMBs are mostly asymptomatic, the total number of CMBs predicts the risk of future symptomatic macro-hemorrhages [6]. CMBs are the hallmark of several diseases, the most common of which are cerebral amyloid angiopathy (CAA) [7] and hypertensive arteriopathy (HTN-A) [8]. Especially when reading SWIs, some typical mimics must be considered. A common problem encountered in daily practice is distinguishing a CMB from a venous blood vessel [9]. While a CMB is confined to a few adjacent slices, a blood vessel is continuously visible through multiple slices. Other mimics, such as hemorrhagic metastasis, diffuse axonal injury, or calcifications [9] or residuals from heart-lung machines during open heart surgery [10] must also be considered. In some locations, e.g., the anterior temporal lobes, it is more demanding to discern single CMBs because they are more prone to seemingly bleed into surrounding structures such as blood vessels and susceptibility artifacts caused by the bone-CSF border. Manual detection of CMBs, especially when there are many, as shown in Figure 1D, can be demanding and time-consuming, e.g., to assess the progression of the number of CMBs.
Furthermore, it is error-prone (“satisfaction of search”, “underreading” [11]) and is associated with great inter-rater variability. Thus, an automated and objective algorithm to detect and count CMB is desirable, especially in clinical routine or in a research setting if large patient cohorts are involved.
Different approaches have been employed in prior research, utilizing both semi-automated and fully automated systems (Table 1). All these approaches focus on increasing sensitivity and decreasing false positives. They, however, differ in the achieved sensitivity, the number of false positives per scan, and the number of scans used for training. Here, we propose an approach that combines advanced image preprocessing as proposed by Koschmieder et al. [12] and Suwalska et al. [13], with a 3D U-Net localization stage and a subsequent 3D CNN discrimination stage as proposed by Dou et al.

2. Materials and Methods

2.1. Methodology

The general structure of the proposed workflow is shown in Figure 2. First, the scans are preprocessed and subsequently passed to the 3D U-Net. The 3D U-Net is applied using a sliding window strategy with a window size of 20 × 20 × 16 voxels and a stride of three for each axis. These prediction volumes are then stitched together by overlaying them based on their positions in the scan, forming a combined prediction volume. After that, we threshold the prediction volume to extract the possible CMB locations. These possible CMB locations are passed to the 3D CNN, which uses a 20 × 20 × 16 voxel window around the possible CMB location to predict the likelihood that this location is a real CMB. The workflow heavily relies on CNNs [13].

2.1.1. 3D U-Net

Marking CMB candidates and maximizing sensitivity are the main objectives of the first stage of our CMB detection process. We use a 3D U-Net (Figure 3) which takes an image of 20 × 20 × 16 voxels as input and outputs a prediction map of the same size, marking possible CMB locations. This, compared to the approach of Dou et al. [17], provides a larger network input size, thereby increasing the field of view. This extended view can help discriminate between blood vessels and microbleeds, since CMBs are mostly surrounded by regions that are less paramagnetic and thus appear bright on the scans (Figure 4). These long blood vessel structures can be identified with our larger field of view. Our larger sliding window enables faster processing by allowing larger strides in the CMB sliding window localization procedure. The convolutional layer applies a filter to the input volume and creates a feature map that sums up the detected features of the input. The batch normalization layer helps the 3D U-Net achieve better sensitivity while reducing training time and stabilizing gradients during training [23,24].

2.1.2. 3D CNN

A 3D CNN is used in the second stage to discriminate CMBs from CMB mimics (see Figure 4). Our net uses a 20 × 20 × 16 input volume and predicts how likely it is for the center of the volume to be a CMB (Figure 5). The net structure is based on the architecture proposed by Dou et al. [17] but additionally uses some batch normalization layers to speed up the training time. These layers prevent undesired high weights from cascading through the neural network during training by normalizing the output of the activation function and then scaling and shifting this normalized output by two trainable parameters. This helps the network’s weight optimization during training achieve smoother, more predictive, and more stable gradient behavior [23,24]. The convolutional layers act as pattern detectors. Small kernels of different sizes stride over the input volumes, transforming them into feature maps. The MaxPooling layer helps create a down-sampled (pooled) feature map and adds translation invariance, so that a small shift in the CMB does not significantly affect the detection performance. After passing through multiple convolutional layers, the feature maps are then flattened, resulting in a single output representing the probability of a CMB presence at the center of the input volume.

2.2. Experiments

2.2.1. Dataset and Preprocessing

For the training and testing of our network architecture, we use a dataset provided by Dou et al. [19]. The dataset consists of 20 SWI scans from 20 different subjects with at least 1 and up to 13 annotated CMBs. However, some scans in this dataset have a different voxel matrix. If we use these images for training, the sensitivity of our neural networks decreases due to differences in CMB volume, as demonstrated in Figure 6. When only using the scans at the same resolution, the mean CMB voxel count is 52 with a standard deviation of 29.5. In comparison, the CMB mean voxel count in the other scans is 82 with a standard deviation of 43. This is why we use only 15 of the 20 datasets, all of which had 512 × 512 × 150 voxels, resulting in 40 marked CMBs in total.
This reduction in the training dataset does not affect our network’s performance, as it better shows the feasibility and performance of our workflow when trained and tested on a small dataset.
The first step in preprocessing is brain masking, which we perform using the HD-BET software (Version 2.0.1) [25]. Although this tool was not specifically designed for SWI scans, its performance is sufficient for this paper. To remove the low-frequency intensity non-uniformity, which is often present in image data, also known as gain field or bias, we use the “N4ITK” (Version 5.3), which is available through the Insight Toolkit of the National Institutes of Health [26]. In the last step, we normalize the intensity values of the data, similar to Koschmieder et al. [21], by detecting the peak hp of the image intensity histogram from the MRI volume (except zero background intensity) and linearly map the data by
histnorm (i) = i/(2 ∗ hp)
where i is the image intensity before normalization. After this normalization, the histogram peak maps to an intensity of 0.5. Figure 7 demonstrates the effect of normalization on the histograms of two example scans. The normalization step effectively compensates for intensity variations across MRI acquisitions and improves the detection performance.

2.2.2. Augmentation

Since deep learning tasks usually require a huge amount of data, we augment the data provided by Dou et al. [17]. Three typical augmentation methods were used to increase the amount of training data. The CMB patches were flipped, rotated 90 degrees in both directions, and shifted by one voxel position to maintain the CMB’s location at the center and incorporate the desired variety into the dataset.

2.2.3. Models and Mimics

The architecture (Figure 4) is used in the training process of our 3D U-Nets. We train two of these 3D U-Nets with the same network structure and apply them in parallel for CBM detection. The first 3D U-Net is trained with a batch size of 20, whereas the second is trained with a batch size of 30. Although this approach is more time-consuming, we are able to achieve a higher sensitivity with a lower average count of false positives. The two prediction maps are normalized and added together to form a final prediction map for each validation scan. By combining these prediction maps, we put more emphasis on the points where the 3D U-Nets predict a CMB. This process achieves a higher sensitivity. For instance, if one 3D U-Net fails to detect a CMB, the other 3D U-Net potentially detects the missed CMB.
After each training step, the 3D U-Nets use two entire scans as a validation set. This scheme corresponds to a leave-two-out cross-validation (LOTCV) with a sliding window strategy, using a stride size of three voxels in each dimension to create a prediction map for each validation set (Figure 8). A LOTCV can be seen as an extension of leave-one-out cross-validation (LOOCV).
To optimize the second-stage training of the 3D CNN, we applied a threshold of 0.15 to the combined U-Net predictions. While this does not affect sensitivity, it increases the false positive count. The coordinates of these false positives are saved in a scan-specific file. This approach was chosen because the 3D U-Net is highly sensitive to low-intensity foci but struggles to differentiate between circular microbleeds (CMB) and blood vessels of similar intensity. By incorporating these ‘mimics’ into the training set, the 3D CNN learns to identify 3D U-Net errors, enhancing its ability to discriminate between real CMBs and artifacts. The training workflow for the 3D CNN follows a similar protocol to the 3D U-Nets, utilizing data augmentation through rotation, flipping, and coordinate shifting, now supplemented by the identified mimics. We trained two 3D CNNs in parallel using different batch sizes of 20 and 30, respectively.

2.3. Hardware and Software

The training and testing of the composed model system were performed on a machine equipped with an AMD (Advanced Micro Devices, Inc., Santa Clara, CA, USA) Ryzen 7 5800x with 16-core parallel processing and 32 GB of RAM. The model system was created and executed with the Tensorflow framework (https://www.tensorflow.org/) (accessed on 20 July 2022).

3. Results

We train two 3D U-Nets with the same network architecture and apply them in parallel for CBM detection. The first 3D U-Net is trained with a batch size of 20, whereas the second is trained with a batch size of 30. Combining two different batch-sized 3D U-Nets improves the sensitivity, as shown in Table 2. This is especially valuable because of the small number of marked CMBs, where a higher sensitivity can be more valuable than a lower time of processing. We obtained a relatively high sensitivity with both solitary 3D U-Nets, which provides the opportunity to choose between a lower processing time or a higher sensitivity.
By thresholding the combined predictions at a value of 0.2, we obtain a first-stage sensitivity of 98.33% with 38.15 false positives per scan.
We also train two 3D CNNs in parallel using different batch sizes of 20 and 30, respectively. The combination of two 3D CNNs in the second stage yields slightly higher sensitivity than either 3D CNN alone (see Figure 9).
While combining the predictions from both models may not significantly increase sensitivity compared to the model trained with a batch size of 20, we nevertheless recommend this ensemble approach. The computational overhead is minimal, with parallel 3D CNN inference adding only a few seconds to the processing time. This combination leads to a marginal increase in false positives compared to the individual models (batch sizes 20 and 30) (see Figure 10).
Given our specific workflow and the extreme class imbalance—where over 99% of voxels are negatives—we have opted to present our results using “sensitivity” and “average false positive rate” (Figure 9 and Figure 10 and Table 2).

4. Discussion

CMBs are recognized as important biomarkers for cerebrovascular diseases. The manual assessment is, however, time-consuming and error-prone. In the present cohort, a median of two CMBs per patient was observed, and counts of more than 10 CMBs are considered high [27]. From the example (Figure 1) given above, it is clear that the total number of CMBs can be considerably higher. This is why there is an ever-growing search for computer-aided detection systems that reliably detect CMBs with a low average count of false positives. Recent research has identified ways to obtain a high sensitivity with low false positive counts [13,21]. The problem with these approaches, however, is that they might not work as well across datasets with different MRI parameters, which is why we propose a method that can be trained on a small dataset with high sensitivity and a low false positive count. Using this two-stage approach with only 15 scans and 40 marked CMBs, we detect 95% of all CMBs while averaging 4.1 false positives per scan. When using a lower threshold, we reached a sensitivity of 97.5% with 14.5 false positives per scan on average. Our results show that it is possible to train neural networks on small datasets with high sensitivity and low false positive counts, which is advantageous in medical settings because radiological facilities can retrain the network using their own data and specifically adapt the network to their MRI sequences. A primary limitation of our current implementation is the increased computational processing time required to run the networks in parallel to generate the final results. With our hardware setup (see Section 2.3), the computation of the 3D U-Net prediction using the sliding window took approximately 20 min. Using both trained 3D U-Nets doubled that time. The processing time of the 3D CNN additionally took a couple of seconds. This processing time is much slower than the 64.3 s reported by Dou et al. [17] and the 20 s achieved by Chen et al. [28] with their workflow. However, this computational time could be improved by using two or more machines with GPU processing units, which would halve the time required for the sliding window 3D U-Net prediction. Further, increasing the 3D U-Net input and output volumes would also lead to faster processing times. Another shortcoming of our approach is the ability to detect CMBs in a cluster. Here, 3 of the 4 CMBs that were missed during validation had some low-intensity tissue near them, which led the 3D CNN to discriminate them. This behavior comes from training the 3D CNN with the U-Net mimics, which were often low-intensity vessels. Through these mimics, the 3D CNN learned to discriminate when points of low intensity were close by. The sensitivity could be improved by reducing the number of mimics during 3D CNN training, but this would also lead to an increase in false positives. It can be assumed that using this approach with a larger dataset would lead to a better detection of CMBs that occur in a cluster or have a low-intensity profile.

5. Conclusions and Future Work

We showed that it is possible to create a workflow with high sensitivity while using only a few training samples. Specified augmentation techniques (see Section 2.2.2) were used to increase the sample size. In order to counteract any problems that might arise as a result, such as the impossibility of a CMB occurring in this form of augmentation, the receptive fields (input layers) were enlarged to enrich the CMB excerpt with more contextual information. Furthermore, the results of the augmentations were constrained to the center of the receptive field, e.g., a shift of 1 to 2 voxels around the center. By ensuring that a CMB was centered, it became easier for the models to recognize a CMB. This approach allows smaller radiological facilities to train the networks using their own datasets, largely independent of their size. Even though the workflow performs well on a small dataset, it still requires further testing across other larger datasets to verify the performance when trained with different MR parameters. A larger dataset allows using only one 3D U-Net and only one 3D CNN rather than combining two different 3D U-Nets and CNNs. This would drastically speed up the process, making it more feasible for use in a clinical environment.

Author Contributions

Conceptualization, O.G., S.S., N.G.M. and U.J.-K.; methodology, L.R., O.G. and S.S.; software, L.R.; validation, L.R., O.G., S.S. and U.J.-K.; formal analysis, L.R. and O.G.; investigation, L.R. and O.G.; resources, S.S. and U.J.-K.; data curation, L.R. and O.G.; writing—original draft preparation, L.R.; writing—review and editing, L.R., O.G., N.G.M., S.S. and U.J.-K.; visualization, L.R. and O.G.; supervision, S.S.; project administration, S.S. and U.J.-K.; funding acquisition, U.J.-K. All authors have read and agreed to the published version of the manuscript.

Funding

This research was in part funded by “Familie Mehdorn Stiftung, Kiel”.

Institutional Review Board Statement

Not applicable.

Informed Consent Statement

Not applicable.

Data Availability Statement

Restrictions apply to the availability of these data. Data were obtained from [Public Repository, Department of Computer Science & Engineering, The Chinese University of Hong Kong] and are available [http://www.cse.cuhk.edu.hk/~qdou/cmb-3dcnn/cmb-3dcnn.html] (accessed 15 July 2022).

Acknowledgments

The authors are indebted to Fiona Bubbers and Sabrina Hillmann-Böhnke for language editing.

Conflicts of Interest

The authors declare no conflicts of interest. The funders had no role in the design of the study; in the collection, analyses, or interpretation of data; in the writing of the manuscript; or in the decision to publish the results.

Abbreviations

The following abbreviations are used in this manuscript:
CMBCerebral microbleeds
CNNConvolutional neural network
ROIRegion of interest
SWISusceptibility-weighted imaging
YOLOYou Only Look Once

References

  1. Duering, M.; Biessels, G.J.; Brodtmann, A.; Chen, C.; Cordonnier, C.; de Leeuw, F.-E.; Debette, S.; Frayne, R.; Jouvent, E.; Rost, N.S.; et al. Neuroimaging standards for research into small vessel disease—Advances since 2013. Lancet Neurol. 2023, 22, 602–618. [Google Scholar] [CrossRef]
  2. Viswanathan, A.; Godin, O.; Jouvent, E.; O’sullivan, M.; Gschwendtner, A.; Peters, N.; Duering, M.; Guichard, J.-P.; Holtmannspötter, M.; Dufouil, C.; et al. Impact of MRI markers in subcortical vascular dementia: A multi-modal analysis in CADASIL. Neurobiol. Aging 2010, 31, 1629–1636. [Google Scholar] [CrossRef]
  3. Werring, D.J.; Frazer, D.W.; Coward, L.J.; Losseff, N.A.; Watt, H.; Cipolotti, L.; Brown, M.M.; Jäger, H.R. Cognitive dysfunction in patients with cerebral microbleeds on T2*-weighted gradient-echo MRI. Brain 2004, 127, 2265–2275. [Google Scholar] [CrossRef]
  4. Linn, J.; Halpin, A.; Demaerel, P.; Ruhland, J.; Giese, A.; Dichgans, M.; van Buchem, M.; Bruckmann, H.; Greenberg, S. Prevalence of superficial siderosis in patients with cerebral amyloid angiopathy. Neurology 2010, 74, 1346–1350. [Google Scholar] [CrossRef]
  5. Tanaka, A.; Ueno, Y.; Nakayama, Y.; Takano, K.; Takebayashi, S. Small Chronic Hemorrhages and Ischemic Lesions in Association with Spontaneous Intracerebral Hematomas. Stroke 1999, 30, 1637–1642. [Google Scholar] [CrossRef] [PubMed]
  6. Greenberg, S.M.; Eng, J.A.; Ning, M.; Smith, E.E.; Rosand, J. Hemorrhage Burden Predicts Recurrent Intracerebral Hemorrhage After Lobar Hemorrhage. Stroke 2004, 35, 1415–1420. [Google Scholar] [CrossRef] [PubMed]
  7. Charidimou, A.; Boulouis, G.; Frosch, M.P.; Baron, J.-C.; Pasi, M.; Albucher, J.F.; Banerjee, G.; Barbato, C.; Bonneville, F.; Brandner, S.; et al. The Boston criteria version 2.0 for cerebral amyloid angiopathy: A multicentre, retrospective, MRI–neuropathology diagnostic accuracy study. Lancet Neurol. 2022, 21, 714–725. [Google Scholar] [CrossRef]
  8. Smith, E.E.; Nandigam, K.R.N.; Chen, Y.-W.; Jeng, J.; Salat, D.; Halpin, A.; Frosch, M.; Wendell, L.; Fazen, L.; Rosand, J.; et al. MRI Markers of Small Vessel Disease in Lobar and Deep Hemispheric Intracerebral Hemorrhage. Stroke 2010, 41, 1933–1938. [Google Scholar] [CrossRef] [PubMed]
  9. Charidimou, A.; Jäger, H.R.; Werring, D.J. Cerebral microbleed detection and mapping: Principles, methodological aspects and rationale in vascular dementia. Exp. Gerontol. 2012, 47, 843–852. [Google Scholar] [CrossRef]
  10. Patel, N.; Banahan, C.; Janus, J.; Horsfield, M.A.; Cox, A.; Li, X.; Cappellugola, L.; Colman, J.; Egan, V.; Garrard, P.; et al. Perioperative Cerebral Microbleeds After Adult Cardiac Surgery. Stroke 2019, 50, 336–343. [Google Scholar] [CrossRef]
  11. Kim, Y.W.; Mansfield, L.T. Fool Me Twice: Delayed Diagnoses in Radiology with Emphasis on Perpetuated Errors. Am. J. Roentgenol. 2014, 202, 465–470. [Google Scholar] [CrossRef]
  12. Koschmieder, K.; Paul, M.M.; van den Heuvel, T.L.A.; van der Eerden, A.; van Ginneken, B.; Manniesing, R. Automated detection of cerebral microbleeds via segmentation in susceptibility-weighted images of patients with traumatic brain injury. NeuroImage Clin. 2022, 35, 103027. [Google Scholar] [CrossRef]
  13. Suwalska, A.; Wang, Y.; Yuan, Z.; Jiang, Y.; Zhu, D.; Chen, J.; Cui, M.; Chen, X.; Suo, C.; Polanska, J. CMB-HUNT: Automatic detection of cerebral microbleeds using a deep neural network. Comput. Biol. Med. 2022, 151, 106233. [Google Scholar] [CrossRef] [PubMed]
  14. Barnes, S.R.S.; Haacke, E.M.; Ayaz, M.; Boikov, A.S.; Kirsch, W.; Kido, D. Semiautomated detection of cerebral microbleeds in magnetic resonance images. Magn. Reson. Imaging 2011, 29, 844–852. [Google Scholar] [CrossRef] [PubMed]
  15. Kuijf, H.J.; de Bresser, J.; Geerlings, M.I.; Conijn, M.M.; Viergever, M.A.; Biessels, G.J.; Vincken, K.L. Efficient detection of cerebral microbleeds on 7.0T MR images using the radial symmetry transform. NeuroImage 2012, 59, 2266–2273. [Google Scholar] [CrossRef]
  16. Ghafaryasl, B.; van der Lijn, F.; Poels, M.; Vrooman, H.; Ikram, M.A.; Niessen, W.J.; van der Lugt, A.; Vernooij, M.; de Bruijne, M. A computer aided detection system for cerebral microbleeds in brain MRI. In 2012 9th IEEE International Symposium on Biomedical Imaging (ISBI); IEEE: Piscataway, NJ, USA, 2012; pp. 138–141. [Google Scholar]
  17. Dou, Q.; Chen, H.; Yu, L.; Zhao, L.; Qin, J.; Wang, D.; Mok, V.C.; Shi, L.; Heng, P.-A. Automatic Detection of Cerebral Microbleeds from MR Images via 3D Convolutional Neural Networks. IEEE Trans. Med. Imaging 2016, 35, 1182–1195. [Google Scholar] [CrossRef]
  18. Bian, W.; Hess, C.P.; Chang, S.M.; Nelson, S.J.; Lupo, J.M. Computer-aided detection of radiation-induced cerebral microbleeds on susceptibility-weighted MR images. NeuroImage Clin. 2013, 2, 282–290. [Google Scholar] [CrossRef] [PubMed]
  19. Fazlollahi, A.; Meriaudeau, F.; Villemagne, V.L.; Rowe, C.C.; Yates, P.; Salvado, O.; Bourgeat, P. Efficient machine learning framework for computer-aided detection of cerebral microbleeds using the Radon transform. In 2014 IEEE 11th International Symposium on Biomedical Imaging (ISBI); IEEE: Piscataway, NJ, USA, 2014; pp. 113–116. [Google Scholar]
  20. van den Heuvel, T.L.A.; van der Eerden, A.W.; Manniesing, R.; Ghafoorian, M.; Tan, T.; Andriessen, T.; Vyvere, T.V.; Hauwe, L.v.D.; Romeny, B.t.H.; Goraj, B.; et al. Automated detection of cerebral microbleeds in patients with traumatic brain injury. NeuroImage Clin. 2016, 12, 241–251. [Google Scholar] [CrossRef]
  21. Liu, S.; Utriainen, D.; Chai, C.; Chen, Y.; Wang, L.; Sethi, S.K.; Xia, S.; Haacke, E.M. Cerebral microbleed detection using Susceptibility Weighted Imaging and deep learning. NeuroImage 2019, 198, 271–282. [Google Scholar] [CrossRef] [PubMed]
  22. Al-Masni, M.A.; Kim, W.-R.; Kim, E.Y.; Noh, Y.; Kim, D.-H. Automated detection of cerebral microbleeds in MR images: A two-stage deep learning approach. NeuroImage Clin. 2020, 28, 102464. [Google Scholar] [CrossRef]
  23. Ioffe, S.; Szegedy, C. Batch Normalization: Accelerating Deep Network Training by Reducing Internal Covariate Shift. In Proceedings of the 32nd International Conference on Machine Learning; JMLR: New York, NY, USA, 2015. [Google Scholar]
  24. Santurkar, S.; Tsipras, D.; Ilyas, A.; Madry, A. How Does Batch Normalization Help Optimization? In Advances in Neural Information Processing Systems 31 (NeurIPS 2018); Curran Associates, Inc.: Red Hook, NY, USA, 2018. [Google Scholar]
  25. Isensee, F.; Schell, M.; Pflueger, I.; Brugnara, G.; Bonekamp, D.; Neuberger, U.; Wick, A.; Schlemmer, H.-P.; Heiland, S.; Wick, W.; et al. Automated brain extraction of multisequence MRI using artificial neural networks. Hum. Brain Mapp. 2019, 40, 4952–4964. [Google Scholar] [CrossRef] [PubMed]
  26. Tustison, N.J.; Avants, B.B.; Cook, P.A.; Zheng, Y.; Egan, A.; Yushkevich, P.A.; Gee, J.C. N4ITK: Improved N3 Bias Correction. IEEE Trans. Med. Imaging 2010, 29, 1310–1320. [Google Scholar] [CrossRef] [PubMed]
  27. Ni, J.; Auriel, E.; Jindal, J.; Ayres, A.; Schwab, K.M.; Martinez-Ramirez, S.; Gurol, E.M.; Greenberg, S.M.; Viswanathan, A. The Characteristics of Superficial Siderosis and Convexity Subarachnoid Hemorrhage and Clinical Relevance in Suspected Cerebral Amyloid Angiopathy. Cerebrovasc. Dis. 2015, 39, 278–286. [Google Scholar] [CrossRef] [PubMed]
  28. Chen, H.; Yu, L.; Dou, Q.; Shi, L.; Mok, V.C.T.; Heng, P.A. Automatic detection of cerebral microbleeds via deep learning based 3D feature representation. In 2015 IEEE 12th International Symposium on Biomedical Imaging (ISBI); IEEE: Piscataway, NJ, USA, 2015; pp. 764–767. [Google Scholar]
Figure 1. Blood sensitive sequences (susceptibility weighted imaging, SWI) demonstrating no CMB (A), a single CMB ((B), circle), a few cerebral microbleeds ((C), circles highlighting some CMBs), and extensive cerebral microbleeds (D).
Figure 1. Blood sensitive sequences (susceptibility weighted imaging, SWI) demonstrating no CMB (A), a single CMB ((B), circle), a few cerebral microbleeds ((C), circles highlighting some CMBs), and extensive cerebral microbleeds (D).
Brainsci 16 00207 g001
Figure 2. Proposed workflow with prediction and discrimination nets. Blue circles indicate detected CMBs from the output of the discrimination nets.
Figure 2. Proposed workflow with prediction and discrimination nets. Blue circles indicate detected CMBs from the output of the discrimination nets.
Brainsci 16 00207 g002
Figure 3. 3D U-Net architecture.
Figure 3. 3D U-Net architecture.
Brainsci 16 00207 g003
Figure 4. Two samples of different CMB sliced along the z-axis (rows CMB1 and CMB2) and negative examples, including CMB mimics, used for training the network weights (last row).
Figure 4. Two samples of different CMB sliced along the z-axis (rows CMB1 and CMB2) and negative examples, including CMB mimics, used for training the network weights (last row).
Brainsci 16 00207 g004
Figure 5. 3D CNN architecture.
Figure 5. 3D CNN architecture.
Brainsci 16 00207 g005
Figure 6. Distribution of voxel values from two different scans pre-normalization.
Figure 6. Distribution of voxel values from two different scans pre-normalization.
Brainsci 16 00207 g006
Figure 7. Distribution of voxel values from two different (blue and orange) scans (A) pre- and (B) post-normalization.
Figure 7. Distribution of voxel values from two different (blue and orange) scans (A) pre- and (B) post-normalization.
Brainsci 16 00207 g007
Figure 8. CMB and their corresponding 3D U-Net predictions.
Figure 8. CMB and their corresponding 3D U-Net predictions.
Brainsci 16 00207 g008
Figure 9. Influence of prediction threshold on the sensitivity across various 3D CNN training strategies and ensemble combinations.
Figure 9. Influence of prediction threshold on the sensitivity across various 3D CNN training strategies and ensemble combinations.
Brainsci 16 00207 g009
Figure 10. Influence of prediction threshold on the average false positive rate across various 3D CNN training strategies and ensemble combinations.
Figure 10. Influence of prediction threshold on the average false positive rate across various 3D CNN training strategies and ensemble combinations.
Brainsci 16 00207 g010
Table 1. Comparison of various published studies. (FP avg: average number of false positives).
Table 1. Comparison of various published studies. (FP avg: average number of false positives).
YearFP Avg.SensitivityScansTechnical ApproachPaper
201110781.70%6statistical thresholding algorithm and support vector machine (SVM) trained by supervised learningBarnes et al. [14]
20121771.20%18radial symmetry transform and manual discriminationKuijf et al. [15]
20124.191.00%237thresholding, ROI labelling and classificationGhafaryasl et al. [16]
20122.7493.16%3203D fully convolutional neural network for candidates and 3D convolutional neural network to discriminate between CMBs and mimicsDou et al. [17]
201344.986.50%152D fast radial symmetry transformation and 3D region growing on CMB candidatesBian et al. [18]
201416.8492.04%41cascade of random forest classifiers analyzing robust Radon-based features and manual discriminationFazlollahi et al. [19]
201613.990.80%33random forest classifier (based on SWI and T1), segmentation and object-based classifierVan den Heuvel et al. [20]
20191.695.80%195candidate detection (3D fast radial symmetry transformation) and discrimination (deep residual neural network)Liu et al. [21]
20201.4293.60%179candidate detection (YOLO) and 3D-CNNAl-Masni et al. [22]
202217.190.00%45classification Patch-CNN, Segmentation-CNN and U-NetKoschmieder et al. [12]
20221.991.50%265segmentation (MiMSeg algorithm), filtration, filter extraction and CNNSuwalska et al. [13]
20264.195.00%15advanced preprocessing, 3D U-Net first stage, 3D CNN discrimination stageOur study (low FP)
202614.597.50%15advanced preprocessing, 3D U-Net first stage, 3D CNN discrimination stageOur study (high sensitivity)
Table 2. 3D U-Net comparison. (FP avg: average number of false positives).
Table 2. 3D U-Net comparison. (FP avg: average number of false positives).
FP Avg.SensitivityBatch Size
47.5395.83%20
38.7396.67%30
38.1598.33%combined
Disclaimer/Publisher’s Note: The statements, opinions and data contained in all publications are solely those of the individual author(s) and contributor(s) and not of MDPI and/or the editor(s). MDPI and/or the editor(s) disclaim responsibility for any injury to people or property resulting from any ideas, methods, instructions or products referred to in the content.

Share and Cite

MDPI and ACS Style

Rau, L.; Granert, O.; Margraf, N.G.; Schneider, S.; Jensen-Kondering, U. A Two-Stage Localization and Refinement Neural Network Structure for Data-Efficient Microbleed Detection. Brain Sci. 2026, 16, 207. https://doi.org/10.3390/brainsci16020207

AMA Style

Rau L, Granert O, Margraf NG, Schneider S, Jensen-Kondering U. A Two-Stage Localization and Refinement Neural Network Structure for Data-Efficient Microbleed Detection. Brain Sciences. 2026; 16(2):207. https://doi.org/10.3390/brainsci16020207

Chicago/Turabian Style

Rau, Lukas, Oliver Granert, Nils G. Margraf, Stephan Schneider, and Ulf Jensen-Kondering. 2026. "A Two-Stage Localization and Refinement Neural Network Structure for Data-Efficient Microbleed Detection" Brain Sciences 16, no. 2: 207. https://doi.org/10.3390/brainsci16020207

APA Style

Rau, L., Granert, O., Margraf, N. G., Schneider, S., & Jensen-Kondering, U. (2026). A Two-Stage Localization and Refinement Neural Network Structure for Data-Efficient Microbleed Detection. Brain Sciences, 16(2), 207. https://doi.org/10.3390/brainsci16020207

Note that from the first issue of 2016, this journal uses article numbers instead of page numbers. See further details here.

Article Metrics

Back to TopTop