Clinical urine microscopy for urinary tract infections

Liou, Natasha; De, Trina; Urbanski, Adrian; Khasriya, Rajvinder; Yakimovich, Artur; Horsley, Harry

doi:10.14278/rodare.2563

September 12, 2023 Dataset Open Access

Clinical urine microscopy for urinary tract infections

Liou, Natasha; De, Trina; Urbanski, Adrian; Khasriya, Rajvinder; Yakimovich, Artur; Horsley, Harry

Urinary tract infections (UTI) are a common disorder. Its diagnosis can be made by microscopic examination of voided urine for cellular markers of infection. We present a dataset containing 300 images and 3,562 manually annotated urinary cells labelled into seven classes of clinically significant urinary content. It is an enriched dataset with samples acquired from the unstained and untreated urine of patients with symptomatic UTI. The aim of the dataset is to facilitate UTI diagnosis in nearly all clinical settings by using a simple imaging system which leverages advanced machine learning techniques.

How to cite us

Liou, Natasha, Trina De, Adrian Urbanski, Catherine Chieng, Qingyang Kong, Anna L. David, Rajvinder Khasriya, Artur Yakimovich, and Harry Horsley. "A clinical microscopy dataset to develop a deep learning diagnostic test for urinary tract infection." Scientific Data 11, no. 1 (2024): 155.

@article{liou2024clinical,
  title={A clinical microscopy dataset to develop a deep learning diagnostic test for urinary tract infection},
  author={Liou, Natasha and De, Trina and Urbanski, Adrian and Chieng, Catherine and Kong, Qingyang and David, Anna L and Khasriya, Rajvinder and Yakimovich, Artur and Horsley, Harry},
  journal={Scientific Data},
  volume={11},
  number={1},
  pages={155},
  year={2024},
  publisher={Nature Publishing Group UK London}
}

Download Timeout Troubleshooting

Use "-C" flag of curl in case you experience timeout of the download:

curl -C - https://rodare...tar.gz_part1\?download\=1 --output ...tar.gz_part1

Data acquisition

300 urine samples were obtained from patients with symptomatic UTI between April and August 2022 from a specialist LUTS outpatient clinic in central London. Urine samples were collected as natural voids and processed on-site within one hour to mitigate cellular degradation. Brightfield microscopic examination (Olympus BX41F microscope frame, U-5RE quintuple nosepiece, U-LS30 LED illuminator, U-AC Abbe condenser) was performed at x20 objective (Olympus PLCN20x Plan C N Achromat 20x/0.4). A disposable haemocytometer (C Chip™) was used for enumeration of red cells (RBC), white cells (WBC), epithelial cells (EPC), and the presence of other cellular content per 1 µl of urine by two experienced microscopists.

Images were acquired using the aforementioned brightfield microscope using a 0.5X C-mount adapter connected to a digital colour camera (Infinity 3S-1UR, Teledyne Lumenera). Images were taken in 16-bit colour in 1392 x 1040 .tif format using Capture and Analyse software. An enriched dataset approach was taken to maximise urinary cellular content in the acquired images. Such data curation was also necessary to overcome class imbalance. Daily Kohler illumination and global white balance was performed to ensure consistency in image acquisition.

Dataset annotation

300 images were acquired and manually annotated by first identifying cells of interest as a binary semantic segmentation task. Individual pixels were dichotomously labelled as either informative cells, foreground, or non-informative background. Non-informative background was further constrained by including unidentifiable cells, such as debris or grossly out-of-focus particles. Binary annotation was initially performed using ilastik, an open-source software using a Random Forest classifier for pixel classification, then manually refined at the pixel level to ensure accurate semantic segmentation. This produced a binary mask in 1392 x 1040 .tif format for each corresponding raw colour image.

Objects of interest were then manually labelled by two expert microscopists into one of seven clinically significant multi-class categories: rods, RBC/WBC, yeast, miscellaneous, single EPC, small EPC sheet, and large EPC sheet. This produced a multi-class mask in 1392 x 1040 .tif format with a label as pixel value from 0-7, where 0 is background (Table 1).

Data structure

The dataset is organised into three root folders: img (image), bin_mask (binary mask), and mult_mask (multi-class mask). Each folder has 300 files in .tif format and labelled with an incremental number.

Table1

Folder         Files        Objects               Count       Pixel Values

img              300        Raw data                                 0-255
bin_mask         300        Background/Foreground                      0/1
mult_mask        300        Background/Class                             0
                            Rod                    1697                  1
                            RBC/WBC                1056                  2
                            Yeast                    41                  3
                            Miscellaneous           550                  4
                            Single EPC              182                  5
                            Small EPC sheet          26                  6
                            Large EPC sheet          10                  7
                                
                            Total                  3562

Preview

Files (2.3 GB)

Name	Size
ds1.zip md5:07d94bb6b35c27bfb478fbeb54a4e992	2.3 GB	Download

3,632

880

views

downloads

See more details...

	All versions	This version
Views	3,632	1,110
Downloads	880	338
Data volume	2.0 TB	765.9 GB
Unique views	3,075	1,028
Unique downloads	450	113

More info on how stats are collected.

Publication date:

September 12, 2023

DOI:

Keyword(s):

clinical microscopy urine microscopy widefield transmission light image segmentation binary segmentation multiclass segmentation

Related identifiers:

Identical to:
https://www.hzdr.de/publications/Publ-37531

Referenced by:
https://www.hzdr.de/publications/Publ-40019

Communities:

License (for files):

Creative Commons Attribution 4.0 International

Versions

Version 1 10.14278/rodare.2563	Sep 12, 2023
Version 1 10.14278/rodare.2562	Sep 12, 2023
Version 1 10.14278/rodare.2473	Sep 12, 2023

Cite all versions? You can cite all versions by using the DOI 10.14278/rodare.2472. This DOI represents all versions, and will always resolve to the latest one. Read more.

Clinical urine microscopy for urinary tract infections

Versions

Share

Cite as

Export

About

Help

Contribute

Follow us

Registered in

Clinical urine microscopy for urinary tract infections

RODARE DOI Badge

DOI

10.14278/rodare.2563

Markdown

[![DOI](https://rodare.hzdr.de/badge/DOI/10.14278/rodare.2563.svg)](https://doi.org/10.14278/rodare.2563)

reStructedText

.. image:: https://rodare.hzdr.de/badge/DOI/10.14278/rodare.2563.svg :target: https://doi.org/10.14278/rodare.2563

HTML

<a href="https://doi.org/10.14278/rodare.2563"><img src="https://rodare.hzdr.de/badge/DOI/10.14278/rodare.2563.svg" alt="DOI"></a>

Image URL

https://rodare.hzdr.de/badge/DOI/10.14278/rodare.2563.svg

Target URL

https://doi.org/10.14278/rodare.2563

Versions

Share

Cite as

Export