-
PoCaP Corpus: A Multimodal Dataset for Smart Operating Room Speech Assistant using Interventional Radiology Workflow Analysis
Authors:
Kubilay Can Demir,
Matthias May,
Axel Schmid,
Michael Uder,
Katharina Breininger,
Tobias Weise,
Andreas Maier,
Seung Hee Yang
Abstract:
This paper presents a new multimodal interventional radiology dataset, called PoCaP (Port Catheter Placement) Corpus. This corpus consists of speech and audio signals in German, X-ray images, and system commands collected from 31 PoCaP interventions by six surgeons with average duration of 81.4 $\pm$ 41.0 minutes. The corpus aims to provide a resource for develo** a smart speech assistant in ope…
▽ More
This paper presents a new multimodal interventional radiology dataset, called PoCaP (Port Catheter Placement) Corpus. This corpus consists of speech and audio signals in German, X-ray images, and system commands collected from 31 PoCaP interventions by six surgeons with average duration of 81.4 $\pm$ 41.0 minutes. The corpus aims to provide a resource for develo** a smart speech assistant in operating rooms. In particular, it may be used to develop a speech controlled system that enables surgeons to control the operation parameters such as C-arm movements and table positions. In order to record the dataset, we acquired consent by the institutional review board and workers council in the University Hospital Erlangen and by the patients for data privacy. We describe the recording set-up, data structure, workflow and preprocessing steps, and report the first PoCaP Corpus speech recognition analysis results with 11.52 $\%$ word error rate using pretrained models. The findings suggest that the data has the potential to build a robust command recognition system and will allow the development of a novel intervention support systems using speech and image processing in the medical domain.
△ Less
Submitted 24 June, 2022;
originally announced June 2022.
-
Deep learning for brain metastasis detection and segmentation in longitudinal MRI data
Authors:
Yixing Huang,
Christoph Bert,
Philipp Sommer,
Benjamin Frey,
Udo Gaipl,
Luitpold V. Distel,
Thomas Weissmann,
Michael Uder,
Manuel A. Schmidt,
Arnd Dörfler,
Andreas Maier,
Rainer Fietkau,
Florian Putz
Abstract:
Brain metastases occur frequently in patients with metastatic cancer. Early and accurate detection of brain metastases is very essential for treatment planning and prognosis in radiation therapy. To improve brain metastasis detection performance with deep learning, a custom detection loss called volume-level sensitivity-specificity (VSS) is proposed, which rates individual metastasis detection sen…
▽ More
Brain metastases occur frequently in patients with metastatic cancer. Early and accurate detection of brain metastases is very essential for treatment planning and prognosis in radiation therapy. To improve brain metastasis detection performance with deep learning, a custom detection loss called volume-level sensitivity-specificity (VSS) is proposed, which rates individual metastasis detection sensitivity and specificity in (sub-)volume levels. As sensitivity and precision are always a trade-off in a metastasis level, either a high sensitivity or a high precision can be achieved by adjusting the weights in the VSS loss without decline in dice score coefficient for segmented metastases. To reduce metastasis-like structures being detected as false positive metastases, a temporal prior volume is proposed as an additional input of DeepMedic. The modified network is called DeepMedic+ for distinction. Our proposed VSS loss improves the sensitivity of brain metastasis detection for DeepMedic, increasing the sensitivity from 85.3% to 97.5%. Alternatively, it improves the precision from 69.1% to 98.7%. Comparing DeepMedic+ with DeepMedic with the same VSS loss, 44.4% of the false positive metastases are reduced in the high sensitivity model and the precision reaches 99.6% for the high specificity model. The mean dice coefficient for all metastases is about 0.81. With the ensemble of the high sensitivity and high specificity models, on average only 1.5 false positive metastases per patient needs further check, while the majority of true positive metastases are confirmed. The ensemble learning is able to distinguish high confidence true positive metastases from metastases candidates that require special expert review or further follow-up, being particularly well-fit to the requirements of expert support in real clinical practice.
△ Less
Submitted 16 September, 2022; v1 submitted 22 December, 2021;
originally announced December 2021.
-
On the field strength dependence of bi- and triexponential intravoxel incoherent motion (IVIM) parameters in the liver
Authors:
Andreas Julian Riexinger,
Jan Martin,
Susanne Rauh,
Andreas Wetscherek,
Mona Pistel,
Tristan Anselm Kuder,
Armin Michael Nagel,
Michael Uder,
Bernhard Hensel,
Lars Müller,
Frederik Bernd Laun
Abstract:
Background: Studies on intravoxel incoherent motion (IVIM) imaging are carried out with different acquisition protocols. Purpose: Investigate the dependence of IVIM parameters on the B_0field strength when using a bi- or triexponential model. Field Strength/Sequence: 20 volunteers were examined at two field strengths (1.5 and 3T). Diffusion-weighted images of the abdomen were acquired at 24 b-valu…
▽ More
Background: Studies on intravoxel incoherent motion (IVIM) imaging are carried out with different acquisition protocols. Purpose: Investigate the dependence of IVIM parameters on the B_0field strength when using a bi- or triexponential model. Field Strength/Sequence: 20 volunteers were examined at two field strengths (1.5 and 3T). Diffusion-weighted images of the abdomen were acquired at 24 b-values ranging from 0.2 to 500 s/mm2. Assessment: ROIs were manually drawn in the liver. Data were fitted with a bi- and a triexponential IVIM model. Resulting parameters were compared between both field strengths. Results: At b-values below 6s/mm2, the triexponential model provided better agreement with the data than the biexponential model. The average tissue diffusivity was D=1.22/1.00 um/ms at 1.5/3T. The average pseudo-diffusion coefficients for the biexponential model were D*=308/260um/ms at 1.5/3T; and for the triexponential model D1*=81.3/65.9um/ms , D2*=2453/2333um/ms at 1.5/3T. The average perfusion fractions for the biexponential model were f=0.286/0.303 at 1.5/3T; and for the triexponential model f1=0.161/0.174 and f2=0.152/0.159 at 1.5/3T. A significant B0 dependence was only found for the biexponential pseudo-diffusion coefficient (ANOVA/KW p=0.037/0.0453) and tissue diffusivity (ANOVA/KW: p<0.001). Conclusion: Our experimental results suggest that triexponential pseudo-diffusion coefficients and perfusion fractions obtained at different field strengths could be compared across different studies using different B0. However, it is recommendable to take the field strength into account when comparing tissue diffusivities or using the biexponential IVIM model. Considering published values for oxygenation-dependent transversal relaxation times of blood, it is unlikely that the two blood compartments of the triexponential model represent venous and arterial blood.
△ Less
Submitted 4 May, 2021;
originally announced May 2021.
-
Contrast-to-noise ratio analysis of microscopic diffusion anisotropy indices in q-space trajectory imaging
Authors:
Jan Martin,
Sebastian Endt,
Andreas Wetscherek,
Tristan Anselm Kuder,
Arnd Doerfler,
Michael Uder,
Bernhard Hensel,
Frederik Bernd Laun
Abstract:
Diffusion anisotropy in diffusion tensor imaging (DTI) is commonly quantified with normalized diffusion anisotropy indices (DAIs). Most often, the fractional anisotropy (FA) is used, but several alternative DAIs have been introduced in attempts to maximize the contrast-to-noise ratio (CNR) in diffusion anisotropy maps. Examples include the scaled relative anisotropy (sRA), the gamma variate anisot…
▽ More
Diffusion anisotropy in diffusion tensor imaging (DTI) is commonly quantified with normalized diffusion anisotropy indices (DAIs). Most often, the fractional anisotropy (FA) is used, but several alternative DAIs have been introduced in attempts to maximize the contrast-to-noise ratio (CNR) in diffusion anisotropy maps. Examples include the scaled relative anisotropy (sRA), the gamma variate anisotropy index (GV), the surface anisotropy (UAsurf), and the lattice index (LI). With the advent of multidimensional diffusion encoding it became possible to determine the presence of microscopic diffusion anisotropy in a voxel, which is theoretically independent of orientation coherence. In accordance with DTI, the microscopic anisotropy is typically quantified by the microscopic fractional anisotropy (uFA). In this work, in addition to the uFA, the four microscopic diffusion anisotropy indices (uDAIs) usRA, uGV, uUAsurf, and uLI are defined in analogy to the respective DAIs by means of the average diffusion tensor and the covariance tensor. Simulations with three representative distributions of microscopic diffusion tensors revealed distinct CNR differences when differentiating between isotropic and microscopically anisotropic diffusion. q-Space trajectory imaging (QTI) was employed to acquire brain in-vivo maps of all indices. For this purpose, a 15 min protocol featuring linear, planar, and spherical tensor encoding was used. The resulting maps were of good quality and exhibited different contrasts, e.g. between gray and white matter. This indicates that it may be beneficial to use more than one uDAI in future investigational studies.
△ Less
Submitted 2 April, 2020;
originally announced April 2020.