Skip to main content

Showing 1–8 of 8 results for author: Tuggener, L

Searching in archive cs. Search in all archives.
.
  1. arXiv:2311.08525  [pdf, other

    cs.CV cs.AI

    Efficient Rotation Invariance in Deep Neural Networks through Artificial Mental Rotation

    Authors: Lukas Tuggener, Thilo Stadelmann, Jürgen Schmidhuber

    Abstract: Humans and animals recognize objects irrespective of the beholder's point of view, which may drastically change their appearances. Artificial pattern recognizers also strive to achieve this, e.g., through translational invariance in convolutional neural networks (CNNs). However, both CNNs and vision transformers (ViTs) perform very poorly on rotated inputs. Here we present artificial mental rotati… ▽ More

    Submitted 14 November, 2023; originally announced November 2023.

  2. Video object detection for privacy-preserving patient monitoring in intensive care

    Authors: Raphael Emberger, Jens Michael Boss, Daniel Baumann, Marko Seric, Shufan Huo, Lukas Tuggener, Emanuela Keller, Thilo Stadelmann

    Abstract: Patient monitoring in intensive care units, although assisted by biosensors, needs continuous supervision of staff. To reduce the burden on staff members, IT infrastructures are built to record monitoring data and develop clinical decision support systems. These systems, however, are vulnerable to artifacts (e.g. muscle movement due to ongoing treatment), which are often indistinguishable from rea… ▽ More

    Submitted 26 June, 2023; originally announced June 2023.

    Comments: 4 pages, 3 figures, 2023 10th Swiss Conference on Data Science (SDS), code available at https://github.com/raember/yolov5r_autodidact and https://github.com/raember/VideoProc

    ACM Class: I.2.10

  3. Is it enough to optimize CNN architectures on ImageNet?

    Authors: Lukas Tuggener, Jürgen Schmidhuber, Thilo Stadelmann

    Abstract: Classification performance based on ImageNet is the de-facto standard metric for CNN development. In this work we challenge the notion that CNN architecture design solely based on ImageNet leads to generally effective convolutional neural network (CNN) architectures that perform well on a diverse set of datasets and application domains. To this end, we investigate and ultimately improve ImageNet a… ▽ More

    Submitted 6 March, 2023; v1 submitted 16 March, 2021; originally announced March 2021.

    Journal ref: Frontiers in Computer Science, Volume 4, 2022

  4. arXiv:1907.08392  [pdf, other

    cs.LG cs.AI stat.ML

    Automated Machine Learning in Practice: State of the Art and Recent Results

    Authors: Lukas Tuggener, Mohammadreza Amirian, Katharina Rombach, Stefan Lörwald, Anastasia Varlet, Christian Westermann, Thilo Stadelmann

    Abstract: A main driver behind the digitization of industry and society is the belief that data-driven model building and decision making can contribute to higher degrees of automation and more informed decisions. Building such models from data often involves the application of some form of machine learning. Thus, there is an ever growing demand in work force with the necessary skill set to do so. This dema… ▽ More

    Submitted 19 July, 2019; originally announced July 2019.

    Comments: Accepted full paper at SDS2019, the 6th Swiss Conference on Data Science

  5. arXiv:1810.05423  [pdf, other

    cs.CV

    DeepScores and Deep Watershed Detection: current state and open issues

    Authors: Ismail Elezi, Lukas Tuggener, Marcello Pelillo, Thilo Stadelmann

    Abstract: This paper gives an overview of our current Optical Music Recognition (OMR) research. We recently released the OMR dataset \emph{DeepScores} as well as the object detection method \emph{Deep Watershed Detector}. We are currently taking some additional steps to improve both of them. Here we summarize current and future efforts, aimed at improving usefulness on real-world task and tackling extreme c… ▽ More

    Submitted 12 October, 2018; originally announced October 2018.

    Comments: Published on WORMS workshop (ISMIR affiliated workshop)

  6. arXiv:1807.04950  [pdf, other

    cs.LG cs.AI cs.CV stat.ML

    Deep Learning in the Wild

    Authors: Thilo Stadelmann, Mohammadreza Amirian, Ismail Arabaci, Marek Arnold, Gilbert François Duivesteijn, Ismail Elezi, Melanie Geiger, Stefan Lörwald, Benjamin Bruno Meier, Katharina Rombach, Lukas Tuggener

    Abstract: Deep learning with neural networks is applied by an increasing number of people outside of classic research environments, due to the vast success of the methodology on a wide range of machine perception tasks. While this interest is fueled by beautiful success stories, practical work in deep learning on novel tasks without existing baselines remains challenging. This paper explores the specific ch… ▽ More

    Submitted 13 July, 2018; originally announced July 2018.

    Comments: Invited paper on ANNPR 2018

  7. arXiv:1805.10548  [pdf, other

    cs.CV cs.AI

    Deep Watershed Detector for Music Object Recognition

    Authors: Lukas Tuggener, Ismail Elezi, Jurgen Schmidhuber, Thilo Stadelmann

    Abstract: Optical Music Recognition (OMR) is an important and challenging area within music information retrieval, the accurate detection of music symbols in digital images is a core functionality of any OMR pipeline. In this paper, we introduce a novel object detection method, based on synthetic energy maps and the watershed transform, called Deep Watershed Detector (DWD). Our method is specifically tailor… ▽ More

    Submitted 26 May, 2018; originally announced May 2018.

    Comments: Accepted on The 19th International Society for Music Information Retrieval Conference 2018

  8. arXiv:1804.00525  [pdf, other

    cs.CV cs.LG

    DeepScores -- A Dataset for Segmentation, Detection and Classification of Tiny Objects

    Authors: Lukas Tuggener, Ismail Elezi, Jürgen Schmidhuber, Marcello Pelillo, Thilo Stadelmann

    Abstract: We present the DeepScores dataset with the goal of advancing the state-of-the-art in small objects recognition, and by placing the question of object recognition in the context of scene understanding. DeepScores contains high quality images of musical scores, partitioned into 300,000 sheets of written music that contain symbols of different shapes and sizes. With close to a hundred millions of sma… ▽ More

    Submitted 26 May, 2018; v1 submitted 27 March, 2018; originally announced April 2018.

    Comments: 6 pages, accepted on IEEE International Conference on Pattern Recognition 2018