Skip to main content

Showing 1–12 of 12 results for author: Hellinckx, P

Searching in archive cs. Search in all archives.
.
  1. Autonomous Port Navigation With Ranging Sensors Using Model-Based Reinforcement Learning

    Authors: Siemen Herremans, Ali Anwar, Arne Troch, Ian Ravijts, Maarten Vangeneugden, Siegfried Mercelis, Peter Hellinckx

    Abstract: Autonomous ship** has recently gained much interest in the research community. However, little research focuses on inland - and port navigation, even though this is identified by countries such as Belgium and the Netherlands as an essential step towards a sustainable future. These environments pose unique challenges, since they can contain dynamic obstacles that do not broadcast their location,… ▽ More

    Submitted 17 November, 2023; originally announced December 2023.

    Comments: Presented at 42nd International Conference on Ocean, Offshore & Arctic Engineering. June 11 - 16, 2023. Melbourne, Australia

    Journal ref: Proceedings of the ASME 2023 42nd International Conference on Ocean, Offshore and Arctic Engineering. Volume 5: Ocean Engineering. Melbourne, Australia. June 11-16, 2023. V005T06A072. ASME

  2. Safety Aware Autonomous Path Planning Using Model Predictive Reinforcement Learning for Inland Waterways

    Authors: Astrid Vanneste, Simon Vanneste, Olivier Vasseur, Robin Janssens, Mattias Billast, Ali Anwar, Kevin Mets, Tom De Schepper, Siegfried Mercelis, Peter Hellinckx

    Abstract: In recent years, interest in autonomous ship** in urban waterways has increased significantly due to the trend of kee** cars and trucks out of city centers. Classical approaches such as Frenet frame based planning and potential field navigation often require tuning of many configuration parameters and sometimes even require a different configuration depending on the situation. In this paper, w… ▽ More

    Submitted 16 November, 2023; originally announced November 2023.

    Comments: \c{opyright} 2022 IEEE. Personal use of this material is permitted. Permission from IEEE must be obtained for all other uses, in any current or future media, including reprinting/republishing this material for advertising or promotional purposes, creating new collective works, for resale or redistribution to servers or lists, or reuse of any copyrighted component of this work in other works

  3. arXiv:2308.04938  [pdf, other

    cs.LG cs.AI cs.MA

    An In-Depth Analysis of Discretization Methods for Communication Learning using Backpropagation with Multi-Agent Reinforcement Learning

    Authors: Astrid Vanneste, Simon Vanneste, Kevin Mets, Tom De Schepper, Siegfried Mercelis, Peter Hellinckx

    Abstract: Communication is crucial in multi-agent reinforcement learning when agents are not able to observe the full state of the environment. The most common approach to allow learned communication between agents is the use of a differentiable communication channel that allows gradients to flow between agents as a form of feedback. However, this is challenging when we want to use discrete messages to redu… ▽ More

    Submitted 9 August, 2023; originally announced August 2023.

    Comments: arXiv admin note: substantial text overlap with arXiv:2204.05669

  4. arXiv:2308.04844  [pdf, other

    cs.LG cs.AI cs.MA

    Scalability of Message Encoding Techniques for Continuous Communication Learned with Multi-Agent Reinforcement Learning

    Authors: Astrid Vanneste, Thomas Somers, Simon Vanneste, Kevin Mets, Tom De Schepper, Siegfried Mercelis, Peter Hellinckx

    Abstract: Many multi-agent systems require inter-agent communication to properly achieve their goal. By learning the communication protocol alongside the action protocol using multi-agent reinforcement learning techniques, the agents gain the flexibility to determine which information should be shared. However, when the number of agents increases we need to create an encoding of the information contained in… ▽ More

    Submitted 9 August, 2023; originally announced August 2023.

    Comments: Paper accepted to the BNAIC/BeNeLearn 2022 conference

  5. Deep set conditioned latent representations for action recognition

    Authors: Akash Singh, Tom De Schepper, Kevin Mets, Peter Hellinckx, Jose Oramas, Steven Latre

    Abstract: In recent years multi-label, multi-class video action recognition has gained significant popularity. While reasoning over temporally connected atomic actions is mundane for intelligent species, standard artificial neural networks (ANN) still struggle to classify them. In the real world, atomic actions often temporally connect to form more complex composite actions. The challenge lies in recognisin… ▽ More

    Submitted 21 December, 2022; originally announced December 2022.

    Comments: Conference VISAPP 2022, 11 pages,5 figures, 2 Tables, 6 plots

    Journal ref: In Proceedings of the 17th International Joint Conference on Computer Vision, Imaging and Computer Graphics Theory and Applications - Volume 5: VISAPP, ISBN 978-989-758-555-5; ISSN 2184-4321, year 2022, pages 456-466

  6. arXiv:2204.05669  [pdf, other

    cs.LG cs.MA

    An Analysis of Discretization Methods for Communication Learning with Multi-Agent Reinforcement Learning

    Authors: Astrid Vanneste, Simon Vanneste, Kevin Mets, Tom De Schepper, Siegfried Mercelis, Steven Latré, Peter Hellinckx

    Abstract: Communication is crucial in multi-agent reinforcement learning when agents are not able to observe the full state of the environment. The most common approach to allow learned communication between agents is the use of a differentiable communication channel that allows gradients to flow between agents as a form of feedback. However, this is challenging when we want to use discrete messages to redu… ▽ More

    Submitted 12 April, 2022; originally announced April 2022.

    Comments: Accepted at Adaptive and Learning Agents Workshop (ALA 2022) https://ala2022.github.io/

  7. Learning to Communicate with Reinforcement Learning for an Adaptive Traffic Control System

    Authors: Simon Vanneste, Gauthier de Borrekens, Stig Bosmans, Astrid Vanneste, Kevin Mets, Siegfried Mercelis, Steven Latré, Peter Hellinckx

    Abstract: Recent work in multi-agent reinforcement learning has investigated inter agent communication which is learned simultaneously with the action policy in order to improve the team reward. In this paper, we investigate independent Q-learning (IQL) without communication and differentiable inter-agent learning (DIAL) with learned communication on an adaptive traffic control system (ATCS). In real world… ▽ More

    Submitted 29 October, 2021; originally announced October 2021.

  8. Mixed Cooperative-Competitive Communication Using Multi-Agent Reinforcement Learning

    Authors: Astrid Vanneste, Wesley Van Wijnsberghe, Simon Vanneste, Kevin Mets, Siegfried Mercelis, Steven Latré, Peter Hellinckx

    Abstract: By using communication between multiple agents in multi-agent environments, one can reduce the effects of partial observability by combining one agent's observation with that of others in the same dynamic environment. While a lot of successful research has been done towards communication learning in cooperative settings, communication learning in mixed cooperative-competitive settings is also impo… ▽ More

    Submitted 29 October, 2021; originally announced October 2021.

  9. arXiv:2110.06742  [pdf, other

    cs.LG

    A Review of the Deep Sea Treasure problem as a Multi-Objective Reinforcement Learning Benchmark

    Authors: Amber Cassimon, Reinout Eyckerman, Siegfried Mercelis, Steven Latré, Peter Hellinckx

    Abstract: In this paper, the authors investigate the Deep Sea Treasure (DST) problem as proposed by Vamplew et al. Through a number of proofs, the authors show the original DST problem to be quite basic, and not always representative of practical Multi-Objective Optimization problems. In an attempt to bring theory closer to practice, the authors propose an alternative, improved version of the DST problem, a… ▽ More

    Submitted 21 May, 2024; v1 submitted 13 October, 2021; originally announced October 2021.

    Comments: 10 pages, 4 figures; Fixed Supplementary Materials PDF

  10. arXiv:2010.03429  [pdf, other

    cs.LG cs.AI

    Exploiting non-i.i.d. data towards more robust machine learning algorithms

    Authors: Wim Casteels, Peter Hellinckx

    Abstract: In the field of machine learning there is a growing interest towards more robust and generalizable algorithms. This is for example important to bridge the gap between the environment in which the training data was collected and the environment where the algorithm is deployed. Machine learning algorithms have increasingly been shown to excel in finding patterns and correlations from data. Determini… ▽ More

    Submitted 7 October, 2020; originally announced October 2020.

  11. arXiv:2006.07200  [pdf, other

    cs.LG cs.MA stat.ML

    Learning to Communicate Using Counterfactual Reasoning

    Authors: Simon Vanneste, Astrid Vanneste, Kevin Mets, Tom De Schepper, Ali Anwar, Siegfried Mercelis, Steven Latré, Peter Hellinckx

    Abstract: Learning to communicate in order to share state information is an active problem in the area of multi-agent reinforcement learning (MARL). The credit assignment problem, the non-stationarity of the communication environment and the creation of influenceable agents are major challenges within this research field which need to be overcome in order to learn a valid communication protocol. This paper… ▽ More

    Submitted 26 April, 2022; v1 submitted 12 June, 2020; originally announced June 2020.

    Comments: Accepted at Adaptive and Learning Agents Workshop (ALA 2022) https://ala2022.github.io/

  12. arXiv:1802.09207  [pdf, other

    cs.DC

    Towards evaluating emergent behavior of the Internet of Things using large scale simulation techniques

    Authors: Stig Bosmans, Siegfried Mercelis Peter Hellinckx, Joachim Denil

    Abstract: With the increase in Internet of Things devices and more decentralized architectures we see a new type of application gain importance, a type where local interactions between individual entities lead to a global emergent behavior, Emergent-based IoT (EBI) Systems. In this position paper we explore techniques to evaluate this emergent behavior in IoT applications. Because of the required scale and… ▽ More

    Submitted 26 February, 2018; originally announced February 2018.