High-Resolution Road Vehicle Collision Prediction for the City of Montreal

Hébert, Antoine; Guédon, Timothée; Glatard, Tristan; Jaumard, Brigitte

doi:10.1109/BigData47090.2019.9006009

Computer Science > Machine Learning

arXiv:1905.08770 (cs)

[Submitted on 21 May 2019 (v1), last revised 11 Nov 2019 (this version, v3)]

Title:High-Resolution Road Vehicle Collision Prediction for the City of Montreal

Authors:Antoine Hébert, Timothée Guédon, Tristan Glatard, Brigitte Jaumard

View PDF

Abstract:Road accidents are an important issue of our modern societies, responsible for millions of deaths and injuries every year in the world. In Quebec only, in 2018, road accidents are responsible for 359 deaths and 33 thousands of injuries. In this paper, we show how one can leverage open datasets of a city like Montreal, Canada, to create high-resolution accident prediction models, using big data analytics. Compared to other studies in road accident prediction, we have a much higher prediction resolution, i.e., our models predict the occurrence of an accident within an hour, on road segments defined by intersections. Such models could be used in the context of road accident prevention, but also to identify key factors that can lead to a road accident, and consequently, help elaborate new policies.
We tested various machine learning methods to deal with the severe class imbalance inherent to accident prediction problems. In particular, we implemented the Balanced Random Forest algorithm, a variant of the Random Forest machine learning algorithm in Apache Spark. Interestingly, we found that in our case, Balanced Random Forest does not perform significantly better than Random Forest.
Experimental results show that 85% of road vehicle collisions are detected by our model with a false positive rate of 13%. The examples identified as positive are likely to correspond to high-risk situations. In addition, we identify the most important predictors of vehicle collisions for the area of Montreal: the count of accidents on the same road segment during previous years, the temperature, the day of the year, the hour and the visibility.

Subjects:	Machine Learning (cs.LG); Machine Learning (stat.ML)
Cite as:	arXiv:1905.08770 [cs.LG]
	(or arXiv:1905.08770v3 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.1905.08770
Journal reference:	2019 IEEE International Conference on Big Data, pp. 1804-1813
Related DOI:	https://doi.org/10.1109/BigData47090.2019.9006009

Submission history

From: Antoine Hébert [view email]
[v1] Tue, 21 May 2019 17:41:23 UTC (122 KB)
[v2] Wed, 23 Oct 2019 17:05:50 UTC (122 KB)
[v3] Mon, 11 Nov 2019 18:50:44 UTC (123 KB)

Computer Science > Machine Learning

Title:High-Resolution Road Vehicle Collision Prediction for the City of Montreal

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:High-Resolution Road Vehicle Collision Prediction for the City of Montreal

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators