Skip to main content

Showing 1–3 of 3 results for author: Foret, P

Searching in archive cs. Search in all archives.
.
  1. arXiv:2012.07976  [pdf, other

    cs.LG stat.ML

    NeurIPS 2020 Competition: Predicting Generalization in Deep Learning

    Authors: Yiding Jiang, Pierre Foret, Scott Yak, Daniel M. Roy, Hossein Mobahi, Gintare Karolina Dziugaite, Samy Bengio, Suriya Gunasekar, Isabelle Guyon, Behnam Neyshabur

    Abstract: Understanding generalization in deep learning is arguably one of the most important questions in deep learning. Deep learning has been successfully adopted to a large number of problems ranging from pattern recognition to complex decision making, but many recent researchers have raised many concerns about deep learning, among which the most important is generalization. Despite numerous attempts, c… ▽ More

    Submitted 14 December, 2020; originally announced December 2020.

    Comments: 20 pages, 2 figures. Accepted for NeurIPS 2020 Competitions Track. Lead organizer: Yiding Jiang

  2. arXiv:2010.01412  [pdf, other

    cs.LG stat.ML

    Sharpness-Aware Minimization for Efficiently Improving Generalization

    Authors: Pierre Foret, Ariel Kleiner, Hossein Mobahi, Behnam Neyshabur

    Abstract: In today's heavily overparameterized models, the value of the training loss provides few guarantees on model generalization ability. Indeed, optimizing only the training loss value, as is commonly done, can easily lead to suboptimal model quality. Motivated by prior work connecting the geometry of the loss landscape and generalization, we introduce a novel, effective procedure for instead simultan… ▽ More

    Submitted 29 April, 2021; v1 submitted 3 October, 2020; originally announced October 2020.

  3. arXiv:2002.02955  [pdf, ps, other

    cs.CL

    A Multilingual View of Unsupervised Machine Translation

    Authors: Xavier Garcia, Pierre Foret, Thibault Sellam, Ankur P. Parikh

    Abstract: We present a probabilistic framework for multilingual neural machine translation that encompasses supervised and unsupervised setups, focusing on unsupervised translation. In addition to studying the vanilla case where there is only monolingual data available, we propose a novel setup where one language in the (source, target) pair is not associated with any parallel data, but there may exist auxi… ▽ More

    Submitted 16 October, 2020; v1 submitted 7 February, 2020; originally announced February 2020.

    Comments: Accepted at Findings of EMNLP 2020 [Fixed processing error.]