Skip to main content

Showing 1–7 of 7 results for author: Szarvas, G

Searching in archive cs. Search in all archives.
.
  1. arXiv:2306.03315  [pdf, other

    cs.CL cs.AI

    Few Shot Rationale Generation using Self-Training with Dual Teachers

    Authors: Aditya Srikanth Veerubhotla, Lahari Poddar, Jun Yin, György Szarvas, Sharanya Eswaran

    Abstract: Self-rationalizing models that also generate a free-text explanation for their predicted labels are an important tool to build trustworthy AI applications. Since generating explanations for annotated labels is a laborious and costly pro cess, recent models rely on large pretrained language models (PLMs) as their backbone and few-shot learning. In this work we explore a self-training approach lever… ▽ More

    Submitted 5 June, 2023; originally announced June 2023.

    Comments: ACL Findings 2023

  2. arXiv:2210.14379  [pdf, other

    cs.CL

    Deploying a Retrieval based Response Model for Task Oriented Dialogues

    Authors: Lahari Poddar, György Szarvas, Cheng Wang, Jorge Balazs, Pavel Danchenko, Patrick Ernst

    Abstract: Task-oriented dialogue systems in industry settings need to have high conversational capability, be easily adaptable to changing situations and conform to business constraints. This paper describes a 3-step procedure to develop a conversational model that satisfies these criteria and can efficiently scale to rank a large set of response candidates. First, we provide a simple algorithm to semi-auto… ▽ More

    Submitted 25 October, 2022; originally announced October 2022.

    Comments: Accepted at EMNLP 2022

  3. arXiv:2205.07633  [pdf, other

    cs.CL cs.AI cs.LG

    Taming Continuous Posteriors for Latent Variational Dialogue Policies

    Authors: Marin Vlastelica, Patrick Ernst, György Szarvas

    Abstract: Utilizing amortized variational inference for latent-action reinforcement learning (RL) has been shown to be an effective approach in Task-oriented Dialogue (ToD) systems for optimizing dialogue success. Until now, categorical posteriors have been argued to be one of the main drivers of performance. In this work we revisit Gaussian variational posteriors for latent-action RL and show that they can… ▽ More

    Submitted 1 June, 2022; v1 submitted 16 May, 2022; originally announced May 2022.

  4. arXiv:2112.13776  [pdf, other

    cs.CL cs.AI

    Transformer Uncertainty Estimation with Hierarchical Stochastic Attention

    Authors: Jiahuan Pei, Cheng Wang, György Szarvas

    Abstract: Transformers are state-of-the-art in a wide range of NLP tasks and have also been applied to many real-world products. Understanding the reliability and certainty of transformer model predictions is crucial for building trustable machine learning applications, e.g., medical diagnosis. Although many recent transformer extensions have been proposed, the study of the uncertainty estimation of transfo… ▽ More

    Submitted 27 December, 2021; originally announced December 2021.

    Comments: AAAI 2022

  5. arXiv:2101.11958  [pdf, other

    cs.CL

    Attention Guided Dialogue State Tracking with Sparse Supervision

    Authors: Shuailong Liang, Lahari Poddar, Gyuri Szarvas

    Abstract: Existing approaches to Dialogue State Tracking (DST) rely on turn level dialogue state annotations, which are expensive to acquire in large scale. In call centers, for tasks like managing bookings or subscriptions, the user goal can be associated with actions (e.g.~API calls) issued by customer service agents. These action logs are available in large volumes and can be utilized for learning dialog… ▽ More

    Submitted 28 January, 2021; originally announced January 2021.

    Comments: 10 pages, 6 figures

  6. arXiv:2010.02573  [pdf, other

    cs.CL cs.IR cs.LG

    The Multilingual Amazon Reviews Corpus

    Authors: Phillip Keung, Yichao Lu, György Szarvas, Noah A. Smith

    Abstract: We present the Multilingual Amazon Reviews Corpus (MARC), a large-scale collection of Amazon reviews for multilingual text classification. The corpus contains reviews in English, Japanese, German, French, Spanish, and Chinese, which were collected between 2015 and 2019. Each record in the dataset contains the review text, the review title, the star rating, an anonymized reviewer ID, an anonymized… ▽ More

    Submitted 6 October, 2020; originally announced October 2020.

    Comments: To appear in EMNLP 2020

  7. arXiv:1910.07333  [pdf, other

    cs.CL cs.LG

    A Probabilistic Framework for Learning Domain Specific Hierarchical Word Embeddings

    Authors: Lahari Poddar, Gyorgy Szarvas, Lea Frermann

    Abstract: The meaning of a word often varies depending on its usage in different domains. The standard word embedding models struggle to represent this variation, as they learn a single global representation for a word. We propose a method to learn domain-specific word embeddings, from text organized into hierarchical domains, such as reviews in an e-commerce website, where products follow a taxonomy. Our s… ▽ More

    Submitted 20 October, 2019; v1 submitted 16 October, 2019; originally announced October 2019.