Skip to main content

Showing 1–4 of 4 results for author: Unni, S

.
  1. arXiv:2311.00807  [pdf, other

    cs.CV cs.LG

    VQA-GEN: A Visual Question Answering Benchmark for Domain Generalization

    Authors: Suraj Jyothi Unni, Raha Moraffah, Huan Liu

    Abstract: Visual question answering (VQA) models are designed to demonstrate visual-textual reasoning capabilities. However, their real-world applicability is hindered by a lack of comprehensive benchmark datasets. Existing domain generalization datasets for VQA exhibit a unilateral focus on textual shifts while VQA being a multi-modal task contains shifts across both visual and textual domains. We propose… ▽ More

    Submitted 1 November, 2023; originally announced November 2023.

  2. arXiv:2307.13757  [pdf, other

    cs.LG cs.HC stat.ME

    UPREVE: An End-to-End Causal Discovery Benchmarking System

    Authors: Suraj Jyothi Unni, Paras Sheth, Kaize Ding, Huan Liu, K. Selcuk Candan

    Abstract: Discovering causal relationships in complex socio-behavioral systems is challenging but essential for informed decision-making. We present Upload, PREprocess, Visualize, and Evaluate (UPREVE), a user-friendly web-based graphical user interface (GUI) designed to simplify the process of causal discovery. UPREVE allows users to run multiple algorithms simultaneously, visualize causal relationships, a… ▽ More

    Submitted 25 July, 2023; originally announced July 2023.

    Comments: 8 pages, Accepted to SBP-BRiMS 2023

  3. arXiv:2307.05006  [pdf, ps, other

    cs.CL cs.LG eess.AS

    Improving RNN-Transducers with Acoustic LookAhead

    Authors: Vinit S. Unni, Ashish Mittal, Preethi Jyothi, Sunita Sarawagi

    Abstract: RNN-Transducers (RNN-Ts) have gained widespread acceptance as an end-to-end model for speech to text conversion because of their high accuracy and streaming capabilities. A typical RNN-T independently encodes the input audio and the text context, and combines the two encodings by a thin joint network. While this architecture provides SOTA streaming accuracy, it also makes the model vulnerable to s… ▽ More

    Submitted 10 July, 2023; originally announced July 2023.

    Comments: 5 pages, 1 fig, 7 tables, Proceedings of Interspeech 2023

  4. arXiv:2111.09243  [pdf, other

    cs.HC

    An Investigation into Keystroke Dynamics and Heart Rate Variability as Indicators of Stress

    Authors: Srijith Unni, Sushma Suryanarayana Gowda, Alan F. Smeaton

    Abstract: Lifelogging has become a prominent research topic in recent years. Wearable sensors like Fitbits and smart watches are now increasingly popular for recording ones activities. Some researchers are also exploring keystroke dynamics for lifelogging. Keystroke dynamics refers to the process of measuring and assessing a persons ty** rhythm on digital devices. A digital footprint is created when a use… ▽ More

    Submitted 17 November, 2021; originally announced November 2021.

    Comments: 12 pages. To appear at MMM 2022, 28th International Conference on Multimedia Modeling, 5-8 April 2022, Phu Quoc, Vietnam