Customer-obsessed science
Research areas
-
September 25, 202612 min readUsing the Neuron Kernel Interface, a Reactor–AWS collaboration tackled the dynamic shapes, memory access patterns, and cache management that make real-time autoregressive diffusion hard—building techniques that generalize across models.
-
September 21, 202611 min read
-
-
August 21, 20269 min read
-
July 30, 20268 min read
Featured news
-
ICLR 2023 Workshop on Deep Learning for Code (DL4C)2023Code understanding and generation require learning the mapping between human and programming languages. As human and programming languages are different in vocabulary, semantic, and, syntax, it is challenging for an autoregressive model to generate a sequence of tokens that is both semantically (i.e., carry the right meaning) and syntactically correct (i.e., in the right sequence order). Inspired by this
-
Web Conference 2023 Workshop on Natural-Language Processing for Social Media2023Language model pre-training has led to state-of-the-art performance in text summarization. While a variety of pre-trained transformer models are available nowadays, they are mostly trained on documents. In this study we introduce self-supervised pre-training to enhance the BERT model’s semantic and structural understanding of dialog texts from social media. We also propose a semisupervised teacher-student
-
ICASSP 20232023This work focuses on modelling a speaker’s accent that does not have a dedicated text-to-speech (TTS) frontend, includ-ing a grapheme-to-phoneme (G2P) module. Prior work on modelling accents assumes a phonetic transcription is avail-able for the target accent, which might not be the case for low-resource, regional accents. In our work, we propose an approach whereby we first augment the target accent data
-
ECIR 20232023AI assistants are gradually becoming embedded in our lives, utilized for everyday tasks like shopping or music. In addition to the everyday utilization of AI assistants, many users engage them with playful shopping requests, gauging their ability to understand – or simply seeking amusement. However, these requests are often not being responded to in the same playful manner, causing dissatisfaction and even
-
EACL 20232023This work focuses on in-context data augmenta-tion for intent detection. Having found that aug-mentation via in-context prompting of large pre-trained language models (PLMs) alone does not improve performance, we introduce a novel approach based on PLMs and pointwise V-information (PVI), a metric that can measure the usefulness of a datapoint for training a model. Our method first fine-tunes a PLM on a
Collaborations
View allWhether you're a faculty member or student, there are number of ways you can engage with Amazon.
View all