Customer-obsessed science
Research areas
-
September 21, 202611 min readThree new papers from Amazon Bio Discovery address bottlenecks in AI-driven antibody engineering, from benchmarking binding predictors to experimentally validating de novo design.
-
-
August 21, 20269 min read
-
July 30, 20268 min read
-
Featured news
-
KDD 2023 Workshop on Online and Adaptive Recommender Systems (OARS)2023Rogue actors employ sophisticated automation techniques to mimic human browsing/click patterns and generate invalid (i.e., fraudulent or robotic) traffic on retail marketplaces to artificially inflate their key performance metrics at the expense of their legitimate competitors. To maintain a clean and fair advertising system, it is essential to identify and mitigate ad traffic that is invalid, i.e., fraudulent
-
KDD 2023 Workshop on Machine Learning in Finance (MLF)2023Subledgers maintain detailed information about specific accounts or transactions in order to substantiate the general ledger. Subledgers provide a granular level of detail for financial reporting and analysis, which is especially essential for accounts receivables and payables. The size of subledgers can vary greatly depending on the complexity and volume of transactions and their size can also increase
-
KDD 2023 Workshop on Robust NLP for Finance (RobustFin)2023In large corporations, millions of cash transactions are booked via cash management software (CMS) per month. Most CMS systems adopt a key-word (search string) based matching logic for booking, which checks if the cash transaction description contains a specific search string and books the transaction to an appropriate general ledger account (GL-account) according to a booking rule. However, due to the
-
Interspeech 20232023In this work, we introduce a diffusion-based text-to-speech (TTS) system for accent modelling. TTS systems have become a natural part of our surroundings. Nevertheless, because of the complexity of accent modelling, recent state-of-the-art solutions mainly focus on the most common variants of each language. In this work, we propose to address this issue with a newly proposed diffusion generative model (
-
ACL Findings 20232023Multilingual information retrieval (IR) is challenging since annotated training data is costly to obtain in many languages. We present an effective method to train multilingual IR systems when only English IR training data and some parallel corpora between English and other languages are available. We leverage parallel and non-parallel corpora to improve the pretrained multilingual language models’ cross-lingual
Collaborations
View allWhether you're a faculty member or student, there are number of ways you can engage with Amazon.
View all