Customer-obsessed science
Research areas
-
August 26, 20265 min readDiscounting the opinions of LLM judges with highly correlated outputs ensures that panels of judges reflect a true diversity of perspectives.
-
August 21, 20269 min read
-
July 30, 20268 min read
-
-
July 9, 202610 min read
Featured news
-
Interspeech 20202020Wakeword detection is responsible for switching on downstream systems in a voice-activated device. To prevent a response when the wakeword is detected by mistake, a secondary network is often utilized to verify the detected wakeword. Published verification approaches are formulated based on Automatic Speech Recognition (ASR) biased towards the wakeword. This approach has several drawbacks, including high
-
SIGIR 2020 Workshop on eCommerce2020In this paper, we study the problem of enabling multi-lingual product search for a global shopping store. In particular, given an existing search system and product catalog in a primary language, and a search query in a secondary language, transform the query into a semantically equivalent one in the primary language in order to retrieve the most relevant products. Direct application of machine translation
-
SIGIR 2020 Workshop on eCommerce2020Large-scale information retrieval systems store documents in different shards. Shard selection enables cost-effective retrieval by searching only relevant shards for the query. Most existing shard selection algorithms focus on web search, and rely on text similarity between the query and shard corpora. In contrast, in e-commerce product search, shards are defined according to product categories, and most
-
ICML 2020 Workshop on HILL2020Human attention is a scarce resource in modern computing. A multitude of microtasks vie for user attention to crowdsource information, perform momentary assessments, personalize services, and execute actions with a single touch. A lot gets done when these tasks take up the invisible free moments of the day. However, an interruption at an inappropriate time degrades productivity and causes annoyance. Prior
-
VLDB 20202020How should we split data among the nodes of a distributed data warehouse in order to boost performance for a forecasted workload? In this paper, we study the effect of different data partitioning schemes on the overall network cost of pairwise joins. We describe a generally-applicable data distribution framework initially designed for Amazon Redshift, a fully-managed petabyte-scale data warehouse in the
Collaborations
View allWhether you're a faculty member or student, there are number of ways you can engage with Amazon.
View all