Customer-obsessed science
Research areas
-
August 26, 20265 min readDiscounting the opinions of LLM judges with highly correlated outputs ensures that panels of judges reflect a true diversity of perspectives.
-
August 21, 20269 min read
-
July 30, 20268 min read
-
-
July 9, 202610 min read
Featured news
-
ACM SIGKDD MiLeTs 20262026Multivariate time series contain two kinds of cross-variable relationships: persistent ones that reflect underlying structure (geographic proximity, shared infrastructure, physical coupling) and dynamic ones that arise from transient conditions in each observation window. Current transformer architectures conflate the two—channel-independent models ignore cross-variable relationships entirely, while cross-variable
-
AutoML Conference 20262026Bayesian hyperparameter optimization typically requires fitting a surrogate model to each new task, incurring per-task training cost that grows with the number of observations and limits deployment flexibility. We show that TabPFN v2 (Hollmann et al., 2025), a pretrained tabular foundation model never trained on Bayesian optimization data, can serve as a drop-in zero-shot BO surrogate, eliminating the per-task
-
ICCCN 20262026The Internet consists of interconnected, independently managed Autonomous Systems (AS) that rely on the Border Gateway Protocol (BGP) for inter-domain routing. BGP anomalies—such as route leaks and hijacks—can divert traffic through unauthorized or inefficient paths, jeopardizing network reliability and security. Although existing rule-based and machine learning methods can detect these anomalies using
-
2026A/B testing remains the standard for rolling out new features in the technology industry. Each experiment, however, consumes real traffic, engineering effort, and weeks of wall-clock time. Can AI agents—conditioned on behavioral profiles and contextual descriptions of the intervention—simulate outcomes accurately enough to vet candidate treatments before committing live traffic? We formalize this question
-
2026Text-to-image models have made significant strides, producing impressive results in generating images from textual descriptions. However, creating a scalable pipeline for deploying these models in production remains a challenge. Achieving the right balance between automation and human feedback is critical to maintain both scale and quality. While automation can handle large volumes, human oversight is still
Collaborations
View allWhether you're a faculty member or student, there are number of ways you can engage with Amazon.
View all