Johns_Hopkins_campus.jpg
Amazon and Johns Hopkins University announced the first recipients of PhD fellowships and faculty research awards as part of the JHU + Amazon Initiative for Interactive AI. The initiative is focused on driving ground-breaking AI advances with an emphasis on machine learning, computer vision, natural language understanding, and speech processing.

Johns Hopkins and Amazon announce six fellows and nine faculty research awards

Inaugural recipients named as part of the JHU + Amazon Initiative for Interactive AI (AI2AI).

Amazon and Johns Hopkins University (JHU) today announced the first recipients of PhD fellowships and faculty research awards as part of the JHU + Amazon Initiative for Interactive AI (AI2AI).

The AI2AI initiative, launched in April and housed in JHU’s Whiting School of Engineering, is focused on driving ground-breaking AI advances with an emphasis on machine learning, computer vision, natural language understanding, and speech processing.

Related content
The JHU + Amazon Initiative for Interactive AI (AI2AI) will be housed in the Whiting School of Engineering.

“We are delighted by the high quality of proposals and PhD fellowship nominations from JHU faculty and students," said Prem Natarajan, vice president of Alexa AI. “There is no question this initiative will drive new advances in the state-of-the-art in interactive and multimodal AI.”

As part of the initiative, annual Amazon fellowships are awarded to PhD students enrolled in the Whiting School of Engineering. Amazon also funds research projects led by JHU faculty in collaboration with post-doctoral researchers, undergraduate and graduate students, and research staff. This year’s recipients mark the inaugural class.

“We are excited that our students and faculty have a chance to partner with Amazon in an area important as interactive AI,” said Larry Nagahara, Johns Hopkins University Whiting School of Engineering’s Vice Dean for Research and Translation. “Leveraging our collective expertise in this area will advance AI and bring many beneficial aspects to our society.”

Below is a list of the fellows, and their research, followed by the faculty award recipients and their research projects.

Amazon Fellows

Top row, left to right, Kelly Marchisio, Arya McCarthy, and Carolina Pacheco Oñate; and bottom row, left to right, Desh Raj, Anshul Shah, and Jeya Maria Jose Valanarasu
Top row, Kelly Marchisio, Arya McCarthy, and Carolina Pacheco Oñate; and bottom row, Desh Raj, Anshul Shah, and Jeya Maria Jose Valanarasu are the inaugural recipients of fellowships awarded to PhD students enrolled in the Whiting School of Engineering.

Kelly Marchisio is pursuing a PhD in computer science, studying under Philipp Koehn, a professor of computer science.

“Word embedding spaces are a critical component of modern natural language processing systems. My work focuses on understanding and exploiting embedding space geometry, with the goal of creating spaces that are smaller, more useful, and more universally applicable across languages and domains.”

Arya McCarthy is pursuing a PhD in computer science, studying under David Yarowsky, a professor of computer science.

“I call my vision for natural language processing, kilolanguage processing: not only modeling thousands of languages but also letting their collective evidence and commonality reinforce each other. To make it happen, I’ve created neural machine translation models; morphological lemmatizers, taggers, and inflectors; and even a thorough analysis of color terminology spanning thousands of languages, aiming to push those frontiers further. This vision is driven by the realities of speaker needs and how NLP fails to meet them today. There are about 7000 identified languages in the world, at least 4000 of which have a book-length digitized written presence. Despite this availability of data, standard NLP tools are available for often far fewer than 100.”

Carolina Pacheco Oñate is pursuing a PhD in biomedical engineering, studying under René Vidal, an Amazon Scholar and the Herschel Seder Professor of Biomedical Engineering.

“I am interested in advancing computer vision to domains with limited availability of data or annotations, which is relevant not only in long-tail events within traditional computer vision tasks, but also in other socially impactful areas such as biomedical sciences. I believe that the combination of deep learning with probabilistic models and domain knowledge can provide the right balance between capacity and structure, enabling learning from limited amounts of data in self- and weakly-supervised regimes.”

Desh Raj is pursuing a PhD in computer science, studying under Sanjeev Khudanpur, associate professor of electrical and computer engineering.

“Since the first automatic speech recognition systems were built more than 30 years ago, improvement in voice technology has enabled applications such as automated customer support and language learning. Through years of research on speech enhancement and robust speech processing, these systems are now deployed in diverse settings such as on home speakers and vehicle controls. Nevertheless, these present systems are passive listeners which transcribe single-speaker utterances and feed into downstream language understanding components. Conversational intelligence of the future is expected to comprise systems that can actively participate in human conversations. While such systems would require intelligence in diverse modalities — dialog systems for context handling, emotion recognition from speech and video, common sense reasoning, to name a few — their ability to recognize free-flowing multi-party conversations is a core component that needs to be solved.”

Anshul Shah is pursuing a PhD in computer science, studying under Rama Chellappa, Bloomberg Distinguished Professor in electrical and computer engineering and biomedical engineering.

“My current research is broadly in the area of pose-based action recognition, video understanding, self-supervised learning and multimodal learning. My research tries to make fundamental contributions to these research areas, obtains new insights and pushes the state of the art. My interests closely align with AI2AI’s focus in areas of interactive AI technologies specifically in the areas of computer vision and multimodal AI.”

Jeya Maria Jose Valanarasu is pursuing a PhD in electrical and computer engineering, studying under Vishal M. Patel, associate professor of electrical and computer engineering.

“Deep learning methods for computer vision have made remarkable progress in field visual recognition. One major reason for its success is the amount of data these models are trained on. Annotating new ground truths for every new problem or application is very inefficient. Also, current vision systems perform poorly on data distribution that it has not seen during training. This problem is called domain adaptation and is important to solve for deploying models in real-time. Also, when the model is adapted to new data during inference, the adaptation needs to be fast and it does not make sense to train the model at test-time. Thus, we need to focus on few-shot or better zero-shot learning for adaptation.”

Faculty research awards

Top row, Mark Dredze, Philipp Koehn, and Kenton Murray; second row, Anqi Liu, Jesus Antonio Villalba López, and Soledad Villar; bottom row, Laureano Moro-Velazquez, Mahsa Yarmohammadi, and Alan Yuille
Top row, Mark Dredze, Philipp Koehn, and Kenton Murray; second row, Anqi Liu, Jesus Antonio Villalba López, and Soledad Villar; bottom row, Laureano Moro-Velazquez, Mahsa Yarmohammadi, and Alan Yuille are inaugural recipients of faculty research awards as part of the JHU + Amazon Initiative for Interactive AI.

Mark Dredze, John C. Malone Associate Professor of Computer Science: “Integrating Knowledge Representation of LLMs with Information Extraction Systems

“In the past few years, new types of AI models that capture patterns in language have become very good at learning information from language. This project explores how we can use information learned by these models to inform practical applications on language data, such as identifying important features or characteristics of products in product reviews. This award will allow us to push the limits of language modeling by exploring how we can use recent advances to help improve various applications of language technologies.”

Philipp Koehn, professor of computer science, and Kenton Murray, research scientist in the Human Language Technology Center of Excellence: “Evaluating the Multilinguality of Multilingual Machine Translation

“The proliferation of deep neural networks into artificial intelligence has allowed researchers and engineers to build systems that can automatically translate between large groups of languages without having to build separate models. However, the limitations of having one large, general model are not well understood. We aim to investigate the cutting-edge frontiers of this class of AI models.”

Anqi Liu, assistant professor of computer science: “Online Domain Adaptation via Distributionally Robust Learning

“This project aims to enable fast and robust adaptation for AI algorithms via modeling uncertainty. This award makes it possible for me to work on fundamental research questions that have the potential for real-world impact.”

Jesus Antonio Villalba López, assistant research professor of electrical and computer engineering, “Generalist Speech Processing Models

“This project will investigate how to efficiently extract the information contained in speech using large-scale AI models. The outcome will be a generalist model able to transcribe speech into text, and determine the speaker’s identity, language, and emotional state, among others.”

Soledad Villar, assistant professor of applied mathematics and statistics: “Green AI: Powerful and Lightweight Machine Learning via Exploiting Symmetries

“In this project we investigate the use of symmetries and low-dimensional structures in the design of machine learning models. Enforcing these mathematical structures will allow us to reduce the energy consumption, time, and amounts of data required for training and evaluating machine learning models while preserving (or even improving) their performance.”

Laureano Moro-Velazquez, assistant research professor, Center for Language and Speech Processing: “Improving Spoken Language Understanding for People with Atypical Speech

“In this project we will create a new dataset and develop new speech technologies meant to improve the lives of individuals with atypical speech and speech impairment. There are almost no publicly available datasets containing atypical speech, and these are necessary to create new assistive technologies for the affected population. This award will allow us to create such dataset which will be useful for us and for many other groups researching atypical speech.”

Mahsa Yarmohammadi, assistant research scientist, Center for Language and Speech Processing: “Rapid Multilingual Dataset Creation with Automatic Projection and Human Supervision

“Artificial intelligence in general, and natural language processing in particular, require a massive scale of data to learn strong models. Such data might not be available in languages other than high-resource ones such as English. In this project, we study the rapid creation of multilingual datasets by automatically translating and aligning an available dataset in one language into multiple other languages. We will also study the impact of human supervision in improving data quality. Once we have created these resources, we intend to use them to co-train single multilingual models for cross-lingual NLP tasks.”

Alan Yuille, Bloomberg Distinguished Professor of Cognitive Science and Computer Science, “Weakly-Supervised Multi-Modal Transformers for Few-Shot Learning with Generalization to Novel Domains and Fine-Grained Tasks

“Self-supervised and weakly supervised transformers have been shown to be highly effective for a variety of vision, language, and vision-language tasks. This proposal targets three challenges. First, to improve performance on standard tasks, particularly on fine-grained tasks (e.g., object attributes and parts), which have received little study. Second, to develop tokenizer approaches to enable few-shot, and ideally zero-shot, learning. Third, to adapt these approaches so that they are able to generalize to novel domains and to out-of-distribution situations. We propose five strategies to achieve these goals which include extending the tokenizer-based approaches, modifying the transformer structure, increasing the text-annotations to help these difficult tasks, and techniques for enabling the algorithms to generalize out-of-domain and out-of-distribution.”

For more information on the JHU and Amazon initiative, including opportunities and events, visit the official site.

Related content

IN, KA, Bangalore
Have you ever ordered a product on Amazon and when that box with the smile arrived you wondered how it got to you so fast? Have you wondered where it came from and how much it cost Amazon to deliver it to you? If so, the WW Amazon Logistics, Business Analytics team is for you. We manage the delivery of tens of millions of products every week to Amazon’s customers, achieving on-time delivery in a cost-effective manner. We are looking for an enthusiastic, customer obsessed, Sr. Applied Scientist with good analytical skills to help manage projects and operations, implement scheduling solutions, improve metrics, and develop scalable processes and tools. The primary role of an Operations Research Scientist within Amazon is to address business challenges through building a compelling case, and using data to influence change across the organization. This individual will be given responsibility on their first day to own those business challenges and the autonomy to think strategically and make data driven decisions. Decisions and tools made in this role will have significant impact to the customer experience, as it will have a major impact on how the final phase of delivery is done at Amazon. Ideal candidates will be a high potential, strategic and analytic graduate with a PhD in (Operations Research, Statistics, Engineering, and Supply Chain) ready for challenging opportunities in the core of our world class operations space. Great candidates have a history of operations research, and the ability to use data and research to make changes. This role requires robust program management skills and research science skills in order to act on research outcomes. This individual will need to be able to work with a team, but also be comfortable making decisions independently, in what is often times an ambiguous environment. Responsibilities may include: - Develop input and assumptions based preexisting models to estimate the costs and savings opportunities associated with varying levels of network growth and operations - Creating metrics to measure business performance, identify root causes and trends, and prescribe action plans - Managing multiple projects simultaneously - Working with technology teams and product managers to develop new tools and systems to support the growth of the business - Communicating with and supporting various internal stakeholders and external audiences
US, NY, New York
We are seeking an Applied Scientist to lead the development of evaluation frameworks and data collection protocols for robotic capabilities. In this role, you will focus on designing how we measure, stress-test, and improve robot behavior across a wide range of real-world tasks. Your work will play a critical role in shaping how policies are validated and how high-quality datasets are generated to accelerate system performance. You will operate at the intersection of robotics, machine learning, and human-in-the-loop systems, building the infrastructure and methodologies that connect teleoperation, evaluation, and learning. This includes developing evaluation policies, defining task structures, and contributing to operator-facing interfaces that enable scalable and reliable data collection. The ideal candidate is highly experimental, systems-oriented, and comfortable working across software, robotics, and data pipelines, with a strong focus on turning ambiguous capability goals into measurable and actionable evaluation systems. Key job responsibilities - Design and implement evaluation frameworks to measure robot capabilities across structured tasks, edge cases, and real-world scenarios - Develop task definitions, success criteria, and benchmarking methodologies that enable consistent and reproducible evaluation of policies - Create and refine data collection protocols that generate high-quality, task-relevant datasets aligned with model development needs - Build and iterate on teleoperation workflows and operator interfaces to support efficient, reliable, and scalable data collection - Analyze evaluation results and collected data to identify performance gaps, failure modes, and opportunities for targeted data collection - Collaborate with engineering teams to integrate evaluation tooling, logging systems, and data pipelines into the broader robotics stack - Stay current with advances in robotics, evaluation methodologies, and human-in-the-loop learning to continuously improve internal approaches - Lead technical projects from conception through production deployment - Mentor junior scientists and engineers About the team Fauna Robotics, an Amazon company, is building capable, safe, and genuinely delightful robots for everyday life. Our goal is simple: make robots people actually want to live and interact with in everyday human spaces. We believe that future won’t arrive until building for robotics becomes far more accessible. Today, too much effort is spent reinventing the fundamentals. We’re changing that by developing tightly integrated hardware and software systems that make it faster, safer, and more intuitive to create real-world robotic products. Our work spans the full stack: mechanical design, control systems, dynamic modeling, and intelligent software. The focus is not just functionality, but experience. We’re building robots that feel responsive, expressive, and genuinely useful. At Fauna, you’ll work at the frontier of this space, helping define how robots move, manipulate, and interact with people in natural environments. It’s an opportunity to solve hard problems across hardware and software with a team focused on making robotics accessible and joyful to build. If you care about making robotics real for everyone and building systems that are as delightful as they are capable, we’re interested in hearing from you.
US, CA, Sunnyvale
Industrial is seeking exceptional talent to help develop the next generation of advanced robotics systems that will transform automation at Amazon's scale. We're building revolutionary robotic systems that combine innovative AI, sophisticated control systems, and advanced mechanical design to create adaptable automation solutions capable of working safely alongside humans in dynamic environments. This is a unique opportunity to shape the future of robotics and automation at unprecedented scale, working with world-class teams pushing the boundaries of what's possible in robotic manipulation, locomotion, and human-robot interaction. This role presents an opportunity to shape the future of robotics through innovative applications of deep learning and large language models. We leverage advanced robotics, machine learning, and artificial intelligence to solve complex operational challenges at unprecedented scale. Our fleet of robots operates across hundreds of facilities worldwide, working in sophisticated coordination to fulfill our mission of customer excellence. We are pioneering the development of robotics foundation models that: - Enable unprecedented generalization across diverse tasks - Integrate multi-modal learning capabilities (visual, tactile, linguistic) - Accelerate skill acquisition through demonstration learning - Enhance robotic perception and environmental understanding - Streamline development processes through reusable capabilities The ideal candidate will contribute to research that bridges the gap between theoretical advancement and practical implementation in robotics. You will be part of a team that's revolutionizing how robots learn, adapt, and interact with their environment. Join us in building the next generation of intelligent robotics systems that will transform the future of automation and human-robot collaboration. As an Applied Scientist, you will develop and improve machine learning systems that help robots perceive, reason, and act in real-world environments. You will leverage state-of-the-art models (open source and internal research), evaluate them on representative tasks, and adapt/optimize them to meet robustness, safety, and performance needs. You will invent new algorithms where gaps exist. You’ll collaborate closely with research, controls, hardware, and product-facing teams, and your outputs will be used by downstream teams to further customize and deploy on specific robot embodiments. Key job responsibilities As an Applied Scientist in the Foundations Model team, you will: - Leverage state-of-the-art models for targeted tasks, environments, and robot embodiments through fine-tuning and optimization. - Execute rapid, rigorous experimentation with reproducible results and solid engineering practices, closing the gap between sim and real environments. - Build and run capability evaluations/benchmarks to clearly profile performance, generalization, and failure modes. - Contribute to the data and training workflow: collection/curation, dataset quality/provenance, and repeatable training recipes. - Write clean, maintainable, well commented and documented code, contribute to training infrastructure, create tools for model evaluation and testing, and implement necessary APIs - Stay current with latest developments in foundation models and robotics, assist in literature reviews and research documentation, prepare technical reports and presentations, and contribute to research discussions and brainstorming sessions. - Work closely with senior scientists, engineers, and leaders across multiple teams, participate in knowledge sharing, support integration efforts with robotics hardware teams, and help document best practices and methodologies.
US, CA, Palo Alto
About Sponsored Products and Brands The Sponsored Products and Brands team at Amazon Ads is re-imagining the advertising landscape through generative AI technologies, revolutionizing how millions of customers discover products and engage with brands across Amazon.com and beyond. We are at the forefront of re-inventing advertising experiences, bridging human creativity with artificial intelligence to transform every aspect of the advertising lifecycle from ad creation and optimization to performance analysis and customer insights. We are a passionate group of innovators dedicated to developing responsible and intelligent AI technologies that balance the needs of advertisers, enhance the shopping experience, and strengthen the marketplace. If you're energized by solving complex challenges and pushing the boundaries of what's possible with AI, join us in shaping the future of advertising. Key job responsibilities As a Machine Learning Applied Scientist, you will: * Conduct deep data analysis to derive insights to the business, and identify gaps and new opportunities * Develop scalable and effective machine-learning models and optimization strategies to solve business problems * Run regular A/B experiments, gather data, and perform statistical analysis * Work closely with software engineers to deliver end-to-end solutions into production * Improve the scalability, efficiency and automation of large-scale data analytics, model training, deployment and serving * Conduct research on new machine-learning modeling and Generative AI solutions to optimize all aspects of Sponsored Products and Brands business About the team The Ad Response Prediction team within Sponsored Products and Brands (SPB) drives personalized shopping experiences for SPB Ads across placements, pages, and devices worldwide. We achieve this through ML and GenAI solutions that include customized shopper response prediction and session-level understanding to optimize every stage of the ad-serving process, from sourcing and bidding to widget discovery and auctions. Our responsibilities include advancing response prediction through model and feature innovations and extending prediction beyond the auction stage to areas such as targeting, sourcing, and bidding.
US, NY, New York
We are seeking a Measurement and Attribution Analytics scientist to further the development and application of analytics methods to examine the complex data flows of measurement, attribution and shopper analytics and to translate deep-dives into actionable insights for our product teams. In this role you will develop new tools to analyze our advertising data to help improve the performance of our bidding algorithms, targeting and relevance systems, help advance our supply strategy, and evaluate the adoption and impact of feature releases. Key job responsibilities - Analyze data trends regarding supply, optimization, ad load, and advertising mix effects that affect advertiser performance and contribute to achieving advertiser goals. - Present papers to senior leaders on issues like feature development impact on identity recognition rates, and changes of ad selection systems to improve fill rate highlighting insights that will inform our business development and engineering roadmaps. - Identify, standardize, and operationalize KPIs to effectively measure the performance of all systems involved in ad serving, and use trend insights to inform business priorities. - Partner with engineering teams to define data logging requirements and getting these prioritized in engineering roadmaps. - Validate financial models through analysis - Develop and own ad revenue and supply intelligence analytics decks that provide ongoing deep-dives A day in the life The Measurement and Attribution Scientist will work closely with business leaders and engineers on developing common data architecture that will optimize our data logging at different grains, and will allow data interoperability from bid flow to optimization to campaign delivery. There will be an emphasis on understanding the impact of shopper journeys to and from Amazon properties and their implication on product development. The candidate will then analyze the data and present papers and ongoing reports on actionable insights. About the team The Ads Science Product Team's Mission: Work alongside those who need product data to apply objective perspective and business logic to uncover insights, advise strategic decisions, and adjust to industry changes.
IN, KA, Bangalore
Does the thought of improving one of the world’s most complex logistic systems inspire you? Is your passion to sift through hundreds of systems, processes, and data sources to solve the puzzle and identify the next big opportunity? Are you a creative big thinker who is passionate about using data to direct decision making and solve complex and large-scale challenges? Are you fascinated by the interactions between operations and strategy? Do you feel like your skills uniquely qualify you to bridge communication between teams with competing priorities? If so, then this position is for you! Come help Amazon create state-of-the-art science-driven technologies for delivering packages to the doorstep of our customers! The Last Mile Routing & Planning organization builds the software, algorithms and tools that make the “magic” of home delivery happen: our flow, sort, dispatch and routing intelligence systems are responsible for the billions of daily decisions needed to plan and execute safe, efficient and frustration-free routes for drivers around the world. Our team supports deliveries (and pickups!) for Amazon Logistics, Same Day, Amazon Grocery, Lockers, and other new initiatives across the world. Key job responsibilities In this role, your main focus will be to apply algorithms, synthesize information, identify business opportunities, provide data-driven insights and communicate business and technical requirements within the team and across stakeholder groups. You will partner closely with other scientists and engineers in a collegial environment with a clear path to business impact. We have an exciting portfolio of research areas including vehicle routing, planning for electric and autonomous vehicles, district and stops planning, ultra-fast deliveries, fleet planning, and forecasting solutions for different delivery programs leveraging the latest OR, ML, and Generative AI methods, at a global scale. Successful candidates will have a deep knowledge of Operations Research and/or Machine/Deep Learning methods, experience in applying these methods to large-scale business problems, the ability to map models into production-worthy code in Python or Java, the communication skills necessary to explain complex technical approaches to a variety of stakeholders and customers, and the excitement to take iterative approaches to tackle big research challenges.
US, CA, San Francisco
Innovators wanted! Are you an entrepreneur? A builder? A dreamer? This role is part of an Amazon Special Projects team that takes the company’s Think Big leadership principle to the extreme. We focus on creating entirely new products and services with a goal of positively impacting the lives of our customers. No industries or subject areas are out of bounds. If you’re interested in innovating at scale to address big challenges in the world, this is the team for you. Here at Amazon, we embrace our differences. We are committed to furthering our culture of inclusion. We have thirteen employee-led affinity groups, reaching 40,000 employees in over 190 chapters globally. We are constantly learning through programs that are local, regional, and global. Amazon’s culture of inclusion is reinforced within our 16 Leadership Principles, which remind team members to seek diverse perspectives, learn and be curious, and earn trust. Our team highly values work-life balance, mentorship and career growth. We believe striking the right balance between your personal and professional life is critical to life-long happiness and fulfillment. We care about your career growth and strive to assign projects and offer training that will challenge you to become your best.
US, WA, Seattle
About Sponsored Products and Brands The Sponsored Products and Brands (SPB) team at Amazon Ads is re-imagining the advertising landscape through generative AI technologies, revolutionizing how millions of customers discover products and engage with brands across Amazon.com and beyond. We are at the forefront of re-inventing advertising experiences, bridging human creativity with artificial intelligence to transform every aspect of the advertising lifecycle from ad creation and optimization to performance analysis and customer insights. We are a passionate group of innovators dedicated to developing responsible and intelligent AI technologies that balance the needs of advertisers, enhance the shopping experience, and strengthen the marketplace. If you're energized by solving complex challenges and pushing the boundaries of what's possible with AI, join us in shaping the future of advertising. About our team The General Shopping Intelligence (GSI) team is a highly motivated, collaborative, and fun-loving group with a strong entrepreneurial spirit and bias for action. We provide advanced real-time machine learning services that connect shoppers with the right ads across all platforms and surfaces worldwide. Through deep understanding of both shoppers and products, we help shoppers discover new products they love, enable advertisers to reach their customers most efficiently, and help Amazon continuously innovate on behalf of all customers. We are seeking a motivated Applied Scientist who loves to innovate at the intersection of customer experience, deep learning, generative AI and high-scale machine learning systems. If you're energized by solving complex challenges and pushing the boundaries of what's possible with AI, join us in shaping the future of advertising. Key job responsibilities As an Applied Scientist, you will: * Leverage Generative AI and Large Language Models (LLMs) to mine complex behavioral data, deriving deep, actionable shopper insights that identify customer experience gaps and unlock new business opportunities. * Design and develop scalable machine learning and GenAI models focused on shopper intent and preference modeling, ensuring a rapid path from prototype to production. * Partner closely with engineering teams to architect and deploy end-to-end GenAI solutions into production, integrating advanced insights directly into real-time, customer-facing systems. * Drive the scalability, efficiency, and automation of large-scale model training and real-time inference systems, pioneering the LLM infrastructure required to support next-generation GenAI workloads at Amazon Ads scale. * Design and run rigorous A/B experiments to quantify the business and customer impact of GenAI-driven shopper insights, performing advanced statistical analysis to guide iterative production rollouts. * Conduct applied research in novel generative AI techniques (e.g., fine-tuning, RAG, agentic workflows) to optimize the shopper experience and drive performance across all aspects of the Sponsored Products and Brands business.
US, WA, Bellevue
As an Applied Scientist in Amazon Fullfilment Technology, you will lead the development of agentic systems to assist with operational decision making and orchestration. You will work building full agentic systems leveraging multi-agent orchestration, tool use, memory, and action execution. You will train LLMs using a combination of rejection sampling approaches, SFT, continual post-training, and Reinforcement Learning (RL). These systems are deployed to Amazon buildings, and you will also work on rigorous offline and online evaluations. Your work will leverage the latest LLMs to develop capabilities for agentic reasoning, coding and analytics. You will also lead research projects to tackle unsolved problems, mentor interns, and author academic papers to summarize your findings for external publication. Key job responsibilities - Generating training and preference data for specific use cases (reasoning trajectories, tool traces) - Reward modeling and policy optimization for LLMs: DPO, IPO, RLHF/RLAIF with PPO/GRPO, rejection sampling. - Supervised fine-tuning on step-by-step trajectories and tool-use traces - Verbal Reinforcement Learning and Continual Learning - RL for LLMs, Offline RL and off-policy evaluation - Agentic memory/state management; episodic and semantic memory; vector search; grounding with RAG. - Evaluation: developing decision quality metrics, scaling LLM-based evaluations. About the team Amazon Fulfillment Technologies (AFT) powers Amazon's global fulfillment network. We invent and deliver software, hardware, and data science solutions that orchestrate processes, robots, machines, and people. We harmonize the physical and virtual world so Amazon customers can get what they want, when they want it. Learn more about AFT: https://tinyurl.com/AFTOverview
US, VA, Arlington
Are you excited about making business decisions using science and data? Are you interested in supporting consumer device concepts from idea inception to launch? Do you want to work on a Science Product team focused on scaling statistics and econometrics with custom tools? If so, this may be the role for you! Amazon.com strives to be Earth's most customer-centric company. The Amazon Devices and Services team focuses on delighting customer by enabling seamless functionality in supplying, entertaining, and managing the home -- and beyond. We seek and hire the world's brightest minds, offering them a fast-paced, technologically-sophisticated, and friendly work environment, where economic theory meets real-world industry. The Decision Science team in Devices owns demand estimates and pricing recommendations of concept devices before customers know they exist. We support devices and services ranging from Echo Frames to Kindle Paperwhite to Blink Video Camera …all prior to launch. We are a cross-functional Product team working to scale Econometrics through Amazon and beyond by incorporating Science into internal facing tools and making it easier for others to do so as well. In this role, you will have input in decision meetings with Amazon senior leadership, which include go/no-go decisions for brand new devices and services and build volume decisions for manufacture prior to receiving any customer signal. You will have direct input to pricing decisions. You will leverage Science and Tools produced by the Decision Science team such as conjoint demand models to produce these recommendations. You will work with Scientists, Economists, Product Managers, and Software Developers to provide meaningful feedback about stakeholder problems to inform business solutions and increase the velocity, quality, and scope behind our recommendations. You will also have the opportunity to work on special projects to both guide the business and advance your own knowledge and understanding of specific topics. Key job responsibilities Applies expertise to develop econometric/machine learning models to measure the demand of devices and the business; Reviews models and results for other scientists, mentors junior scientists; Generates economic insights for the Devices and Services business and work with stakeholders to run the business for effectively; Describes strategic importance of vision inside and outside of team; and, Identifies business opportunities, defines the problem and how to solve it; Engages with senior scientists, business leadership outside Devices and Services to understand interplay between different business units.