This picture is an overhead shot inside an Amazon center, workers can be seen moving amidst hundreds of boxes which sit on conveyor belts and carts, in the upper left foreground, a yellow railing extends into the distance.
When faced with the need to evolve Amazon’s supply chain to meet customer needs, a team of scientists, developers, and other professionals worked together to create an inventory planning system that would help Amazon fulfill its delivery promises.
F4D Studios

The evolution of Amazon’s inventory planning system

How Amazon’s scientists developed a first-of-its-kind multi-echelon system for inventory buying and placement.

For every order placed on the Amazon Store, mathematical models developed by Amazon’s Supply Chain Optimization Technologies organization (SCOT) work behind the scenes to ensure that product inventories are best positioned to fulfill the order. 

Forecasting models developed by SCOT predict the demand for every product. Buying systems determine the right level of product to purchase from different suppliers, while large-scale placement systems determine the optimal location for products across the hundreds of facilities belonging to Amazon’s global fulfillment network.

“With hundreds of millions of products sold across multiple geographies, developing automated models to make inventory planning decisions at Amazon scale is one of the most challenging and rewarding parts of our work,” said Deepak Bhatia, vice president of Supply Chain Optimization Technologies at Amazon.

We made the decision to redesign Amazon’s supply chain systems from the ground up.
Deepak Bhatia

In the first half of the past decade, Amazon transitioned from a largely manual supply chain management system to an automated one. However, when faced with the need to evolve Amazon’s supply chain to meet customer needs, and the introduction of same day delivery services like Prime Now, the team moved to replace that system with a new one that would better help Amazon fulfill delivery promises made to customers.

“As far back as 2016, we were able to see that the automated system we had at the time wouldn’t help us meet the ever-growing expectations of our customers,” Bhatia recalled. “As a result, we made the decision to redesign Amazon’s supply chain systems from the ground up.”

A global company catering to local needs

“In 2016, Amazon’s supply chain network was designed for scenarios where inventory from any fulfillment center could be shipped to any customer to meet a two-day promise,” said Salal Humair, senior principal research scientist at Amazon who has been with the company for seven years.

This design was inadequate for the new world in which Amazon was operating; one shaped by what Humair calls the “globalization-localization imperative.” Amazon’s expansion included an increasing number of international locations — at the time, the company had 175 fulfillment centers serving customers in 185 countries around the world.

“Meeting the needs of our customer base meant that we needed to serve those customers in multiple geographies,” Humair said.

As Amazon continued to expand internationally, the company also launched one-day and same day delivery windows in local regions for services like Amazon Prime and Amazon Prime Now.

“We quickly realized that in addition to serving customers around the globe, we also had to pivot from functioning as a national network to a local one, where we could position inventory close to our customers,” Humair says.

A row of five profile photos shows, left to right, Deepak Bhatia, vice president of Supply Chain Optimization Technologies at Amazon; Salal Humair, senior principal research scientist; Alp Muharremoglu, a senior principal scientist; Jeff Maurer, a vice president; and Yan Xia, principal applied scientist.
Left to right, Deepak Bhatia, vice president of Supply Chain Optimization Technologies at Amazon; Salal Humair, senior principal research scientist; Alp Muharremoglu, a senior principal scientist; Jeff Maurer, a vice president in SCOT; and Yan Xia, principal applied scientist, were among those instrumental in migrating Amazon to the multi-echelon system.

In addition to the ‘globalization-localization imperative,’ the growing complexity of Amazon’s supply chain network further complicated matters. To meet the increased customer demand for a diverse variety of shipping speeds, Amazon’s fulfillment network was expanding to include an increasing number of building types and sizes: from fulfillment centers (for everyday products) and non-sortable fulfillment centers (for larger items), to smaller fulfillment centers catering to same-day orders, and distribution centers that supplied products to downstream fulfillment centers. The network was increasingly becoming layered, and fulfillment centers in one layer (or echelon) were acting as suppliers to other layers.

“We had to reimagine every aspect of our system to account for this increasing number of echelons,” Humair said.

The science behind multi-echelon inventory planning

The sheer scale of Amazons operations posed a significant challenge from a scientific perspective. Amazon Store orders are fulfilled through complex dynamic optimization processes — where a real-time order assignment system can choose to fulfill an order from the optimal fulfillment center that can meet the customer promise. This real-time order assignment makes inventory planning an incredibly complex problem to solve.

Other inventory-related dependencies further complicate matters: the same pool of inventory is frequently used to serve demand for orders with different shipping speeds. Consider a box of diapers: it can be used to fulfill an order for a two-day Prime delivery. It can also be used to ease the life of harried parents who have placed an order on Prime Now, and need diapers for their baby delivered in a two-hour window.

Amazon’s scientists also have to contend with a high degree of uncertainty. Customer demand for products cannot be perfectly predicted even with the most advanced machine learning models. In addition, lead times from vendors are subject to natural variation due to manufacturing capacity, transportation times, weather, etc., adding another layer of uncertainty.

This required building a custom solution, one that relies on sound scientific principles and rigor, and borrowing ideas from academic literature as building blocks, but with ground-breaking in-house invention.
Alp Muharremoglu

Humair notes that the scale of Amazon’s operations, the complexity of the network, and the uncertainties associated with the company’s dynamic ordering system make it impossible to even write down a closed-form objective function for the optimization problem the team was trying to solve.

While multi-echelon inventory optimization is a well-researched field, the bulk of literature focused on single-product models, proposed solutions for much simpler networks, or used greatly simplified assumptions for replenishing inventory.

“There is a large body of academic literature on multi-echelon inventory management, and papers typically focus on one or two main aspects of the problem,” noted Alp Muharremoglu, a senior principal scientist in SCOT who spent 15 years as a faculty member at Columbia University and the University of Texas at Dallas. “Amazon’s scale and complexity meant no existing solution was a perfect fit. This required building a custom solution, one that relies on sound scientific principles and rigor, and borrowing ideas from academic literature as building blocks, but with ground-breaking in-house invention to push the boundaries of academic research. It is a thrill to see multi-echelon inventory theory truly in action in such a large scale and dynamic supply chain.”

As a result, the system developed by SCOT (a project whose roots stretch back to 2016) is a significant break from the past. The heart of the model is a multi-product, multi-fulfillment center, capacity-constrained model for optimizing inventory levels for multiple delivery speeds, under a dynamic fulfillment policy. The framework then uses a Lagrangian-type decomposition framework to control and optimize inventory levels across Amazon’s network in near real-time.

Broadly speaking, decomposition is a mathematical technique that breaks a large, complex problem up into smaller and simpler ones. Each of these problems is then solved in parallel or sequentially. The Lagrangian method of decomposition factors complicated constraints into the solution, while providing a ‘cost’ for violating these constraints. This cost makes the problem easier to solve by providing an upper bound to the maximization problem, which is critical when planning for inventory levels at Amazon’s scale. 

“We computed opportunity costs for storage and flows at every fulfillment center,” Humair said. “Using Lagrangean decomposition, we then used these costs to calculate the related inventory positions at these locations. Crucially, we incorporated a stochastic dynamic fulfillment policy in a scalable optimization model, allowing Amazon to calculate inventory levels not at just one location, but at every layer in our fulfillment network.”

Mobilizing the organization

While creating the new multi-echelon system was an imposing scientific challenge, it also represented a significant organizational accomplishment, one that required collaboration across multiple teams.

“Moving multi-echelon from concept to implementation was one of the most difficult organizational challenges we’ve worked through; we had many potential implementations that looked radically different in terms of model capabilities, interfaces, engineering challenges, and long-term implications for how our teams would interact with each other,” said Jeff Maurer, a SCOT vice president who has been instrumental in rolling out the automation of Amazon’s supply chain and oversaw the roll out of the multi-echelon system.

“This was also a case where there wasn’t a great way to decide between them without building and exploring one or more approaches in production. Ultimately, that’s what we did — we picked the best options we could identify, built them out, learned from them, then repeated that process. We learned things by experimenting with real production implementations that we could never have learned from simplified models or simulations alone, given the complexity of the real-world dynamics of our supply chain. But it was hard on the teams — it wasn’t always obvious that the systems the teams were iterating on were the best path, given the high directional ambiguity.”

Packages moving through a fulfillment center

“Sometimes, the only way to make a massive change is to realize that you have no option but to make that change,” said Yan Xia, principal applied scientist at Amazon. Humair noted that Xia played “a pivotal role” over the four years it took the company to migrate to the new multi-echelon system.

Xia recalled that teams within SCOT were keenly aware of the limitations of the existing system.  However, there was skepticism that the multi-echelon system was the right solution.

“The skepticism was understandable,” Xia said. “It’s one thing to have a big idea. But you also have to be able to present the benefits of your idea in a coherent way.”

Xia gave an example of how he helped convince members from the buying and placement teams about the benefits of the new model.

“One team decides optimal suppliers to source products from, while another team makes decisions on where these products should be placed,” Xia explained. “I was able to show them how the two functions would essentially be unified in the multi-echelon system. Sure, it would change how they worked on a day-to-day basis — but it would do so in a way that made their lives simpler.”

To help ensure that resources were made available for the development of the multi-echelon system, Xia also focused on driving alignment among leaders in SCOT. He developed a simulation based on real-world data. The results clearly demonstrated that the proposed solution for inventory forecasting, buying, and placement would result in a steep decline in shipping costs, which in turn would allow Amazon to keep prices lower for customers.

Teams involved in multi-echelon planning discussions were galvanized after seeing the results of the simulation.

“Everyone bought into the vision,” Xia said. “We began to collaborate in near real-time. If we ran into a problem, we didn’t wait around for a weekly sprint meeting. We just got together in a room, or stood next to a whiteboard and solved it.”

Xia said that this was also when things began to get more complex. 

“An awareness of the complexity of the existing setup began to dawn on us,” says Xia. “We began to realize how every component in the system had multiple dependencies. For example, the buying platforms were tightly integrated with older legacy systems – we now had to factor these dependencies into our solutions.”

Solving a multi-item, multi-echelon with stochastic demand and lead-time and aggregated capacity constraints and differentiated customer service levels. That sort of thing is just unheard of in the academia and the industry.
Deepak Bhatia

The team iterated on the multi-echelon solution in a sequence of three in-production experiments (or labs) that spanned 2018 to 2020. The first lab incorporated components of the new system coupled with the old platform. It was a resounding success in terms of reducing costs, even while fulfilling orders associated with higher shipping speeds. The team moved on to testing the subsequent version of the multi-echelon system in the second lab. 

“That wasn’t nearly as good,” Xia recalled. “Most things didn’t work as expected.”

However, the team was encouraged by leadership to keep going. This wasn’t SCOT’s first attempt at taking on big and ambitious projects. The organization had taken three years to deploy the first automated supply chain management system where they overcame various challenges.

“Sure, the failure of the second lab was demotivating,” Xia says. “But we knew from experience that this failure was only to be expected. It was part of the process.”

The team fixed the bugs, and moved on to testing new features in the third lab. These included critical system capabilities, such the ability to model order cut-off times for deliveries within a particular time window.

The system went live in 2020, and over the past year, the multi-echelon system has had a large and statistically significant impact in positioning products closer to customers.

“On a personal level, I am incredibly proud of our team. Having worked in the area of multi-echelon inventory optimization before I joined Amazon, I have a deep appreciation of how difficult it was,” Bhatia noted. “There is a strong sense of pride for the work the team is doing — such as solving a multi-item, multi-echelon with stochastic demand and lead-time and aggregated capacity constraints and differentiated customer service levels. That sort of thing is just unheard of in academia and industry. This is why I find it gratifying to work as a scientist and a leader at Amazon. It gives me a lot of pride, and none of this could have been achieved without the people and the culture we have.”

Related content

US, WA, Bellevue
The Artificial General Intelligence (AGI) team is looking for a passionate, talented, and inventive Senior Applied Scientist to work on methodologies for Generative Artificial Intelligence (GenAI) models. As a Senior Applied Scientist, you will be responsible for leading the development of novel algorithms and modeling techniques to advance the state of the art. Your work will directly impact our customers and will leverage Amazon’s heterogeneous data sources and large-scale computing resources to accelerate development with multi-modal Large Language Models (LLMs) and GenAI. You will have significant influence on our overall strategy by working at the intersection of engineering and applied science to scale pre-training and post-training workflows and build efficient models. You will support the system architecture and the best practices that enable a quality infrastructure. Key job responsibilities Join us to work as an integral part of a team that has experience with GenAI models in this space. We work on these areas: - Pre-training and post-training multimodal LLMs - Scale training, optimization methods, and learning objectives - Utilize, build, and extend upon industry-leading frameworks - Work with other team members to investigate design approaches, prototype new technology, scientific techniques and evaluate technical feasibility - Deliver results independently in a self-organizing Agile environment while constantly embracing and adapting new scientific advances About the team The AGI team has a mission to push the envelope in GenAI with Large Language Models (LLMs) and multimodal systems, in order to provide the best-possible experience for our customers.
CA, BC, Vancouver
Join our Amazon Private Brands Selection Guidance organization in building science and tech solutions at scale to delight our customers with products across our leading private brands such as Amazon Basics, Amazon Essentials, and by Amazon. The Selection Guidance team applies Generative AI, Machine Learning, Statistics, and Economics solutions to drive our private brands product assortment, strategic business decisions, and product inputs such as title, price, merchandising and ordering. We are an interdisciplinary team of Scientists, Economists, Engineers, and Product Managers incubating and building day one solutions using novel technology, to solve some of the toughest business problems at Amazon. As a Sr. Data Scientist you will invent novel solutions and prototypes, and directly contribute to bringing your ideas to life through production implementation. Current research areas include entity resolution, agentic AI, large language models, and product substitutes. You will review and guide scientists across the team on their designs and implementations, and raise the team bar for science research and prototypes. This is a unique, high visibility opportunity for someone who wants to develop ambitious science solutions and have direct business and customer impact. Key job responsibilities - Partner with business stakeholders to deeply understand APB business problems and frame ambiguous business problems as science problems and solutions. - Invent novel science solutions, develop prototypes, and deploy production software to solve business problems. - Review and guide science solutions across the team. - Publish and socialize your and the team's research across Amazon and external avenues as appropriate - Leverage industry best practices to establish repeatable applied science practices, principles & processes.
US, WA, Seattle
We are looking for a passionate Applied Scientist to help pioneer the next generation of agentic AI applications for Amazon advertisers. In this role, you will design agentic architectures, develop tools and datasets, and contribute to building systems that can reason, plan, and act autonomously across complex advertiser workflows. You will work at the forefront of applied AI, developing methods for fine-tuning, reinforcement learning, and preference optimization, while helping create evaluation frameworks that ensure safety, reliability, and trust at scale. You will work backwards from the needs of advertisers—delivering customer-facing products that directly help them create, optimize, and grow their campaigns. Beyond building models, you will advance the agent ecosystem by experimenting with and applying core primitives such as tool orchestration, multi-step reasoning, and adaptive preference-driven behavior. This role requires working independently on ambiguous technical problems, collaborating closely with scientists, engineers, and product managers to bring innovative solutions into production. Key job responsibilities - Design and build agents to guide advertisers in conversational and non-conversational experience. - Design and implement advanced model and agent optimization techniques, including supervised fine-tuning, instruction tuning and preference optimization (e.g., DPO/IPO). - Curate datasets and tools for MCP. - Build evaluation pipelines for agent workflows, including automated benchmarks, multi-step reasoning tests, and safety guardrails. - Develop agentic architectures (e.g., CoT, ToT, ReAct) that integrate planning, tool use, and long-horizon reasoning. - Prototype and iterate on multi-agent orchestration frameworks and workflows. - Collaborate with peers across engineering and product to bring scientific innovations into production. - Stay current with the latest research in LLMs, RL, and agent-based AI, and translate findings into practical applications. About the team The Sponsored Products and Brands team at Amazon Ads is re-imagining the advertising landscape through the latest generative AI technologies, revolutionizing how millions of customers discover products and engage with brands across Amazon.com and beyond. We are at the forefront of re-inventing advertising experiences, bridging human creativity with artificial intelligence to transform every aspect of the advertising lifecycle from ad creation and optimization to performance analysis and customer insights. We are a passionate group of innovators dedicated to developing responsible and intelligent AI technologies that balance the needs of advertisers, enhance the shopping experience, and strengthen the marketplace. If you're energized by solving complex challenges and pushing the boundaries of what's possible with AI, join us in shaping the future of advertising. The Advertiser Guidance team within Sponsored Products and Brands is focused on guiding and supporting 1.6MM advertisers to meet their advertising needs of creating and managing ad campaigns. At this scale, the complexity of diverse advertiser goals, campaign types, and market dynamics creates both a massive technical challenge and a transformative opportunity: even small improvements in guidance systems can have outsized impact on advertiser success and Amazon’s retail ecosystem. Our vision is to build a highly personalized, context-aware agentic advertiser guidance system that leverages LLMs together with tools such as auction simulations, ML models, and optimization algorithms. This agentic framework, will operate across both chat and non-chat experiences in the ad console, scaling to natural language queries as well as proactively delivering guidance based on deep understanding of the advertiser. To execute this vision, we collaborate closely with stakeholders across Ad Console, Sales, and Marketing to identify opportunities—from high-level product guidance down to granular keyword recommendations—and deliver them through a tailored, personalized experience. Our work is grounded in state-of-the-art agent architectures, tool integration, reasoning frameworks, and model customization approaches (including tuning, MCP, and preference optimization), ensuring our systems are both scalable and adaptive.
US, CA, Pasadena
The Amazon Web Services (AWS) Center for Quantum Computing (CQC) is a multi-disciplinary team of scientists, engineers, and technicians on a mission to develop a fault-tolerant quantum computer. You will be joining a team located in Pasadena, CA that conducts materials research to improve the performance of superconducting quantum processors. We seek a Quantum Research Scientist to investigate how material defects affect qubit performance. In this role, you will combine expertise in numerical simulations and materials characterization to study materials loss mechanisms such as two-level systems, quasiparticles, vortices, etc. Key job responsibilities Provide subject matter expertise on integrated experimental and computational studies of materials defects Develop and use computational tools for large-scale simulations of disordered structures Develop and implement multi-technique materials characterization workflows for thin films and devices, with a focus on the surfaces and interfaces Identify material properties that can be a reliable proxy for the performance of superconducting resonators and qubits Communicate findings to teammates, the broader CQC team and, when appropriate, publish findings in scientific journals A day in the life At the AWS CQC, we understand that developing quantum computing technology is a marathon, not a sprint. The work/life integration within our team encourages a culture where employees work hard and also have ownership over their downtime. We are committed to the growth and development of every employee at the AWS CQC, and that includes our research scientists. You will receive management and mentorship from within the team that is geared toward career growth, and also have the opportunity to participate in Amazon's mentorship programs for scientists and engineers. Working closely with other quantum research scientists in other disciplines – like design, measurement and cryogenic hardware – will provide opportunities to dive deep into an education on quantum computing. About the team Our team contributes to the fabrication of processors and other hardware that enable quantum computing technologies. Doing that necessitates the development of materials with tailored properties for superconducting circuits. Research Scientists and Engineers on the Materials team operate deposition and characterization systems in order to develop and optimize thin film processes for use in these devices. They work alongside other Research Scientists and Engineers to help deliver the fabricated devices for quantum computing experiments. Export Control Requirement: Due to applicable export control laws and regulations, candidates must be either a U.S. citizen or national, U.S. permanent resident (i.e., current Green Card holder), or lawfully admitted into the U.S. as a refugee or granted asylum, or be able to obtain a U.S export license. If you are unsure if you meet these requirements, please apply and Amazon will review your application for eligibility. About the team Diverse Experiences AWS values diverse experiences. Even if you do not meet all of the preferred qualifications and skills listed in the job description, we encourage candidates to apply. If your career is just starting, hasn’t followed a traditional path, or includes alternative experiences, don’t let it stop you from applying. Why AWS? Amazon Web Services (AWS) is the world’s most comprehensive and broadly adopted cloud platform. We pioneered cloud computing and never stopped innovating — that’s why customers from the most successful startups to Global 500 companies trust our robust suite of products and services to power their businesses. Inclusive Team Culture AWS values curiosity and connection. Our employee-led and company-sponsored affinity groups promote inclusion and empower our people to take pride in what makes us unique. Our inclusion events foster stronger, more collaborative teams. Our continual innovation is fueled by the bold ideas, fresh perspectives, and passionate voices our teams bring to everything we do. Mentorship & Career Growth We’re continuously raising our performance bar as we strive to become Earth’s Best Employer. That’s why you’ll find endless knowledge-sharing, mentorship and other career-advancing resources here to help you develop into a better-rounded professional. Work/Life Balance We value work-life harmony. Achieving success at work should never come at the expense of sacrifices at home, which is why we strive for flexibility as part of our working culture. When we feel supported in the workplace and at home, there’s nothing we can’t achieve. Export Control Requirement: Due to applicable export control laws and regulations, candidates must be either a U.S. citizen or national, U.S. permanent resident (i.e., current Green Card holder), or lawfully admitted into the U.S. as a refugee or granted asylum, or be able to obtain a U.S export license. If you are unsure if you meet these requirements, please apply and Amazon will review your application for eligibility.
US, CA, Sunnyvale
Prime Video is a first-stop entertainment destination offering customers a vast collection of premium programming in one app available across thousands of devices. Prime members can customize their viewing experience and find their favorite movies, series, documentaries, and live sports – including Amazon MGM Studios-produced series and movies; licensed fan favorites; and programming from Prime Video add-on subscriptions such as Apple TV+, Max, Crunchyroll and MGM+. All customers, regardless of whether they have a Prime membership or not, can rent or buy titles via the Prime Video Store, and can enjoy even more content for free with ads. Are you interested in shaping the future of entertainment? Prime Video's technology teams are creating best-in-class digital video experience. As a Prime Video technologist, you’ll have end-to-end ownership of the product, user experience, design, and technology required to deliver state-of-the-art experiences for our customers. You’ll get to work on projects that are fast-paced, challenging, and varied. You’ll also be able to experiment with new possibilities, take risks, and collaborate with remarkable people. We’ll look for you to bring your diverse perspectives, ideas, and skill-sets to make Prime Video even better for our customers. With global opportunities for talented technologists, you can decide where a career Prime Video Tech takes you! Key job responsibilities - Develop ML models for various recommendation & search systems using deep learning, online learning, and optimization methods - Work closely with other scientists, engineers and product managers to expand the depth of our product insights with data, create a variety of experiments to determine the high impact projects to include in planning roadmaps - Stay up-to-date with advancements and the latest modeling techniques in the field - Publish your research findings in top conferences and journals A day in the life We're using advanced approaches such as foundation models to connect information about our videos and customers from a variety of information sources, acquiring and processing data sets on a scale that only a few companies in the world can match. This will enable us to recommend titles effectively, even when we don't have a large behavioral signal (to tackle the cold-start title problem). It will also allow us to find our customer's niche interests, helping them discover groups of titles that they didn't even know existed. We are looking for creative & customer obsessed machine learning scientists who can apply the latest research, state of the art algorithms and ML to build highly scalable page personalization solutions. You'll be a research leader in the space and a hands-on ML practitioner, guiding and collaborating with talented teams of engineers and scientists and senior leaders in the Prime Video organization. You will also have the opportunity to publish your research at internal and external conferences. About the team Prime Video Recommendation Science team owns science solution to power recommendation and personalization experience on various Prime Video surfaces and devices. We work closely with the engineering teams to launch our solutions in production.
US, CA, Sunnyvale
Prime Video is a first-stop entertainment destination offering customers a vast collection of premium programming in one app available across thousands of devices. Prime members can customize their viewing experience and find their favorite movies, series, documentaries, and live sports – including Amazon MGM Studios-produced series and movies; licensed fan favorites; and programming from Prime Video add-on subscriptions such as Apple TV+, Max, Crunchyroll and MGM+. All customers, regardless of whether they have a Prime membership or not, can rent or buy titles via the Prime Video Store, and can enjoy even more content for free with ads. Are you interested in shaping the future of entertainment? Prime Video's technology teams are creating best-in-class digital video experience. As a Prime Video technologist, you’ll have end-to-end ownership of the product, user experience, design, and technology required to deliver state-of-the-art experiences for our customers. You’ll get to work on projects that are fast-paced, challenging, and varied. You’ll also be able to experiment with new possibilities, take risks, and collaborate with remarkable people. We’ll look for you to bring your diverse perspectives, ideas, and skill-sets to make Prime Video even better for our customers. With global opportunities for talented technologists, you can decide where a career Prime Video Tech takes you! Key job responsibilities - Develop ML models for various recommendation & search systems using deep learning, online learning, and optimization methods - Work closely with other scientists, engineers and product managers to expand the depth of our product insights with data, create a variety of experiments to determine the high impact projects to include in planning roadmaps - Stay up-to-date with advancements and the latest modeling techniques in the field - Publish your research findings in top conferences and journals A day in the life We're using advanced approaches such as foundation models to connect information about our videos and customers from a variety of information sources, acquiring and processing data sets on a scale that only a few companies in the world can match. This will enable us to recommend titles effectively, even when we don't have a large behavioral signal (to tackle the cold-start title problem). It will also allow us to find our customer's niche interests, helping them discover groups of titles that they didn't even know existed. We are looking for creative & customer obsessed machine learning scientists who can apply the latest research, state of the art algorithms and ML to build highly scalable page personalization solutions. You'll be a research leader in the space and a hands-on ML practitioner, guiding and collaborating with talented teams of engineers and scientists and senior leaders in the Prime Video organization. You will also have the opportunity to publish your research at internal and external conferences. About the team Prime Video Recommendation Science team owns science solution to power recommendation and personalization experience on various Prime Video surfaces and devices. We work closely with the engineering teams to launch our solutions in production.
US, CA, Sunnyvale
Prime Video is a first-stop entertainment destination offering customers a vast collection of premium programming in one app available across thousands of devices. Prime members can customize their viewing experience and find their favorite movies, series, documentaries, and live sports – including Amazon MGM Studios-produced series and movies; licensed fan favorites; and programming from Prime Video add-on subscriptions such as Apple TV+, Max, Crunchyroll and MGM+. All customers, regardless of whether they have a Prime membership or not, can rent or buy titles via the Prime Video Store, and can enjoy even more content for free with ads. Are you interested in shaping the future of entertainment? Prime Video's technology teams are creating best-in-class digital video experience. As a Prime Video technologist, you’ll have end-to-end ownership of the product, user experience, design, and technology required to deliver state-of-the-art experiences for our customers. You’ll get to work on projects that are fast-paced, challenging, and varied. You’ll also be able to experiment with new possibilities, take risks, and collaborate with remarkable people. We’ll look for you to bring your diverse perspectives, ideas, and skill-sets to make Prime Video even better for our customers. With global opportunities for talented technologists, you can decide where a career Prime Video Tech takes you! We are looking for a self-motivated, passionate and resourceful Applied Scientist to bring diverse perspectives, ideas, and skill-sets to make Prime Video even better for our customers. You will spend your time as a hands-on machine learning practitioner and a research leader. You will play a key role on the team, building and guiding machine learning models from the ground up. At the end of the day, you will have the reward of seeing your contributions benefit millions of Amazon.com customers worldwide. Key job responsibilities - Develop AI solutions for various Prime Video Search systems using Deep learning, GenAI, Reinforcement Learning, and optimization methods; - Work closely with engineers and product managers to design, implement and launch AI solutions end-to-end; - Design and conduct offline and online (A/B) experiments to evaluate proposed solutions based on in-depth data analyses; - Effectively communicate technical and non-technical ideas with teammates and stakeholders; - Stay up-to-date with advancements and the latest modeling techniques in the field; - Publish your research findings in top conferences and journals. About the team Prime Video Search Science team owns science solution to power search experience on various devices, from sourcing, relevance, ranking, to name a few. We work closely with the engineering teams to launch our solutions in production.
US, CA, Sunnyvale
Prime Video is a first-stop entertainment destination offering customers a vast collection of premium programming in one app available across thousands of devices. Prime members can customize their viewing experience and find their favorite movies, series, documentaries, and live sports – including Amazon MGM Studios-produced series and movies; licensed fan favorites; and programming from Prime Video add-on subscriptions such as Apple TV+, Max, Crunchyroll and MGM+. All customers, regardless of whether they have a Prime membership or not, can rent or buy titles via the Prime Video Store, and can enjoy even more content for free with ads. Are you interested in shaping the future of entertainment? Prime Video's technology teams are creating best-in-class digital video experience. As a Prime Video technologist, you’ll have end-to-end ownership of the product, user experience, design, and technology required to deliver state-of-the-art experiences for our customers. You’ll get to work on projects that are fast-paced, challenging, and varied. You’ll also be able to experiment with new possibilities, take risks, and collaborate with remarkable people. We’ll look for you to bring your diverse perspectives, ideas, and skill-sets to make Prime Video even better for our customers. With global opportunities for talented technologists, you can decide where a career Prime Video Tech takes you! We are looking for a self-motivated, passionate and resourceful Applied Scientist to bring diverse perspectives, ideas, and skill-sets to make Prime Video even better for our customers. You will spend your time as a hands-on machine learning practitioner and a research leader. You will play a key role on the team, building and guiding machine learning models from the ground up. At the end of the day, you will have the reward of seeing your contributions benefit millions of Amazon.com customers worldwide. Key job responsibilities - Develop AI solutions for various Prime Video Search systems using Deep learning, GenAI, Reinforcement Learning, and optimization methods; - Work closely with engineers and product managers to design, implement and launch AI solutions end-to-end; - Design and conduct offline and online (A/B) experiments to evaluate proposed solutions based on in-depth data analyses; - Effectively communicate technical and non-technical ideas with teammates and stakeholders; - Stay up-to-date with advancements and the latest modeling techniques in the field; - Publish your research findings in top conferences and journals. About the team Prime Video Search Science team owns science solution to power search experience on various devices, from sourcing, relevance, ranking, to name a few. We work closely with the engineering teams to launch our solutions in production.
US, CA, Cupertino
We are seeking a highly skilled Data Scientist to join our Machine Learning Architecture team, focusing on power and performance optimization for ML acceleration workloads across Amazon's global data center infrastructure. This role combines advanced data science techniques with deep technical understanding of ML hardware acceleration to drive efficiency improvements in training and inference workloads at massive scale. Key job responsibilities ata Analysis & Optimization * Analyze power consumption and performance metrics across all Amazon data centers for machine learning acceleration workloads * Develop predictive models and statistical frameworks to identify optimization opportunities and performance bottlenecks * Create automated monitoring and alerting systems for power and performance anomalies Strategic Planning & Deployment Guidance * Provide data-driven recommendations for server deployments and capacity planning decisions across Amazon's global data center network * Develop optimization scenarios and business cases to improve capacity delivery efficiency to customers worldwide * Support strategic decision-making through comprehensive analysis of power, performance, and cost trade-offs Cross-Functional Collaboration * Partner with software engineering teams to optimize ML frameworks, drivers, and runtime systems * Collaborate with hardware engineering teams to influence chip design, server architecture, and cooling system optimization * Work closely with data center operations teams to implement and validate optimization strategies Research & Development * Conduct applied research on emerging ML acceleration technologies and their power/performance characteristics * Develop novel methodologies for measuring and improving energy efficiency in large-scale ML workloads * Publish findings and contribute to industry best practices in sustainable ML infrastructure
IN, KA, Bengaluru
Amazon Devices is an inventive research and development company that designs and engineer high-profile devices like the Kindle family of products, Fire Tablets, Fire TV, Health Wellness, Amazon Echo & Astro products. This is an exciting opportunity to join Amazon in developing state-of-the-art techniques that bring Gen AI on edge for our consumer products. We are looking for exceptional scientists to join our Applied Science team and help develop the next generation of edge models, and optimize them while doing co-designed with custom ML HW based on a revolutionary architecture. Work hard. Have Fun. Make History. Key job responsibilities What will you do? - Quantize, prune, distill, finetune Gen AI models to optimize for edge platforms - Fundamentally understand Amazon’s underlying Neural Edge Engine to invent optimization techniques - Analyze deep learning workloads and provide guidance to map them to Amazon’s Neural Edge Engine - Use first principles of Information Theory, Scientific Computing, Deep Learning Theory, Non Equilibrium Thermodynamics - Train custom Gen AI models that beat SOTA and paves path for developing production models - Collaborate closely with compiler engineers, fellow Applied Scientists, Hardware Architects and product teams to build the best ML-centric solutions for our devices - Publish in open source and present on Amazon's behalf at key ML conferences - NeurIPS, ICLR, MLSys.