Digital justice
Credit: Pitiphothivichit / iStock

3 questions about the Amazon–National Science Foundation collaboration on fairness in AI

NSF deputy assistant director Erwin Gianchandani on the challenges addressed by funded projects.

A year ago, Amazon and the National Science Foundation announced a $20 million collaboration to fund academic research on fairness in AI over a three-year period. A month ago, NSF announced the first ten recipients of the program’s grants. Erwin Gianchandani, deputy assistant director for Computer and Information Science and Engineering at NSF, took some time to answer three questions about the program for amazon.science.

1. What is the challenge of fairness in AI?

Four things come to mind.

The first is trying to get to an understanding of what fairness really means. If you think about a mathematical definition of fairness, you could look at two different population types, and you could look at some statistical metric, such as success rate, when you run an algorithm or a classifier on each population. One notion of fairness is that you are trying to ensure that the metric is consistent across both of those population types.

There are other definitions of fairness, though. Philosophers have debated the different notions of fairness for ages. So at the heart of what we’re trying to do with this effort is to better understand what fairness means in the abstract sense so that we can understand how we can design our systems to build fairness into them.

Erwin Gianchandani
Erwin Gianchandani, deputy assistant director for Computer and Information Science and Engineering at NSF.

A second challenge that we’ve identified is who is responsible if you have an AI system that makes unfair decisions. This is where it’s important to think about accountability and how we empower the user of an AI system to have confidence in their ability to take what’s coming out of the AI system and make an informed decision.

You’re trying to provide the user with as much information as possible to minimize the likelihood of unfairness in the outcome — or at least provide an understanding of the types and levels of unfairness that may be inherent to the prediction from the AI system. In other words, this is about trying to present to the end user all of the data that the system used to derive a recommendation to give the user a certain degree of confidence about that recommendation.

A third challenge area that we like to think about is taking this issue of fairness and turning it on its head: how can I harness AI to improve fairness and equity in society? You can think about, for example, equitable distribution of scarce resources like food, of access to health-care, of interventions that might be able to prevent homelessness, and so on. How do we take the vast array of data that are out there and apply AI systems to those data to extract meaningful insights that can allow us to yield improvements in equity in society?

A fourth and final challenge is, how do we construct AI systems so that their benefits are available to everyone? For example, facial-recognition systems should work equally well for people of all races; currently, they do not. Similarly, speech and natural-language systems should work for users from different socioeconomic, ethnic, age, cultural, and geographic groups; that poses significant challenges for current techniques.

2. How do the funded projects address these challenges?

Let me walk through a few examples. Before I do, I want to emphasize that these are just that — examples — and I don’t mean to imply any kind of preference, either toward these funded projects or toward the topics that they are pursuing.

The first challenge is to develop a definition of fairness. One project that we’ve funded in this space is looking at developing a robust theory and methodology for trying to assess and ensure fairness in settings where fairness metrics are currently hard to pin down. You could either specify a particular metric for fairness for a task or domain, or you could look at a particular set of input-output combinations and try to associate fairness characteristics to those.

Take a particular use case, like whether someone has the finances to open a bank account. There might be a set of inputs into the algorithm — one’s monthly or weekly income, current level of debt, and so forth. For every input characteristic or output characteristic, can we define a range within which we feel confident in the accuracy, so that we can essentially try to bound the degree of fairness or unfairness that might exist in that algorithm?

The team of researchers in this case is looking at a particular use case — recidivism in the criminal justice system.

The second challenge is to understand how an AI system produces a given result. We’ve funded a project that is seeking to develop techniques to facilitate better understanding of the entire life cycle of deep neural networks — the preparation of the data, the identification of features, the objectives when it comes to optimization of the system — so that the steps that led to a given output, along with that output, are presented to the user to inform their decision making.

So it’s about really being able to engineer into the outputs a sense of what the system is doing each step of the way so that the human user can see the various decision points. In other words, this is about making it easier to decipher the inner workings of the AI system and, in the process, allowing the user to appreciate any biases.

The third and fourth challenges are somewhat related — harnessing AI to improve equity in society and designing AI systems such that their benefits are equitably available to everyone. One of the projects we’ve funded in this space is looking at racial disparities following cardiac surgery.

We’ve known for quite some time, for example, that certain ethnic groups have higher rates of heart disease than others and are also known to suffer higher rates of postoperative issues — issues that occur after surgical interventions for heart disease. But what we don’t have a sense of is how much of that disparity is due to biological factors, how much of it is due do socioeconomic factors, how much of it is due to the differences in care depending on where people go for treatment, and so on.

We’ve funded a project that is to trying to bring AI tools to a rich electronic-health-record data set to try to understand conceptually and practically the source points for the disparities that we see.

Again, these are just a few examples illustrating the broad research areas, and I expect future awards through this collaboration may be outside these specific topics.

3. What are the advantages of a public-private partnership in addressing these challenges?

We see a significant value proposition in bringing the public and private sectors together.

First, it’s valuable for our academic community to understand the kinds of challenges that industry is seeing. We often call such research “use-inspired”: we have an ability to look at concrete problems and use those to motivate the research questions themselves.

Beyond that, we all know that today’s AI revolution is grounded in large quantities of data that are readily available, along with compute resources to leverage those data sets. In general, access to both of these — for example, access to cloud computing resources — can be really valuable to our academic researchers.

Third, academic researchers benefit from companies’ experience with accelerating the transition of research results out of the laboratory environment and into practice.

Finally, another dimension that’s really important to us is training the next generation of researchers and practitioners. I think we all agree that we’re going to see a real need for competencies in data science, machine learning, and AI across all sectors of our economy. Providing our students who are studying fairness in AI with exposure to industry — to the problems that industry is facing — is a means to nurture the talent that our research ecosystem is going to need going forward. It would be great if some of the students funded on these joint projects benefit from this exposure when they graduate and go on to start their careers.

See a complete list of the projects funded through the new NSF-Amazon collaboration.

Research areas

Related content

IN, KA, Bengaluru
Every product a customer returns is a moment where Amazon either recovers value or writes it off — and India's ReCommerce business is on a multi-million-dollar mission to recover more of it, more intelligently, at scale. Machine learning is the core lever: predicting whether a returned unit is sellable without a human touching it, detecting damage and fraud inside sealed packaging from images, routing each unit to its highest-value disposition, and pricing recovered inventory dynamically. India's returns network is large, fast-growing, and structurally different from other geographies — a rich, high-impact environment for an Applied Scientist to build models that move real financial and customer-experience metrics. We are hiring an Applied Scientist to build and adapt the ML that powers India ReCommerce. You will work at the intersection of two mandates: building India-first models for problems unique to our market, and adapting proven Worldwide models to India's data, catalog, and operational reality — recalibrating them where distribution, language, and process differ. You will own problems end-to-end, from framing and data through modeling, evaluation, and production deployment, partnering closely with engineering, product, and operations. Key job responsibilities Build ML models for automated returns grading — predicting the salability of returned units from structured and unstructured signals so units can be evaluated with zero or minimal human touch, improving speed, accuracy, and recovery value. Develop computer-vision models for defect detection, condition assessment, and anomaly/fraud identification (including inside sealed packaging), and for establishing chain-of-custody and damage attribution across the returns journey. Build disposition-prediction and routing models that direct each unit to its highest-value recovery path (resale, repair, liquidation, donation, recycle) as early as possible in the network. Develop pricing and recovery-optimization models for liquidation and resale, moving from flat rates toward dynamic, grade- and condition-aware pricing. Adapt Worldwide ML models to India — retraining, recalibrating, and re-evaluating for India's return distribution, catalog, languages, and operational constraints, and closing the gaps that prevent a direct lift-and-shift. Own the full model lifecycle — problem framing, data pipelines, feature engineering, training, offline/online evaluation, monitoring, and retraining — with rigorous attention to calibration, drift, and business-metric impact. Partner cross-functionally with engineering (to productionize), product (to frame problems and measure impact), and operations (to ground models in how the network actually runs), and use modern GenAI/LLM tooling to accelerate research and delivery. A day in the life You start by reviewing the performance of a grading model in production — checking calibration and drift against last week's returns, and confirming the recovery-value lift is holding. Mid-morning, you dig into a computer-vision problem: improving detection of a damage type that's driving write-offs, using images captured across the returns journey. In the afternoon you work with a Worldwide science team to bring one of their models to India — scoping what retraining and recalibration India's data requires — then pair with an engineer to move your latest model toward production behind a clean evaluation gate. You close by framing a new problem with a product partner: quantifying the opportunity, defining the label and success metric, and sketching the modeling approach. About the team India ReCommerce owns the systems and science that turn returned and unsellable inventory into recovered value and a better customer experience. You will join a team building an increasingly automated, ML-driven returns network — leveraging Worldwide platforms where they fit and building India-first capabilities where they don't. It is a high-ownership environment with a direct line from your models to measurable business and customer outcomes.
US, VA, Arlington
How do you measure what makes a great leader? How do you evaluate a development program when outcomes take years to materialize and clean experimental conditions are rarely available? How do you take a scientific methodology that a researcher validated carefully in one context and turn it into a system that any HR team across a company of over a million employees can run on their own? These are the kinds of questions the Senior Talent and Transformation Science team works on inside Amazon's People eXperience and Technology organization, and they are questions that matter: the systems this team builds shape how Amazon identifies, develops, and invests in its most senior leaders. As an Applied Scientist on this team you are the person who closes the gap between a validated scientific methodology and a system that runs in production without a scientist standing next to it. The architectural decisions about how scientific methods get encoded into software, the engineering quality bar for the code that implements them, and the reliability of the pipelines that other teams depend on are yours to own. You will work alongside Senior and Principal Research Scientists, an Amazon Scholar, Product Management, and a Senior Applied Scientist who bring deep expertise in behavioral science, psychometrics, and causal inference, and you will be the driving force behind turning that expertise into working, deployable systems for our Amazon executives. The problems you will be building for are genuinely hard and largely unsolved. Scoring a simulation-based leadership assessment with an LLM requires both measurement rigor and a production system that behaves consistently at scale. Estimating the effect of a talent program on leader outcomes requires both a defensible identification strategy and an analytical pipeline someone else can run and trust. Building a self-serve tool that lets a PXT team evaluate a new feature without calling a scientist requires both sound methodology and software that is robust enough to operate without expert supervision. If you want to do work that is technically demanding, scientifically cutting edge, and consequential for real leaders in a large organization, this is that role. Key job responsibilities • Own the production implementation of the team's scientific systems from end to end. When the team validates a new assessment methodology, evaluation framework, or causal identification strategy, you are the scientist who translates it into code that runs reliably, scales, and does not require a scientist standing next to it to operate. • Make the architectural and tooling decisions that determine how scientific methods get encoded into software on this team, choosing abstractions, data structures, and system designs that make the team's scientific components testable, maintainable, and extensible over time. • Define and hold the engineering quality bar for scientific code across the team, establishing and modeling best practices for testing, documentation, reproducibility, and peer review of code in a research team that does not have dedicated software development engineers. • Build the LLM-powered pipelines that operationalize the team's people science, including prompt orchestration, retrieval grounding, automated scoring, and LLM-as-judge evaluation harnesses, writing the implementation yourself and owning the quality and reliability of those systems once deployed. • Extend and adapt scientific techniques at the product level when established approaches fall short. When scoring a simulation-based assessment, estimating a program effect under unusual identification constraints, or evaluating a novel AI feature requires a methodological contribution that does not yet exist, you devise and implement that solution. • Partner with the Research Scientists during methodology design to surface implementation feasibility and trade-offs early, contributing your own scientific judgment on what can be built rigorously within real production constraints before design decisions become expensive to reverse. • Build reusable scientific components, services, and templates that encode methodology once and allow downstream teams to run it without scientist involvement, making the team's research operational infrastructure rather than a bespoke consulting engagement. • Contribute to the design and execution of quasi-experimental evaluations of people programs, owning the analytical implementation and the code pipelines that produce defensible causal evidence from observational and field data. • Mentor scientists on the team on software engineering practices and applied implementation, and participate actively in peer review of experiment designs, analytical approaches, and scientific code written by others. • Communicate implementation trade-offs and system design decisions clearly to product and HR partners in written documents that connect technical choices to business outcomes. A day in the life Your day is anchored in building and testing. You might spend the morning working through a thorny implementation problem, figuring out how to encode a psychometric scoring model into a pipeline that holds up under the messiness of real production data, debugging an LLM evaluation harness that is behaving inconsistently across assessment scenarios, or refactoring a causal estimation component so that another team can run it without calling you first. In the afternoon a Research Scientist might pull you into a methodology design conversation, and your job in that room is not just to follow along but to push back on approaches that would be difficult or brittle to implement, and to propose alternatives that preserve scientific rigor while actually being buildable. You might then shift to reviewing a colleague's code, writing documentation that makes a deployed pipeline understandable to someone who was not in the room when it was designed, or working through a data pipeline problem that is blocking the team's ability to evaluate a new product feature. At the end of most days something that was not working is now working, and the science the team does is a little more durable and a little more independent of any one person than it was in the morning.
US, CA, Culver City
Prime Video is an industry leading, high-growth business and a critical driver of Amazon Prime subscriptions, which contributes to customer loyalty and lifetime value. Prime Video is a digital video streaming and download service that offers Amazon customers the ability to rent, purchase or subscribe to a huge catalog of videos. In addition, Prime Video offers a variety of live sport streaming services in multiple locales. The Prime Video Economist team is looking for an Economist to support PV content valuation. As an economist focusing on Prime Video, you will be responsible for understanding the value that the business creates for our customers and to develop new, disruptive innovations to grow global Prime Video usage and customer value. This role requires an individual with strong quantitative modeling skills and the ability to apply statistical/machine learning, structural models, and experimental design methods to large amount of individual level data. The candidate should have strong communication skills, be able to work closely with stakeholders and translate data-driven findings into actionable insights. The successful candidate will be a self-starter comfortable with ambiguity, with strong attention to detail and ability to work in a fast-paced and ever-changing environment. Key job responsibilities The candidate's responsibilities will include: - Build scalable analytic solutions using state of the art tools based on large datasets - Build causal inference models, conduct statistical/machine learning analyses, or design experiments to measure the value of the business and its many features - Partner closely with Business, Finance, Science, and Tech partners to build prototypes and implement production solutions - Independently identify new opportunities for leveraging economic insights and models in the Video business - Develop and execute product workplans from concept, prototype to production incorporating feedback from customers, scientists and business leaders - Write both technical white papers and business-facing documents to clearly explain complex technical concepts to audiences with diverse business/scientific backgrounds
US, NY, New York
We are seeking a Human-Robot Interaction (HRI) Applied Scientist to develop cutting-edge interactions that make robots feel alive, personal, and fun. In this role, you will focus on verbal and non-verbal conversational systems, social dynamics, memory, and long-term relationship formation between robots, their environments, and the people they interact with. Your contributions will be essential in advancing robotics by enabling expressive, socially intelligent, and trustworthy interactions between robots and humans. Key job responsibilities - Develop interactive systems that leverage large language models, multimodal inputs and outputs, reinforcement learning from human feedback, or other advanced techniques to achieve fluid, engaging, and socially appropriate robot behavior - Design and implement intelligent conversational systems that handle turn-taking, grounding, interruption, and incorporates context drawn from a robot's physical environment and shared history with a user - Integrate perceptual sensor streams including gaze, facial expression, gesture, posture, and more to understand social context and produce coherent, lifelike interactions. - Develop memory and personalization systems that allow robots to form lasting relationships with individual users, learn their environments, and adapt their behavior over weeks and months - Stay updated on advancements in HRI, NLP, multimodal AI, and cognitive and social science to apply cutting-edge techniques to robot interaction challenges - Lead technical projects from conception through production deployment - Mentor junior scientists and engineers - Bridge research initiatives with practical engineering implementation
US, CA, San Francisco
Amazon is on a mission to redefine the future of automation — and we're looking for exceptional talent to help lead the way. We are building the next generation of advanced robotic systems that seamlessly blend cutting-edge AI, sophisticated control systems, and novel mechanical design to create adaptable, intelligent automation solutions capable of operating safely alongside humans in dynamic, real-world environments. At Amazon, we leverage the power of machine learning, artificial intelligence, and advanced robotics to solve some of the most complex operational challenges at a scale unlike anywhere else in the world. Our fleet of robots spans hundreds of facilities globally, working in sophisticated coordination to deliver on our promise of customer excellence — and we're just getting started. As a Scientist in Robot Navigation, you will be at the forefront of this transformation — architecting and delivering navigation systems that are intelligent, safe, and scalable. You will bring deep expertise in learning-based planning and control, a strong understanding of foundation models and their application to embodied agents, and as well as have in-depth understanding of control-theoretic approaches such as model predictive control (MPC)-based trajectory planning. You will develop navigation solutions that seamlessly blend data-driven intelligence with principled control-theoretic guarantees. Our vision is bold: to build navigation systems that allow robots to move fluidly and safely through dynamic environments — understanding context, anticipating change, and adapting in real time. You will lead research that bridges the gap between cutting-edge academic advances and production grade deployment, collaborating with world-class teams pushing the boundaries of robotic autonomy, manipulation, and human-robot interaction. Join us in building the next generation of intelligent navigation systems that will define the future of autonomous robotics at scale. Key job responsibilities - Design, develop, and deploy perception algorithms for robotics systems, including object detection, segmentation, tracking, depth estimation, and scene understanding - Lead research initiatives in computer vision, sensor fusion and 3D perception - Collaborate with cross-functional teams including robotics engineers, software engineers, and product managers to define and deliver perception capabilities - Drive end-to-end ownership of ML models — from data collection and labeling strategy to training, evaluation, and deployment - Mentor junior scientists and engineers; contribute to a culture of technical excellence - Define and track key metrics to measure perception system performance in real-world environments - Publish research findings in top-tier venues (CVPR, ICCV, ECCV, ICRA, NeurIPS, etc.) and contribute to patents A day in the life - Train ML models for deployment in simulation and real-world robots, identify and document their limitations post-deployment - Drive technical discussions within your team and with key stakeholders to develop innovative solutions to address identified limitations - Actively contribute to brainstorming sessions on adjacent topics, bringing fresh perspectives that help peers grow and succeed — and in doing so, build lasting trust across the team - Mentor team members while maintaining significant hands-on contribution to technical solutions About the team Our team is a group is a diverse group of scientists and engineers passionate about building intelligent machines. We value curiosity, rigor, and a bias for action. We believe in learning from failure and iterating quickly toward solutions that matter.
US, NY, New York
We are seeking a Human-Robot Interaction (HRI) Applied Scientist to develop cutting-edge interactions that make robots feel alive, personal, and fun. In this role, you will focus on verbal and non-verbal conversational systems, social dynamics, memory, and long-term relationship formation between robots, their environments, and the people they interact with. Your contributions will be essential in advancing robotics by enabling expressive, socially intelligent, and trustworthy interactions between robots and humans. Key job responsibilities - Develop interactive systems that leverage large language models, multimodal inputs and outputs, reinforcement learning from human feedback, or other advanced techniques to achieve fluid, engaging, and socially appropriate robot behavior - Design and implement intelligent conversational systems that handle turn-taking, grounding, interruption, and incorporates context drawn from a robot's physical environment and shared history with a user - Integrate perceptual sensor streams including gaze, facial expression, gesture, posture, and more to understand social context and produce coherent, lifelike interactions. - Develop memory and personalization systems that allow robots to form lasting relationships with individual users, learn their environments, and adapt their behavior over weeks and months - Stay updated on advancements in HRI, NLP, multimodal AI, and cognitive and social science to apply cutting-edge techniques to robot interaction challenges - Lead technical projects from conception through production deployment - Mentor junior scientists and engineers - Bridge research initiatives with practical engineering implementation
US, WA, Seattle
Our organization in Amazon Robotics builds robots that perform contact-rich manipulation safely and reliably in complex, unstructured environments, at Amazon scale. Our scientists and engineers push the boundaries of robotic manipulation to handle enormous object diversity, bringing deep expertise across planning, control, perception, and machine learning. We learn from real-world data at a scale that few teams in robotics can access. We are seeking an experienced Senior Applied Scientist to help guide a small team advancing reinforcement learning for manipulation. We are creating robots that learn how to push, flip, rearrange, and dexterously insert items with unparalleled robustness, speed, and reliability. Our goal is to deploy robots that will work across Amazon's global network and can handle the full diversity of items that Amazon sells. You will set the technical direction for how we learn these behaviors, from simulation training through reliable execution on physical robots, and you will demonstrate new manipulation capabilities on real hardware at scale. This team's mission reaches beyond any single product: to invent and apply manipulation capabilities that generalize to many future robotics applications. The robots our organization already deploys at scale give you a rare proving ground to collect data, run experiments, and get new policies onto real hardware faster than almost anywhere in the field. You will raise the bar for scientific rigor and engineering quality, and mentor other scientists as the team grows. Key job responsibilities - Set the technical direction for learning non-prehensile and contact-rich manipulation policies, from testing the latest advances in the field through demonstrated capability on hardware. - Oversee the development of reinforcement learning approaches that address the long tail of diverse, demanding manipulation conditions. - Own the path from simulation training to reliable, real-time execution on physical robots, making evidence based calls on where learned approaches should replace engineered ones. - Demonstrate new manipulation capabilities on real robots at scale, and turn one-off results into repeatable methods. - Establish the standards, evaluation practices, and data-informed improvement loops that the team builds on. - Mentor scientists and engineers, and raise the bar for applied science rigor and engineering quality. - Partner across control, perception, and hardware to integrate learned behaviors into working systems. - Represent Amazon in academia through publications and scientific presentations. A day in the life Amazon offers a full range of benefits that support you and eligible family members, including domestic partners. Benefits can vary by location, the number of regularly scheduled hours you work, length of employment, and job status such as seasonal or temporary employment. The benefits that generally apply to regular, full-time employees include: 1. Medical, Dental, and Vision Coverage 2. Maternity and Parental Leave Options 3. Paid Time Off (PTO) 4. 401(k) Plan If you are not sure that every qualification on the list above describes you exactly, we'd still love to hear from you! At Amazon, we value people with unique backgrounds, experiences, and skillsets. If you’re passionate about this role and want to make an impact on a global scale, please apply
IN, MH, Mumbai
Amazon Science gives you insight into the company’s approach to customer-obsessed scientific innovation. Amazon fundamentally believes that scientific innovation is essential to being the most customer-centric company in the world. It’s the company’s ability to have an impact at scale that allows us to attract some of the brightest minds in artificial intelligence and related fields. Our scientists continue to publish, teach, and engage with the academic community, in addition to utilizing our working backwards method to enrich the way we live and work. Please visit https://www.amazon.science for more information. About Amazon Prime Video “Many of the problems we face have no textbook solution, and so we-happily-invent new ones.” – Jeff Bezos
 The Amazon Prime Video team is shaping the future of digital video entertainment. We are seeking a Data Scientist to uncover key insights on how consumers watch videos on Amazon. The ideal candidate will be an expert in the areas of data science, machine learning and statistics, having hands-on experience with multiple improvement initiatives as well as balancing technical and business judgment to make the right decisions about technology, models and methodologies. As consumers increasingly consume digital video, we need to make agile decisions based on what content appeals to our customers. As a Data Scientist at Amazon Prime Video APAC and ANZ analytics team, you will have the opportunity to work on one of the world's largest consumer data sets, influence the long term evolution of our analytics capability and support the expansion of Amazon's digital video business. The Data Scientist will work closely with other research scientists, machine-learning experts, and economists to design and run experiments, research new algorithms, and find new ways to improve optimization across all our associate facing tools. 
 A successful candidate will be able to understand and manage key operational and technical concepts. They will have excellent project and communication skills, and motivation to achieve results in a fast-paced environment. Candidates should demonstrate a passion for working on behalf of customers, have a record of accomplishment of timely delivery of large-scale projects, and have the ability to influence multiple global teams. Autonomy, judgment, influence, and leadership skills are essential. This person will be responsible for ensuring we meet our key deliverables, on time with high quality, and communicating status to internal and external stakeholders. Key Responsibilities - Support the Content team on business reporting, ad hoc analysis, statistical inference and predictive modelling for all Prime Video APAC and ANZ. - Mine and analyze data pertaining to customers viewing experiences to identify critical business insight and make recommendations to optimize content selection. - Proactively develop new ML models using streaming, video, audio and textual data to understand and predict customer streaming behaviour - Translate analytic insights into concrete, actionable recommendations for business or product improvement. Develop and present these as papers to senior stakeholders. - Liaise with your peers in other prime video territories to develop solutions that greatly benefit our global customers - This role will be based in Mumbai, India
US, WA, Seattle
Join us at the forefront of Amazon's sustainability initiatives to work on environmental and social advancements that support Amazon's long-term worldwide sustainability strategy. At Amazon, we're working to be the most customer-centric company on earth. To get there, we need exceptionally talented, bright, and driven people. We are looking for a Senior Research Scientist to join our growing Sustainability team to drive the science behind value chain decarbonization. This role will establish Amazon's scientific methodologies for sector- and cross-sectoral decarbonization mechanisms and establish benchmarks for automated validation and risk assessment. As a Senior Research Scientist, you will be responsible for independently leading assessments of environmental issues across the full spectrum of Amazon businesses and evaluating sustainability impacts across the value chain. You will independently develop quality frameworks and methodologies that enable Amazon to scale procurement of high-quality environmental interventions while maintaining scientific rigor and environmental integrity. Key job responsibilities - Develop quality assessment frameworks for complex environmental interventions, baseline-setting approaches, and measurement methodologies - Build quantitative benchmark and statistical models that enable scalable evaluation across heterogeneous data sources - Create attribution methodologies for supply chain interventions across Amazon's diverse footprint - Develop social and environmental safeguard criteria that integrate community impact assessments - Collaborate with cross-functional teams including procurement, sustainability operations, and business units to translate scientific methodologies into operational requirements - Work under the direction of senior business leaders while acting as lead Subject Matter Expert for value chain decarbonization science, including designing and leading research, data collection, modeling, documentation, interpretation, and validation About the team Diverse Experiences: Worldwide Sustainability values diverse experiences. Even if you do not meet all of the qualifications and skills listed in the job description, we encourage candidates to apply. If your career is just starting, hasn’t followed a traditional path, or includes alternative experiences, don’t let it stop you from applying. Inclusive Team Culture: It’s in our nature to learn and be curious. Our employee-led affinity groups foster a culture of inclusion that empower us to be proud of our differences. Ongoing events and learning experiences, including our Conversations on Race and Ethnicity (CORE) and AmazeCon (inclusive diversity) conferences, inspire us to never stop embracing our uniqueness. Mentorship & Career Growth: We’re continuously raising our performance bar as we strive to become Earth’s Best Employer. That’s why you’ll find endless knowledge-sharing, mentorship and other career-advancing resources here to help you develop into a better-rounded professional.
CN, 31, Shanghai
Worldwide Global Selling has been helping individuals and businesses increase sales and reach new customers around the globe. Today, more than 50% of Amazon's total unit sales come from third-party selection. The Global Selling team in China is responsible for recruiting local businesses to sell on Amazon's 19+ overseas marketplaces and supporting local Sellers' success and growth on Amazon. Our vision is to be the first choice for all types of Chinese business to go globally. The Worldwide Global Selling Analytics, Intelligence, and Technology (WWGS-AIT) team serves as the research, automation, and insight arm of the International Seller Service data hub, enabling rapid delivery of growth insights through strategic investments in regional data foundations, self-service business intelligence solutions, and artificial intelligence tools. The WWGS-AIT team is positioned to establish AI-ready foundational capabilities across the WWGS organization while maintaining excellence in business insight generation, and self-service BI/AI application development. WWGS-AIT is looking for a Data Scientist to design and build seller-facing AI agents that turn our AI-ready data foundation into intelligent, conversational experiences for Amazon's global sellers. You will own the intelligence layer of these agents end-to-end, from modeling and retrieval to evaluation and launch, working alongside applied scientists, data engineers, and the Seller Assistant platform team to put trustworthy AI directly into sellers' hands. Key job responsibilities - Design, build, and iterate seller-facing AI agents (LLM-powered) that help Chinese sellers grow globally, reasoning over WWGS-AIT's AI-ready data foundation and knowledge base. - Develop the intelligence layer of agents: retrieval-augmented generation (RAG) over our knowledge management system, tool-use / function-calling orchestration, prompt engineering, and model fine-tuning or adaptation where needed. - Ground agent responses in standardized metrics and unified seller profiles to guarantee consistency and accuracy across agents; design and enforce guardrails that prevent hallucination and protect sensitive, compliance-restricted data. - Build rigorous evaluation frameworks (golden datasets, offline evaluation, and online experimentation) to measure and continuously improve agent quality, safety, and seller impact. - Develop seller-intelligence models (segmentation, entity resolution / One-ID, ranking and recommendation) that power personalized agent experiences. - Partner with WWGS Tech and the Seller Assistant platform team to productionize agents and tools (e.g., via MCP), defining the model and intelligence contract while engineering operates the runtime. - Collaborate with business, product, and cross-functional partners to translate seller pain points into agent capabilities and measurable business outcomes. - Stay current with advances in GenAI and agentic systems, and bring applied research into production.