SimBot Challenge FAQs

Frequently asked questions about the challenge.
General
What is the Alexa Prize?
Alexa is Amazon’s cloud-based voice service available on over 100 million devices from Amazon and third-party device manufacturers. With Alexa, you can build natural voice experiences that offer customers a more intuitive way to interact with the technology they use every day. Our collection of tools, APIs, reference solutions, and documentation makes it easy for anyone to build with Alexa.
Why did Amazon create the Alexa Prize?
The Alexa Prize, an annual university competition dedicated to accelerating the field of conversational artificial intelligence (AI), was created to recognize students from around the globe who are changing the way we interact with technology. The goal is to advance several areas of conversational AI including natural language understanding (NLU), context modeling, dialog management, commonsense reasoning, natural language generation (NLG), and knowledge acquisition.
How does the Alexa Prize support research?
The Alexa Prize is a research testbed for university students to experiment with and advance conversational AI at scale.

Research teams own the intellectual property (IP) in their systems and are encouraged to publish scientific articles on their work. As described in the Official Rules, participating teams grant Amazon a non-exclusive license to any technology or software they develop in connection with the competition.
What new datasets will I have access to as part of the Alexa Prize?
The TEACh dataset as well as a training dataset proprietary to the live interaction portion of the SimBot Challenge will be available to competitors in each phase.
SimBot
What is the SimBot Challenge?
SimBot, a competition focused on helping advance development of next-generation virtual assistants that will assist humans in completing real-world tasks by continuously learning, and gaining the ability to perform commonsense reasoning.

The SimBot Challenge will have two phases: A public benchmark phase, and a live interactions phase. Participants in both phases will build machine-learning models for natural language understanding, human-robot interaction, and robotic task completion. Artificial intelligence challenges addressed in the competition relate to reasoning on language and scene understanding, learning from demonstration, self-learning, and task completion utilizing natural language.

Unlike previous Alexa Prize competitions, the public benchmark challenge phase will be open to university teams, as well as individuals in academia and industry interested in advancing the science of AI and engaging top researchers from around the globe. The SimBot Challenge public benchmark phase is like existing visual language navigation competitions.
What is the difference between the Public Benchmark Challenge and the Live Interaction Challenge?
The Public Benchmark Challenge is open to any individual or team, academic or industry, who wants to complete and submit a model for evaluation and rating during the evaluation period. It is based on the TEACh dataset offered to the research community in October, 2021. The Live Interaction Period’s participation is limited solely to the university-based teams selected for the SimBot Challenge in November 2021 and June, 2022. These teams will each receive Amazon sponsorship to build a SimBot that will compete in a challenge from July, 2022 to September 2022 where they will receive real time ratings and feedback from Alexa Users.
Why is the SimBot - Live Interaction Challenge only limited to university teams?
The SimBot Challenge - live interaction phase is part of the Alexa Prize which is currently limited to universities and university students in Amazon initiatives to advance AI. However, the Public Benchmarking Challenge is open to both academic and university-based teams.
What will my SimBot do?
SimBot will navigate in a virtual environment to complete challenges guided by Alexa users. SimBot will interact with objects in the environment including but not limited to those common to offices and homes. There will be obstacles and hazards introduced to make game play fun and challenging.
How will I build my SimBot?
SimBots will use images from the game and instructions from Alexa users to navigate in a virtual world. Teams will build AI models which recognize objects and scenes as well as understand natural language commands from users. We plan to provide baseline data and a baseline model as a reference but we expect teams to develop new and novel approaches as well as augment the provided data to improve performance.
Eligibility: Public Benchmark Challenge
Phase 1, SimBot challenge
Who is eligible to participate in the public benchmark challenge?
Any individual or team, academic or industry-based.
How do I sign up to participate in the public benchmark challenge?
A registration form will open on November 15, 2021 at alexaprize.com.
Is funding available to university teams competing in the public benchmark challenge?
No, funding is only available to university teams selected to participate in the SimBot Challenge. In addition to the teams selected in 2021 up to four new, high-performing teams from the public benchmark challenge will be invited to apply for the SimBot challenge, live interaction phase for which funding is available. Learn more here.
Who can apply to participate?
Any individual or team can register and participate in the public benchmark phase with exception of parties from Cuba, Iran, North Korea, Sudan, Syria, and the region of Crimea.

Registration for the public benchmark challenge will open November 15, 2021. The challenge will begin on January 10, 2022.
Do I need to be a certain age?
Participants must be at or above the age of majority in the country, state, province or jurisdiction of residence at the time of registration.
Can I enroll if a family member is an Amazon employee?
Immediate family members and household members of Amazon employees, directors and contractors are not eligible to participate.
Eligibility: SimBot Challenge
Includes sponsored teams for both Public Benchmark and Live Interaction phases
Who can apply to participate?
The Alexa Prize is open to full-time students enrolled in an accredited university, with the exception of universities in Cuba, Iran, North Korea, Sudan, Syria, and the region of Crimea (see Official Rules). Proof of enrollment will be required to participate.
Can I participate if I don’t attend a university?
No. The Alexa Prize is open only to full-time enrolled university students.
Do I need to be enrolled in a university program throughout the duration of the competition?
All participating team members must remain full-time students in good standing at their university while participating in the competition.
Do I need to be a certain age?
Participants must be at or above the age of majority in the country, state, province or jurisdiction of residence at the time of entry.
Can I enroll if a family member is an Amazon employee?
Immediate family members and household members of Amazon employees, directors and contractors are not eligible to participate. See Official Rules for additional restrictions.
If my team fails to apply for the SimBot Challenge now, will there be any future opportunities to compete in the Live Interaction challenge?
Yes, the initial application period run from October 4-31, 2021. In addition to the teams selected in 2021 up to four new, high-performing university teams from the public benchmark challenge will be invited to apply for the SimBot challenge, live interaction phase for which funding is available. Learn more here.
Will university teams selected post-Public Benchmark Challenge be eligible for funding and for what amount?
Teams selected for the SimBot Challenge in June, 2022 will be eligible for funding.
If my team was not selected during the first application period are we eligible to re-apply in may if we qualify?
Yes. University-based teams that participate in the public benchmark phase are eligible to re-apply for the SimBot challenge between May 9 - 23, 2022. Four new, high-performing university teams from the public benchmark challenge will be invited to apply for the SimBot challenge, live interaction phase for which funding is available. Learn more here.
My team is not completely comprised of university students but we perform well in the public benchmark challenge? Is there an opportunity for us to continue in the competition?
Yes, we plan to continue the public benchmark competition through the end of 2022.
My company does collaborative work with several universities, are we eligible to compete in the live interaction challenge alongside them?
No, only full-time students are eligible to participate in the SimBot challenge, which includes the live interaction period.
My university and another frequently collaborate on projects - is there any exception to the students all enrolled in one university rule?
Applications are limited to students all from the same, single, university.
Teams
How many teams will be selected to participate?
All applications will be reviewed and evaluated by a panel of Amazon experts. Up to ten teams will be selected and sponsored by Amazon. All teams selected in November, 2021 will receive a $250,000 grant while those selected in June, 2022 will receive a $200,000 grant intended to support two full-time students, a month of faculty time, free Alexa devices, and free AWS hosting including access to CPU and GPU based machines, SQL and NoSQL databases, and object storage. See the Official Rules for details.
How many team members can our team have?
There are no minimum or maximum team member requirements. All team members must be enrolled in their university throughout the duration of the competition. All teams will receive a $250,000 grant if selected in November 2021 or a $200,000 grant if selected in June, 2022 regardless of how many members are on the team. We recommend a team with 4-6 students with diverse fields of study or areas of expertise.
Can students from different universities be on the same team?
Teams must be comprised of students attending the same university.
Can one university have more than one team?
Yes, universities may have more than one team.
Can I participate on two separate teams?
You can only be a part of one team for the duration of the competition.
Can undergraduate and graduate students work together?
Yes, teams may be comprised of undergraduate and graduate students.
Do I need a faculty advisor?
All teams must nominate a faculty advisor and include the faculty advisor’s consent in the applications.
What is the role of the faculty advisor?
Faculty advisors will advise students on technical directions and be a sounding board for new ideas, similar to a graduate school advisor. They will also act as the official representative from the university for this competition.
Can we add or remove team members during the competition?
During the competition, faculty advisors may request to remove or add members to the team, subject to approval by Amazon.
Can we discuss our SimBot with faculty or students who aren’t on our team?
Only team members may work on their SimBot. However, the faculty advisor and other students and faculty members at your university may provide support and advice to your team and may co-author technical publications and research papers.
Application process
How do we apply?
Check the SimBot page for the latest update on applications.
What do we need to apply?
Once you have selected your team members, team leader, and faculty sponsor, you are ready to begin the application process.
Do all team members have to apply?
Each team must have a team lead, who should apply on behalf of the whole team. Your application must include all of your team members’ information.
Is there an application fee?
There is no application fee.
How will teams be selected to participate?
All applications will be reviewed by a panel of Amazon employees. Teams will be selected based on the following criteria: (1) the potential scientific contribution to the field; (2) the technical merit of the approach; (3) the novelty of the idea; and (4) an assessment of the team’s ability to execute against their plan. Please be sure to provide enough detail in your application to enable our experts to evaluate your proposal.
Competition details
What is the goal of the challenge?
To develop AI models which advance the state of the art and allow users to naturally interact with a robotic assistant in a virtual world to successfully complete a range of challenges.
How will winners be selected?
Winners will be determined based on the final standings at the completion of the finals period.
Can we use other funding to help us participate in this challenge?
Yes, you may use other funding to support your team, subject to the terms described in the Official Rules. External funding will need to be disclosed by January 1, 2023.
Will Alexa customers be able to engage with our SimBot?
Your team will be required to submit its SimBot for certification and publication by the Amazon Alexa team. After certification, you will enter the Internal Amazon Beta Period, where Amazon employees will test your SimBot and provide feedback. After the Internal Amazon Beta Period, we will allow Amazon Alexa customers to try your SimBot and provide feedback to you. Amazon may impose Availability Criteria, or requirements the SimBot must meet before it will be made available to Alexa users. Availability Criteria may include criteria such as a minimum average customer rating, uptime requirements, and an ability to consistently filter offensive content.
Amazon launched the Echo in the UK, Germany, India, Japan, and other countries. Will localized languages be supported?
Your team must build its SimBot using U.S. English. Your SimBot will be available to Alexa customers in the U.S. Customers in other countries may also access it by setting their Amazon PFM (Preferred Marketplace) to U.S.
Will we publish our research from the Alexa Prize?
Yes. Publishing research papers as an outcome of your work on the Alexa Prize is required for all teams participating in the competition, although teams should not publish Amazon confidential information, as described in the Official Rules. The Alexa Prize requires all teams to submit a technical paper for the Alexa Prize proceedings. Your SimBot will not be selected for the finals if your team does not submit a technical paper for Alexa Prize proceedings. Papers will be published at the end of the competition in an online Proceedings of the Alexa Prize, which will be publicly available.

Teams may also publish research papers in third-party publications and conferences, as long as all papers are provided to Amazon for review at least two weeks before the submission deadlines and no research papers are published before the Alexa Prize proceedings are published without Amazon’s prior approval.
Who will own the intellectual property rights in my submission?
You will retain ownership over your SimBot. Amazon will have a non-exclusive license to any technology or software you develop in connection with the competition. See the Official Rules for details.
Prizes
What are the prizes for winning the competition?
A prize of $500,000 will be awarded to the team that creates the best SimBot. The second-place and third-place finalist teams will receive a $100,000 and a $50,000 prize, respectively. See the contest rules for details.
Do we get a stipend and devices to participate in the Alexa Prize?
Up to ten teams will be sponsored to participate in the Alexa Prize in November, 2021. These teams’ universities will receive a $250,000 research grant to fund the team members’ work over the year. Additional teams selected to participate in the SimBot Live Interaction challenge in June, 2022 will receive a $200,000 research grant to fund the team members’ work over the year.

The sponsorship includes one Alexa-enabled device per team member for up to a total of three devices per team and one Alexa-enabled device per faculty advisor, free AWS services to support the development of their SimBot, and support from the Alexa team.
How can the grant be spent?
The grants will be awarded with the intention that they will support two full-time students for the duration of the Competition and one month of the Faculty Advisor’s salary. No more than 35% of the research grants may be allocated to administrative fees. If your team would like to use the funds in another manner, your faculty advisor must receive approval from Amazon before doing so.
What happens if we are selected and receive a stipend but can no longer participate?
Stipends will be awarded in installments payable to the university. If your team withdraws before any of the installments, remaining funds will not be transferred to the university.
How will the prizes be distributed among a team?
The first, second, and third place prizes will be distributed equally among all registered team members. The official list of registered team members must be confirmed by January 30, 2023.
Timeline
What are the key milestones of the competition?
Teams must submit their applications between October 1, 2021 and October 31, 2021. Between November 1 and November 10, 2021, we will announce teams selected to participate. In the summer of 2022 following the public benchmark phase and the second application period teams will be invited to an Alexa Prize Bootcamp at Amazon where they will receive training on the resources made available to all competing teams. The finals will be scheduled between January 30 and March 27, 2023 and will determine the winning teams in the first annual SimBot Challenge.
If selected, when will we receive the stipend, devices, access to the Alexa Prize SimBot toolkit, our AWS account, and be introduced to our point of contact?
We will reach out to all teams no later than November 10, 2021 with instructions on next steps. Up to ten teams will be selected to receive a $250,000 stipend, Alexa-enabled devices, free AWS services to support their development efforts, and support from the Alexa team.
More information
See the full competition rules or submit your questions. Need assistance? Email: alexaprizesupport@amazon.com

Latest news

The latest updates, stories, and more about Alexa Prize.
IN, KA, Bengaluru
Do you want to join an innovative team of scientists who use machine learning and statistical techniques to create state-of-the-art solutions for providing better value to Amazon’s customers? Do you want to build and deploy advanced algorithmic systems that help optimize millions of transactions every day? Are you excited by the prospect of analyzing and modeling terabytes of data to solve real world problems? Do you like to own end-to-end business problems/metrics and directly impact the profitability of the company? Do you like to innovate and simplify? If yes, then you may be a great fit to join the Machine Learning and Data Sciences team for India Consumer Businesses. If you have an entrepreneurial spirit, know how to deliver, love to work with data, are deeply technical, highly innovative and long for the opportunity to build solutions to challenging problems that directly impact the company's bottom-line, we want to talk to you. Major responsibilities - Use machine learning and analytical techniques to create scalable solutions for business problems - Analyze and extract relevant information from large amounts of Amazon’s historical business data to help automate and optimize key processes - Design, development, evaluate and deploy innovative and highly scalable models for predictive learning - Research and implement novel machine learning and statistical approaches - Work closely with software engineering teams to drive real-time model implementations and new feature creations - Work closely with business owners and operations staff to optimize various business operations - Establish scalable, efficient, automated processes for large scale data analyses, model development, model validation and model implementation - Mentor other scientists and engineers in the use of ML techniques Key job responsibilities Use machine learning and analytical techniques to create scalable solutions for business problems Analyze and extract relevant information from large amounts of Amazon’s historical business data to help automate and optimize key processes Design, develop, evaluate and deploy, innovative and highly scalable ML models Work closely with software engineering teams to drive real-time model implementations Work closely with business partners to identify problems and propose machine learning solutions Establish scalable, efficient, automated processes for large scale data analyses, model development, model validation and model maintenance Work proactively with engineering teams and product managers to evangelize new algorithms and drive the implementation of large-scale complex ML models in production Leading projects and mentoring other scientists, engineers in the use of ML techniques About the team International Machine Learning Team is responsible for building novel ML solutions that attack India first (and other Emerging Markets across MENA and LatAm) problems and impact the bottom-line and top-line of India business. Learn more about our team from https://www.amazon.science/working-at-amazon/how-rajeev-rastogis-machine-learning-team-in-india-develops-innovations-for-customers-worldwide
IN, KA, Bengaluru
RBS (Retail Business Services) Tech team works towards enhancing the customer experience (CX) and their trust in product data by providing technologies to find and fix Amazon CX defects at scale. Our platforms help in improving the CX in all phases of customer journey, including selection, discoverability & fulfilment, buying experience and post-buying experience (product quality and customer returns). As a Sciences team in RBS Tech, we focus on foundational ML research and develop scalable state-of-the-art ML solutions to solve the problems covering customer experience (CX) and Selling partner experience (SPX). We work to solve problems related to multi-modal understanding (text and visual), supervised and unsupervised techniques, multi-task learning, multi-label classification, aspect and topic extraction for Customer Anecdote Mining, product similarity, using GenAI, LLMs, NLP and Computer Vision. Key job responsibilities As an Applied Science Manager, you will be responsible to design and deploy scalable GenAI, NLP and Computer Vision solutions that will impact the content visible to millions of customer and solve key customer experience issues. You will Lead scientists on the team and oversee research and development projects at various stages ranging from initial exploration to deployment into production systems. You will partner with business and engineering teams to identify and solve large and significantly complex problems that require scientific innovation. You will help the team leverage your expertise, by coaching and mentoring. You will contribute to the professional development of colleagues, improving their technical knowledge and the engineering practices. You will create the environment in the team to file for patents and/or publish research work where opportunities arise. You will impact the large product strategy, identifies new business opportunities and provides strategic direction to the team.
ES, M, Madrid
Are you interested in changing how Amazon does marketing — moving beyond platform-optimized broad reach to campaigns that find the right customer, at the right moment, using Amazon's unmatched 1P data? We are seeking an Applied Scientist to join PRIMAS (Prime & Marketing Analytics and Science). In this role, you will design and run the experiments that answer the foundational question for EU marketing: does adding 1P audience signal on top of Value-Based Optimization (VBO) improve marketing efficiency — and if so, for which customer cohorts, on which surfaces, and at what scale? Amazon's current marketing model is largely platform-led: we set objectives and let platforms optimize toward conversion. This approach works well for broad acquisition but systematically underserves lifecycle goals — it cannot distinguish between a Bargain Hunter who will never pay full price and a high-potential customer one nudge away from becoming a Prime member. This role sits at the center of changing that. You will build the 1P audiences, design the experiments that test them, and generate the evidence that guides how Amazon allocates hundreds of millions in marketing spend. Year 1 is an experimentation year. You will deploy 1P audiences across multiple surfaces and channels — Meta, Google, Amazon Display Ads — and measure incrementally against VBO baselines. The goal is not to replace platform optimization but to understand when and where the combination of 1P signal + VBO outperforms VBO alone, and to build the experimental infrastructure that makes this learning scalable. Key job responsibilities 1P Audience Development & Experimentation: - Build and validate 1P audience segments from Amazon behavioral, transactional, and lifecycle data - Design experiments that isolate the incremental effect of 1P audience signal over platform VBO baselines - Deploy audiences across activation surfaces and establish measurement standards that make cross-surface comparison valid Causal Measurement & Incrementality: - Apply causal inference methods to measure the true incremental lift of audience-based targeting vs. VBO - Develop power analysis frameworks and guardrails that enable rapid experimentation without underpowered or conflated tests - Deliver optimization recommendations grounded in experimental evidence: which cohorts respond, which surfaces deliver, which creative strategies drive behavior change Scaling the Learning: - Build reusable audience and measurement frameworks that can be deployed across campaigns and channels — year 1 experiments should produce infrastructure, not one-off analyses - Document experimental learnings in a way that informs both the 2026 roadmap and the business case for investing further in 1P audience capabilities in 2027+ - Partner with engineering and PMT to translate validated audience prototypes into production-ready solutions that scale beyond the experimentation phase About the team The PRIMAS team, is part of a larger tech tech team of 100+ people called WIMSI (WW Integrated Marketing Systems and Intelligence). WIMSI core mission is to accelerate marketing technology capabilities that enable de-averaged customer experiences across the marketing funnel: awareness, consideration, and conversion.
US, WA, Seattle
Are you interested in building Agentic AI solutions that solve complex builder experience challenges with significant global impact? The Security Tooling team designs and builds high-performance AI systems using LLMs and machine learning that identify builder bottlenecks, automate security workflows, and optimize the software development lifecycle—empowering engineering teams worldwide to ship secure code faster while maintaining the highest security standards. As a Data Scientist on our Security Tool team, you will focus on building state-of-the-art ML models to enhance builder experience and productivity. You will identify builder bottlenecks and pain points across the software development lifecycle, design and apply experiments to study developer behavior, and measure the downstream impacts of security tooling on engineering velocity and code quality. Our team rewards curiosity while maintaining a laser-focus on bringing products to market that empower builders while maintaining security excellence. Competitive candidates are responsive, flexible, and able to succeed within an open, collaborative, entrepreneurial, startup-like environment. At the forefront of both academic and applied research in builder experience and security automation, you have the opportunity to work together with a diverse and talented team of scientists, engineers, and product managers and collaborate with other teams. This role offers a unique opportunity to work on projects that could fundamentally transform how builders interact with security tools and how organizations balance security requirements with developer productivity. Key job responsibilities Design and run rigorous experiments to evaluate and improve security tooling performance, builder experience, and adoption across hundreds of thousands of builders, multiple security tools, and diverse business verticals. Lead the end-to-end lifecycle of data science and ML models — from research and experimentation through production launch — including defining success metrics, obtaining stakeholder sign-off, and managing rollout. Conduct online and offline analyses to measure the real-world impact of security tooling improvements beyond adoption metrics, including downstream effects on vulnerability resolution, builder productivity, and organizational security posture. Develop and deploy production-grade machine learning and statistical models using Python, SQL, and related tools to automate insights, detect patterns, and drive decision-making across STF's security tool ecosystem. Perform large-scale exploratory data analysis on builder feedback, ticket resolution, tool usage, and customer satisfaction data to uncover patterns, identify opportunities, and inform product and tooling decisions. Translate complex research findings into clear insights and recommendations for technical and non-technical stakeholders at all levels, including STF leadership metric reporting and customer satisfaction publications. Contribute to Amazon's scientific community and the broader research field through collaboration and publication in top-tier venues. A day in the life Morning - Review overnight pipeline health — nudge systems, ticket classification models, and adoption dashboards running as expected - Join daily standup with the SDI team to align on priorities and flag blockers - Dive into exploratory analysis — investigating a spike in unresolved tickets or segmenting builder feedback to understand adoption gaps Midday - Partner with security tool owners (e.g., Shepherd, Talos, Scorecard) to review experiment results — did the latest nudge improve resolution rates? - Translate findings into actionable recommendations for leadership reviews or WBR updates - Analyze CSAT survey data to surface emerging dissatisfaction themes Afternoon - Write production code — building features for the classification pipeline, optimizing SQL for the metrics scorecard, or iterating on a model for predicting resolution timelines - Collaborate with STF stakeholders to define success metrics for an upcoming model launch - Document findings, update trackers, and queue next steps About the team Diverse Experiences Amazon Security values diverse experiences. Even if you do not meet all of the qualifications and skills listed in the job description, we encourage candidates to apply. If your career is just starting, hasn’t followed a traditional path, or includes alternative experiences, don’t let it stop you from applying. Why Amazon Security? At Amazon, security is central to maintaining customer trust and delivering delightful customer experiences. Our organization is responsible for creating and maintaining a high bar for security across all of Amazon’s products and services. We offer talented security professionals the chance to accelerate their careers with opportunities to build experience in a wide variety of areas including cloud, devices, retail, entertainment, healthcare, operations, and physical stores. Inclusive Team Culture In Amazon Security, it’s in our nature to learn and be curious. Ongoing DEI events and learning experiences inspire us to continue learning and to embrace our uniqueness. Addressing the toughest security challenges requires that we seek out and celebrate a diversity of ideas, perspectives, and voices. Training & Career Growth We’re continuously raising our performance bar as we strive to become Earth’s Best Employer. That’s why you’ll find endless knowledge-sharing, training, and other career-advancing resources here to help you develop into a better-rounded professional. Work/Life Balance We value work-life harmony. Achieving success at work should never come at the expense of sacrifices at home, which is why flexible work hours and arrangements are part of our culture. When we feel supported in the workplace and at home, there’s nothing we can’t achieve.
US, WA, Seattle
At Amazon Selection and Catalog Systems (ASCS), our mission is to power the online buying experience for customers worldwide so they can find, discover, and buy any product they want. We innovate on behalf of our customers to ensure uniqueness and consistency of product identity and to infer relationships between products in Amazon Catalog to drive the selection gateway for the search and browse experiences on the website. We're solving a fundamental AI challenge: establishing product relevant information at unprecedented scale with Frontier Models and Agents. The scale is staggering: billions of products, petabytes of multimodal data, millions of sellers, dozens of languages, and infinite product diversity ranging from electronics to groceries to digital content. The research challenges are immense. GenAI and VLMs hold transformative promise for catalog understanding, but we operate where traditional methods fail: ambiguous problem spaces, incomplete and noisy data, inherent uncertainty, reasoning across both images and textual data, and explaining decisions at scale. Enriching product information requires sophisticated models that reason across text, images, and structured data, all while maintaining accuracy and trust for high-stakes business decisions affecting millions of customers daily. Amazon's Catalog System Services Science team is looking for an innovative and customer-focused applied scientist to help us make the world's best product catalog even better. In this role, you will partner with technology and business leaders to build new state-of-the-art algorithms, models, and services. You will pioneer advanced GenAI solutions that power next-generation agentic shopping experiences, working in a collaborative environment where you can experiment with massive data from the world's largest product catalog, tackle problems at the frontier of AI research, rapidly implement and deploy your algorithmic ideas at scale, across millions of customers. Key job responsibilities * Formulate novel research problems at the intersection of GenAI, multimodal learning, and large-scale information retrieval. In essence, translating ambiguous business challenges into tractable scientific frameworks * Design and implement leading models leveraging frontier models, and agentic architectures to enrich catalog information at billion-product scale * Pioneer explainable AI methodologies that balance model performance with scalability requirements for production systems impacting millions of daily customer decisions * Own end-to-end ML pipelines from research ideation to production deployment, processing petabytes of multimodal data with rigorous evaluation frameworks * Define research roadmaps aligned with business priorities, balancing foundational research with incremental product improvements * Mentor peer scientists and engineers on advanced ML techniques, experimental design, and scientific rigor, building organizational capability in GenAI and multimodal AI * Represent the team in the broader science community - publishing findings, delivering tech talks, and staying at the forefront of GenAI, VLM, and agentic system research
US, VA, Arlington
The AWS Certification team is seeking a Psychometrician with experience working with criterion-referenced assessment programs to support a large global AWS Certification and Credentialing program. In this role, you will support all psychometric aspects of exam development and operation, including job analyses, standard setting, automated test assembly, item and test analyses, optimal item bank design, quality assurance, and project planning. You will work closely with a team of psychometricians, subject matter experts, certification exam program managers, publishing, delivery, security, and product management teams to support ongoing analyses of exam and credential data. To be successful in this position, you must be highly motivated, creative, detail oriented, and a self-starter who is able to think big, execute, ensure high quality, yet stay focused on the details. Key job responsibilities • Conduct Job Task Analysis (JTA) workshops and post-JTA survey analyses to define the blueprint and test specifications for new certifications or updates to existing certifications • Conduct standard setting studies to set the passing score for exams and credentials • Run item analysis to evaluate quality and performance of exam items • Use automated test assembly procedures to assemble forms or item pools • Work with content development to track item bank trends and optimize the health of item banks • Support the development of a cloud-based analytics and reporting system • Partake in development and performance analysis of credentials • Interpret and clearly communicate the results of analyses to stakeholders through written and oral reports • Follow the accreditation standards set by ISO/IEC:2012 17024 and the National Council for Certifying Agencies (NCCA) as they relate to valid psychometric practices • Contribute to the development and execution of the strategic goals regarding the AWS certification and credentialing program. • Consult with leadership, internal staff, external consultants, and industry leaders regarding advancement of current offerings
JP, 13, Tokyo
Every day, Amazon Japan delivers millions of packages to customers' doors. Behind every routing decision, capacity plan, and network design is a modeling problem — and the science behind how those models learn, generalize, and improve in production is where you come in. JP OPS STAR Foundation is the applied science and engineering team that builds the decision systems powering Amazon Japan's transportation operations. We sit at the intersection of generative AI, large-scale optimization, and graph-based learning — developing models that turn complex operational structure into actionable intelligence. As an applied scientist, you will formulate and solve research problems that directly shape operational outcomes. Your work will span: - LLM-based agent systems — designing agentic architectures with structured evaluation frameworks that measure reasoning quality, tool use accuracy, and task completion under real operational constraints - GPU-optimized model serving and training — developing pipelines for large model inference and fine-tuning, with rigorous benchmarking of latency, throughput, and cost trade-offs - Graph representation learning — building graph neural network models that capture logistics network topology and learn node/edge representations for downstream prediction and optimization tasks You will own problems end-to-end: from formulation and experimentation through deployment and continuous measurement in production. We expect scientific rigor — well-designed experiments, proper baselines, quantified uncertainty — and the engineering judgment to make your methods work reliably at scale. This is a high-autonomy, high-impact role. You will collaborate with data engineers, BIEs, and operations leaders, and you will have the freedom to define your research agenda within our problem space. We value scientists who publish and share their work, and who treat production performance as ground truth for their ideas. At Amazon, you'll work alongside the latest AI and GenAI tools that are increasingly woven into how teams operate: from AI-powered capabilities that accelerate decision-making, to Generative AI that helps you focus on work that truly matters. You'll have opportunities and resources to develop AI fluency at your own pace, with continuous learning built into the culture. Key job responsibilities - Formulate and solve modeling problems for LLM-powered agent systems (MCP servers, text-to-SQL, RAG), designing validation methods and confidence scoring for reliable outputs under operational constraints. - Design evaluation frameworks for AI agents: define metrics for reasoning quality and task completion, build backtesting pipelines, and develop automated regression detection. - Research and implement graph neural network architectures to learn structural representations of logistics networks; develop training methodology and evaluation protocols for production deployment. - Develop GPU-optimized pipelines for model training and inference, with rigorous benchmarking of latency, throughput, and cost trade-offs. - Own the scientific lifecycle end-to-end: problem formulation, experimental design, deployment, and continuous measurement against real operational outcomes. - Collaborate with data engineers and operations stakeholders to validate model outputs and quantify business impact. A day in the life - Diagnose a confidence-score drift in the agent eval dashboard, trace the root cause to a schema change, and redesign the validation logic to be schema-invariant. - Experiment with a graph attention architecture on the logistics network; benchmark against the baseline embedding on a downstream prediction task. - Profile a GPU training job to unblock a scaling experiment, identify a memory bottleneck, and restructure the data loader to halve iteration time. - Present model accuracy trends to operations leaders and recommend threshold adjustments based on the data.
US, TX, Austin
Applied Scientists in AWS Automated Reasoning are dedicated to making AWS the best computing service in the world for customers who require advanced and rigorous solutions for automated reasoning, privacy, and sovereignty. Key job responsibilities The successful candidate will: - Solve large or significantly complex problems that require deep knowledge and understanding of your domain and scientific innovation. - Own strategic problem solving, and take the lead on the design, implementation, and delivery for solutions that have a long-term quantifiable impact. - Provide cross-organizational technical influence, increasing productivity and effectiveness by sharing your deep knowledge and experience. - Develop strategic plans to identify fundamentally new solutions for business problems. - Assist in the career development of others, actively mentoring individuals and the community on advanced technical issues. A day in the life This is a unique and rare opportunity to get in early on a fast-growing segment of AWS and help shape the technology, product and the business. You will have a chance to utilize your deep technical experience within a fast moving, start-up environment and make a large business and customer impact. About the team Diverse Experiences Amazon Automated Reasoning values diverse experiences. Even if you do not meet all of the qualifications and skills listed in the job description, we encourage candidates to apply. If your career is just starting, hasn't followed a traditional path, or includes alternative experiences, don't let it stop you from applying. Why Amazon Automated Reasoning? At Amazon, automated reasoning is central to maintaining customer trust and delivering delightful customer experiences. Our organization is responsible for creating and maintaining a high bar for automated reasoning across all of Amazon's products and services. We offer talented automated reasoning professionals the chance to accelerate their careers with opportunities to build experience in a wide variety of areas including cloud, devices, retail, entertainment, healthcare, operations, and physical stores. Inclusive Team Culture In Amazon Automated Reasoning, it's in our nature to learn and be curious. Ongoing DEI events and learning experiences inspire us to continue learning and to embrace our uniqueness. Addressing the toughest automated reasoning challenges requires that we seek out and celebrate a diversity of ideas, perspectives, and voices. Training & Career Growth We're continuously raising our performance bar as we strive to become Earth's Best Employer. That's why you'll find endless knowledge-sharing, training, and other career-advancing resources here to help you develop into a better-rounded professional. Work/Life Balance We value work-life harmony. Achieving success at work should never come at the expense of sacrifices at home, which is why flexible work hours and arrangements are part of our culture. When we feel supported in the workplace and at home, there's nothing we can't achieve.
US, WA, Seattle
Applied Scientists in AWS Automated Reasoning are dedicated to making AWS the best computing service in the world for customers who require advanced and rigorous solutions for automated reasoning, privacy, and sovereignty. Key job responsibilities The successful candidate will: - Solve large or significantly complex problems that require deep knowledge and understanding of your domain and scientific innovation. - Own strategic problem solving, and take the lead on the design, implementation, and delivery for solutions that have a long-term quantifiable impact. - Provide cross-organizational technical influence, increasing productivity and effectiveness by sharing your deep knowledge and experience. - Develop strategic plans to identify fundamentally new solutions for business problems. - Assist in the career development of others, actively mentoring individuals and the community on advanced technical issues. A day in the life This is a unique and rare opportunity to get in early on a fast-growing segment of AWS and help shape the technology, product and the business. You will have a chance to utilize your deep technical experience within a fast moving, start-up environment and make a large business and customer impact. About the team Diverse Experiences Amazon Automated Reasoning values diverse experiences. Even if you do not meet all of the qualifications and skills listed in the job description, we encourage candidates to apply. If your career is just starting, hasn't followed a traditional path, or includes alternative experiences, don't let it stop you from applying. Why Amazon Automated Reasoning? At Amazon, automated reasoning is central to maintaining customer trust and delivering delightful customer experiences. Our organization is responsible for creating and maintaining a high bar for automated reasoning across all of Amazon's products and services. We offer talented automated reasoning professionals the chance to accelerate their careers with opportunities to build experience in a wide variety of areas including cloud, devices, retail, entertainment, healthcare, operations, and physical stores. Inclusive Team Culture In Amazon Automated Reasoning, it's in our nature to learn and be curious. Ongoing DEI events and learning experiences inspire us to continue learning and to embrace our uniqueness. Addressing the toughest automated reasoning challenges requires that we seek out and celebrate a diversity of ideas, perspectives, and voices. Training & Career Growth We're continuously raising our performance bar as we strive to become Earth's Best Employer. That's why you'll find endless knowledge-sharing, training, and other career-advancing resources here to help you develop into a better-rounded professional. Work/Life Balance We value work-life harmony. Achieving success at work should never come at the expense of sacrifices at home, which is why flexible work hours and arrangements are part of our culture. When we feel supported in the workplace and at home, there's nothing we can't achieve.
US, CA, Sunnyvale
We are seeking an Applied Scientist to focus on Robotics Localization, Mapping, and Geometric Reconstruction. In this role, you'll research and build advanced SLAM and World Reconstruction algorithms that enable robots to know its pose in the world and fulfill various tasks in complex and dynamic home environments. You'll work across a broad spectrum of visual inertial odometry, 3D pose estimation, mapping, and geometric reconstruction problems to build robust solutions that support autonomous decision making, task planning, navigation, and manipulation. Key job responsibilities - Develop and implement robust robot state estimation, pose estimation, mapping, and reconstruction algorithms. - Build simulation-based and on-robot evaluation frameworks with comprehensive benchmarks and metrics for systematic evaluation of our spatial intelligence stack. - Lead scientific and technical projects across the spatial intelligence team. - Collaborate with navigation, manipulation, and other teams to ensure seamless integration of localization, tracking, and mapping capabilities. - Mentor junior scientists and engineers to solve real-world robotics problems. About the team Fauna Robotics, an Amazon company, is building capable, safe, and genuinely delightful robots for everyday life. Our goal is simple: make robots people actually want to live and interact with in everyday human spaces. We believe that future won’t arrive until building for robotics becomes far more accessible. Today, too much effort is spent reinventing the fundamentals. We’re changing that by developing tightly integrated hardware and software systems that make it faster, safer, and more intuitive to create real-world robotic products. Our work spans the full stack: mechanical design, control systems, dynamic modeling, and intelligent software. The focus is not just functionality, but experience. We’re building robots that feel responsive, expressive, and genuinely useful. At Fauna, you’ll work at the frontier of this space, helping define how robots move, manipulate, and interact with people in natural environments. It’s an opportunity to solve hard problems across hardware and software with a team focused on making robotics accessible and joyful to build. If you care about making robotics real for everyone and building systems that are as delightful as they are capable, we’re interested in hearing from you.