SimBot Challenge FAQs

Frequently asked questions about the challenge.
General
What is the Alexa Prize?
Alexa is Amazon’s cloud-based voice service available on over 100 million devices from Amazon and third-party device manufacturers. With Alexa, you can build natural voice experiences that offer customers a more intuitive way to interact with the technology they use every day. Our collection of tools, APIs, reference solutions, and documentation makes it easy for anyone to build with Alexa.
Why did Amazon create the Alexa Prize?
The Alexa Prize, an annual university competition dedicated to accelerating the field of conversational artificial intelligence (AI), was created to recognize students from around the globe who are changing the way we interact with technology. The goal is to advance several areas of conversational AI including natural language understanding (NLU), context modeling, dialog management, commonsense reasoning, natural language generation (NLG), and knowledge acquisition.
How does the Alexa Prize support research?
The Alexa Prize is a research testbed for university students to experiment with and advance conversational AI at scale.

Research teams own the intellectual property (IP) in their systems and are encouraged to publish scientific articles on their work. As described in the Official Rules, participating teams grant Amazon a non-exclusive license to any technology or software they develop in connection with the competition.
What new datasets will I have access to as part of the Alexa Prize?
The TEACh dataset as well as a training dataset proprietary to the live interaction portion of the SimBot Challenge will be available to competitors in each phase.
SimBot
What is the SimBot Challenge?
SimBot, a competition focused on helping advance development of next-generation virtual assistants that will assist humans in completing real-world tasks by continuously learning, and gaining the ability to perform commonsense reasoning.

The SimBot Challenge will have two phases: A public benchmark phase, and a live interactions phase. Participants in both phases will build machine-learning models for natural language understanding, human-robot interaction, and robotic task completion. Artificial intelligence challenges addressed in the competition relate to reasoning on language and scene understanding, learning from demonstration, self-learning, and task completion utilizing natural language.

Unlike previous Alexa Prize competitions, the public benchmark challenge phase will be open to university teams, as well as individuals in academia and industry interested in advancing the science of AI and engaging top researchers from around the globe. The SimBot Challenge public benchmark phase is like existing visual language navigation competitions.
What is the difference between the Public Benchmark Challenge and the Live Interaction Challenge?
The Public Benchmark Challenge is open to any individual or team, academic or industry, who wants to complete and submit a model for evaluation and rating during the evaluation period. It is based on the TEACh dataset offered to the research community in October, 2021. The Live Interaction Period’s participation is limited solely to the university-based teams selected for the SimBot Challenge in November 2021 and June, 2022. These teams will each receive Amazon sponsorship to build a SimBot that will compete in a challenge from July, 2022 to September 2022 where they will receive real time ratings and feedback from Alexa Users.
Why is the SimBot - Live Interaction Challenge only limited to university teams?
The SimBot Challenge - live interaction phase is part of the Alexa Prize which is currently limited to universities and university students in Amazon initiatives to advance AI. However, the Public Benchmarking Challenge is open to both academic and university-based teams.
What will my SimBot do?
SimBot will navigate in a virtual environment to complete challenges guided by Alexa users. SimBot will interact with objects in the environment including but not limited to those common to offices and homes. There will be obstacles and hazards introduced to make game play fun and challenging.
How will I build my SimBot?
SimBots will use images from the game and instructions from Alexa users to navigate in a virtual world. Teams will build AI models which recognize objects and scenes as well as understand natural language commands from users. We plan to provide baseline data and a baseline model as a reference but we expect teams to develop new and novel approaches as well as augment the provided data to improve performance.
Eligibility: Public Benchmark Challenge
Phase 1, SimBot challenge
Who is eligible to participate in the public benchmark challenge?
Any individual or team, academic or industry-based.
How do I sign up to participate in the public benchmark challenge?
A registration form will open on November 15, 2021 at alexaprize.com.
Is funding available to university teams competing in the public benchmark challenge?
No, funding is only available to university teams selected to participate in the SimBot Challenge. In addition to the teams selected in 2021 up to four new, high-performing teams from the public benchmark challenge will be invited to apply for the SimBot challenge, live interaction phase for which funding is available. Learn more here.
Who can apply to participate?
Any individual or team can register and participate in the public benchmark phase with exception of parties from Cuba, Iran, North Korea, Sudan, Syria, and the region of Crimea.

Registration for the public benchmark challenge will open November 15, 2021. The challenge will begin on January 10, 2022.
Do I need to be a certain age?
Participants must be at or above the age of majority in the country, state, province or jurisdiction of residence at the time of registration.
Can I enroll if a family member is an Amazon employee?
Immediate family members and household members of Amazon employees, directors and contractors are not eligible to participate.
Eligibility: SimBot Challenge
Includes sponsored teams for both Public Benchmark and Live Interaction phases
Who can apply to participate?
The Alexa Prize is open to full-time students enrolled in an accredited university, with the exception of universities in Cuba, Iran, North Korea, Sudan, Syria, and the region of Crimea (see Official Rules). Proof of enrollment will be required to participate.
Can I participate if I don’t attend a university?
No. The Alexa Prize is open only to full-time enrolled university students.
Do I need to be enrolled in a university program throughout the duration of the competition?
All participating team members must remain full-time students in good standing at their university while participating in the competition.
Do I need to be a certain age?
Participants must be at or above the age of majority in the country, state, province or jurisdiction of residence at the time of entry.
Can I enroll if a family member is an Amazon employee?
Immediate family members and household members of Amazon employees, directors and contractors are not eligible to participate. See Official Rules for additional restrictions.
If my team fails to apply for the SimBot Challenge now, will there be any future opportunities to compete in the Live Interaction challenge?
Yes, the initial application period run from October 4-31, 2021. In addition to the teams selected in 2021 up to four new, high-performing university teams from the public benchmark challenge will be invited to apply for the SimBot challenge, live interaction phase for which funding is available. Learn more here.
Will university teams selected post-Public Benchmark Challenge be eligible for funding and for what amount?
Teams selected for the SimBot Challenge in June, 2022 will be eligible for funding.
If my team was not selected during the first application period are we eligible to re-apply in may if we qualify?
Yes. University-based teams that participate in the public benchmark phase are eligible to re-apply for the SimBot challenge between May 9 - 23, 2022. Four new, high-performing university teams from the public benchmark challenge will be invited to apply for the SimBot challenge, live interaction phase for which funding is available. Learn more here.
My team is not completely comprised of university students but we perform well in the public benchmark challenge? Is there an opportunity for us to continue in the competition?
Yes, we plan to continue the public benchmark competition through the end of 2022.
My company does collaborative work with several universities, are we eligible to compete in the live interaction challenge alongside them?
No, only full-time students are eligible to participate in the SimBot challenge, which includes the live interaction period.
My university and another frequently collaborate on projects - is there any exception to the students all enrolled in one university rule?
Applications are limited to students all from the same, single, university.
Teams
How many teams will be selected to participate?
All applications will be reviewed and evaluated by a panel of Amazon experts. Up to ten teams will be selected and sponsored by Amazon. All teams selected in November, 2021 will receive a $250,000 grant while those selected in June, 2022 will receive a $200,000 grant intended to support two full-time students, a month of faculty time, free Alexa devices, and free AWS hosting including access to CPU and GPU based machines, SQL and NoSQL databases, and object storage. See the Official Rules for details.
How many team members can our team have?
There are no minimum or maximum team member requirements. All team members must be enrolled in their university throughout the duration of the competition. All teams will receive a $250,000 grant if selected in November 2021 or a $200,000 grant if selected in June, 2022 regardless of how many members are on the team. We recommend a team with 4-6 students with diverse fields of study or areas of expertise.
Can students from different universities be on the same team?
Teams must be comprised of students attending the same university.
Can one university have more than one team?
Yes, universities may have more than one team.
Can I participate on two separate teams?
You can only be a part of one team for the duration of the competition.
Can undergraduate and graduate students work together?
Yes, teams may be comprised of undergraduate and graduate students.
Do I need a faculty advisor?
All teams must nominate a faculty advisor and include the faculty advisor’s consent in the applications.
What is the role of the faculty advisor?
Faculty advisors will advise students on technical directions and be a sounding board for new ideas, similar to a graduate school advisor. They will also act as the official representative from the university for this competition.
Can we add or remove team members during the competition?
During the competition, faculty advisors may request to remove or add members to the team, subject to approval by Amazon.
Can we discuss our SimBot with faculty or students who aren’t on our team?
Only team members may work on their SimBot. However, the faculty advisor and other students and faculty members at your university may provide support and advice to your team and may co-author technical publications and research papers.
Application process
How do we apply?
Check the SimBot page for the latest update on applications.
What do we need to apply?
Once you have selected your team members, team leader, and faculty sponsor, you are ready to begin the application process.
Do all team members have to apply?
Each team must have a team lead, who should apply on behalf of the whole team. Your application must include all of your team members’ information.
Is there an application fee?
There is no application fee.
How will teams be selected to participate?
All applications will be reviewed by a panel of Amazon employees. Teams will be selected based on the following criteria: (1) the potential scientific contribution to the field; (2) the technical merit of the approach; (3) the novelty of the idea; and (4) an assessment of the team’s ability to execute against their plan. Please be sure to provide enough detail in your application to enable our experts to evaluate your proposal.
Competition details
What is the goal of the challenge?
To develop AI models which advance the state of the art and allow users to naturally interact with a robotic assistant in a virtual world to successfully complete a range of challenges.
How will winners be selected?
Winners will be determined based on the final standings at the completion of the finals period.
Can we use other funding to help us participate in this challenge?
Yes, you may use other funding to support your team, subject to the terms described in the Official Rules. External funding will need to be disclosed by January 1, 2023.
Will Alexa customers be able to engage with our SimBot?
Your team will be required to submit its SimBot for certification and publication by the Amazon Alexa team. After certification, you will enter the Internal Amazon Beta Period, where Amazon employees will test your SimBot and provide feedback. After the Internal Amazon Beta Period, we will allow Amazon Alexa customers to try your SimBot and provide feedback to you. Amazon may impose Availability Criteria, or requirements the SimBot must meet before it will be made available to Alexa users. Availability Criteria may include criteria such as a minimum average customer rating, uptime requirements, and an ability to consistently filter offensive content.
Amazon launched the Echo in the UK, Germany, India, Japan, and other countries. Will localized languages be supported?
Your team must build its SimBot using U.S. English. Your SimBot will be available to Alexa customers in the U.S. Customers in other countries may also access it by setting their Amazon PFM (Preferred Marketplace) to U.S.
Will we publish our research from the Alexa Prize?
Yes. Publishing research papers as an outcome of your work on the Alexa Prize is required for all teams participating in the competition, although teams should not publish Amazon confidential information, as described in the Official Rules. The Alexa Prize requires all teams to submit a technical paper for the Alexa Prize proceedings. Your SimBot will not be selected for the finals if your team does not submit a technical paper for Alexa Prize proceedings. Papers will be published at the end of the competition in an online Proceedings of the Alexa Prize, which will be publicly available.

Teams may also publish research papers in third-party publications and conferences, as long as all papers are provided to Amazon for review at least two weeks before the submission deadlines and no research papers are published before the Alexa Prize proceedings are published without Amazon’s prior approval.
Who will own the intellectual property rights in my submission?
You will retain ownership over your SimBot. Amazon will have a non-exclusive license to any technology or software you develop in connection with the competition. See the Official Rules for details.
Prizes
What are the prizes for winning the competition?
A prize of $500,000 will be awarded to the team that creates the best SimBot. The second-place and third-place finalist teams will receive a $100,000 and a $50,000 prize, respectively. See the contest rules for details.
Do we get a stipend and devices to participate in the Alexa Prize?
Up to ten teams will be sponsored to participate in the Alexa Prize in November, 2021. These teams’ universities will receive a $250,000 research grant to fund the team members’ work over the year. Additional teams selected to participate in the SimBot Live Interaction challenge in June, 2022 will receive a $200,000 research grant to fund the team members’ work over the year.

The sponsorship includes one Alexa-enabled device per team member for up to a total of three devices per team and one Alexa-enabled device per faculty advisor, free AWS services to support the development of their SimBot, and support from the Alexa team.
How can the grant be spent?
The grants will be awarded with the intention that they will support two full-time students for the duration of the Competition and one month of the Faculty Advisor’s salary. No more than 35% of the research grants may be allocated to administrative fees. If your team would like to use the funds in another manner, your faculty advisor must receive approval from Amazon before doing so.
What happens if we are selected and receive a stipend but can no longer participate?
Stipends will be awarded in installments payable to the university. If your team withdraws before any of the installments, remaining funds will not be transferred to the university.
How will the prizes be distributed among a team?
The first, second, and third place prizes will be distributed equally among all registered team members. The official list of registered team members must be confirmed by January 30, 2023.
Timeline
What are the key milestones of the competition?
Teams must submit their applications between October 1, 2021 and October 31, 2021. Between November 1 and November 10, 2021, we will announce teams selected to participate. In the summer of 2022 following the public benchmark phase and the second application period teams will be invited to an Alexa Prize Bootcamp at Amazon where they will receive training on the resources made available to all competing teams. The finals will be scheduled between January 30 and March 27, 2023 and will determine the winning teams in the first annual SimBot Challenge.
If selected, when will we receive the stipend, devices, access to the Alexa Prize SimBot toolkit, our AWS account, and be introduced to our point of contact?
We will reach out to all teams no later than November 10, 2021 with instructions on next steps. Up to ten teams will be selected to receive a $250,000 stipend, Alexa-enabled devices, free AWS services to support their development efforts, and support from the Alexa team.
More information
See the full competition rules or submit your questions. Need assistance? Email: alexaprizesupport@amazon.com

Latest news

The latest updates, stories, and more about Alexa Prize.
US, WA, Redmond
At Amazon, we’re inventing on behalf of customers, and with Amazon Leo, we’re redefining what global connectivity looks like. Our mission is to deliver fast, affordable broadband to unserved and underserved communities around the world through a constellation of low Earth orbit (LEO) satellites. Every system we build helps connect people to education, healthcare, opportunity, and each other. As a Data Scientist, you will be responsible for developing advanced analytics and machine learning solutions for user terminals. You will develop predictive models to proactively identify possible user terminal failures in the field. You will work in a collaborative environment with a multi-disciplinary team, including constellation, RF, antenna, silicon, algorithm, and software engineers. Key job responsibilities As a Data Scientist, you will develop analytic tools for a team developing current and future user terminals. Your responsibilities include: • Develop statistical and analytical tool to enable the regression decision from on-orbit and lab measurement of user terminals • Publish documents and create compelling visualizations and presentations to communicate insights to stakeholders • Create and manage datasets for continued pre-training and supervised fine-tuning of LLMs • Develop scalable visualizations for analysis of user terminal performance • Work closely with constellation, RF, antenna, silicon, algorithm, and software engineers to root-cause the failures using data as the primary tool • Drive consensus on metrics and analysis approaches to support product development strategy Export Control Requirement: Due to applicable export control laws and regulations, candidates must be a U.S. citizen or national, U.S. permanent resident (i.e., current Green Card holder), or lawfully admitted into the U.S. as a refugee or granted asylum. A day in the life As a Data Scientist in the LEO Customer Terminal Team, you will work daily with satellite constellation, algorithm, RF, antenna, silicon, hardware, and software teams in a collaborative environment. Your focus will be using data as an intelligence source to enable design decisions for the team. About the team The LEO Customer Terminal team is responsible for developing both outdoor and indoor devices that enable customers to access internet service via the LEO satellite network. We own the entire process from early prototypes through mass production, including requirements documentation, architecture definition, hardware development, algorithm development, and all integration and verification testing.
US, NY, New York
We are seeking an Applied Scientist to lead the development of evaluation frameworks and data collection protocols for robotic capabilities. In this role, you will focus on designing how we measure, stress-test, and improve robot behavior across a wide range of real-world tasks. Your work will play a critical role in shaping how policies are validated and how high-quality datasets are generated to accelerate system performance. You will operate at the intersection of robotics, machine learning, and human-in-the-loop systems, building the infrastructure and methodologies that connect teleoperation, evaluation, and learning. This includes developing evaluation policies, defining task structures, and contributing to operator-facing interfaces that enable scalable and reliable data collection. The ideal candidate is highly experimental, systems-oriented, and comfortable working across software, robotics, and data pipelines, with a strong focus on turning ambiguous capability goals into measurable and actionable evaluation systems. Key job responsibilities - Design and implement evaluation frameworks to measure robot capabilities across structured tasks, edge cases, and real-world scenarios - Develop task definitions, success criteria, and benchmarking methodologies that enable consistent and reproducible evaluation of policies - Create and refine data collection protocols that generate high-quality, task-relevant datasets aligned with model development needs - Build and iterate on teleoperation workflows and operator interfaces to support efficient, reliable, and scalable data collection - Analyze evaluation results and collected data to identify performance gaps, failure modes, and opportunities for targeted data collection - Collaborate with engineering teams to integrate evaluation tooling, logging systems, and data pipelines into the broader robotics stack - Stay current with advances in robotics, evaluation methodologies, and human-in-the-loop learning to continuously improve internal approaches - Lead technical projects from conception through production deployment - Mentor junior scientists and engineers
US, WA, Redmond
Amazon Leo is building a constellation of thousands of low Earth orbit satellites to deliver fast, reliable internet beyond the reach of existing networks. As a Guidance, Navigation & Control engineer, you will design the algorithms, simulations, and onboard autonomy that keep every satellite precisely pointed and stable so the payload can serve customers below, and keep the fleet flying safely. This is satellite GNC with demanding pointing and performance requirements, and you will see your work go from concept to orbit on a constellation that is launching and growing today. Amazon Leo will delight customers with reliable connectivity regardless of where they are. In this role, you will: - Take your GNC algorithms from concept all the way to orbit, and stay close to them through integration, test, and on-orbit operations - Solve hard precise-pointing, payload-cueing, and fleet-autonomy problems against demanding performance requirements - Work across the full lifecycle — design, simulation, test, and flight operations — rather than a narrow slice - Join a small team of some of the top GNC engineers in the industry, building a product people will love to use This position is part of the Satellite Attitude Determination and Control team, where we own the full lifecycle of our work — from algorithm design and flight software, through lab integration, to flying our own satellites on orbit. You will design and analyze the control system and algorithms, support development of our flight hardware and software, help integrate the satellite in our labs, and operate the GNC system in flight as a constellation of satellites flows through the production line into orbit. **Export Control Requirement:** Due to applicable export control laws and regulations, candidates must be a U.S. citizen or national, U.S. permanent resident (i.e., current Green Card holder), or lawfully admitted into the U.S. as a refugee or granted asylum. Key job responsibilities - Design and analyze algorithms for state estimation, attitude determination and control, precise pointing, and payload cueing against demanding performance requirements. - Write the production C++ flight code that runs these algorithms on the satellite. Our GNC engineers own their algorithms all the way into flight software, working alongside software development engineers rather than handing prototypes off to be reimplemented by another team. - Design concepts of operation that balance pointing and payload cueing, momentum management, power, thermal, and communications across all phases of the mission, including off-nominal and contingency scenarios. - Develop momentum management strategies using reaction wheels and magnetic torquers, and help size and lay out the GNC sensor and actuator suite — star trackers, magnetometers, sun sensors, IMUs, and GPS receivers — and the algorithms that consume them. - Develop onboard autonomy, including Fault Detection, Isolation, and Recovery, so the fleet operates safely at scale with minimal intervention. - Develop and validate models and simulations of the satellite and constellation across a range of fidelity levels, including 6-DOF time-domain simulation and Monte Carlo analysis, and run hardware-in-the-loop testing to verify closed-loop performance. - Support component-level environmental testing, functional and performance checkout, and subsystem and satellite integration in our labs. - Analyze fleet telemetry in bulk and build the metrics, dashboards, and tooling that let us monitor and continuously improve pointing performance and payload operations time as the constellation grows. - Operate the GNC system on orbit — reviewing fleet data, supporting maneuvers, and resolving anomalies when on-orbit behavior diverges from the ground model. - Contribute to architecture, design, and code reviews, and document your algorithms, analyses, and results clearly. A day in the life This role spans a diverse set of problems across the spacecraft and network system. You will derive and prototype estimation and control algorithms, turn them into production flight code, and prove them out through 6-DOF and Monte Carlo simulation and hardware-in-the-loop testing in the lab. You will support satellite integration, and when something on orbit does not behave the way the ground model predicted, you will dig into fleet telemetry to find out why and drive the fix. You will own problems end to end and have the focus to go deep on each one. Because GNC sits at the center of safe flight, you will work across the whole program. The payload, power, thermal, and operations teams all depend on GNC, so you will balance their needs against your pointing and momentum budgets. To get your algorithms running reliably on the spacecraft, you will partner closely with avionics, flight software, and our embedded software development engineers. You will help the mechanical teams shape the next generation of satellite hardware, from reaction wheels to star trackers. And as satellites move from design through manufacturing into orbit, you will work alongside systems engineering and Assembly, Integration & Test. The problems are genuinely hard — demanding pointing and performance requirements, a large constellation to keep flying safely, and the chance to invent advanced capabilities that have not been built before. The work is staffed with some of the strongest engineers in the industry, and the constant collaboration across teams makes it a place to keep learning and growing. About the team Our team brings deep experience across many satellite systems and other flight vehicles, with real bench strength in both our mission and the core GNC disciplines. We design, prototype, test, iterate, and learn together, and because GNC is central to safe flight, we tend to drive Concepts of Operation and many of the system-level analyses on the program. We also have a lot of freedom. We trust our engineers to pursue the problems they face in whatever way they find best, to choose their own tools and approaches, and to own the results. If you want the autonomy to solve hard problems your way, and a team of strong engineers to do it with, you will feel at home here.
US, NY, New York
Amazon Advertising is one of Amazon's fastest growing and most profitable businesses, responsible for defining and delivering a collection of advertising products that drive discovery and sales. Our products are strategically important to our businesses driving long term growth. We deliver billions of ad impressions and millions of clicks and break fresh ground in product and technical innovations every day! Advertiser Growth Engine (AGE) team owns and builds services and applications across Amazon World-Wide Advertising that make advertising across multi-marketplaces as easy as flipping on a switch. We are focused on: (1) expanding Amazon Ads advertiser base, and (2) eliminating localization, operational, and marketplace knowledge gap barrier for Advertisers who advertise across multiple marketplaces. Our products and solutions are strategically important to enable our Retail and Marketplace businesses to drive long-term growth globally. We're looking for an experienced Applied Scientist with exceptional technical, analytical, and innovative capabilities to research, design, and create elegant machine learning solutions. The solutions will help our advertisers with multi-media and multi-lingual advertising offerings. You will use ideas from various domains of machine learning, including supervised and unsupervised methods, Deep Neural Networks, Natural Language Processing (NLP), and Computer Vision (CV) to build ML models that localizes multi-media advertising contents, including text, images and videos. You will also identify opportunities to leverage ML beyond localization, including, international expansion and global campaigns. Your work will directly impact our customers in the form of products and services used directly by our advertisers as well as our third-party integrators. As an Applied Scientist on this team, you will: - Build and deliver end-to-end machine learning solutions; build ML models and perform data analysis to deliver scalable solutions to business problems. - Perform hands-on analysis and modeling with enormous data sets to develop insights that increase traffic monetization and merchandise sales without compromising shopper experience. - Work closely with software engineers on detailed requirements to productionize the ML models you build. - Run A/B experiments that affect hundreds of millions of customers, evaluate the impact of your optimizations and communicate your results to various business stakeholders. - Establish scalable, efficient, automated processes for large-scale data analysis, machine-learning model development, model validation and serving. - Research new innovate machine learning approaches. Why you will love this opportunity: Amazon is investing heavily in building a world-class advertising business. This team defines and delivers a collection of advertising products that drive discovery and sales. Our solutions generate billions in revenue and drive long-term growth for Amazon’s Retail and Marketplace businesses. We deliver billions of ad impressions, millions of clicks daily, and break fresh ground to create world-class products. We are a highly motivated, collaborative, and fun-loving team with an entrepreneurial spirit - with a broad mandate to experiment and innovate. Impact and Career Growth: You will invent new experiences and influence customer-facing shopping experiences to help suppliers grow their retail business and the auction dynamics that leverage native advertising; this is your opportunity to work within the fastest-growing businesses across all of Amazon! Define a long-term science vision for our advertising business, driven from our customers' needs, translating that direction into specific plans for research and applied scientists, as well as engineering and product teams. This role combines science leadership, organizational ability, technical strength, product focus, and business understanding. Team video https://youtu.be/zD_6Lzw8raE
US, NY, New York
We are seeking an Applied Scientist to lead the development of evaluation frameworks and data collection protocols for robotic capabilities. In this role, you will focus on designing how we measure, stress-test, and improve robot behavior across a wide range of real-world tasks. Your work will play a critical role in shaping how policies are validated and how high-quality datasets are generated to accelerate system performance. You will operate at the intersection of robotics, machine learning, and human-in-the-loop systems, building the infrastructure and methodologies that connect teleoperation, evaluation, and learning. This includes developing evaluation policies, defining task structures, and contributing to operator-facing interfaces that enable scalable and reliable data collection. The ideal candidate is highly experimental, systems-oriented, and comfortable working across software, robotics, and data pipelines, with a strong focus on turning ambiguous capability goals into measurable and actionable evaluation systems. Key job responsibilities - Design and implement evaluation frameworks to measure robot capabilities across structured tasks, edge cases, and real-world scenarios - Develop task definitions, success criteria, and benchmarking methodologies that enable consistent and reproducible evaluation of policies - Create and refine data collection protocols that generate high-quality, task-relevant datasets aligned with model development needs - Build and iterate on teleoperation workflows and operator interfaces to support efficient, reliable, and scalable data collection - Analyze evaluation results and collected data to identify performance gaps, failure modes, and opportunities for targeted data collection - Collaborate with engineering teams to integrate evaluation tooling, logging systems, and data pipelines into the broader robotics stack - Stay current with advances in robotics, evaluation methodologies, and human-in-the-loop learning to continuously improve internal approaches - Lead technical projects from conception through production deployment - Mentor junior scientists and engineers
IL, Tel Aviv
Are you a scientist interested in pushing the state of the art in Information Retrieval, Large Language Models and Recommendation Systems? Are you interested in innovating on behalf of millions of customers, helping them accomplish their every day goals? Do you wish you had access to large datasets and tremendous computational resources? Do you want to join a team of capable scientist and engineers, building the future of e-commerce? Answer yes to any of these questions, and you will be a great fit for our team at Amazon. Our team is part of Amazon’s Personalization organization, a high-performing group that leverages Amazon’s expertise in machine learning, generative AI, large-scale data systems, and user experience design to deliver the best shopping experiences for our customers. Our team is building next-generation personalization systems powered by Large Language Models. We are tackling novel research challenges to help customers discover products they'll love - at Amazon scale and latency requirements. We are a team uniquely placed within Amazon, to have a direct window of opportunity to influence how customers will think about their shopping journey in the future. As an Applied Science Manager, you will lead a team of scientists working at the frontier of LLM-based personalization. You will set the technical vision, drive the research agenda, and ensure your team delivers production-ready solutions. You will hire, mentor, and develop world-class scientists while fostering a culture of innovation and scientific rigor. You will partner closely with engineering and product teams to translate ambitious research into customer-facing impact, and represent your team's work to senior leadership. Please visit https://www.amazon.science for more information.
US, WA, Seattle
AWS Applied AI Solutions (AAIS) is building toward a future where every business innovates with Amazon AI teammates. To get there, we build AI solutions that improve human capabilities and transform entire business functions. We create end-to-end products that surprise and delight out-of-the-box, making complex things easy and hard things possible, with no cloud experience required. We start with customers who embrace the future and build bridges to meet the rest where they are. We pursue ambitious opportunities with conviction, and we are looking for builders who share that mindset. The Team Join the next science revolution at AWS Life Sciences Applied AI Solutions where you'll work alongside world-class scientists to build AI that transforms how therapeutics are discovered, developed, and brought to patients. We're out to revolutionize how medicines are discovered, developed, and brought to patients powered by a new generation of AI. Our team tackles some of the hardest open problems at the intersection of frontier AI and life sciences. We apply biological foundation models large language models and agentic reasoning systems to life sciences problems then put them into the hands of customers as applications and managed services they can fine-tune tailor and deploy on their own data. The science challenges are deep: how do you design agentic systems that reason correctly over complex biological regulatory and clinical logic? How do you enable customers to tailor foundation models to their proprietary data and get better outputs with less effort? How do you adapt models to reason faithfully in high-stakes scientific and regulatory domains? Today we're focused on two areas. In clinical trials we're building AI that automates and optimizes regulatory and clinical development workflows. In drug design our products (including Amazon Bio Discovery) accelerate discovery by giving bench scientists AI-guided protein engineering and antibody design capabilities. We combine frontier research with production-scale delivery to put breakthrough science into the hands of customers solving humanity's hardest problems. We value scientific rigor encourage publication and support conference participation. If you want to do research that ships this is the team. The Role We are seeking an Applied Scientist to build the models and methods behind our life sciences AI products with a primary focus on clinical trial operations and agentic reasoning. You will design train and evaluate systems that reason over complex clinical and operational logic and ship them into products customers use directly. You will work closely with senior and principal scientists on well-scoped research problems own your results end to end and see your work reach production. This role combines expertise in LLM reasoning and agentic AI with applied impact in life sciences. You will work on how large language models reason plan and act in complex scientific domains while applying domain knowledge to ensure models produce scientifically valid outputs. The problems span multiple fronts: • How do you build LLM-based agentic systems that correctly reason over clinical protocols regulatory standards and complex multi-step operational workflows? • How do you evaluate agent reliability and faithfulness rigorously enough to trust in high-stakes clinical settings? • How do you develop model customization methods (fine-tuning retrieval augmentation domain adaptation) that let customers get strong results from foundation models on their own data? You will focus on clinical trial operations (agentic automation structured reasoning evaluation domain adaptation) with opportunities to contribute across drug discovery (protein engineering antibody design) as the portfolio grows. You will own end-to-end scientific solutions from research through production and your work will directly shape the tools that scientists use daily. Key job responsibilities • Design train fine-tune and evaluate LLM-based agentic systems that reason over clinical protocols regulatory standards and operational workflows • Build rigorous evaluation harnesses and benchmarks to measure agent reliability faithfulness and failure modes in high-stakes domains • Develop model customization methods (fine-tuning RLHF retrieval augmentation domain adaptation) that help customers get better outputs on their own data with less effort • Contribute to graph-based and causal modeling approaches for clinical trial operations • Partner with Life Sciences domain experts product and engineering to translate scientific challenges into shipped capabilities • Own experiments end to end: problem framing implementation evaluation iteration and hand-off to production • Publish at top-tier venues where the work supports it • Contribute to drug discovery efforts (protein engineering antibody design) as opportunities arise A day in the life • Design and run an experiment to validate a new agentic reasoning or fine-tuning method then ship it as a capability customers can use • Diagnose why a model is failing on a new class of inputs and implement a fix to unblock a delivery milestone • Build or extend an evaluation benchmark to measure how faithfully an agent reasons over clinical logic • Meet with domain experts to scope what the next model release needs to do • Review results with a senior scientist sharpen the approach and get it over the finish line • Prototype a new idea that could become the next capability in the product About the team Amazon values diverse experiences. Even if you do not meet all of the preferred qualifications and skills listed in the job description, we encourage candidates to apply. If your career is just starting, hasn't followed a traditional path, or includes alternative experiences, don't let it stop you from applying. Amazon Web Services (AWS) is the world's most comprehensive and broadly adopted cloud platform. We pioneered cloud computing and never stopped innovating — that's why customers from the most successful startups to Global 500 companies trust our robust suite of products and services to power their businesses. We value work-life harmony. Achieving success at work should never come at the expense of sacrifices at home, which is why flexible work hours and arrangements are part of our culture. When we feel supported in the workplace and at home, there's nothing we can't achieve in the cloud. Here at AWS, it's in our nature to learn and be curious. Our employee-led affinity groups foster a culture of inclusion that empower us to be proud of our differences. Ongoing events and learning experiences, including our Conversations on Race and Ethnicity and AmazeCon conferences, inspire us to never stop embracing our uniqueness. We're continuously raising our performance bar as we strive to become Earth's Best Employer. That's why you'll find endless knowledge-sharing, mentorship and other career-advancing resources here to help you develop into a better-rounded professional.
BR, SP, Sao Paulo
Do you feel the challenge and the adrenaline kick when a huge data-set stares you in the face and you know that somewhere inside are hidden very important business insights that can fundamentally alter the way top business leaders think and act? Do you enjoy presenting strong data backed insights to business leaders; insights that can topple their long held beliefs and compel them to change their direction completely? If yes, then you are the one we are looking for. We are looking to invite passionate leaders, with expertise in generate power business insights from very large datasets, on a journey where the primary aim would be to enable needle moving business impacts through statistical analysis. We are looking for leaders who can envision the design and development of analytical infrastructure which can support strategic and tactical decision-making. Those who join this high visibility team would have to navigate through significant ambiguity in defining business problems and converting them to analytical problems. This role requires additional exposure and experience to Machine Learning. Key job responsibilities Use machine learning and analytical techniques to create scalable solutions for business problems • Analyze and extract relevant information from large amounts of Amazon’s historical business data to help automate and optimize key processes • Design, development, evaluate and deploy innovative and highly scalable ml models such as risk scorecards, income models, fraud models for predictive learning in credit risk applications • Research and implement novel machine learning and statistical approaches • Work closely with software engineering teams to drive real-time model implementations and new feature creations • Work closely with business owners and operations staff to optimize various business operations • Establish scalable, efficient, automated processes for large scale data analyses, model development, model validation and model implementation • Mentor other scientists and engineers in the use of ML techniques • Innovate with the latest GenAI technology to build highly automated solutions for efficient customer promotions • Design, develop and deploy end-to-end machine learning solutions in the Amazon production environment to delight Amazon customers • Collaborate with cross-functional teams to develop comprehensive ML/statistical models that can scale to millions of customers to multiple countries Understand the credit risk data and evaluate the best ml model/ solution for dynamic business problems. About the team Brazil Payments is part of the International Emerging Stores Payments team and focuses on supporting the launch of new payment and financial products to our customers in Brazil.
BR, SP, Sao Paulo
Do you feel the challenge and the adrenaline kick when a huge data-set stares you in the face and you know that somewhere inside are hidden very important business insights that can fundamentally alter the way top business leaders think and act? Do you enjoy presenting strong data backed insights to business leaders; insights that can topple their long held beliefs and compel them to change their direction completely? If yes, then you are the one we are looking for. We are looking to invite passionate leaders, with expertise in generate power business insights from very large datasets, on a journey where the primary aim would be to enable needle moving business impacts through statistical analysis. We are looking for leaders who can envision the design and development of analytical infrastructure which can support strategic and tactical decision-making. Those who join this high visibility team would have to navigate through significant ambiguity in defining business problems and converting them to analytical problems. This role requires additional exposure and experience to Machine Learning. Key job responsibilities • Use machine learning and analytical techniques to create scalable solutions for business problems • Analyze and extract relevant information from large amounts of Amazon’s historical business data to help automate and optimize key processes • Design, development, evaluate and deploy innovative and highly scalable models for predictive learning • Research and implement novel machine learning and statistical approaches • Work closely with software engineering teams to drive real-time model implementations and new feature creations • Work closely with business owners and operations staff to optimize various business operations • Establish scalable, efficient, automated processes for large scale data analyses, model development, model validation and model implementation • Mentor other scientists and engineers in the use of ML techniques • Innovate with the latest GenAI technology to build highly automated solutions for efficient customer promotions • Design, develop and deploy end-to-end machine learning solutions in the Amazon production environment to delight Amazon customers • Collaborate with cross-functional teams to develop comprehensive ML/statistical models that can scale to millions of customers to multiple countries About the team Brazil Payments is part of the International Emerging Stores Payments team and focuses on supporting the launch of new payment and financial products to our customers in Brazil.
BR, SP, Sao Paulo
Do you feel the challenge and the adrenaline kick when a huge data-set stares you in the face and you know that somewhere inside are hidden very important business insights that can fundamentally alter the way top business leaders think and act? Do you enjoy presenting strong data backed insights to business leaders; insights that can topple their long held beliefs and compel them to change their direction completely? If yes, then you are the one we are looking for. We are looking to invite passionate leaders, with expertise in generate power business insights from very large datasets, on a journey where the primary aim would be to enable needle moving business impacts through statistical analysis. We are looking for leaders who can envision the design and development of analytical infrastructure which can support strategic and tactical decision-making. Those who join this high visibility team would have to navigate through significant ambiguity in defining business problems and converting them to analytical problems. This role requires additional exposure and experience to Machine Learning. Key job responsibilities Use machine learning and analytical techniques to create scalable solutions for business problems • Analyze and extract relevant information from large amounts of Amazon’s historical business data to help automate and optimize key processes • Design, development, evaluate and deploy innovative and highly scalable ml models such as risk scorecards, income models, fraud models for predictive learning in credit risk applications • Research and implement novel machine learning and statistical approaches • Work closely with software engineering teams to drive real-time model implementations and new feature creations • Work closely with business owners and operations staff to optimize various business operations • Establish scalable, efficient, automated processes for large scale data analyses, model development, model validation and model implementation • Mentor other scientists and engineers in the use of ML techniques • Innovate with the latest GenAI technology to build highly automated solutions for efficient customer promotions • Design, develop and deploy end-to-end machine learning solutions in the Amazon production environment to delight Amazon customers • Collaborate with cross-functional teams to develop comprehensive ML/statistical models that can scale to millions of customers to multiple countries Understand the credit risk data and evaluate the best ml model/ solution for dynamic business problems. About the team Brazil Payments is part of the International Emerging Stores Payments team and focuses on supporting the launch of new payment and financial products to our customers in Brazil.