Ten university teams selected for Alexa Prize TaskBot Challenge 2

Second iteration features five new teams.

Amazon today announced that ten teams from around the globe have been selected to participate in the Alexa Prize TaskBot Challenge year 2, a university challenge focused on developing multimodal (voice and vision) conversational agents that assist customers in completing tasks requiring multiple steps and decisions.

Alexa Prize is a flagship industry-academic collaboration dedicated to accelerating the science of conversational artificial intelligence (AI) and multimodal human-AI interactions.

“Prize competitions provide an agile science experimentation framework for researchers and students encouraging them to explore transformational ideas at the boundaries of what is achievable,” said Reza Ghanadan, senior principal scientist with Alexa AI and head of Alexa Prize. “We have developed the CoBot platform and tools to lower the barriers to AI innovation for both the academic research community and students interested in conversational AI assistants. These tools allow students to quickly deploy their solutions at scale in the real world with Alexa, then observe, evaluate, and enhance their research results using feedback from Alexa customers.”

Photo of Participants in the Alexa Prize TaskBot Challenge Bootcamp
The Alexa Prize TaskBot Bootcamp was held in Seattle, Washington, with representatives from all ten university teams.

The teams selected for the challenge, which began in January, feature five returning entrants — including the top three finishers in the most recent challenge — and five new universities.

Team

University

Faculty advisor

Returning

TWIZ

NOVA School of Science and Technology

João Magalhães

EvoquerBOT

Penn State University

Rui Zhang

Taco 2.0

The Ohio State University

Huan Sun

GRILL

University of Glasgow

Jeff Dalton

Maruna

University of Massachusetts Amherst

Hamed Zamani

New

BoilerBot

Purdue University

Julia Rayz

DiWBot

Rutgers University

Matthew Stone

Sage

University of California, Santa Cruz

Xin (Eric) Wang

ISABEL

University of Pittsburgh

Malihe Alikhani

PLAN-Bot

Virginia Tech

Ismini Lourentzou

The prizes for overall performance in the competition will be $500,000 for the first-place team, $100,000 for second, and $50,000 for third. Those prizes will be paid out to the students on the teams with the best overall performance.

“I am delighted to see that new teams are joining the second year of the competition together with returning teams, who, by competing again, are signaling to us that they found value in the TaskBot challenge, said Yoelle Maarek, vice president research and science for Amazon Shopping.  

“We expect these talented graduate students to continue surprising us, as well as Amazon customers, this year. Connecting academia, Amazonians, and actual customers experimenting with taskbots, is a winning combination to keep pushing the boundaries of science in conversational AI for Alexa to delight and ease the lives of millions of customers.”

The Alexa Prize is a competition for university students dedicated to advancing the field of conversational AI. Launched in 2016, the program was created to recognize students from around the globe who are changing the way we interact with technology.

TaskBot Challenge 2 teams are working to address one of the hardest problems in conversational AI — creating next-generation conversational AI experiences that delight customers by addressing their changing needs as they complete complex tasks. This challenge builds upon the Alexa Prize’s foundation of providing universities a unique opportunity to test cutting-edge machine learning models with actual customers at scale.

The Alexa Prize TaskBot challenge provides a realistic scenario with real-user multimodal interactions, making this the perfect setting to observe and measure human-bot conversations and AI algorithms in a groundbreaking setting.
rafael_ferreira_twiz.jpg
Rafael Ferreira, NOVA School of Science and Technology, Team TWIZ
Our vision of EvoquerBOT combines improving task completion rates and elevating user satisfaction. To this end, we deliver innovative solutions to fundamental NLP challenges.
haoran_zhang.jpeg
Haoran Zhang, Penn State University, Team EvoquerBOT
We are especially interested in developing innovative ways to achieve successful coordination of multiple modalities, such as visual and verbal elements, and create a more engaging and intuitive user experience.
Lingbo_Mo.JPG
Lingbo Mo, The Ohio State University, Team Taco 2.0
The GRILL team is excited to continue bringing cutting-edge AI research to improve people’s lives. Our research team works on new capabilities of foundation models that understand text, images, and the surrounding world.
Sophie_portrait.jpg
Sophie Fischer, University of Glasgow, Team GRILL
The competition lets us create interfaces for the general public in a production environment – it’s a unique opportunity to connect our research with our career goals.
Baber (Rutgers).jpeg
Baber Khalid, Rutgers University, Team DiWBot
We are very excited to be part of the community and look forward to working with the Alexa team and other teams.
Anthony_Sicilia.jpg
Anthony Sicilia, University of Pittsburgh, Team ISABEL
The Alexa Prize TaskBot Challenge combines a vast range of tasks over multiple domains with multimodal outputs. This is the ultimate test for any moonshot concept, and we can't wait to see what the real world has in store for us.
purdue 2.jpg
Rey (Alex) Gonzalez, Purdue University, Team BoilerBot
Participating in this competition is an incredible opportunity that will allow us to do applied research and ship it to real users.
ChrisSamarinas_DSC02670.jpg
Chris Samarinas, University of Massachusetts Amherst, Team Maruna
Although artificial intelligence has experienced explosive development in the past decade, there is still a gap between research and real-world application. The TaskBot Challenge provides us with a unique opportunity to explore multimodal AI in practical situations.
UCSC Kaishi TB2.png
Kaizhi Zheng Univerisity of California, Santa Cruz-Amherst, Team Sage
Our bot will make adaptable conversation a reality by allowing customers to follow personalized decisions through the completion of multiple, sequential sub-tasks and adapt to the tools, materials, or ingredients available to the user by proposing appropriate substitutes and alternatives.
Afrina Tabassum
Afrina Tabassum

TaskBot is the first conversational AI challenge to incorporate multimodal customer experiences, so in addition to receiving verbal instructions, customers with Echo Show or Fire TV devices, can also be presented with step-by-step instructions, images, or diagrams that enhance task guidance.

This year’s challenge has been expanded to include more hobbies and at-home activities. Participating teams were asked to propose interesting ways to incorporate visual aids into every conversation turn when a screen is available. Innovative ideas on improving the presentation of visual aids, as well as the coordination of visual and verbal modalities, were part of the team selection criteria.

Each university selected for the challenge receives a $250,000 research grant, Alexa-enabled devices, free Amazon Web Services (AWS) cloud computing services to support their research and development efforts, access to Amazon scientists, the CoBot (conversational bot) toolkit and other tools such as automated speech recognition through Alexa, neural detection and generation models, conversational data sets, and design guidance and development support from the Alexa Prize team.

"Alexa, let's work together"

The university teams’ taskbots will be available for Alexa customers to engage with in May 2023 with a finals event being held in September, and winners announced later that month.

As with the previous challenge, Alexa customers can engage in conversation with teams’ taskbots when they become available in May by saying, “Alexa, let’s work together.” Until then, “Alexa, let’s work together” will direct you to conversations with the previous challenge winners of 2022 and the Alexa Prize TaskBot.

After initiating the interaction, Alexa customers then receive a brief message informing them that they are interacting with an Alexa Prize university taskbot before being randomly connected to one of the participating taskbots.

After exiting the conversation with the taskbot, which customers can do at any time, the customer is prompted for a verbal rating, followed by an option to provide additional feedback. The interactions, ratings, and feedback are shared with the teams to help them improve their taskbots. Customer ratings are also used to determine which university teams will move on to the semifinals and finals.

Our goal is to contribute to the multimodal conversational AI field and move it closer to the way humans perceive, reason, and communicate through multimodal information.
joao_magalhaes_twiz.jpg
João Magalhães, associate professor, NOVA School of Science and Technology, Team TWIZ
We look forward to the Challenge because it is the perfect platform to create multimodal, tasked-oriented dialogue systems that elevate user experience and engagement.
rui_zhang.jpeg
Rui Zhang, assistant professor, Penn State University, Team EvoquerBOT
Through this TaskBot Challenge, we hope our work can expand the horizon of conversational AI along dimensions like dialogue depth, multi-modal coordination, commonsense reasoning, and learning from use.
Huan_Sun.png
Huan Sun, associate professor, The Ohio State University, Team Taco 2.0
The GRILL team is creating the next generation of open assistants that understand and use knowledge about the world and can communicate effectively to inform and educate.
jeff.jpeg
Jeff Dalton, associate professor, University of Glasgow, Team GRILL
Our TaskBot will help people get things done through personalized, adaptive, and context-aware conversational interaction by combining our research results with the state-of-the-art capabilities of Alexa devices.
Matthew Stone (Rutgers).jpg
Matthew Stone, professor, Rutgers University, Team DiWBot
We work towards making conversational AI technology more inclusive and collaborative. Inclusive Alexa can collaborate with users from diverse cultures and with different communication capabilities and preferences.
Malihe_Alikhani.jpg
Malihe Alikhani, assistant professor, University of Pittsburgh, Team ISABEL
We hope to develop a task-oriented system that can interact with users based on their level of knowledge, experience, and communication preference.
purdue 1.jpg
Julia Rayz, professor, Purdue University, Team BoilerBot

Success in the previous TaskBot Challenge required teams to address many difficult AI obstacles. The challenge required the fusion of multiple AI techniques including knowledge representation and inference, commonsense and causal reasoning, and language understanding and generation.

The “GRILLBot” team from University of Glasgow won the TaskBot 1 Challenge, earning a $500,000 prize for its performance. Teams from NOVA School of Science and Technology (Portgual) and The Ohio State University earned second- and third-place prizes, respectively.

Research papers from Amazon’s Alexa Prize team, and each of the competing teams, can be viewed and downloaded here.

Alexa Prize Taskbot Challenge Finals | Amazon Science

Research areas

Latest news

The latest updates, stories, and more about Alexa Prize.
US, CA, San Francisco
The Artificial General Intelligence (AGI) team is looking for a passionate, talented, and inventive Member of Technical Staff with a strong deep learning background, to build industry-leading Generative Artificial Intelligence (GenAI) technology with Large Language Models (LLMs) and multimodal systems. Key job responsibilities As a Member of Technical Staff with the AGI team, you will lead the development of algorithms and modeling techniques, to advance the state of the art with LLMs. You will lead the foundational model development in an applied research role, including model training, dataset design, and pre- and post-training optimization. Your work will directly impact our customers in the form of products and services that make use of GenAI technology. You will leverage Amazon’s heterogeneous data sources and large-scale computing resources to accelerate advances in LLMs. About the team The AGI team has a mission to push the envelope in GenAI with LLMs and multimodal systems, in order to provide the best-possible experience for our customers.
US, CA, San Francisco
The Artificial General Intelligence (AGI) team is looking for a passionate, talented, and inventive Member of Technical Staff with a strong deep learning background, to build industry-leading Generative Artificial Intelligence (GenAI) technology with Large Language Models (LLMs) and multimodal systems. Key job responsibilities As a Member of Technical Staff with the AGI team, you will lead the development of algorithms and modeling techniques, to advance the state of the art with LLMs. You will lead the foundational model development in an applied research role, including model training, dataset design, and pre- and post-training optimization. Your work will directly impact our customers in the form of products and services that make use of GenAI technology. You will leverage Amazon’s heterogeneous data sources and large-scale computing resources to accelerate advances in LLMs. About the team The AGI team has a mission to push the envelope in GenAI with LLMs and multimodal systems, in order to provide the best-possible experience for our customers.
US, CA, San Francisco
The Artificial General Intelligence (AGI) team is looking for a passionate, talented, and inventive Member of Technical Staff with a strong deep learning background, to build industry-leading Generative Artificial Intelligence (GenAI) technology with Large Language Models (LLMs) and multimodal systems. Key job responsibilities As a Member of Technical Staff with the AGI team, you will lead the development of algorithms and modeling techniques, to advance the state of the art with LLMs. You will lead the foundational model development in an applied research role, including model training, dataset design, and pre- and post-training optimization. Your work will directly impact our customers in the form of products and services that make use of GenAI technology. You will leverage Amazon’s heterogeneous data sources and large-scale computing resources to accelerate advances in LLMs. About the team The AGI team has a mission to push the envelope in GenAI with LLMs and multimodal systems, in order to provide the best-possible experience for our customers.
US, CA, San Francisco
The Artificial General Intelligence (AGI) team is looking for a passionate, talented, and inventive Member of Technical Staff with a strong deep learning background, to build industry-leading Generative Artificial Intelligence (GenAI) technology with Large Language Models (LLMs) and multimodal systems. Key job responsibilities As a Member of Technical Staff with the AGI team, you will lead the development of algorithms and modeling techniques, to advance the state of the art with LLMs. You will lead the foundational model development in an applied research role, including model training, dataset design, and pre- and post-training optimization. Your work will directly impact our customers in the form of products and services that make use of GenAI technology. You will leverage Amazon’s heterogeneous data sources and large-scale computing resources to accelerate advances in LLMs. About the team The AGI team has a mission to push the envelope in GenAI with LLMs and multimodal systems, in order to provide the best-possible experience for our customers.
US, CA, San Francisco
The Artificial General Intelligence (AGI) team is looking for a passionate, talented, and inventive Member of Technical Staff with a strong deep learning background, to build industry-leading Generative Artificial Intelligence (GenAI) technology with Large Language Models (LLMs) and multimodal systems. Key job responsibilities As a Member of Technical Staff with the AGI team, you will lead the development of algorithms and modeling techniques, to advance the state of the art with LLMs. You will lead the foundational model development in an applied research role, including model training, dataset design, and pre- and post-training optimization. Your work will directly impact our customers in the form of products and services that make use of GenAI technology. You will leverage Amazon’s heterogeneous data sources and large-scale computing resources to accelerate advances in LLMs. About the team The AGI team has a mission to push the envelope in GenAI with LLMs and multimodal systems, in order to provide the best-possible experience for our customers.
US, CA, Sunnyvale
Prime Video is a first-stop entertainment destination offering customers a vast collection of premium programming in one app available across thousands of devices. Prime members can customize their viewing experience and find their favorite movies, series, documentaries, and live sports – including Amazon MGM Studios-produced series and movies; licensed fan favorites; and programming from Prime Video add-on subscriptions such as Apple TV+, Max, Crunchyroll and MGM+. All customers, regardless of whether they have a Prime membership or not, can rent or buy titles via the Prime Video Store, and can enjoy even more content for free with ads. Are you interested in shaping the future of entertainment? Prime Video's technology teams are creating best-in-class digital video experience. As a Prime Video technologist, you’ll have end-to-end ownership of the product, user experience, design, and technology required to deliver state-of-the-art experiences for our customers. You’ll get to work on projects that are fast-paced, challenging, and varied. You’ll also be able to experiment with new possibilities, take risks, and collaborate with remarkable people. We’ll look for you to bring your diverse perspectives, ideas, and skill-sets to make Prime Video even better for our customers. With global opportunities for talented technologists, you can decide where a career Prime Video Tech takes you! We are looking for a self-motivated, passionate and resourceful Sr. Applied Scientists with Recommender System or Search Ranking or Ads Ranking experience to bring diverse perspectives, ideas, and skill-sets to make Prime Video even better for our customers. You will spend your time as a hands-on machine learning practitioner and a research leader. You will play a key role on the team, building and guiding machine learning models from the ground up. At the end of the day, you will have the reward of seeing your contributions benefit millions of Amazon.com customers worldwide. Key job responsibilities - Develop AI solutions for various Prime Video Recommendation/Search systems using Deep learning, GenAI, Reinforcement Learning, and optimization methods; - Work closely with engineers and product managers to design, implement and launch AI solutions end-to-end; - Design and conduct offline and online (A/B) experiments to evaluate proposed solutions based on in-depth data analyses; - Effectively communicate technical and non-technical ideas with teammates and stakeholders; - Stay up-to-date with advancements and the latest modeling techniques in the field; - Publish your research findings in top conferences and journals. About the team Prime Video Recommendation/Search Science team owns science solution to power search experience on various devices, from sourcing, relevance, ranking, to name a few. We work closely with the engineering teams to launch our solutions in production.
US, WA, Seattle
Amazon's Price Perception and Evaluation team is seeking a driven Principal Applied Scientist to harness planet scale multi-modal datasets, and navigate a continuously evolving competitor landscape, in order to build and scale an advanced self-learning scientific price estimation and product understanding system, regularly generating fresh customer-relevant prices on billions of Amazon and Third Party Seller products worldwide. We are looking for a talented, organized, and customer-focused technical leader with a charter to derive deep neural product relationships, quantify substitution and complementarity effects, and publish trust-preserving probabilistic price ranges on all products listed on Amazon. This role requires an individual with excellent scientific modeling and system design skills, bar-raising business acumen, and an entrepreneurial spirit. We are looking for an experienced leader who is a self-starter comfortable with ambiguity, demonstrates strong attention to detail, and has the ability to work in a fast-paced and ever-changing environment. Key job responsibilities - Develop the team. Mentor a highly talented group of applied machine learning scientists & researchers. - See the big picture. Shape long term vision for Amazon's science-based competitive, perception-preserving pricing techniques - Build strong collaborations. Partner with product, engineering, and science teams within Pricing & Promotions to deploy machine learning price estimation and error correction solutions at Amazon scale - Stay informed. Establish mechanisms to stay up to date on latest scientific advancements in machine learning, neural networks, natural language processing, probabilistic forecasting, and multi-objective optimization techniques. Identify opportunities to apply them to relevant Pricing & Promotions business problems - Keep innovating for our customers. Foster an environment that promotes rapid experimentation, continuous learning, and incremental value delivery. - Deliver Impact. Develop, Deploy, and Scale Amazon's next generation foundational price estimation and understanding system
US, WA, Seattle
Here at Amazon, we embrace our differences. We are committed to furthering our culture of diversity and inclusion of our teams within the organization. How do you get items to customers quickly, cost-effectively, and—most importantly—safely, in less than an hour? And how do you do it in a way that can scale? Our teams of hundreds of scientists, engineers, aerospace professionals, and futurists have been working hard to do just that! We are delivering to customers, and are excited for what’s to come. Check out more information about Prime Air on the About Amazon blog (https://www.aboutamazon.com/news/transportation/amazon-prime-air-delivery-drone-reveal-photos). If you are seeking an iterative environment where you can drive innovation, apply state-of-the-art technologies to solve real world delivery challenges, and provide benefits to customers, Prime Air is the place for you. Come work on the Amazon Prime Air Team! We are seeking a highly skilled Navigation Scientist to help develop advanced algorithms and software for our Prime Air delivery drone program. In this role, you will conduct comprehensive navigation analysis to support cross-functional decision-making, define system architecture and requirements, contribute to the development of flight algorithms, and actively identify innovative technological opportunities that will drive significant enhancements to meet our customers' evolving demands. Export Control License: This position may require a deemed export control license for compliance with applicable laws and regulations. Placement is contingent on Amazon’s ability to apply for and obtain an export control license on your behalf.
IN, KA, Bengaluru
Alexa+ is Amazon’s next-generation, AI-powered virtual assistant. Building on the original Alexa, it uses generative AI to deliver a more conversational, personalized, and effective experience. As an Applied Scientist II on the Alexa Sensitive Content Intelligence (ASCI) team, you'll be part of an elite group developing industry-leading technologies in attribute extraction and sensitive content detection that work seamlessly across all languages and countries. In this role, you'll join a team of exceptional scientists pushing the boundaries of Natural Language Processing. Working in our dynamic, fast-paced environment, you'll develop novel algorithms and modeling techniques that advance the state of the art in NLP. Your innovations will directly shape how millions of customers interact with Amazon Echo, Echo Dot, Echo Show, and Fire TV devices every day. What makes this role exciting is the unique blend of scientific innovation and real-world impact. You'll be at the intersection of theoretical research and practical application, working alongside talented engineers and product managers to transform breakthrough ideas into customer-facing experiences. Your work will be crucial in ensuring Alexa remains at the forefront of AI technology while maintaining the highest standards of trust and safety. We're looking for a passionate innovator who combines strong technical expertise with creative problem-solving skills. Your deep understanding of NLP models (including LSTM and transformer-based architectures) will be essential in tackling complex challenges and identifying novel solutions. You'll leverage your exceptional technical knowledge, strong Computer Science fundamentals, and experience with large-scale distributed systems to create reliable, scalable, and high-performance products that delight our customers. Key job responsibilities In this dynamic role, you'll design and implement GenAI solutions that define the future of AI interaction. You'll pioneer novel algorithms, conduct ground breaking experiments, and optimize user experiences through innovative approaches to sensitive content detection and mitigation. Working alongside exceptional engineers and scientists, you'll transform theoretical breakthroughs into practical, scalable solutions that strengthen user trust in Alexa globally. You'll also have the opportunity to mentor rising talent, contributing to Amazon's culture of scientific excellence while helping build high-performing teams that deliver swift, impactful results. A day in the life Imagine starting your day collaborating with brilliant minds on advancing state-of-the-art NLP algorithms, then moving on to analyze experiment results that could reshape how Alexa understands and responds to users. You'll partner with cross-functional teams - from engineers to product managers - to ensure data quality, refine policies, and enhance model performance. Your expertise will guide technical discussions, shape roadmaps, and influence key platform features that require cross-team leadership. About the team The mission of the Alexa Sensitive Content Intelligence (ASCI) team is to (1) minimize negative surprises to customers caused by sensitive content, (2) detect and prevent potential brand-damaging interactions, and (3) build customer trust through appropriate interactions on sensitive topics. The term “sensitive content” includes within its scope a wide range of categories of content such as offensive content (e.g., hate speech, racist speech), profanity, content that is suitable only for certain age groups, politically polarizing content, and religiously polarizing content. The term “content” refers to any material that is exposed to customers by Alexa (including both 1P and 3P experiences) and includes text, speech, audio, and video.
IN, KA, Bengaluru
Alexa+ is Amazon’s next-generation, AI-powered virtual assistant. Building on the original Alexa, it uses generative AI to deliver a more conversational, personalized, and effective experience. As an Applied Scientist II on the Alexa Sensitive Content Intelligence (ASCI) team, you'll be part of an elite group developing industry-leading technologies in attribute extraction and sensitive content detection that work seamlessly across all languages and countries. In this role, you'll join a team of exceptional scientists pushing the boundaries of Natural Language Processing. Working in our dynamic, fast-paced environment, you'll develop novel algorithms and modeling techniques that advance the state of the art in NLP. Your innovations will directly shape how millions of customers interact with Amazon Echo, Echo Dot, Echo Show, and Fire TV devices every day. What makes this role exciting is the unique blend of scientific innovation and real-world impact. You'll be at the intersection of theoretical research and practical application, working alongside talented engineers and product managers to transform breakthrough ideas into customer-facing experiences. Your work will be crucial in ensuring Alexa remains at the forefront of AI technology while maintaining the highest standards of trust and safety. We're looking for a passionate innovator who combines strong technical expertise with creative problem-solving skills. Your deep understanding of NLP models (including LSTM and transformer-based architectures) will be essential in tackling complex challenges and identifying novel solutions. You'll leverage your exceptional technical knowledge, strong Computer Science fundamentals, and experience with large-scale distributed systems to create reliable, scalable, and high-performance products that delight our customers. Key job responsibilities In this dynamic role, you'll design and implement GenAI solutions that define the future of AI interaction. You'll pioneer novel algorithms, conduct ground breaking experiments, and optimize user experiences through innovative approaches to sensitive content detection and mitigation. Working alongside exceptional engineers and scientists, you'll transform theoretical breakthroughs into practical, scalable solutions that strengthen user trust in Alexa globally. You'll also have the opportunity to mentor rising talent, contributing to Amazon's culture of scientific excellence while helping build high-performing teams that deliver swift, impactful results. A day in the life Imagine starting your day collaborating with brilliant minds on advancing state-of-the-art NLP algorithms, then moving on to analyze experiment results that could reshape how Alexa understands and responds to users. You'll partner with cross-functional teams - from engineers to product managers - to ensure data quality, refine policies, and enhance model performance. Your expertise will guide technical discussions, shape roadmaps, and influence key platform features that require cross-team leadership. About the team The mission of the Alexa Sensitive Content Intelligence (ASCI) team is to (1) minimize negative surprises to customers caused by sensitive content, (2) detect and prevent potential brand-damaging interactions, and (3) build customer trust through appropriate interactions on sensitive topics. The term “sensitive content” includes within its scope a wide range of categories of content such as offensive content (e.g., hate speech, racist speech), profanity, content that is suitable only for certain age groups, politically polarizing content, and religiously polarizing content. The term “content” refers to any material that is exposed to customers by Alexa (including both 1P and 3P experiences) and includes text, speech, audio, and video.