SocialBot Grand Challenge FAQs

Frequently asked questions about the challenge.
General
What is a SocialBot?
In the context of Alexa Prize, a SocialBot is an Alexa skill that can engage customers in a conversation about popular topics such as entertainment, sports, politics, technology, and fashion.
Can I choose to build any kind of conversational bot?
No, this competition focuses solely on SocialBots. Your SocialBot will need to converse about a wide range of popular societal topics and current events, and not only topics of your choosing. You will develop your own technologies relevant to your approach, such as natural language understanding, dialog management, knowledge acquisition, commonsense reasoning, conversational planning, machine learning, natural language generation, etc. Far-field automatic speech recognition (ASR) and text-to-speech (TTS) will be provided by Amazon, along with additional data, models, and tools.
What will my SocialBot do?
Your SocialBot will be an Alexa skill that can converse coherently and engagingly with humans on popular topics and current events. If your SocialBot advances to the pertinent stage of the competition, Alexa users in the U.S. may ask to converse with your SocialBot about topics of interest to them, such as baseball playoffs, celebrity gossip, or scientific breakthroughs. For example:
  • User: Let’s chat about the Mars Mission.
  • SocialBot: There are multiple Mars missions, some public and some private.
  • User: Who do you think will succeed?
  • SocialBot: I think more than one will succeed. The first one to get to Mars will open the doors for others.
  • User: I’d love to go to Mars.
  • SocialBot: Me too. Luckily I can hitch a ride on any computer.
  • User: That’s very funny.
  • SocialBot: The biggest challenge isn’t technology, it’s funding. It will cost $500 billion to send humans to Mars.
Your SocialBot will continue turn-by-turn interaction, starting with a topic the user asked for, until the user chooses to stop. Like an everyday human conversation, the interaction may shift naturally to related topics, as in the above example, but the conversation should remain coherent, relevant, and engaging. Your SocialBot may suggest topics to keep the conversation flowing. The goal is to keep the conversation from deteriorating to the point where the user loses interest.
How will I build my SocialBot?
You will use the Alexa Skills Kit (ASK) to build an Alexa skill, hosted on AWS Lambda, that will create the end-to-end conversational experience for a user. Using the provided APIs, your skill will receive as input the text of the user’s utterance, and produce as output a text sentence that will be spoken to the user. You do not need to tackle ASR (automatic speech recognition) or TTS (text to speech). You will also be provided with the CoBot Toolkit (a conversational bot toolkit), a software development kit that works with ASK and was built specifically for Alexa Prize teams to reduce the involved engineering in setting up a SocialBot and allow teams to focus on the science.


Your skill will need to determine an appropriate response at each turn of the conversation. It will also need to keep up with current news and events using the provided data sources. You may use additional data sources or libraries if you wish, subject to the terms described in the Official Rules.
What is the Alexa Skills Kit (ASK)?
The Alexa Skills Kit (ASK) is a collection of free, self-service APIs, tools, documentation, and code samples that make it fast and easy for you to add skills to Alexa. Your team will use ASK to build, deploy, and test a SocialBot that is capable of conversing with millions of Alexa users.
Competition details
What is the goal of the challenge?
The goal of the SocialBot Grand challenge is to advance several areas of conversational AI including natural language understanding (NLU), context modeling, dialog management, commonsense reasoning, natural language generation (NLG), and knowledge acquisition. The grand challenge objective is to create a SocialBot that converses coherently and engagingly with humans on popular topics for 20 minutes while achieving a user rating of at least 4.0/5.0.
How will winners be selected?
Through various phases of the competition, SocialBots will be evaluated based on feedback from Alexa users and assessment by Amazon.


Following the initial feedback period, SocialBots that have been certified and published will be evaluated on criteria such as the average interaction rating, uptime requirements, and ability to filter offensive content in order to advance to the Semifinals Interaction Period.



During the semifinals interaction period, Alexa customers will evaluate the semifinalist SocialBots. The two SocialBots with the highest Semifinals Interaction Rating Average and up to three more SocialBots selected by Amazon will advance to the finals.



Teams that advance to and complete the Semifinals Interaction Period, regardless of whether they advanced to the Finals Event, will be eligible to compete for Scientific Invention and Innovation Prizes based on the level of scientific invention and innovation demonstrated by each Entrant Team throughout the Competition.



The Teams whose SocialBots attain the three highest Composite Scores during the finals event will be the winners of the Overall Performance Prizes.
Will this competition be judged like a Turing Test?
No. The goal of the Alexa Prize is to create SocialBots that engage in interesting, human-like conversations, not to make them indistinguishable from a human when compared side-by-side. While the SocialBots built for the Alexa Prize will be human-like in some respects, they will be very different in others, and could easily reveal themselves in a Turing Test. For example, SocialBots may have ready access to much more information than a human. Asking the SocialBots to act like a human could diminish the customer experience and hinder the efforts of the participants to build the best SocialBot to further conversational AI.
When and where is the finals event?
The finals event will be held in July 2023 at a location to be determined, with a science invention and innovation presentation review to follow. The competition results will be announced in August 2023.
Can we use other funding to help us participate in this challenge?
Yes, you may use other funding to support your team, subject to the terms described in the Official Rules. External funding must be disclosed to Amazon.
Will Alexa customers be able to engage with our SocialBot?
Your team will be required to submit its SocialBot for certification and publication by the Amazon Alexa team. After certification, you will enter the Internal Amazon Beta Period, where Amazon employees will test your SocialBot and provide feedback. After the Internal Amazon Beta Period, we will allow Alexa users to try your SocialBot and provide feedback to you. Amazon may impose requirements that the SocialBots must meet before they will be made available to Alexa users. Such requirements may include, among other things, a minimum average customer rating, uptime requirements, or the ability to consistently filter offensive content.
Which Alexa users will be able to interact with the SocialBots, and what languages must they support?
SocialBots will be made available to Alexa users in the United States or who select the United States as their preferred marketplace. Your team must build its SocialBot using U.S. English.
Will we publish our research from the Alexa Prize?
Yes. Publishing research papers as an outcome of your work on Alexa Prize is required for all teams participating in the competition, although teams may not publish Amazon confidential information, as described in the Official Rules. The Alexa Prize requires all teams to submit a technical paper for the Alexa Prize proceedings. Your SocialBot will not be selected for the finals if your team does not submit a technical paper for Alexa Prize proceedings. Papers will be published online at the end of the competition and made publicly available.

Teams may also publish research papers in third-party publications and conferences, as long as all papers are provided to Amazon for review at least two weeks before the submission deadlines and no research papers are published before the Alexa Prize proceedings are published, unless Amazon approves otherwise in writing.
Who will own the intellectual property rights in my submission?
You will retain ownership over your SocialBot. Amazon will have a non-exclusive license to any technology or software you develop in connection with the competition. See the Official Rules for details.
Eligibility
Who can apply to participate?
The Alexa Prize is open to full-time students enrolled in an accredited university, with the exception of universities in Cuba, Iran, Syria, North Korea, Sudan, the region of Crimea, and where prohibited by law (see Official Rules). Proof of enrollment will be required to participate.
Can I participate if I don’t attend a university?
No. The Alexa Prize is open only to full-time enrolled university students.
Do I need to be enrolled in a university program throughout my participation in the competition?
All participating team members must remain full-time students in good standing at their university while participating in the competition.
Do I need to be a certain age?
Participants must be at or above the age of majority in the country, state, province, or jurisdiction of residence at the time of entry.
Can I enroll if a family member is an Amazon employee?
Immediate family members and household members of Amazon employees, directors, and contractors are not eligible to participate. See Official Rules for additional restrictions.
Teams
How many teams will be selected to participate?
All applications will be reviewed and evaluated by Amazon. Up to ten teams will be selected and sponsored by Amazon. All teams will receive a $250,000 grant intended to support two full-time students and a month of faculty time, free Alexa devices, and free AWS hosting including access to CPU and GPU based machines, SQL and NoSQL databases, and object storage. See Official Rules for details.
How many team members can our team have?
There is no minimum or maximum number of team members. All team members must be enrolled in their university throughout their participation. All teams will receive a $250,000 grant regardless of how many members are on the team. We recommend a team with four to six students with diverse fields of study or areas of expertise.
Can students from different universities be on the same team?
No. Teams must be comprised of students attending the same university.
Can one university have more than one team?
Yes, universities may have more than one team. Multiple teams cannot have the same faculty advisor.
Can I participate on two separate teams?
No. You can only be a part of one team for the duration of the competition.
Can undergraduate and graduate students work together?
Yes, teams may be comprised of undergraduate and graduate students.
Do I need a faculty advisor?
All teams must nominate a faculty advisor and include the faculty advisor’s consent in the applications.
What is the role of the faculty advisor?
Faculty advisors will advise students on technical directions and be a sounding board for new ideas, similar to a graduate school advisor. They will also act as the official representative from the university for this competition.
Can we add or remove team members during the competition?
During the competition, there will be a period of time during which faculty advisors may request to remove or add members to the team, subject to approval by Amazon. See Official Rules for details.
Can we discuss our SocialBot with faculty or students who aren’t on our team?
Only team members may work on their SocialBots. However, the faculty advisor and other students and faculty members at your university may provide support and advice to your team and may co-author technical publications and research papers.
Application process
How do we apply?
Begin the application via YouNoodle.
What do we need to apply?
Once you have selected your team members, team leader, and faculty sponsor, you are ready to begin the application process.
Do all team members have to apply?
Each team must have a team lead, who should submit only one application on behalf of the whole team. Your application must include all of your team members’ information.
Is there an application fee?
There is no application fee.
How will teams be selected to participate?
All applications will be reviewed. Teams will be selected by Amazon based on the following criteria: (1) the potential scientific contribution to the field; (2) the technical merit of the approach; (3) the novelty of the idea; and (4) an assessment of the team’s ability to execute against their plan. Please be sure to provide enough detail in your application to enable evaluation of your proposal.
Prizes
What are the prizes for winning the competition?
Overall Performance Prize: For the three teams that build the SocialBot with the highest overall performance, the first-place team will win $250,000, the second-place team will win $50,000, and the third-place team will win $25,000. These prizes will be paid directly to the students on each winning team.


Scientific Invention and Innovation Prize: For the three teams that demonstrate the most scientific invention and innovation throughout the competition, the first-place team will win $250,000, the second-place team will win $50,000, and the third-place team will win $25,000. These prizes will be paid directly to the students on each winning team.



Grand Prize: If and only if the SocialBot of the team that wins the first-place Overall Performance Prize also achieves the grand challenge of conversing coherently and engagingly with humans for 20 minutes in at least two-thirds of its conversations at the finals event and achieves a 4.0 or higher composite score, that team’s university will be awarded a $1 million research grant.



See Official Rules for details.
Do we get a stipend and devices to participate in the Alexa Prize?
Up to ten teams will be sponsored to participate in the competition. Each sponsored team’s university will receive a $250,000 research grant to help fund the team’s participation.


The sponsorship includes Alexa-enabled devices, free AWS services to support the development of the team’s SocialBot, and support from the Alexa Prize team.
How can the grant be spent?
The grant is intended to support two full-time students for the duration of the competition and one month of the faculty advisor’s salary. No more than 35% of the research grant may be allocated to administrative fees. If your team would like to use the funds in another manner, your faculty advisor must receive approval from Amazon before doing so.
How will the prizes be distributed among a team?
Each Overall Performance Prize and the Scientific Invention and Innovation Prize will be distributed equally among the members of each winning team.
Timeline
What are the key milestones of the competition?
Teams must submit their applications by October 5, 2022. Teams selected to participate in the competition will be notified in October of November 2022. The competition will run from about November 2022 through August 2023. See Official Rules for details.

Latest news

The latest updates, stories, and more about Alexa Prize.
  • Behnam Hedayatnia
    March 5, 2019
    The 2018 Alexa Prize featured eight student teams from four countries, each of which adopted distinctive approaches to some of the central technical questions in conversational AI. We survey those approaches in a paper we released late last year, and the teams themselves go into even greater detail in the papers they submitted to the latest Alexa Prize Proceedings. Here, we touch on just a few of the teams’ innovations.
  • Anushree Venkatesh
    February 27, 2019
    To ensure that Alexa Prize contestants can concentrate on dialogue systems — the core technology of socialbots — Amazon scientists and engineers built a set of machine learning modules that handle fundamental conversational tasks and a development environment that lets contestants easily mix and match existing modules with those of their own design.
US, WA, Seattle
Come be a part of a rapidly expanding $35 billion-dollar global business. At Amazon Business, a fast-growing startup passionate about building solutions, we set out every day to innovate and disrupt the status quo. We stand at the intersection of tech & retail in the B2B space developing innovative purchasing and procurement solutions to help businesses and organizations thrive. At Amazon Business, we strive to be the most recognized and preferred strategic partner for smart business buying. Bring your insight, imagination and a healthy disregard for the impossible. Join us in building and celebrating the value of Amazon Business to buyers and sellers of all sizes and industries. Unlock your career potential. Amazon Business Data Insights and Analytics team is looking for a Data Scientist to lead the research and thought leadership to drive our data and insights strategy for Amazon Business. This role is central in shaping the definition and execution of the long-term strategy for Amazon Business. You will be responsible for researching, experimenting and analyzing predictive and optimization models, designing and implementing advanced detection systems that analyze customer behavior at registration and throughout their journey. You will work on ambiguous and complex business and research science problems with large opportunities. You'll leverage diverse data signals including customer profiles, purchase patterns, and network associations to identify potential abuse and fraudulent activities. You are an analytical individual who is comfortable working with cross-functional teams and systems, working with state-of-the-art machine learning techniques and AWS services to build robust models that can effectively distinguish between legitimate business activities and suspicious behavior patterns You must be a self-starter and be able to learn on the go. Excellent written and verbal communication skills are required as you will work very closely with diverse teams. Key job responsibilities - Interact with business and software teams to understand their business requirements and operational processes - Frame business problems into scalable solutions - Adapt existing and invent new techniques for solutions - Gather data required for analysis and model building - Create and track accuracy and performance metrics - Prototype models by using high-level modeling languages such as R or in software languages such as Python. - Familiarity with transforming prototypes to production is preferred. - Create, enhance, and maintain technical documentation
US, TX, Austin
Amazon Leo is an initiative to launch a constellation of Low Earth Orbit satellites that will provide low-latency, high-speed broadband connectivity to unserved and underserved communities around the world. As a Systems Engineer, this role is primarily responsible for the design, development and integration of communication payload and customer terminal systems. The Role: Be part of the team defining the overall communication system and architecture of Amazon Leo’s broadband wireless network. This is a unique opportunity to innovate and define groundbreaking wireless technology at global scale. The team develops and designs the communication system for Leo and analyzes its overall system level performance such as for overall throughput, latency, system availability, packet loss etc. This role in particular will be responsible for leading the effort in designing and developing advanced technology and solutions for communication system. This role will also be responsible developing advanced physical layer + protocol stacks systems as proof of concept and reference implementation to improve the performance and reliability of the LEO network. In particular this role will be responsible for using concepts from digital signal processing, information theory, wireless communications to develop novel solutions for achieving ultra-high performance LEO network. This role will also be part of a team and develop simulation tools with particular emphasis on modeling the physical layer aspects such as advanced receiver modeling and abstraction, interference cancellation techniques, FEC abstraction models etc. This role will also play a critical role in the integration and verification of various HW and SW sub-systems as a part of system integration and link bring-up and verification. Export Control Requirement: Due to applicable export control laws and regulations, candidates must be a U.S. citizen or national, U.S. permanent resident (i.e., current Green Card holder), or lawfully admitted into the U.S. as a refugee or granted asylum.
US, MA, N.reading
Amazon Industrial Robotics Group is seeking exceptional talent to help develop the next generation of advanced robotics systems that will transform automation at Amazon's scale. We're building revolutionary robotic systems that combine cutting-edge AI, sophisticated control systems, and advanced mechanical design to create adaptable automation solutions capable of working safely alongside humans in dynamic environments. This is a unique opportunity to shape the future of robotics and automation at an unprecedented scale, working with world-class teams pushing the boundaries of what's possible in robotic dexterous manipulation, locomotion, and human-robot interaction. This role presents an opportunity to shape the future of robotics through innovative applications of deep learning and large language models. At Amazon Industrial Robotics Group, we leverage advanced robotics, machine learning, and artificial intelligence to solve complex operational challenges at an unprecedented scale. Our fleet of robots operates across hundreds of facilities worldwide, working in sophisticated coordination to fulfill our mission of customer excellence. We are pioneering the development of dexterous manipulation system that: - Enables unprecedented generalization across diverse tasks - Enables contact-rich manipulation in different environments - Seamlessly integrates low-level skills and high-level behaviors - Leverage mechanical intelligence, multi-modal sensor feedback and advanced control techniques. The ideal candidate will contribute to research that bridges the gap between theoretical advancement and practical implementation in robotics. You will be part of a team that's revolutionizing how robots learn, adapt, and interact with their environment. Join us in building the next generation of intelligent robotics systems that will transform the future of automation and human-robot collaboration. A day in the life - Work on design and implementation of methods for Visual SLAM, navigation and spatial reasoning - Leverage simulation and real-world data collection to create large datasets for model development - Develop a hierarchical system that combines low-level control with high-level planning - Collaborate effectively with multi-disciplinary teams to co-design hardware and algorithms for dexterous manipulation
US, NY, New York
We are seeking an Applied Scientist to lead the development of evaluation frameworks and data collection protocols for robotic capabilities. In this role, you will focus on designing how we measure, stress-test, and improve robot behavior across a wide range of real-world tasks. Your work will play a critical role in shaping how policies are validated and how high-quality datasets are generated to accelerate system performance. You will operate at the intersection of robotics, machine learning, and human-in-the-loop systems, building the infrastructure and methodologies that connect teleoperation, evaluation, and learning. This includes developing evaluation policies, defining task structures, and contributing to operator-facing interfaces that enable scalable and reliable data collection. The ideal candidate is highly experimental, systems-oriented, and comfortable working across software, robotics, and data pipelines, with a strong focus on turning ambiguous capability goals into measurable and actionable evaluation systems. Key job responsibilities - Design and implement evaluation frameworks to measure robot capabilities across structured tasks, edge cases, and real-world scenarios - Develop task definitions, success criteria, and benchmarking methodologies that enable consistent and reproducible evaluation of policies - Create and refine data collection protocols that generate high-quality, task-relevant datasets aligned with model development needs - Build and iterate on teleoperation workflows and operator interfaces to support efficient, reliable, and scalable data collection - Analyze evaluation results and collected data to identify performance gaps, failure modes, and opportunities for targeted data collection - Collaborate with engineering teams to integrate evaluation tooling, logging systems, and data pipelines into the broader robotics stack - Stay current with advances in robotics, evaluation methodologies, and human-in-the-loop learning to continuously improve internal approaches - Lead technical projects from conception through production deployment - Mentor junior scientists and engineers
US, WA, Bellevue
We are seeking a passionate, talented, and inventive individual to join the Applied AI team and help build industry-leading technologies that customers will love. This team offers a unique opportunity to make a significant impact on the customer experience and contribute to the design, architecture, and implementation of a cutting-edge product. The mission of the Applied AI team is to enable organizations within Worldwide Amazon.com Stores to accelerate the adoption of AI technologies across various parts of our business. We are looking for a Senior Applied Scientist to join our Applied AI team to work on LLM-based solutions. On our team you will push the boundaries of ML and Generative AI techniques to scale the inputs for hundreds of billions of dollars of annual revenue for our eCommerce business. If you have a passion for AI technologies, a drive to innovate and a desire to make a meaningful impact, we invite you to become a valued member of our team. You will be responsible for developing and maintaining the systems and tools that enable us to accelerate knowledge operations and work in the intersection of Science and Engineering. You will push the boundaries of ML and Generative AI techniques to scale the inputs for hundreds of billions of dollars of annual revenue for our eCommerce business. If you have a passion for AI technologies, a drive to innovate and a desire to make a meaningful impact, we invite you to become a valued member of our team. We are seeking an experienced Scientist who combines superb technical, research, analytical and leadership capabilities with a demonstrated ability to get the right things done quickly and effectively. This person must be comfortable working with a team of top-notch developers and collaborating with our research teams. We’re looking for someone who innovates, and loves solving hard problems. You will be expected to have an established background in building highly scalable systems and system design, excellent project management skills, great communication skills, and a motivation to achieve results in a fast-paced environment. You should be somebody who enjoys working on complex problems, is customer-centric, and feels strongly about building good software as well as making that software achieve its operational goals.
IN, KA, Bengaluru
Do you want to lead the development of advanced machine learning systems that protect millions of customers and power a trusted global eCommerce experience? Are you passionate about modeling terabytes of data, solving highly ambiguous fraud and risk challenges, and driving step-change improvements through scientific innovation? If so, the Amazon Buyer Risk Prevention (BRP) Machine Learning team may be the right place for you. We are seeking a Senior Applied Scientist to define and drive the scientific direction of large-scale risk management systems that safeguard millions of transactions every day. In this role, you will lead the design and deployment of advanced machine learning solutions, influence cross-team technical strategy, and leverage emerging technologies—including Generative AI and LLMs—to build next-generation risk prevention platforms. Key job responsibilities Lead the end-to-end scientific strategy for large-scale fraud and risk modeling initiatives Define problem statements, success metrics, and long-term modeling roadmaps in partnership with business and engineering leaders Design, develop, and deploy highly scalable machine learning systems in real-time production environments Drive innovation using advanced ML, deep learning, and GenAI/LLM technologies to automate and transform risk evaluation Influence system architecture and partner with engineering teams to ensure robust, scalable implementations Establish best practices for experimentation, model validation, monitoring, and lifecycle management Mentor and raise the technical bar for junior scientists through reviews, technical guidance, and thought leadership Communicate complex scientific insights clearly to senior leadership and cross-functional stakeholders Identify emerging scientific trends and translate them into impactful production solutions
US, CA, Palo Alto
The Sponsored Products and Brands (SPB) team at Amazon Ads is re-imagining the advertising landscape through state-of-the-art generative AI technologies, revolutionizing how millions of customers discover products and engage with brands across Amazon.com and beyond. We are at the forefront of re-inventing advertising experiences, bridging human creativity with artificial intelligence to transform every aspect of the advertising lifecycle from ad creation and optimization to performance analysis and customer insights. We are a passionate group of innovators dedicated to developing responsible and intelligent AI technologies that balance the needs of advertisers, enhance the shopping experience, and strengthen the marketplace. If you're energized by solving complex challenges and pushing the boundaries of what's possible with AI, join us in shaping the future of advertising. The Off-Search team within Sponsored Products and Brands (SPB) is focused on building delightful ad experiences across various surfaces beyond Search on Amazon—such as product detail pages, the homepage, and store-in-store pages—to drive monetization. Our vision is to deliver highly personalized, context-aware advertising that adapts to individual shopper preferences, scales across diverse page types, remains relevant to seasonal and event-driven moments, and integrates seamlessly with organic recommendations such as new arrivals, basket-building content, and fast-delivery options. To execute this vision, we work in close partnership with Amazon Stores stakeholders to lead the expansion and growth of advertising across Amazon-owned and -operated pages beyond Search. We operate full stack—from backend ads-retail edge services, ads retrieval, and ad auctions to shopper-facing experiences—all designed to deliver meaningful value. Curious about our advertising solutions? Discover more about Sponsored Products and Sponsored Brands to see how we’re helping businesses grow on Amazon.com and beyond! Key job responsibilities This role will be pivotal in redesigning how ads contribute to a personalized, relevant, and inspirational shopping experience, with the customer value proposition at the forefront. Key responsibilities include, but are not limited to: - Contribute to the design and development of GenAI, deep learning, multi-objective optimization and/or reinforcement learning empowered solutions to transform ad retrieval, auctions, whole-page relevance, and/or bespoke shopping experiences. - Collaborate cross-functionally with other scientists, engineers, and product managers to bring scalable, production-ready science solutions to life. - Stay abreast of industry trends in GenAI, LLMs, and related disciplines, bringing fresh and innovative concepts, ideas, and prototypes to the organization. - Contribute to the enhancement of team’s scientific and technical rigor by identifying and implementing best-in-class algorithms, methodologies, and infrastructure that enable rapid experimentation and scaling. - Mentor and grow junior scientists and engineers, cultivating a high-performing, collaborative, and intellectually curious team. A day in the life As an Applied Scientist on the Sponsored Products and Brands Off-Search team, you will contribute to the development in Generative AI (GenAI) and Large Language Models (LLMs) to revolutionize our advertising flow, backend optimization, and frontend shopping experiences. This is a rare opportunity to redefine how ads are retrieved, allocated, and/or experienced—elevating them into personalized, contextually aware, and inspiring components of the customer journey. You will have the opportunity to fundamentally transform areas such as ad retrieval, ad allocation, whole-page relevance, and differentiated recommendations through the lens of GenAI. By building novel generative models grounded in both Amazon’s rich data and the world’s collective knowledge, your work will shape how customers engage with ads, discover products, and make purchasing decisions. If you are passionate about applying frontier AI to real-world problems with massive scale and impact, this is your opportunity to define the next chapter of advertising science. About the team The Off-Search team within Sponsored Products and Brands (SPB) is focused on building delightful ad experiences across various surfaces beyond Search on Amazon—such as product detail pages, the homepage, and store-in-store pages—to drive monetization. Our vision is to deliver highly personalized, context-aware advertising that adapts to individual shopper preferences, scales across diverse page types, remains relevant to seasonal and event-driven moments, and integrates seamlessly with organic recommendations such as new arrivals, basket-building content, and fast-delivery options. To execute this vision, we work in close partnership with Amazon Stores stakeholders to lead the expansion and growth of advertising across Amazon-owned and -operated pages beyond Search. We operate full stack—from backend ads-retail edge services, ads retrieval, and ad auctions to shopper-facing experiences—all designed to deliver meaningful value. Curious about our advertising solutions? Discover more about Sponsored Products and Sponsored Brands to see how we’re helping businesses grow on Amazon.com and beyond!
US, MA, Boston
The Artificial General Intelligence (AGI) team is seeking a dedicated, skilled, and innovative Applied Scientist with a robust background in machine learning, statistics, quality assurance, auditing methodologies, and automated evaluation systems to ensure the highest standards of data quality, to build industry-leading technology with Large Language Models (LLMs) and multimodal systems. Key job responsibilities As part of the AGI team, an Applied Scientist will collaborate closely with core scientist team developing Amazon Nova models. They will lead the development of comprehensive quality strategies and auditing frameworks that safeguard the integrity of data collection workflows. This includes designing auditing strategies with detailed SOPs, quality metrics, and sampling methodologies that help Nova improve performances on benchmarks. The Applied Scientist will perform expert-level manual audits, conduct meta-audits to evaluate auditor performance, and provide targeted coaching to uplift overall quality capabilities. A critical aspect of this role involves developing and maintaining LLM-as-a-Judge systems, including designing judge architectures, creating evaluation rubrics, and building machine learning models for automated quality assessment. The Applied Scientist will also set up the configuration of data collection workflows and communicate quality feedback to stakeholders. An Applied Scientist will also have a direct impact on enhancing customer experiences through high-quality training and evaluation data that powers state-of-the-art LLM products and services. A day in the life An Applied Scientist with the AGI team will support quality solution design, conduct root cause analysis on data quality issues, research new auditing methodologies, and find innovative ways of optimizing data quality while setting examples for the team on quality assurance best practices and standards. Besides theoretical analysis and quality framework development, an Applied Scientist will also work closely with talented engineers, domain experts, and vendor teams to put quality strategies and automated judging systems into practice.
US, MA, Boston
The Artificial General Intelligence (AGI) team is seeking a dedicated, skilled, and innovative Applied Scientist with a robust background in machine learning, statistics, quality assurance, auditing methodologies, and automated evaluation systems to ensure the highest standards of data quality, to build industry-leading technology with Large Language Models (LLMs) and multimodal systems. Key job responsibilities As part of the AGI team, an Applied Scientist will collaborate closely with core scientist team developing Amazon Nova models. They will lead the development of comprehensive quality strategies and auditing frameworks that safeguard the integrity of data collection workflows. This includes designing auditing strategies with detailed SOPs, quality metrics, and sampling methodologies that help Nova improve performances on benchmarks. The Applied Scientist will perform expert-level manual audits, conduct meta-audits to evaluate auditor performance, and provide targeted coaching to uplift overall quality capabilities. A critical aspect of this role involves developing and maintaining LLM-as-a-Judge systems, including designing judge architectures, creating evaluation rubrics, and building machine learning models for automated quality assessment. The Applied Scientist will also set up the configuration of data collection workflows and communicate quality feedback to stakeholders. An Applied Scientist will also have a direct impact on enhancing customer experiences through high-quality training and evaluation data that powers state-of-the-art LLM products and services. A day in the life An Applied Scientist with the AGI team will support quality solution design, conduct root cause analysis on data quality issues, research new auditing methodologies, and find innovative ways of optimizing data quality while setting examples for the team on quality assurance best practices and standards. Besides theoretical analysis and quality framework development, an Applied Scientist will also work closely with talented engineers, domain experts, and vendor teams to put quality strategies and automated judging systems into practice.
US, MA, Boston
The Artificial General Intelligence (AGI) team is seeking a dedicated, skilled, and innovative Applied Scientist with a robust background in machine learning, statistics, quality assurance, auditing methodologies, and automated evaluation systems to ensure the highest standards of data quality, to build industry-leading technology with Large Language Models (LLMs) and multimodal systems. Key job responsibilities As part of the AGI team, an Applied Scientist will collaborate closely with core scientist team developing Amazon Nova models. They will lead the development of comprehensive quality strategies and auditing frameworks that safeguard the integrity of data collection workflows. This includes designing auditing strategies with detailed SOPs, quality metrics, and sampling methodologies that help Nova improve performances on benchmarks. The Applied Scientist will perform expert-level manual audits, conduct meta-audits to evaluate auditor performance, and provide targeted coaching to uplift overall quality capabilities. A critical aspect of this role involves developing and maintaining LLM-as-a-Judge systems, including designing judge architectures, creating evaluation rubrics, and building machine learning models for automated quality assessment. The Applied Scientist will also set up the configuration of data collection workflows and communicate quality feedback to stakeholders. An Applied Scientist will also have a direct impact on enhancing customer experiences through high-quality training and evaluation data that powers state-of-the-art LLM products and services. A day in the life An Applied Scientist with the AGI team will support quality solution design, conduct root cause analysis on data quality issues, research new auditing methodologies, and find innovative ways of optimizing data quality while setting examples for the team on quality assurance best practices and standards. Besides theoretical analysis and quality framework development, an Applied Scientist will also work closely with talented engineers, domain experts, and vendor teams to put quality strategies and automated judging systems into practice.