Animation shows a flow of dots (historical data) flowing through a CloudTune forecasting icon to generate forecasts, it also includes some detailed shots of pretend peak event forecasts for the US and India.
CloudTune Forecasting, which uses past data to generate forecasts, was initially intended to help US service teams know how much computational capacity they needed for peak events. Since then, improvements have focused on differentiating across teams and regions around the world.

How CloudTune generates forecasts for the Amazon Store

The system has expanded from generating peak computation-load forecasts one year in advance to a series of forecasts that include per-minute forecasts several months into the future.

On what are known as game days to teams inside Amazon, millions of virtual “customers” log on to the Amazon Store to search for items, browse product pages, load shopping carts, and check out as if they were real customers hunting for bargains during a sale such as Prime Day.

Jeff Barr, chief evangelist for AWS, shares what he calls some of the "most interesting and/or mind-blowing metrics" from Prime Day.

“It’s like a fire drill, a planned practice,” said Molly McElheny, a principal technical program manager in Central Reliability Engineering at Amazon. McElheny is responsible for helping to oversee those game days, which her organization runs at strategically chosen times in advance of big sales. Their goal? Make sure the Amazon Store and the many teams who help it run smoothly are ready ahead of time for potentially massive spikes in traffic.

That planned practice draws on forecasts of traffic and loads on Amazon services generated by CloudTune, a system that serves as a communications vehicle between the teams who plan events such as Prime Day and service teams that own infrastructure components and help run the Amazon Store.

Related content
The SCOT science team used lessons from the past — and improved existing tools — to contend with “a peak that lasted two years”.

CloudTune Forecasting emanated from Amazon’s central economics team back in 2015 as an improved methodology for capacity planning to handle major events such as Prime Day and Black Friday, explained Oleksiy Mnyshenko, a senior manager and economist at Amazon.

“These events have large peak-to-mean spreads,” he noted. “This means we need to proactively model the expected peak load and continuously assess our AWS capacity needs to support it.”

Demand forecasting

The CloudTune Forecasting system has expanded over the years from generating peak computation-load forecasts one year in advance in the United States to a series of forecasts that range from per-week forecasts up to two years out to per-minute forecasts several months into the future. In addition, those forecasts — which are continually refreshed with new data — are now also generated for a wide variety of Amazon teams and regions around the world.

While the need for specific regional forecasts may be obvious — a Mother’s Day sale forecast in the United States will not be relevant for a Diwali sale in India — many unique service teams that support the Amazon Store also rely on these forecasts.

When you go to the Amazon Store, ... in the background, there are thousands of software systems that together constitute what the experience is, and all of these systems and teams owning them need to be ready for these peak events.
Oleksiy Mnyshenko

One team may be responsible for the home page in a specific region, whereas another team is responsible for the shopping cart experience there, and yet another handles the checkout process. Each team experiences traffic differently and, necessarily, consumes AWS computing power differently. Over time, teams at Amazon have collaborated to improve CloudTune forecasts to be useful for each of those teams and their specific concerns.

“When you go to the Amazon Store, it feels very seamless as you go from searching for something to navigating to details about the product to then checking out, but in the background, there are thousands of software systems that together constitute what the experience is, and all of these systems and teams owning them need to be ready for these peak events,” Mnyshenko said.

In the early years, CloudTune forecasts were geared primarily to help service teams know how much computational capacity they needed for peak events. Since then, improvements have focused on differentiating across teams and regions. As the Amazon Store continued to grow, it became important to extend demand outlook to a two-years-out aggregate forecast per region to help inform decisions for AWS related to computing power, networking, and data center planning.

Related content
The story of a decade-plus long journey toward a unified forecasting model.

“A data center is not built in a day,” noted Chunpeng Wang, a senior applied scientist at Amazon who works on the CloudTune forecast team. “Our forecasts are an important input into long-term capacity planning for AWS.”

What’s more, the Amazon Store is not alone in contending with peak events, noted Ben Mildenhall, a senior manager in cloud computing and auto scaling.

“Many AWS external customers have Black Friday and Cyber Monday events as well,” Mildenhall said. “So it’s important we optimize to give all of our customers a great experience.”

CloudTune forecasts provide inputs to AWS to help size infrastructure in a way that maximizes utilization efficiency, noted Mnyshenko. “The way CloudTune specifically helps here is continuously getting better at anticipating the mix of capacity we’re using by generation, by type, by location, so that we can have those conversations and provide this feedback to AWS,” he said.

Granular, flexible, and explainable

Like many demand-forecasting applications, CloudTune is a time-series forecasting system. What’s unique about it is the ability to predict demand at one-minute granularity, noted Mnyshenko. This level of granularity provides insight into patterns such as short-duration spikes in website traffic. Teams use the forecasts as inputs to determine their computing capacity not just for peak events like back to school but also peak times during any given day, week, or month.

“Our comparative advantage is intra-day load predictions at one-minute granularity, allowing us to track actuals during peak events, highlighting these sharp edges where checkout spikes way beyond the natural peak for the period,” Mnyshenko said.

In addition, CloudTune forecasts need to be flexible to accommodate changes in the day and duration of events, such as the evolution of Prime Day from a 24-hour event to a 48-hour event on different days each year.

Related content
Part-time sabbatical plan turns into full-time role for author of five books and more than 170 research articles.

At other times, CloudTune needs to make forecasts for special events such as the launch of popular gaming consoles, which may sell out in a matter of minutes.

“That can create a huge spike, and we have to predict the traffic spike and the order spike,” explained Ebrahim Nasrabadi, a senior manager of applied science who leads the CloudTune Forecasting science team.

The team responsible for CloudTune Forecasting has developed modular and configurable models to address these and other challenges, he noted.

For example, built-in functionality allows the removal of outliers — due to things such as a spike in robot traffic that can decrease or increase actual website traffic and order rate unexpectedly — from predictable seasonal behavior and known calendar events. Since these interruptions do not regularly occur, the tool allows forecast teams to exclude those outliers from data used in the forecast.

“Our models are simple and quite flexible to include additional variables and seasonality,” noted Nasrabadi. The models also take into account significant changes in a trend within a dataset, also known as a slope break.

The CloudTune team also emphasizes forecast models that are explainable.

“We have to be very crisp about what we are doing, very transparent about our expectations,” said Wang.

Hundreds of Amazon Store software teams use these forecasts to help determine their AWS capacity needs for peak events. The better these teams understand the forecasts, the more trust they have in them, noted Mnyshenko.

“We need to be able to explain what goes into the ingredients and, more importantly, what we are doing to reduce the spread in errors,” he said.

Continuous automation

Currently, service teams not yet using automation enhancements take the CloudTune forecasts and translate them into capacity orders for servers through the Amazon Elastic Compute Cloud (Amazon EC2) using many different manual tools and processes, said Doug Smith, a senior technical program manager responsible for delivering improvements and features to the CloudTune toolset.

A key future direction for CloudTune is to continuously enhance these tools and automate as many manual processes as possible, Smith noted.

The world we’re envisioning between our team and CloudTune is one where services teams don’t have to worry about scaling at all.
Molly McElheny

“We’re moving into automation so that we can take our CloudTune forecasts as inputs into these new products that we’re building to provide a hands-off experience,” he said.

And while the game days McElheny’s team runs in advance of these major events will continue apace, she has a vision for the future there as well. Today, she said, the forecasts enable simulations of high-level customer journeys. She’d like to get to a forecast that allows her team to simulate an event down to the types of products customers are ordering when and where.

“This matters because different services get called depending on a lot of different factors. The closer we can simulate the real traffic the better, because we’re actually hitting services with the traffic they expect to see during the event,” McElheny said.

To get there, McElheny, Smith, and their colleagues work together to make sure the forecasts provide the best data for the most realistic simulations.

“The world we’re envisioning between our team and CloudTune is one where services teams don’t have to worry about scaling at all,” McElheny said. “CloudTune does it for them, and then we run a game day, and as we find issues during game day, CloudTune goes and places orders to scale things up for those customers.”

Research areas

Related content

US, NY, New York
We are seeking a Robotics/AI Motor Control Scientist to develop cutting-edge machine learning algorithms for motor control systems in robots. In this role, you will focus on creating and optimizing intelligent motor control strategies to enable robots to perform complex, whole-body tasks. Your contributions will be essential in advancing robotics by enabling fluid, reliable, and safe interactions between robots and their environments. Key job responsibilities - Develop controllers that leverage reinforcement learning, imitation learning, or other advanced AI techniques to achieve natural, robust, and adaptive motor behaviors - Collaborate with multi-disciplinary teams to integrate motor control systems with robotic hardware, ensuring alignment with real-world constraints such as actuator dynamics and energy efficiency - Use simulation and real-world testing to refine and validate control algorithms - Stay updated on advancements in robotics, AI, and control systems to apply advanced techniques to robotic motion challenges - Lead technical projects from conception through production deployment - Mentor junior scientists and engineers - Bridge research initiatives with practical engineering implementation About the team Fauna Robotics, an Amazon company, is building capable, safe, and genuinely delightful robots for everyday life. Our goal is simple: make robots people actually want to live and interact with in everyday human spaces. We believe that future won’t arrive until building for robotics becomes far more accessible. Today, too much effort is spent reinventing the fundamentals. We’re changing that by developing tightly integrated hardware and software systems that make it faster, safer, and more intuitive to create real-world robotic products. Our work spans the full stack: mechanical design, control systems, dynamic modeling, and intelligent software. The focus is not just functionality, but experience. We’re building robots that feel responsive, expressive, and genuinely useful. At Fauna, you’ll work at the frontier of this space, helping define how robots move, manipulate, and interact with people in natural environments. It’s an opportunity to solve hard problems across hardware and software with a team focused on making robotics accessible and joyful to build. If you care about making robotics real for everyone and building systems that are as delightful as they are capable, we’re interested in hearing from you. an opportunity to solve hard problems across hardware and software with a team focused on making robotics accessible and joyful to build. If you care about making robotics real for everyone and building systems that are as delightful as they are capable, we’re interested in hearing from you.
US, WA, Seattle
Prime Video is a first-stop entertainment destination offering customers a vast collection of premium programming in one app available across thousands of devices. Prime members can customize their viewing experience and find their favorite movies, series, documentaries, and live sports – including Amazon MGM Studios-produced series and movies; licensed fan favorites; and programming from Prime Video subscriptions such as Apple TV+, HBO Max, Peacock, Crunchyroll and MGM+. All customers, regardless of whether they have a Prime membership or not, can rent or buy titles via the Prime Video Store, and can enjoy even more content for free with ads. Are you interested in shaping the future of entertainment? Prime Video's technology teams are creating best-in-class digital video experience. As a Prime Video team member, you’ll have end-to-end ownership of the product, user experience, design, and technology required to deliver state-of-the-art experiences for our customers. You’ll get to work on projects that are fast-paced, challenging, and varied. You’ll also be able to experiment with new possibilities, take risks, and collaborate with remarkable people. We’ll look for you to bring your diverse perspectives, ideas, and skill-sets to make Prime Video even better for our customers. With global opportunities for talented technologists, you can decide where a career Prime Video Tech takes you! Key job responsibilities As a highly experienced and seasoned science leader, you will apply state of the art natural language processing and computer vision research to video centric digital media, while also responsible for creating and maintaining the best environment for applied science in order to recruit, retain and develop top talent. You will lead the research direction for a team of deeply talented applied scientists, creating the roadmaps for forward-looking research and communicate them effectively to senior leadership. You will also hire and develop applied scientists - growing the team to meet the evolving needs of our customers. About the team This team's mission is to deeply understand all content and empower all customers with relevant language options, innovative accessibility assists, and rich title-information across all their content-experiences on Prime Video. We create and publish content on-time that's meaningful, accurate, and accessible to every customer globally. We delight our customers by pushing the boundaries of content understanding and enrichment. Through inclusion and innovation, we do the most fulfilling work of our career.
GB, MLN, Edinburgh
Do you want to make a real difference to real people's lives? Want to design and build fair and explainable systems which automate recruitment processes across Amazon? Come and be part of a team that develops new machine learning (ML) technologies, which help Amazon scale for its customers by recruiting diverse teams. Join our Recommendations team within Intelligent Talent Acquisition (ITA) where you’ll build machine learning products that transform how job seekers find opportunities and recruiters discover talent. You’ll develop sophisticated recommendation systems powering both Amazon Jobs and internal hiring platforms, operating at global scale to match the right people with the right positions. Using techniques including representation learning, reinforcement learning, and probabilistic modeling, your work will directly improve efficiency for recruiters and help candidates find their ideal roles. This position offers the chance to solve complex problems with significant impact by creating systems that make Amazon’s entire hiring ecosystem more effective while collaborating with scientists across the organization. Key job responsibilities - Design and implement machine learning models that power recommendation systems for job seekers and recruiters, ensuring high performance, scalability, and reliability at global scale. Our ideal candidate has a strong scientific foundation and experience of statistical analysis and model building and has a passion for fairness and explainability in ML systems. - Collaborate with engineers, scientists, and product managers to define requirements, create solutions, and deliver products that improve the hiring experience. - Participate in the full software development lifecycle including scoping, design, coding, testing, documentation, deployment, and maintenance of recommendation systems and ML models. - Solve complex ML problems using optimal data structures and algorithms, making thoughtful trade-offs between efficiency and maintainability. - Stay current with scientific literature and develop novel approaches that address business challenges in talent acquisition. You will have the opportunity to provide feedback on scientific work across the organization helping the entire Intelligent Talent Acquisition organization improve. A day in the life You might spend the morning reviewing a colleague’s code for a new recommendation algorithm feature, then collaborate with product managers to refine requirements for an upcoming enhancement. After lunch, you’ll dive into model development, analyzing performance metrics from recent A/B tests and implementing improvements to the job-seeker recommendation pipeline. Throughout the day, you’ll participate in scientific discussions with peers across the organization, providing valuable feedback while continuing to refine your expertise. About the team The Recommendations team is a hybrid group of software engineers and applied scientists located in Edinburgh. We build tools that match people to jobs and jobs to people, optimizing experiences for both recruiters and candidates. Our work directly impacts Amazon’s ability to find and hire exceptional talent globally. The team maintains a collaborative environment with regular knowledge sharing and mentorship opportunities. We work closely with our product teams to understand business needs and develop innovative scientific solutions that improve hiring outcomes across both industry and student requisitions worldwide.
US, CA, Pasadena
We are seeking an Applied Scientist to join the SAF Lab. In this role, you will lead the effort in safe reinforcement learning (RL) including the development of legged locomotion algorithms that internalize safety and are deployable on physical hardware—enabling highly dynamic robots to walk, run, avoid collisions and recover from disturbances with agility and robustness. You will develop RL architectures that interface with physics-based models (for dynamic retargeting and reward shaping), internalize safety constraints in training, sim-to-real transfer and interface with safety filters at run-time. Therefore, your work will sit at the intersection of safety-critical control and learning, and you will collaborate with others in the SAF Lab and Amazon working on perception, planning, whole-body and safety-critical control. This is an opportunity to shape the foundations of safe learning on emerging platforms that will remove bottlenecks to deployment and enable these robots to safely operate around humans. Key job responsibilities • Collaborate with product teams and science leaders to set a science roadmap (with eventual impact on real robots). • Design, train, and deploy reinforcement learning (RL) policies for dynamic legged locomotion including walking, running, stair climbing, and fall recovery on physical robots • Develop sim-to-real transfer pipelines that produce policies robust to the reality gap, including domain randomization, system identification, and adaptive strategies • Integrate control-based methods with RL, as inputs to the RL (dynamic retargeting and control-guided rewards), in training (internalizing safety constraints in training), and as the RL feeds into safety layers and whole-body control • Develop and maintain large-scale training infrastructure for locomotion policy learning, including physics simulation environments, domain randomization and GPU parallelization • Investigate the distillation of locomotion policies, integration with whole-body control, foundation models, VLAs, world models, perception and full-stack autonomy • Evaluate policy performance rigorously through simulation benchmarks, hardware experiments, and failure-mode analysis • Publish research at top-tier robotics and ML venues and contribute to Amazon's scientific reputation in advanced robotics • Collaborate with perception and planning teams to enable terrain-aware and goal-conditioned locomotion behaviors A day in the life Amazon offers a full range of benefits that support you and eligible family members, including domestic partners and their children. Benefits can vary by location, the number of regularly scheduled hours you work, length of employment, and job status such as seasonal or temporary employment. The benefits that generally apply to regular, full-time employees include: 1. Medical, Dental, and Vision Coverage 2. Maternity and Parental Leave Options 3. Paid Time Off (PTO) 4. 401(k) Plan If you are not sure that every qualification on the list above describes you exactly, we'd still love to hear from you! At Amazon, we value people with unique backgrounds, experiences, and skillsets. If you’re passionate about this role and want to make an impact on a global scale, please apply! About the team Work with the inventor of control barrier functions in the Safe Autonomy Frontiers (SAF) Lab. The first industry research lab in safe autonomy, developing a universal safety layer for the next generation of robotic systems: mobile robots, manipulators, mobile manipulators, and future platforms with dynamic stability. You will push the frontiers of performant safety for highly dynamic robots: CBF theory integrated with perception and learning, evaluated on next-generation robots. Your work will underpin robots operating alongside people at Amazon's unprecedented scale.
US, WA, Redmond
We are searching for a talented candidate with expertise in orbital mechanics and spaceflight navigation, including LEO Satellite Orbit Determination. This position requires experience in simulation and analysis of spacecraft orbital mechanics and sequential orbit determination methods, including Extended Kalman Filters (EKF) and/or Unscented Kalman Filter (UKF). Strong analysis skills are required to develop engineering studies of complex large-scale dynamical systems. This position requires demonstrated expertise in computational analysis automation and tool development. Key job responsibilities - Perform spacecraft maneuver or navigation analysis in support of multi-disciplinary trades within the Amazon Leo team. - Contribute to prototype software development of flight algorithms. - Test and assess navigation software for integration into flight systems. - Assess and trouble-shoot the performance of Leo on-board GNSS hardware and software systems. - Work closely with GNC engineers to manage on-orbit performance and develop flight dynamics operations processes. Export Control Requirement: Due to applicable export control laws and regulations, candidates must be a U.S. citizen or national, U.S. permanent resident (i.e., current Green Card holder), or lawfully admitted into the U.S. as a refugee or granted asylum. A day in the life - Interacting with GNC teams to evaluate and troubleshoot satellite issues. - Working within the Flight Dynamics Research team to prioritize tasks. - Performing analysis, simulation, testing and documentation to address assigned tasks.
BR, SP, Sao Paulo
Are you passionate about helping customers achieve business transformation through AI? Do you want to lead forward-deployed teams that embed directly into the enterprise and unlock real business outcomes? And are you ready to operate as a general manager across engineering, science, and commercial strategy in the fastest-moving space in AI and infrastructure? The AWS Generative AI Innovation Center (GenAIIC) is on a mission to accelerate enterprise AI transformation across global customers going from beyond isolated use cases to holistic, C-suite-sponsored initiatives that reshape how organizations operate. We combine deep AI expertise across science, strategy, and business transformation. We start with the customer's most critical operational challenges and work backwards, and deploy multidisciplinary teams that embed with the customer, prove impact in 45-day sprints, and expand across the enterprise. We are a fast-moving, entrepreneurial team that values leaders who can operate across technical depth and commercial breadth. You will lead a team of ML engineers, AI scientists, and AI strategists who work alongside customers to architect and deliver AI solutions that move and stay in production, realizing value. You will regularly engage with CFOs, CIOs, and C-suite executives. You must bring equal fluency in engineering, data science, go-to-market, and customer delivery. You are ready to roll up your sleeves alongside the team, whether that means scoping an agentic AI architecture, presenting to a board, or operationalizing a repeatable delivery motion. You will partner with customers, AWS Sales, AWS service teams, AWS industry teams and AWS Professional Services delivery teams to meet the specific needs of the customer, and extend that use to other customers. The successful candidate will possess both technical and customer-facing skills that will allow you to be the technical “face” of AWS within our solution providers’ ecosystem/environment as well as directly to end customers. You will be able to drive discussions with senior technical and management personnel within customers and partners, as well as the technical background that enables them to interact with and give guidance to data/research/applied scientists and software developers. The ideal candidate will also have a demonstrated ability to think strategically about business, product, and technical issues. Finally, and of critical importance, the candidate will be an excellent technical team manager, someone who knows how to hire, develop, and retain high quality technical talent. About the team Diverse Experiences AWS values diverse experiences. Even if you do not meet all of the qualifications and skills listed in the job description, we encourage candidates to apply. If your career is just starting, hasn’t followed a traditional path, or includes alternative experiences, don’t let it stop you from applying. Why AWS? Amazon Web Services (AWS) is the world’s most comprehensive and broadly adopted cloud platform. We pioneered cloud computing and never stopped innovating — that’s why customers from the most successful startups to Global 500 companies trust our robust suite of products and services to power their businesses. Inclusive Team Culture Here at AWS, it’s in our nature to learn and be curious. Our employee-led affinity groups foster a culture of inclusion that empower us to be proud of our differences. Mentorship & Career Growth We’re continuously raising our performance bar as we strive to become Earth’s Best Employer. That’s why you’ll find endless knowledge-sharing, mentorship and other career-advancing resources here to help you develop into a better-rounded professional. Work/Life Balance We value work-life harmony. Achieving success at work should never come at the expense of sacrifices at home, which is why flexible work hours and arrangements are part of our culture. When we feel supported in the workplace and at home, there’s nothing we can’t achieve in the cloud.
US, WA, Seattle
Join the AWS Perimeter Protection team as a Senior Applied Scientist, where you will bring your deep ML engineering expertise to design, build, and scale AI-driven security solutions that protect AWS customers worldwide. This role is ideal for someone who has already built and shipped production ML systems at industry scale and is looking to apply that experience to high-impact security challenges. You will own the full ML lifecycle — from research and prototyping to production deployment and optimization — powering services including Web Application Firewall, DDoS Protection, Bot Management, and Infrastructure Protection. With services spanning all AWS regions and handling trillions of requests per week, you will solve complex engineering and science problems where model performance, system reliability, and low-latency inference are critical. Key job responsibilities - Design, build, and deploy production-grade ML models and systems for real-time threat detection, mitigation, and protection against evolving cyber threats at cloud scale. - Own the full ML lifecycle end-to-end — from problem formulation, data engineering, and model development through to production deployment, monitoring, and continuous improvement. - Architect and optimize ML pipelines, training infrastructure, and serving systems to meet strict latency, throughput, and reliability requirements at AWS scale. - Bridge the gap between research and production by translating novel ML approaches into robust, scalable, and maintainable systems that operate in real-time security environments. - Design and implement feature engineering workflows and large-scale data processing pipelines to support rapid experimentation and reliable model iteration. - Collaborate closely with software engineering teams to integrate ML models into distributed, low-latency security services, driving engineering decisions around model serving, infrastructure, and system design. - Analyze large-scale production data to identify patterns, anomalies, and emerging threat vectors, and translate findings into measurable improvements to detection and mitigation capabilities. - Establish and improve best practices for ML system design, model evaluation, A/B testing, and production monitoring across the team. - Mentor junior scientists and engineers, raising the bar on both scientific rigor and engineering quality.
US, WA, Seattle
The Annapurna ML team is looking for a Senior Applied Scientist to work on the intersection of Artificial Intelligence and program analysis to raise the code quality bar in our state-of-the-art deep learning compiler stack. This stack is designed to optimize application models across diverse domains, including Large Language and Vision, originating from leading frameworks such as PyTorch, TensorFlow, and JAX. Your role will involve working closely with our custom-built Machine Learning accelerators, Inferentia and Trainium, which represent the forefront of Annapurna innovation for advanced ML capabilities, and is the underpinning of Generative AI. As a Senior Applied Scientist, you'll be instrumental in designing, developing, and deploying analyzers for ML compiler stages and compiler IRs. You will architect and implement business-critical tooling, publish research, and mentor a brilliant team of experienced scientists and engineers. You will need to be technically capable, credible, and curious in your own right as a trusted scientist, innovating on behalf of our customers. Your responsibilities will involve tackling crucial challenges alongside a talented engineering team, contributing to leading-edge design and research in compiler technology and deep-learning systems software. Strong experience in programming languages, compilers, program analyzers, and program synthesis engines will be a benefit in this role. A background in machine learning and AI accelerators is preferred but not required. A day in the life Diverse Experiences Amazon values diverse experiences. Even if you do not meet all of the qualifications and skills listed in the job description, we encourage candidates to apply. If your career is just starting, hasn’t followed a traditional path, or includes alternative experiences, don’t let it stop you from applying. Why Amazon? We pioneered cloud computing and never stopped innovating — that’s why customers from the most successful startups to Global 500 companies trust our robust suite of products and services to power their businesses. Inclusive Team Culture Here at Amazon, it’s in our nature to learn and be curious. Our employee-led affinity groups foster a culture of inclusion that empower us to be proud of our differences. Ongoing events and learning experiences, including our Conversations on Race and Ethnicity (CORE) and AmazeCon (diversity) conferences, inspire us to never stop embracing our uniqueness. Mentorship & Career Growth We’re continuously raising our performance bar as we strive to become Earth’s Best Employer. That’s why you’ll find endless knowledge-sharing, mentorship and other career-advancing resources here to help you develop into a better-rounded professional. Work/Life Balance We value work-life harmony. Achieving success at work should never come at the expense of sacrifices at home, which is why we strive for flexibility as part of our working culture. When we feel supported in the workplace and at home, there’s nothing we can’t achieve in the cloud.
US, WA, Seattle
The Automated Reasoning Group in the Amazon Neuron team is looking for an Applied Scientist to work on the intersection of Artificial Intelligence and program analysis to raise the code quality bar in our state-of-the-art deep learning compiler stack. This stack is designed to optimize application models across diverse domains, including Large Language and Vision, originating from leading frameworks such as PyTorch and JAX. Your role will involve working closely with our custom-built Machine Learning accelerator, Trainium, which represents the forefront of innovation for advanced ML capabilities, and is the underpinning of Generative AI. In this role as an Applied Scientist, you'll be instrumental in designing, developing, and deploying analyzers for ML compiler stages and compiler IRs. You will architect and implement business-critical tooling, publish research, and mentor a brilliant team of experienced scientists and engineers. You will need to be technically capable, credible, and curious in your own right as a trusted AWS Neuron engineer, innovating on behalf of our customers. Your responsibilities will involve tackling crucial challenges alongside a talented engineering team, contributing to leading-edge design and research in compiler technology and deep-learning systems software. Strong experience in programming languages, compilers, program analyzers, theorem provers, and program synthesis engines will be a benefit in this role. A background in machine learning and AI accelerators is preferred but not required.
US, WA, Seattle
The Shopping Convo Foundations Team - Pre-purchases Science is looking for an Senior Applied Scientist with expertise in Artificial Intelligence and Machine Learning to drive scientific innovation that expands Amazon's product catalogue. Our goal is to leverage AI/ML solutions to enhance catalogue coverage with high precision. In this role, you will lead the research and development of novel machine learning approaches to solve complex catalogue expansion and product attribute challenges. You lead the design and develop state-of-the-art ML models, conduct rigorous experimentation, translate scientific breakthroughs into production-ready solutions, and guide a set of junior scientists. You will work closely with ML Engineers and Software Development Engineers to optimize model performance, ensure scalability, and deploy low-latency solutions at Amazon scale. About the team Our team is a horizontal applied science group that works across the full lifecycle of brand and product data extraction and quality. We span seven workstreams — from products sourcing and Brand entitlement, to designing and evaluating extraction strategies for ASINs, offers, and brand attributes at scale, as well as relevance modeling and search. We drive root cause analysis through human-in-the-loop evaluation, improve how catalog data surfaces in search, optimize business metrics tied to data quality, and build brand intelligence capabilities. This cross-cutting scope positions the team as a connective layer across product, engineering, and science — ensuring that improvements in one area compound across the system rather than remain siloed.