Amazon Halo Rise advances the future of sleep

Built-in radar technology, deep domain adaptation for sleep stage classification, and low-latency incremental sleep tracking enable Halo Rise to deliver a seamless, no-contact way to help customers improve sleep.

The benefits of quality sleep are well documented, and sleep affects nearly every aspect of our physical and emotional well-being. Yet one in three adults doesn’t get enough sleep. Given Amazon’s expertise in machine learning and radar technology innovation, we wanted to invent a device that would help customers improve their sleep by looking holistically at the factors that contribute to a good night’s rest.

That’s why we’re excited to announce that Amazon has unveiled its first dedicated sleep device — Halo Rise, a combined bedside sleep tracker, wake-up light, and smart alarm. Powered by custom machine learning algorithms and a suite of built-in sensors, Halo Rise accurately determines users’ sleep stages and provides valuable insights that can be used to optimize their sleep, including information about their sleep environments. Halo Rise has no sensors to wear, batteries to charge, or apps to open. And since a good wake-up experience is core to good sleep, Halo Rise features a wake-up light and smart alarm, designed to help customers start the day feeling rested and alert.

Halo Rise in action
A built-in radar sensor uses ultralow-power radio signals to sense respiration and movement patterns and determine sleep stages.

Designing with customer trust as our foundation

Customer privacy and safety are foundational to Halo Rise, and that's evident in both the hardware design and the technologies used to power the experience. Halo Rise features neither a camera nor a microphone and instead relies on ambient radar technology and machine learning to accurately determine sleep stages: deep, light, REM (rapid eye movement), and awake.

The technology at the core of Halo Rise is a built-in radar sensor that safely emits and receives an ultralow-power radio signal. The sensor uses phase differences between reflected signals at different antennas to measure movement and distance. Through on-chip signal processing, Halo Rise produces a discrete waveform corresponding to the user’s respiration. The device cannot detect noise or visual identifiers associated with an individual user, such as body images.

Using built-in radar technology enables us to prioritize customer privacy while still delivering accurate measurements and useful results. Customers have the option to manually put Halo Rise into Standby mode, which turns off the device’s ability to detect someone’s presence or track sleep.

Halo Rise hardware design
Halo Rise features a suite of sensors to accurately track your sleep and measure your room’s temperature, humidity, and light levels. 

Intuitive and accurate experience

To design the sleep-tracking algorithm that powers Halo Rise, we thought about the most common bedtime behaviors and the ways in which customers and their families (pets included) might engage with the bedroom. This led us to innovate on five main technological fronts:

  • Presence detection: Halo Rise activates its sleep detection only when someone is in range of the sensor. Otherwise, the device remains in a monitoring mode, where no data is transmitted to the cloud.
  • Primary-user tracking: Halo Rise distinguishes the sleep of the primary user (the user closest to the device) from that of other people or pets in the same bed, even though the respiration signal cannot be associated with individual users.
  • Sleep intent detection: Halo Rise detects when the user first starts trying to sleep and distinguishes that attempt from other in-bed activities — such as reading or watching TV — to accurately measure the time it takes to fall asleep, an important indicator of sleep health.
  • Sleep stage classification: Halo Rise reliably correlates respiration-driven movement signals with sleep stages.
  • Smart-alarm integration: During the user’s alarm window, the Halo Rise smart alarm checks the user’s sleep stage every few minutes to detect light sleep, while also maximizing sleep duration.
Halo-Vienna-MM_Wave-Chart.png
A combination of breathing and movement patterns enables Halo Rise to determine the primary user for the sleep session and to measure that person’s sleep throughout the night.

Presence detection

Halo Rise has an easy setup process. To get started, a customer will place Halo Rise on their bedside table facing their chest and note in the Amazon Halo app what side of the bed they sleep on — and that’s it: Halo Rise is ready to go. The radar sensor detects motion within a 3-D geometric volume that fans out from the sensor, an area called the detection zone. Within this zone, the presence detection algorithm estimates the location of the bed and an “out-of-bed” area between the bed and the device.

On-chip algorithms detect the motion and location of respiration events within the detection zone. In both cases — motion and respiration — the algorithm evaluates the quality of the signals. On that basis, it computes a score indicating its confidence that the readings are reliable and a user is present. Only if the confidence score crosses a reliability threshold does Halo Rise begin streaming sensor data to the cloud, where it is processed by the primary-user-tracking algorithm.

Radar Fan.png
The Halo Rise detection zone is the region within which the radar sensor senses motion and location.

Primary-user tracking

We know that many of our customers share their beds, be it with other people or with pets, so our algorithms are designed to track the sleep of only the primary user. Halo Rise starts a sleep session after it detects someone’s presence within the detection zone for longer than five minutes. From there, the primary-user-tracking algorithm runs continuously in the background, sensing the closest user’s sleep stages. As long as the user sleeps on their side of the bed, and their partner sleeps on the other side, Halo Rise will track the primary user’s sleep quality irrespective of who comes to bed first and who leaves the bed last.

During the sleep session, Halo Rise dynamically monitors changes in the user’s distance from the sensor, the respiration signal quality, and abrupt changes in respiration patterns that indicate another person’s presence. These changes cause the algorithm to reassess whether it’s actually sensing the intended user and to ignore the data unrelated to the primary user. For instance, if the user gets into bed after their partner has already fallen asleep, or if they use the restroom in the middle of the night, Halo Rise detects that and adjusts the sleep results accordingly.

Sleep intent detection

Another big algorithmic challenge we faced was determining when a user is quietly sitting in bed reading their Kindle or watching TV rather than trying to fall asleep. The time it takes to fall asleep (also known as sleep latency) is an important indicator of sleep health. Too short of a time may result from sleep deprivation, while too long of a time may be due to difficulty winding down.

To address this problem, we used a combination of presence and primary-user tracking along with a machine-learning model trained and evaluated on tens of thousands of hours of sleep diaries to accurately identify when the user is trying to sleep. The model uses sensor data streamed from the device — including respiration, movement, and distance — to generate a sleep intent score. The score is then post-processed by a regularized change-point detection algorithm to determine when the user is trying to fall asleep or wake up.

Halo Rise Sleep Intent v2.png
A machine learning model trained on thousands of hours of sleep uses respiration, movement, and distance data to generate a sleep intent score.

Sleep stage classification

Wearable health trackers like Halo Band and Halo View use heart rate and motion signals to determine sleep stages during the night, but Halo Rise uses respiration. To learn how to reliably recognize those stages, we needed to develop new machine learning models.

We pretrained a deep-learning model to predict sleep stages using a rich and diverse clinical dataset that included tens of thousands of hours of sleep collected by academic and research sources. The research included sleep data measured using the clinical gold standard, polysomnography (PSG). PSG studies use a large array of sensors attached to the body to measure sleep, including respiratory inductance plethysmography (RIP) sensors, whose output is analogous to the respiration data measured by Halo Rise.

Pretraining the model to predict sleep stages from RIP sensors enabled it to develop meaningful representations of the relationship between respiration and sleep prior to additional training on radar datasets collected alongside PSG. To collect radar training data for the models, we partnered with sleep clinics to conduct thousands of hours of PSG studies. Ultimately, this enables our models to classify sleep stages using just a built-in radar in the comfort of a customer’s home.

Halo_hypnogram.png
In the morning, customers can access a sleep hypnogram that provides a detailed breakdown of time spent in each sleep stage throughout the night.

A smarter wake-up experience

When woken naturally during a light sleep stage, people are most likely to feel rested, refreshed, and ready to tackle the day. Consequently, Halo Rise features a wake-up light, which gently simulates the colors and gradual brightening of a sunrise, and a smart alarm. Customers can also set an audible smart alarm that’s integrated with our sleep stage classification algorithms, optimizing their wake experience. Ahead of their scheduled wake-up time, the audible smart alarm monitors their sleep stages and wakes them up at their ideal time for getting up. This combination of wake-up light and smart alarm is shown to increase cognitive and physical performance throughout the day.

The smart-alarm algorithms are trained around two factors: sensing when the user is in light sleep and maximizing the user’s sleep duration. For the first component, Halo Rise needs to continuously monitor sleep stages during the alarm window — the 30 minutes before a user’s scheduled alarm — to identify when the user has entered a light sleep stage, known as the “wake window.”

At this phase, our algorithms work to sense “wakeable events,” such as a change in motion or breathing. This requires incrementally computing sleep stages to trigger the alarm with low latency. Unlike many sleep algorithms, Halo Rise does not require data from the entirety of the sleep session to classify sleep stages, allowing predictions to be used directly for alarm triggers as data is streamed.

For the second component, the system’s models are trained to predict the latest moment to trigger the alarm during the wake window. This ensures that as the user drifts between sleep stages, they are getting those crucial minutes of additional sleep before the alarm goes off.

The Halo Rise wake-up light
Halo Rise identifies a “wake window” when the user is in light sleep, while also maximizing sleep duration before activating an audible smart alarm.

A solution you can trust

To evaluate our machine learning algorithms, we collected thousands of hours of sleep studies comparing Halo Rise to PSG for over a hundred sleepers, developed with input from leading sleep labs. While sleep studies are typically conducted in sleep labs, we performed in-home PSG studies at participants’ homes under supervision of registered PSG technologists to test the device in naturalistic settings.

We used three different registered PSG technologists to reliably annotate ground truth sleep stages per the American Academy of Sleep Medicine’s scoring rules. We then compared Halo Rise’s outputs to the ground truth sleep data across 14 different sleep metrics — including time asleep, time awake, time to fall asleep, and accuracy for every 30 seconds — following analysis guidelines from a standardized framework for sleep stage classification assessment. This evaluation was supplemented by thousands of sleep diaries from our beta trials, expanding our evaluation to a diverse population of adults to account for variations in preferred sleep postures, age, body shapes, and other background conditions.

What’s next?

As we look to invent new products that help our customers live better longer, Halo Rise is an important step in giving our customers greater agency over their health and well-being. By looking holistically at the end-to-end sleep experience — not just going to sleep but also getting up in the morning — Halo Rise unlocks an entirely new way for customers to understand and manage sleep. We’re excited to help them make sense of valuable sleep data, from the quality and quantity of their sleep to their room’s environment, and deliver actionable insights and resources to improve it in the future. Halo Rise is just getting started, and we are going to learn from our customers how this technology can continue to evolve and become even more personalized to better meet their needs.

Research areas

Related content

US, MA, N.reading
Amazon is seeking exceptional talent to help develop the next generation of advanced robotics systems that will transform automation at Amazon's scale. We're building revolutionary robotic systems that combine cutting-edge AI, sophisticated control systems, and advanced mechanical design to create adaptable automation solutions capable of working safely alongside humans in dynamic environments. This is a unique opportunity to shape the future of robotics and automation at an unprecedented scale, working with world-class teams pushing the boundaries of what's possible in robotic dexterous manipulation, locomotion, and human-robot interaction. This role presents an opportunity to shape the future of robotics through innovative applications of deep learning and large language models. At Amazonwe leverage advanced robotics, machine learning, and artificial intelligence to solve complex operational challenges at an unprecedented scale. Our fleet of robots operates across hundreds of facilities worldwide, working in sophisticated coordination to fulfill our mission of customer excellence. The ideal candidate will contribute to research that bridges the gap between theoretical advancement and practical implementation in robotics. You will be part of a team that's revolutionizing how robots learn, adapt, and interact with their environment. Join us in building the next generation of intelligent robotics systems that will transform the future of automation and human-robot collaboration. Key job responsibilities - Collaborate with simulation and robotics experts to translate physical modeling needs into robust, scalable, and maintainable simulation solutions. - Design and implement high-performance simulation modeling and tools for rigid and deformable body simulation. - Identify and optimize performance bottlenecks in simulation pipelines to support real-time and batch simulation workflows. - Help build validation and unit testing pipelines to ensure correctness and physical fidelity of simulation results. - Identify potential sources of sim-to-real gaps and propose modeling and numerical approximations to reduce them. - Stay current with the latest advances in numerical methods, parallel computing, and GPU architectures, and incorporate them into our tools.
US, MA, North Reading
robotics systems that will transform automation at Amazon's scale. We're building revolutionary robotic systems that combine cutting-edge AI, sophisticated control systems, and advanced mechanical design to create adaptable automation solutions capable of working safely alongside humans in dynamic environments. This is a unique opportunity to shape the future of robotics and automation at unprecedented scale, working with world-class teams pushing the boundaries of what's possible in robotic manipulation, locomotion, and human-robot interaction. This role presents an opportunity to shape the future of robotics through innovative applications of deep learning and large language models. At Amazon Industrial Robotics we leverage advanced robotics, machine learning, and artificial intelligence to solve complex operational challenges at unprecedented scale. Our fleet of robots operates across hundreds of facilities worldwide, working in sophisticated coordination to fulfill our mission of customer excellence. We are pioneering the development of robotics foundation models that: Enable unprecedented generalization across diverse tasks Enable unprecedented robustness and reliability, industry-ready Integrate multi-modal learning capabilities (visual, tactile, linguistic) Accelerate skill acquisition through demonstration learning Enhance robotic perception and environmental understanding Streamline development processes through reusable capabilities The ideal candidate will contribute to research that bridges the gap between theoretical advancement and practical implementation in robotics. You will be part of a team that's revolutionizing how robots learn, adapt, and interact with their environment. Join us in building the next generation of intelligent robotics systems that will transform the future of automation and human-robot collaboration. As an Applied Science Manager in the Foundation Model team, you will build and lead a team that develops and improves machine learning systems that help robots perceive, reason, and act in real-world environments. You will set the technical direction for leveraging state-of-the-art models (open source and internal research), evaluating them on representative tasks, and adapting/optimizing them to meet robustness, safety, and performance needs. You will drive the capability roadmap and the evaluation strategy that defines “what the robot brain can do,” and you will sponsor targeted innovation when gaps remain. You’ll collaborate closely with research, controls, hardware, and product teams, and ensure the team’s outputs can be further customized and deployed by downstream teams on specific robot embodiments.
IN, KA, Bengaluru
Are you passionate about giving customers the richest, most inspiring experience in their shopping journey? Do you like to dive deep to understand how customer-centric solutions drive measurable results? Do you enjoy working closely with the business and software engineers to design rigorous experiments, build the data infrastructure behind them, and translate results into decisions? You are in the right place! Come join our Prime & Marketing Analytics and Science (PRIMAS) team, where your work will directly impact millions of customers. The EU Marketing & Prime organization is looking for a Data Scientist to join the PRIMAS team. This role sits at the intersection of applied statistics and large-scale analytics — you'll design experiments and causal models, and also own the data pipelines, metrics, and reporting infrastructure that make those results usable across the business. The PRIMAS team provides a comprehensive understanding of customer segments, affinities, and lifetime value. We use data science tools and advanced statistical techniques to study customer purchase and engagement behaviors, and generate actionable insights on where, when, and how we deliver products and programs to customers. We help increase customer engagement, sales, and marketing efficiency, and our systems are built entirely in-house on automated large-scale analytics infrastructure. You will design, launch, and measure experiments across marketing channels (SEM/SEO, Affiliates, Display, Social, Mobile, Email, Onsite, etc.), engagement products, and customer segments. You will improve our understanding of customer behavior, run rigorous power and minimum detectable effect (MDE) analyses to size experiments correctly, and build the causal and conversion models that value and target our marketing — then build the pipelines and dashboards that keep those signals flowing reliably to stakeholders and downstream systems. You will work at the forefront of consumer analytics, tackling some of the hardest measurement problems in the industry alongside strong scientists, statisticians, and software engineers. Key job responsibilities 1. Design and implement scalable, statistically rigorous experiments (A/B, geo, holdout, quasi-experiments) to measure marketing incrementality across channels. 2. Perform power analysis and minimum detectable effect (MDE) calculations to determine experiment sample sizes, durations, and design trade-offs before launch. 3. Build causal and treatment-effect models that produce conversion and valuation signals consumed by downstream bidding and budgeting systems. 4. Building the ETL, metric definitions, and datasets that make results scalable, extensible, and repeatable rather than one-off analyses. 5. Develop measurement frameworks that quantify the true, platform-independent contribution of marketing over time, and build the dashboards and reporting that keep those metrics visible to the business. 6. Apply statistical, mathematical, and machine learning techniques to solve ambiguous business problems where the right approach isn't obvious. 7. Analyze experiment results for validity — inspecting distributions, checking for sample ratio mismatch, exploring covariate balance, and tracking down the source of anomalies. 8. Communicate experiment design, results, and trade-offs clearly to business and leadership audiences, including inputs into business reviews, and influence decisions and technical direction across teams. 9. Establish scalable, repeatable processes and best practices for experiment design, data modeling, and analysis.
IN, KA, Bengaluru
Are you passionate about giving customers the richest, most inspiring experience in their shopping journey? Do you like to dive deep to understand how customer-centric solutions drive measurable results? Do you enjoy working closely with the business and software engineers to design rigorous experiments, build the data infrastructure behind them, and translate results into decisions? You are in the right place! Come join our Prime & Marketing Analytics and Science (PRIMAS) team, where your work will directly impact millions of customers. The EU Marketing & Prime organization is looking for a Data Scientist to join the PRIMAS team. This role sits at the intersection of applied statistics and large-scale analytics — you'll design experiments and causal models, and also own the data pipelines, metrics, and reporting infrastructure that make those results usable across the business. The PRIMAS team provides a comprehensive understanding of customer segments, affinities, and lifetime value. We use data science tools and advanced statistical techniques to study customer purchase and engagement behaviors, and generate actionable insights on where, when, and how we deliver products and programs to customers. We help increase customer engagement, sales, and marketing efficiency, and our systems are built entirely in-house on automated large-scale analytics infrastructure. You will design, launch, and measure experiments across marketing channels (SEM/SEO, Affiliates, Display, Social, Mobile, Email, Onsite, etc.), engagement products, and customer segments. You will improve our understanding of customer behavior, run rigorous power and minimum detectable effect (MDE) analyses to size experiments correctly, and build the causal and conversion models that value and target our marketing — then build the pipelines and dashboards that keep those signals flowing reliably to stakeholders and downstream systems. You will work at the forefront of consumer analytics, tackling some of the hardest measurement problems in the industry alongside strong scientists, statisticians, and software engineers. Key job responsibilities 1. Design and implement scalable, statistically rigorous experiments (A/B, geo, holdout, quasi-experiments) to measure marketing incrementality across channels. 2. Perform power analysis and minimum detectable effect (MDE) calculations to determine experiment sample sizes, durations, and design trade-offs before launch. 3. Build causal and treatment-effect models that produce conversion and valuation signals consumed by downstream bidding and budgeting systems. 4. Building the ETL, metric definitions, and datasets that make results scalable, extensible, and repeatable rather than one-off analyses. 5. Develop measurement frameworks that quantify the true, platform-independent contribution of marketing over time, and build the dashboards and reporting that keep those metrics visible to the business. 6. Apply statistical, mathematical, and machine learning techniques to solve ambiguous business problems where the right approach isn't obvious. 7. Analyze experiment results for validity — inspecting distributions, checking for sample ratio mismatch, exploring covariate balance, and tracking down the source of anomalies. 8. Communicate experiment design, results, and trade-offs clearly to business and leadership audiences, including inputs into business reviews, and influence decisions and technical direction across teams. 9. Establish scalable, repeatable processes and best practices for experiment design, data modeling, and analysis.
IN, KA, Bengaluru
Are you passionate about giving customers the richest, most inspiring experience in their shopping journey? Do you like to dive deep to understand how customer-centric solutions drive measurable results? Do you enjoy working closely with the business and software engineers to design rigorous experiments, build the data infrastructure behind them, and translate results into decisions? You are in the right place! Come join our Prime & Marketing Analytics and Science (PRIMAS) team, where your work will directly impact millions of customers. The EU Marketing & Prime organization is looking for a Data Scientist to join the PRIMAS team. This role sits at the intersection of applied statistics and large-scale analytics — you'll design experiments and causal models, and also own the data pipelines, metrics, and reporting infrastructure that make those results usable across the business. The PRIMAS team provides a comprehensive understanding of customer segments, affinities, and lifetime value. We use data science tools and advanced statistical techniques to study customer purchase and engagement behaviors, and generate actionable insights on where, when, and how we deliver products and programs to customers. We help increase customer engagement, sales, and marketing efficiency, and our systems are built entirely in-house on automated large-scale analytics infrastructure. You will design, launch, and measure experiments across marketing channels (SEM/SEO, Affiliates, Display, Social, Mobile, Email, Onsite, etc.), engagement products, and customer segments. You will improve our understanding of customer behavior, run rigorous power and minimum detectable effect (MDE) analyses to size experiments correctly, and build the causal and conversion models that value and target our marketing — then build the pipelines and dashboards that keep those signals flowing reliably to stakeholders and downstream systems. You will work at the forefront of consumer analytics, tackling some of the hardest measurement problems in the industry alongside strong scientists, statisticians, and software engineers. Key job responsibilities 1. Design and implement scalable, statistically rigorous experiments (A/B, geo, holdout, quasi-experiments) to measure marketing incrementality across channels. 2. Perform power analysis and minimum detectable effect (MDE) calculations to determine experiment sample sizes, durations, and design trade-offs before launch. 3. Build causal and treatment-effect models that produce conversion and valuation signals consumed by downstream bidding and budgeting systems. 4. Building the ETL, metric definitions, and datasets that make results scalable, extensible, and repeatable rather than one-off analyses. 5. Develop measurement frameworks that quantify the true, platform-independent contribution of marketing over time, and build the dashboards and reporting that keep those metrics visible to the business. 6. Apply statistical, mathematical, and machine learning techniques to solve ambiguous business problems where the right approach isn't obvious. 7. Analyze experiment results for validity — inspecting distributions, checking for sample ratio mismatch, exploring covariate balance, and tracking down the source of anomalies. 8. Communicate experiment design, results, and trade-offs clearly to business and leadership audiences, including inputs into business reviews, and influence decisions and technical direction across teams. 9. Establish scalable, repeatable processes and best practices for experiment design, data modeling, and analysis.
US, WA, Seattle
Want to apply data science to one of Amazon's fastest-growing payment businesses, serving millions of sellers and buyers across 20+ global marketplaces? As a Data Scientist I on the cross-border payments science team, you will build and deploy production models that power real-time FX risk monitoring, generative AI seller chatbots, and multi-agent AI tools. You will work end-to-end—from problem framing through production deployment on AWS—delivering solutions that directly influence billion-dollar payment flows. This is a high-visibility, small-team environment where your work informs senior leadership decisions and creates measurable impact for customers worldwide. Key job responsibilities - Build, tune, and evaluate large language models and generative AI applications, including seller-facing chatbots and multi-agent AI tools for the cross-border payments business. - Develop and deploy real-time statistical and machine learning models for FX risk monitoring, validating your data, assumptions, and results throughout the process. - Gather and use large datasets from multiple sources across global marketplaces to design production-ready solutions that meet customer needs and team goals. - Write accurate, clear, and mathematically rigorous technical documents that communicate model performance and business impact to both technical and non-technical audiences. - Partner with engineering, business, and science teams to translate payment-domain problems into data science solutions and ensure smooth deployment on AWS. A day in the life You will spend your morning reviewing model outputs from production FX risk systems and tuning generative AI prototypes for seller chatbots. In the afternoon you might pair with an engineer to deploy a new model version on AWS, then present preliminary results to senior leadership. Throughout, you will seek feedback from senior scientists on your methodology and collaborate with cross-functional partners to refine problem framing. About the team We are the science team behind Amazon's cross-border payments business, supporting products that serve millions of sellers and buyers across 20+ global marketplaces. Our team is small and at an inflection point—we are expanding our generative AI capabilities, building multi-agent systems, and strengthening real-time risk models. You will join a group that values end-to-end ownership, from research to production, and whose work directly shapes decisions on large-scale payment flows.
US, TX, Austin
Are You Ready to Redefine How the World Receives Its Packages? What if your algorithms defined the most efficient path for millions of deliveries — every single day? At Amazon, we're building the science that makes that possible, and we're looking for exceptional scientists to help lead the way. The Last Mile Routing & Planning organization develops the software, algorithms, and tools that power the "magic" of home delivery. Our planning and routing intelligence systems drive billions of daily decisions — enabling safe, efficient, and frustration-free routes for drivers across the globe. What You'll Do In this role, you'll sit at the intersection of state-of-the-art research and real-world impact. You will: - Design and build algorithms that solve large-scale, complex logistics problems - Synthesize data from diverse sources to identify high-value business opportunities - Provide research direction and data-driven insights to guide strategic decisions - Translate complex technical approaches into clear communication for scientists, engineers, and business stakeholders - Partner closely with scientists and engineers in a collaborative, high-impact environment What You'll Work On We have an exciting and growing portfolio of research areas, including: - Routing for same-day and grocery deliveries - Planning for electric and autonomous vehicles - District-level and stop-level planning - Forecasting solutions for diverse delivery programs All of this is powered by the latest methods in Operations Research (OR), Machine Learning (ML), and Generative AI — at a truly global scale. Successful candidates will lead one or more of these problem spaces. What We're Looking For - Deep expertise in Operations Research and/or Machine Learning methods - Proven experience applying these methods to large-scale, real-world business problems - Ability to translate models into production-ready code in Python or Java - Strong communication skills — you can explain complex technical concepts to diverse audiences - A bias for action and an iterative mindset when tackling ambitious research challenges Why Amazon We're passionate about your growth. Whether you want to explore emerging technologies, take on broader scope, or accelerate your career trajectory, we'll invest in helping you get there. Our business is scaling fast — and so are the opportunities for the people who build it. If you're driven by the challenge of optimizing one of the world's most complex logistics systems and excited to see your work impact millions of customers daily, we'd love to hear from you. Key job responsibilities - Invent and design novel solutions for scientifically complex problem areas, and identify opportunities for invention within existing and new business initiatives - Deliver large-scale, high-impact solutions to complex problems in support of medium-to-large business goals - Shape the design of scientifically complex software systems, personally contributing significant portions of the critical scientific novelty - Apply mathematical optimization, machine learning, and Generative AI techniques to develop solution methodologies for in-house decision support tools and software - Research, prototype, simulate, and experiment with models — and actively participate in their production-level deployment in Python or Java - Engage with the broader scientific community by publishing research articles and participating in leading research conferences
US, WA, Bellevue
The Amazon GDS-MOP (modeling, Optimization and Planning) Science team is seeking an exceptional Applied Scientist with strong operations research and optimization expertise to develop production solutions for one of the most complex systems in the world: Amazon's Fulfillment Network labor capacity planning. At MOP Science, we design, build, and deploy optimization, statistics, machine learning, and GenAI/LLM solutions that power Amazon Labor Planning systems (ALPS) running across Amazon Fulfillment Centers worldwide. We solve a wide range of challenges encountered throughout the network, including labor planning and staffing, pick scheduling, stow guidance, and capacity risk management. We are tasked with developing innovative, scalable, and reliable science-driven production solutions that exceed the published state of the art, enabling systems to run frequently (ranging from every few minutes to every few hours per use case) and continuously across our large-scale network. Key job responsibilities As an Applied Scientist, you will collaborate with other scientists, software engineers, product managers, and operations leaders to develop optimization-driven solutions using a variety of tools and observe direct impact on process efficiency and associate experience in the fulfillment network. Key responsibilities include: • Develop understanding and domain knowledge of operational processes, system architecture and functions, and business requirements • Deep dive into data and code to identify opportunities for continuous improvement and/or disruptive new approaches • Develop scalable mathematical models for production systems to derive optimal or near-optimal solutions for existing and new challenges • Create prototypes and simulations for agile experimentation of devised solutions • Advocate for technical solutions with business stakeholders, engineering teams, and senior leadership • Partner with engineers to integrate prototypes into production systems • Design experiments to test new or incremental solutions launched in production and build metrics to track performance About the team Amazon offers a full range of benefits that support you and eligible family members, including domestic partners and their children. Benefits can vary by location, the number of regularly scheduled hours you work, length of employment, and job status such as seasonal or temporary employment. The benefits that generally apply to regular, full-time employees include: • Medical, Dental, and Vision Coverage • Maternity and Parental Leave Options • Paid Time Off (PTO) • 401(k) Plan
US, CA, Sunnyvale
We are looking for a Senior Inference Engineer to own inference for real-time multimodal conversational AI. This is a full-stack inference role: you will work across the entire path a model takes from research to production — shaping model architecture so it is servable, building the real-time runtime that serves it within hard latency budgets, and building the offline systems that train and reinforce it. You will operate at the boundary of Science and Inference, taking frontier-scale speech and audio models and making them run within real-time latency budgets on production hardware. You will co-design architectures with scientists to make them inference-friendly from inception, own the low-latency streaming serving path, and build the training and reinforcement-learning infrastructure that closes the loop. You will have the compute, data, and runway to solve problems that few teams in the world are positioned to tackle. As a Senior Engineer, you will own a significant area of the inference stack end to end, drive its technical execution, contribute to the team's roadmap, and work closely with scientists and hardware partners to ensure our models run fast enough to feel human in real time — and at a cost that makes them viable at scale. You may go deep in one of the areas below while contributing across the others. Key job responsibilities Model Architecture & Inference Co-Design • Partner with research scientists to make model architectures servable from inception — surfacing the latency, memory, and cost implications of architecture choices before they are locked in • Implement and optimize the inference path for large-scale multimodal models — attention and KV-cache mechanisms, multimodal/autoregressive decoding, and the compute primitives on the critical path Apply efficiency techniques across the stack — quantization (per-tensor/per-channel/per- group, INT8/FP8/BF16), speculative decoding, operator fusion, and paged KV-cache — and quantify their quality/latency trade-offs • Develop and tune high-performance kernels for critical operations where off-the-shelf implementations leave performance on the table, integrating them into production serving with minimal overhead • Profile end-to-end performance with tools such as Nsight Compute/Systems and roofline analysis to identify and eliminate bottlenecks in large-scale inference workloads Real-Time & Interactive Runtime • Own the real-time serving path for streaming multimodal conversational AI, meeting sub- second, streaming latency budgets under concurrent session load • Build and tune continuous batching, scheduling, and preemption to balance throughput against per-request latency SLAs for interactive workloads • Customize production serving frameworks (e.g., vLLM, PyTorch) for real-time streaming generative models that fall outside standard LLM serving patterns — sustained low-latency output under concurrent session load • Implement multi-GPU inference (tensor parallelism, collective communication) for latency- critical paths, and drive cost toward parity with existing production baselines • Establish latency, throughput, and cost benchmarking, and publish the operational metrics that gate deployment Offline Systems: Training, RL & Evaluation Infrastructure • Build and scale the offline inference systems behind post-training — high-throughput rollout generation and reward-model serving for reinforcement learning (RL/RLHF/RLAIF) • Ensure train/serve consistency — that the inference path used in RL and evaluation faithfully matches production online behavior (e.g., parity across sampling and logit processing) • Work with the evaluation team to enable offline inference that captures the quality dimensions unique to real-time conversation — latency sensitivity, audio quality, and interaction naturalness
US, WA, Redmond
We are searching for a talented candidate with experience in orbital mechanics, orbit determination, launch vehicle trajectories, and launch vehicle mission planning. In this position, a successful candidate would serve as a Research Scientist in support of Amazon Leo’s constellation with particular focus on Launch Vehicle support. Strong analysis skills are required to develop engineering studies of complex large-scale dynamical systems. This position requires demonstrated expertise in computational analysis automation and tool development. Export Control Requirement: Due to applicable export control laws and regulations, candidates must be a U.S. citizen or national, U.S. permanent resident (i.e., current Green Card holder), or lawfully admitted into the U.S. as a refugee or granted asylum. Key job responsibilities Working with the Leo GNC team, you will: • Perform spacecraft maneuver or navigation analysis in support of multi-disciplinary trades within the Amazon Leo team. • Contribute to prototype software development of flight algorithms. • Test and assess navigation software for integration into flight systems. • Assess and trouble-shoot the performance of Leo on-board GNSS hardware and software systems. • Work closely with GNC engineers to manage on-orbit performance and develop flight dynamics operations processes. • Manage engineering trades as needed for various launch vehicle mission designs. • Support Leo’s Launch Vehicle Mission Management team with technical expertise in Launch Vehicle trajectory requirements specification • Evaluate Launch Vehicle performance and compliance with mission requirements • Develop tools to support Mission Management planning for over 80 launches! • Work collaboratively with launch vehicle system technical teams About the team The Flight Dynamics team is responsible for the guidance, navigation, control and safety of the spaceflight of the Amazon Leo constellation. This team provides solutions to spaceflight challenges in constellation design, orbit selection, launch vehicle insertion requirements, navigation, trajectory design, space situational awareness, and space traffic coordination.