openqareer

Machine Learning Engineer, Sequence Models (project Sequoia)

Amazon.com Services LLC · New York, New York, USA

# Machine Learning Engineer, Sequence Models (project Sequoia) **Amazon.com Services LLC** · New York, New York, USA · `On-site` · `full-time` 🕒 **Статус:** *Опубликовано: вчера* · *Источник: Amazon* --- ### Top Skills & Match 🎯 **Ключевой стек роли:** `[Software Development]` `[Software Development]` --- ### About the Role The Sequence Models team serves a centralized role developing sequence models to fundamentally understand Amazon's customers' journeys across all Amazon products (including Stores, Prime Video, Audible, Music, Twitch, etc.). These models will unlock new, actionable insights to optimize the customer experience, including when and how we serve ads. We leverage a host of scientific technologies to accomplish this mission, including Generative AI, classical ML, Causal Inference, Natural Language Processing, and Computer Vision. As the Machine Learning Engineer on the team, you will deliver on our engineering vision to streamline the model development lifecycle from research to production, building ML infrastructure and establishing MLOps practices that enables rapid experimentation and deployment of ML models. You will invent and design new solutions to solve complex challenges that come with petabyte scale storage. Key job responsibilities - Build and scale ML infrastructure across data processing, distributed training, and model serving. Optimize GPU utilization, training throughput, serving latency, and Infra costs. - Own the data pipelines that feed model training, including ingestion of structured and unstructured inputs, schema evolution, backfills, and data quality checks across upstream sources. - Partner with Applied Scientists to shorten the time from experiment to production. - Evolve model serving and feature delivery to support continuous experimentation. - Establish automated, repeatable processes for large-scale data analysis, model training, validation, and deployment. - Own operational excellence for high-volume, low-latency production systems, including monitoring, alarming, troubleshooting, and on-call. - 3+ years of non-internship professional software development experience - 1+ years of designing and developing large-scale, multi-tiered, multi-threaded, embedded or distributed software applications, tools, systems, and services using: C#, C++, Java, or Perl experience - Bachelor's degree or foreign equivalent in Computer Science, Engineering, Mathematics, or a related field - Experience with Machine Learning and Large Language Model fundamentals, including architecture, training/inference lifecycles, and optimization of model execution - Experience in developing and deploying LLMs in production on GPUs, Neuron, TPU or other AI acceleration hardware

Наблюдалась 2026-10-01, впервые 2026-09-30, источник — Amazon.

Открыть у работодателя