Netflix

Software Engineer L5, Offline Inference, Machine Learning Platform

Posted 6 Days Ago

Remote

Hiring Remotely in USA

100K-720K

Senior level

Remote

Hiring Remotely in USA

100K-720K

Senior level

Design and develop systems for batch inference workloads, build developer-friendly tools, and ensure operational excellence for ML practices.

The summary above was generated by AI

Netflix is one of the world's leading entertainment services, with over 300 million paid memberships in over 190 countries enjoying TV series, films and games across a wide variety of genres and languages. Members can play, pause and resume watching as much as they want, anytime, anywhere, and can change their plans at any time.

Machine Learning (ML) is core to that experience. From personalizing the home page to optimizing studio operations and powering new types of content, ML helps us entertain the world faster and better.

The Machine Learning Platform (MLP) organization builds the scalable, reliable infrastructure that accelerates every ML practitioner at Netflix. Within MLP, the Offline Inference team owns the batch-prediction layer—enabling practitioners to generate, store, and serve predictions for various models, including LLMs, computer-vision systems, and other foundation models. One of our most critical customer groups today is the content and studio ML practitioners in the company, whose work influences what we create and how we produce movies and shows you see when you log into the Netflix app.

The Opportunity:

We’re looking for a talented Software Engineer L5 to join the newly formed Offline Inference team. You will design, build, and operate next-generation systems that run large-scale batch inference workloads—from minutes to multi-day jobs—while delivering a friction-free, self-service experience for ML practitioners across Netflix. Success in this role means not only building robust distributed systems, but also deeply understanding the ML development lifecycle to build platforms that truly accelerate our users.

What You’ll Do

Build developer-friendly APIs, SDKs, and CLIs that let researchers and engineers—experts and non-experts alike—submit and manage batch inference jobs with minimal effort, particularly in the domain of content and media
Design, implement, and operate distributed services that package, schedule, execute, and monitor batch inference workflows at massive scale.
Instrument the platform for reliability, debuggability, observability, and cost control; define SLOs and share an equitable on-call rotation
Foster a culture of engineering excellence through design reviews, mentorship, and candid, constructive feedback

Minimum Qualifications:

Hands-on experience with ML engineering or production systems involving training or inference of deep-learning models.
Proven track record of operating scalable infrastructure for ML workloads (batch or online).
Proficiency in one or more modern backend languages (e.g. Python, Java, Scala).
Production experience with containerization & orchestration (Docker, Kubernetes, ECS, etc.) and at least one major cloud provider (AWS preferred).
Comfortable with ambiguity and working across multiple layers of the tech stack to execute on both 0-to-1 and 1-to-100 projects
Commitment to operational best practices—observability, logging, incident response, and on-call excellence.
Excellent written and verbal communication skills; effective collaboration across distributed teams and time zones.
Comfortable working in a team with peers and partners distributed across (US) geographies & time zones.

Preferred Qualifications:

Deep understanding of real-world ML development workflows and close partnership with ML researchers or modeling engineers.
Familiarity with cloud-based AI/ML services (e.g., SageMaker, Bedrock, Databricks, OpenAI, Vertex) or open-source stacks (Ray, Kubeflow, MLflow).
Experience optimizing inference for large language models, computer-vision pipelines, or other foundation models (e.g., FSDP, tensor/pipeline parallelism, quantization, distillation).
Open-source contributions, patents, or public speaking/blogging on ML-infrastructure topics.

What We Offer:

Our compensation structure consists solely of an annual salary; we do not have bonuses. You choose each year how much of your compensation you want in salary versus stock options. To determine your personal top of market compensation, we rely on market indicators and consider your specific job family, background, skills, and experience to determine your compensation in the market range. The range for this role is $100,000 - $720,000.

Netflix provides comprehensive benefits including Health Plans, Mental Health support, a 401(k) Retirement Plan with employer match, Stock Option Program, Disability Programs, Health Savings and Flexible Spending Accounts, Family-forming benefits, and Life and Serious Injury Benefits. We also offer paid leave of absence programs. Full-time hourly employees accrue 35 days annually for paid time off to be used for vacation, holidays, and sick paid time off. Full-time salaried employees are immediately entitled to flexible time off. See more detail about our Benefits here.

Netflix is a unique culture and environment. Learn more here.

Inclusion is a Netflix value and we strive to host a meaningful interview experience for all candidates. If you want an accommodation/adjustment for a disability or any other reason during the hiring process, please send a request to your recruiting partner.

We are an equal-opportunity employer and celebrate diversity, recognizing that diversity builds stronger teams. We approach diversity and inclusion seriously and thoughtfully. We do not discriminate on the basis of race, religion, color, ancestry, national origin, caste, sex, sexual orientation, gender, gender identity or expression, age, disability, medical condition, pregnancy, genetic makeup, marital status, or military service.

Job is open for no less than 7 days and will be removed when the position is filled.

Top Skills

AWS

Bedrock

Databricks

Docker

Ecs

Java

Kubeflow

Kubernetes

Mlflow

Openai

Python

Ray

Sagemaker

Scala

Vertex

5808 W Sunset Blvd, Los Angeles, CA, United States, 90028

Similar Jobs

Resident

Supply Planning Manager

An Hour Ago

Remote

USA

82K-90K

Mid level

82K-90K

Mid level

eCommerce • Retail

Responsible for supply planning across domestic and international products, optimizing inventory, production plans, and supplier performance, while ensuring cross-functional alignment and efficiency improvements.

Top Skills: LookerNetSuiteOraclePower BISAPTableau

Resident

Operations Associate

An Hour Ago

Remote

USA

65K-72K

Mid level

65K-72K

Mid level

eCommerce • Retail

The Marketing Operations Associate supports digital marketing campaigns by managing project timelines, creative assets, and analyzing performance metrics.

Top Skills: AsanaFacebook Ads ManagerFigmaGoogle AdsGoogle WorkspaceLookerSlackTiktok Ads ManagerYoutube

Hiya Inc.

Front-end Engineer

2 Hours Ago

Remote or Hybrid

USA

Mid level

Artificial Intelligence • Cloud • Mobile • Security • Software

The Front-end Engineer will implement features for web applications, collaborate with teams, and ensure pixel-perfect designs are executed, while using tools like React and Next.js.

Top Skills: Auth0AWSGitKubernetesNext.JsReact

What you need to know about the Los Angeles Tech Scene

Los Angeles is a global leader in entertainment, so it’s no surprise that many of the biggest players in streaming, digital media and game development call the city home. But the city boasts plenty of non-entertainment innovation as well, with tech companies spanning verticals like AI, fintech, e-commerce and biotech. With major universities like Caltech, UCLA, USC and the nearby UC Irvine, the city has a steady supply of top-flight tech and engineering talent — not counting the graduates flocking to Los Angeles from across the world to enjoy its beaches, culture and year-round temperate climate.

Key Facts About Los Angeles Tech

Number of Tech Workers: 375,800; 5.5% of overall workforce (2024 CompTIA survey)
Major Tech Employers: Snap, Netflix, SpaceX, Disney, Google
Key Industries: Artificial intelligence, adtech, media, software, game development
Funding Landscape: $11.6 billion in venture capital funding in 2024 (Pitchbook)
Notable Investors: Strong Ventures, Fifth Wall, Upfront Ventures, Mucker Capital, Kittyhawk Ventures
Research Centers and Universities: California Institute of Technology, UCLA, University of Southern California, UC Irvine, Pepperdine, California Institute for Immunology and Immunotherapy, Center for Quantum Science and Engineering

Netflix

Software Engineer L5, Offline Inference, Machine Learning Platform

Top Skills

Netflix Los Angeles, California, USA Office

Similar Jobs

Supply Planning Manager

Operations Associate

Front-end Engineer

What you need to know about the Los Angeles Tech Scene

Key Facts About Los Angeles Tech