Netflix Logo

Netflix

ML Engineer L4, Consumer Inference

Posted 15 Days Ago
In-Office
Los Gatos, CA, USA
100K-464K Annually
Mid level
In-Office
Los Gatos, CA, USA
100K-464K Annually
Mid level
Develop and maintain customer-facing libraries and online inference services for low-latency, reliable real-time predictions. Optimize and deploy LLMs for efficient GPU inference, maintain a model registry, and improve ML Platform CI/CD, observability, and incident management. Collaborate cross-functionally to productize models and accelerate research-to-production velocity for Netflix ML practitioners.
The summary above was generated by AI

Netflix is one of the world’s leading entertainment services with 278 million paid memberships in over 190 countries enjoying TV series, films and games across a wide variety of genres and languages. Members can play, pause and resume watching as much as they want, anytime, anywhere, and can change their plans at any time.

The Role

With more than 230 million members in over 190 countries, Netflix continues to shape the future of entertainment around the world. Machine Learning/Artificial Intelligence is powering innovation in all areas of the business, from helping members choose the right title for them through personalization, to better understanding our audience and our content slate, to optimizing our payment processing and other revenue-focused initiatives. The Machine Learning Platform (MLP) provides the foundation for all of this innovation. It offers ML/AI practitioners across Netflix the means to achieve the highest possible impact with their work by making it easy to develop, deploy and improve their machine learning models.  As part of our mission to support the infrastructure for machine learning across the company, we are hiring for a Machine Learning Engineer to join our team to contribute to the team's mission of bridging the gap between ML research and productization. In this role, you will: -Develop customer facing libraries and services to productize machine learning models for efficient and scalable inference.  -Develop and maintain online inference services that provide real-time predictions with low latency and high reliability. -Optimize and deploy large language models (LLMs) for efficient, scalable inference, ensuring high performance and low latency in production environments. -Maintain and improve a model registry to facilitate the discovery, versioning, and governance of machine learning models. -Participate in and improve ML Platform incident management and support workflows. What we offer Opportunity for impact. You will work on cutting edge ML infrastructure use cases and technologies. This role will develop into strategic ownership opportunities for defining MLP’s path from research to production, specifically focusing on building services and tools to accelerate research to production velocity for Netflix ML practitioners.. Responsibility. Netflix offers true transparency and autonomy. Our culture is unique and is key to how we innovate. From day one, your expertise and opinion will be respected and valued by the team and you’ll be given autonomy in deciding the best direction to set for optimizing research to the production path for ML practitioners at Netflix. Learning. You will be developing libraries, tools and services to ensure an efficient and reliable journey of productizing ML models. You will have the opportunity to work with stunning colleagues who value collaboration and have a wealth of experience you can tap into. A work environment where you can grow your career. ML Platform offers a wide variety of projects that can help find the areas you are passionate about. Who will be successful in this role?
  • You are highly customer-driven / developer-driven and empathic. You strive to always focus on delivering customer / user value with an excellent customer service mentality.
  • You have a strong understanding of building scalable and efficient model serving solutions to support large-scale inference for generative models and large language models (LLMs). You create solutions that your stakeholders love and you drive development success from planning to implementation to delivery.
  • You can successfully execute changes within a team's systems, including developing, testing, deploying, and revising solutions.
  • You can communicate and collaborate effectively  (e.g. project meetings, team meetings, code reviews) with immediate team peers and cross-functional project teams.
  • You are eager to both go deep and wide on ML-facing projects. When a project needs deep technical expertise in a domain area you are able to get up to speed quickly. When projects require breadth of focus you are eager to do what’s needed to deliver value even if it means going outside of your comfort zone.
Skills:
  • Strong programming skills, particularly in languages such as Python and Java, and familiarity with ML libraries and frameworks like TensorFlow, PyTorch.
  • Familiarity tools and techniques for deploying machine learning models into production environments, with a particular emphasis on GPU inference optimization (e.g., Triton Inference Server, TensorRT), as well as containerization (e.g., Docker) and orchestration (e.g., Kubernetes).
  • Experience designing with data handling, preprocessing, and transformation techniques to prepare data for model inference.
  • Demonstrated industry-leading experience in large-scale build, release, CI/CD and observability techniques, with particular emphasis on multi-language environments including Scala, Java, and Python.
  • Adopt and promote best practices in operations, including observability, logging, reporting, and on-call processes to ensure engineering excellence.
Our compensation structure consists solely of an annual salary; we do not have bonuses. You choose each year how much of your compensation you want in salary versus stock options. To determine your personal top of market compensation, we rely on market indicators and consider your specific job family, background, skills, and experience to determine your compensation in the market range. The range for this role is $100,000 - $464,000. Netflix provides comprehensive benefits including Health Plans, Mental Health support, a 401(k) Retirement Plan with employer match, Stock Option Program, Disability Programs, Health Savings and Flexible Spending Accounts, Family-forming benefits, and Life and Serious Injury Benefits. We also offer paid leave of absence programs.  Full-time hourly employees accrue 35 days annually for paid time off to be used for vacation, holidays, and sick paid time off. Full-time salaried employees are immediately entitled to flexible time off. See more detail about our Benefits here. Netflix is a unique culture and environment.  Learn more here.

We are an equal-opportunity employer and celebrate diversity, recognizing that diversity of thought and background builds stronger teams. We approach diversity and inclusion seriously and thoughtfully. We do not discriminate on the basis of race, religion, color, ancestry, national origin, caste, sex, sexual orientation, gender, gender identity or expression, age, disability, medical condition, pregnancy, genetic makeup, marital status, or military service.

HQ

Netflix Los Gatos, California, USA Office

100 Winchester Circle, Los Gatos, CA, United States

Netflix San Jose, California, USA Office

San Jose, United States, 0

Netflix Santa Clara, California, USA Office

Santa Clara, United States, 0

Similar Jobs

An Hour Ago
Easy Apply
Remote or Hybrid
USA
Easy Apply
16K-200K Annually
Senior level
16K-200K Annually
Senior level
Food • Software
Lead strategy and execution for automated marketing on behalf of thousands of restaurants. Own campaign creation/automation (email & SMS), diner database and consent model, agentic layer to convert data into marketing actions, marketing measurement standards, and activation/adoption of campaign tools.
Top Skills: AgentsAIAutomationCRMEmailProduct AnalyticsSegmentationSmsSQL
An Hour Ago
In-Office
175K-285K Annually
Expert/Leader
175K-285K Annually
Expert/Leader
Aerospace • Artificial Intelligence • Hardware • Machine Learning • Software • Defense • Manufacturing
Lead and scale a team of People Business Partners to act as trusted advisors to executives. Build operating models, run talent calibrations, develop managers, translate people data into recommendations, handle complex employee relations, and partner with Talent, Total Rewards, and People Ops to align people programs with business needs.
An Hour Ago
Easy Apply
In-Office
Easy Apply
141K-211K Annually
Senior level
141K-211K Annually
Senior level
Aerospace • Hardware • Robotics • Software • Manufacturing
Lead sourcing and supplier management for complex machined aerospace components, develop and execute supply chain strategies, negotiate contracts, manage supplier performance and risk, support NPI efforts, and mentor team members to ensure cost, quality, and delivery targets for Terran R production.
Top Skills: 5-Axis MillingAs9100Cnc MachiningEdmErpLnNadcapOraclePrecision GrindingSAPTurning

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account