Netflix Logo

Netflix

Research Engineer L5 - Machine Learning Efficiency

Posted 7 Hours Ago
Be an Early Applicant
In-Office
Los Gatos, CA, USA
100K-720K Annually
Senior level
In-Office
Los Gatos, CA, USA
100K-720K Annually
Senior level
Design and implement efficiency improvements for large-scale deep neural networks and LLM training/serving. Develop optimizations (quantization, pruning, distillation, efficient fine-tuning), work on ML hardware/software accelerator integration, and build scalable training and serving infrastructure in collaboration with scientists and cross-functional teams.
The summary above was generated by AI

Netflix is one of the world’s leading entertainment services with 278 million paid memberships in over 190 countries enjoying TV series, films and games across a wide variety of genres and languages. Members can play, pause and resume watching as much as they want, anytime, anywhere, and can change their plans at any time.

The Role

Fast-paced innovation in the theory and practice of large language models (LLMs) and other foundation models is greatly helping to advance state-of-the-art in personalization and discovery experiences. However, cost-effective and efficient training and serving of the models at Netflix’s scale is a technical challenge. Hence we are looking for an exceptional applied research engineer to help us develop the technology that would enable efficient training and serving of these models. 

In this role, you will aid applied research and product development by conceptualizing, designing, and implementing engineering improvements related to large-scale deep neural networks. You would have proven expertise in efficiency optimizations using techniques such as quantization, model pruning, distillation, compute-efficient finetuning, etc. You have to be deeply knowledgeable in ML hardware and software to be successful in this role. Additionally, you need solid software development skills, a love of learning, a passion for solving problems, a bias to action, and effective collaboration with scientists. 

What we are looking for:

  • 5+ years of software engineering experience with a track record of delivering quality results.
  • Proven expertise in training and serving infrastructure for LLMs and other large foundation models.
  • Strong problem-solving skills with knowledge of statistical methods.
  • Strong software development experience in languages such as Python and Java.
  • Deep understanding of TensorFlow and/or PyTorch.
  • Familiarity with hardware and software accelerators and GPU-based optimizations
  • Great interpersonal skills.
  • Strong communication skills - written and verbal.
  • Graduate degree in Computer Science, Statistics, or a related field.

Preferred, but not required, additional areas of experience:

  • Experience as a technical leader.
  • Experience working with cross-functional teams.
  • Experience in Search, Recommendations, Natural Language Processing, Knowledge Graphs, Conversational Agents, and Personalization.
  • Experience with Spark or other distributed computed platforms.
  • Experience with cloud computing platforms and large web-scale distributed systems.
  • Experience in applied research in industrial settings.
  • Open source contributions.
  • Research publications at peer-reviewed journals and conferences on relevant topics.

Links to some of our published work:

  • Synergistic Signals: Exploiting Co-Engagement and Semantic Links via GNN - Under review.
  • Lessons Learnt From Consolidating ML Models in a Large-Scale Recommendation System
  • Search Personalization at Netflix - PaRiS Workshop - WebConf 2023.
  • Augmenting Netflix Search with In-Session Adapted Recommendations - RecSys 2022
  • Query Facet Mapping and its Applications in Streaming Services - SIGIR 2022 
  • Recommendations and Results Organization in Netflix Search - RecSys 2021
  • Challenges in Search on Streaming Services: Netflix Case Study - SIGIR 2019
  • Netflix Research site

Our compensation structure consists solely of an annual salary; we do not have bonuses. You choose each year how much of your compensation you want in salary versus stock options. To determine your personal top of market compensation, we rely on market indicators and consider your specific job family, background, skills, and experience to determine your compensation in the market range. The range for this role is $100,000 - $720,000.

Netflix provides comprehensive benefits including Health Plans, Mental Health support, a 401(k) Retirement Plan with employer match, Stock Option Program, Disability Programs, Health Savings and Flexible Spending Accounts, Family-forming benefits, and Life and Serious Injury Benefits. We also offer paid leave of absence programs. Full-time hourly employees accrue 35 days annually for paid time off to be used for vacation, holidays, and sick paid time off. Full-time salaried employees are immediately entitled to flexible time off. See more detail about our Benefits here.

Netflix is a unique culture and environment. Learn more here.

We are an equal-opportunity employer and celebrate diversity, recognizing that diversity of thought and background builds stronger teams. We approach diversity and inclusion seriously and thoughtfully. We do not discriminate on the basis of race, religion, color, ancestry, national origin, caste, sex, sexual orientation, gender, gender identity or expression, age, disability, medical condition, pregnancy, genetic makeup, marital status, or military service.

HQ

Netflix Los Gatos, California, USA Office

100 Winchester Circle, Los Gatos, CA, United States

Netflix San Jose, California, USA Office

San Jose, United States, 0

Netflix Santa Clara, California, USA Office

Santa Clara, United States, 0

Similar Jobs

An Hour Ago
In-Office
Sunnyvale, CA, USA
182K-242K Annually
Senior level
182K-242K Annually
Senior level
Cloud • Information Technology • Machine Learning
Lead and grow a MetalDev RAS team responsible for the reliability, availability, and serviceability of CoreWeave's bare-metal infrastructure. Set technical direction, KPIs/SLOs, observability (Prometheus/Grafana), automation and CI/CD for server lifecycle, run incident management and RCA practices, and collaborate cross-functionally and with vendors to drive operational excellence and reduce on-call toil.
Top Skills: BmcCi/CdGoGrafanaKubernetes (K8S)PrometheusRedfish
An Hour Ago
In-Office
2 Locations
182K-242K Annually
Senior level
182K-242K Annually
Senior level
Cloud • Information Technology • Machine Learning
Design, build, and operate a scalable multi-tenant control plane for high-performance AI storage. Implement and optimize S3-compatible object storage and distributed filesystems, integrate dedicated storage clusters, improve reliability and observability, and collaborate with platform, compute, and operations teams to monitor and tune performance. Mentor engineers and contribute to scalable, durable storage architecture.
Top Skills: CCephClickhouseDaosGoGpu Direct StorageGrafanaInfinibandKubernetesNfsPrometheusRdmaRoceRustS3Spdk
An Hour Ago
In-Office
2 Locations
157K-210K Annually
Senior level
157K-210K Annually
Senior level
Cloud • Information Technology • Machine Learning
Design and build scalable data pipelines, models, and dashboards for supply chain KPIs. Partner with business stakeholders to enable forecasting, scenario planning, data governance, automation, and actionable reporting for leadership and projects.
Top Skills: AirflowAlationAtlanBigQueryCollibraDbtFivetranInformaticaLlm/Ai AgentsLookerOracle NetsuitePower BIPythonRedshiftSap EccSap S/4HanaSnowflakeSQLTableau

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account