Wayve Logo

Wayve

Machine Learning Engineer, Performance Tooling

Posted 3 Days Ago
In-Office
Sunnyvale, CA, USA
Entry level
In-Office
Sunnyvale, CA, USA
Entry level
Own and extend Wayve’s machine learning compilation pipeline from model checkpoint to deployable embedded bundles across NVIDIA TensorRT and Qualcomm QNN targets. Design compiler passes, quantization and precision-typing systems, graph partitioning, legalization, regression testing, and benchmarking infrastructure. Partner with model and training teams to ensure compilability, accuracy, and latency across architectures and SoCs. Provide technical direction and mentorship for compiler engineering.
The summary above was generated by AI
About us   

Founded in 2017, Wayve is the leading developer of Embodied AI technology.  Our advanced AI software and foundation models enable vehicles to perceive, understand, and navigate any complex environment, enhancing the usability and safety of automated driving systems.

Our vision is to create autonomy that propels the world forward.  Our intelligent, mapless, and hardware-agnostic AI products are designed for automakers, accelerating the transition from assisted to automated driving. 
In our fast-paced environment big problems ignite us—we embrace uncertainty, leaning into complex challenges to unlock groundbreaking solutions. We aim high and stay humble in our pursuit of excellence, constantly learning and evolving as we pave the way for a smarter, safer future.

At Wayve, your contributions matter.  We value diversity, embrace new perspectives, and foster an inclusive work environment; we back each other to deliver impact.  

Make Wayve the experience that defines your career!  

The role

Wayve is building autonomous driving technology that runs on real vehicles. Getting our models onto embedded hardware — correctly, quickly, and reproducibly — is one of the hardest problems between research and product.

As a ML Compiler Engineer, you will own the compilation pipeline that makes that possible. You will build and extend Wayve's ML compiler end-to-end: designing passes, integrating with vendor toolchains like NVIDIA TensorRT and Qualcomm QNN, and delivering deployable bundles that meet our accuracy and latency requirements on every target platform.

Each stage in the pipeline — capture, decomposition, precision assignment, legalisation, partitioning — can affect accuracy, latency, or whether a vendor backend accepts the graph. Your work spans the full lowering stack, building compiler passes and infrastructure that scale across architectures and target platforms.

Key responsibilities
  • Own the ML compilation pipeline end-to-end — from checkpoint to deployable bundle on NVIDIA (TensorRT) and Qualcomm (QNN) targets.
  • Design and implement compiler passes with accuracy and latency gates, so bad compiles are caught before they reach hardware.
  • Build compilation infrastructure that scales across platforms, model architectures, and SoCs — without re-engineering for each new target.
  • Partner with model and training teams on compilability; build regression and benchmarking to validate changes across releases.
  • Set technical direction and raise the bar for compiler engineering across the team.
About you
  • You have built or significantly extended ML compilation or graph-lowering pipelines.
  • You understand multi-stage lowering (capture, decomposition, precision assignment, legalisation) and can debug what breaks at each stage.
  • Strong proficiency with at least one relevant stack (e.g. MLIR, ONNX, TensorRT, Qualcomm QNN, PyTorch export/capture) and confidence learning adjacent frameworks quickly.
  • Experience with quantisation in compilation — precision typing, PTQ integration, and tracking down accuracy loss from compiler transforms.
  • Comfortable from high-level model graphs down to vendor backend constraints; strong Python, with C++ a plus.
  • Clear communicator who can align cross-functional teams on compilation trade-offs.
  • Real compiler ownership — full lowering pipeline from checkpoint to deployable bundle, working deeply with TensorRT and QNN.
  • Hard problems — quantisation preservation through decomposition, cross-SoC precision typing, graph partitioning under speed/accuracy trade-offs, legalisation that does not silently break earlier passes.
  • Vehicle impact — compiler passes determine what runs on embedded hardware in Wayve's driving product.
  • Greenfield at Staff level — small team, high leverage, shaping the compilation stack from early stages.
  • Scalable infrastructure — building pipelines that work across platforms and architectures without starting from scratch each time.

Day-to-day / scope of the role

  • Own the ML compilation pipeline end-to-end on NVIDIA (TensorRT) and Qualcomm (QNN) targets.
  • Design and implement compiler passes with accuracy and latency gates.
  • Extend precision typing and graph-splitting logic for new architectures and SoCs.
  • Partner with model and training teams on compilability.
  • Build regression and benchmarking to validate changes across releases.
  • Set technical direction and mentor on compiler design.

Top hard requirements (skills/experience)

  1. Built or owned significant parts of an ML compilation or graph-lowering pipeline.
  2. Deep experience with quantisation in compilation — precision typing, PTQ integration, debugging accuracy loss from compiler transforms.
  3. Strong Python; comfortable building and testing compiler infrastructure in production codebases.
  4. Proficiency with at least one of: MLIR, ONNX, TensorRT, Qualcomm QNN, PyTorch graph capture/export.
  5. Experience with multi-target compilation or graph partitioning across hardware backends.
  6. Ability to reason about correctness and performance trade-offs at each compiler stage.

Wayve is committed to creating an inclusive interview experience. If you require any accommodations or adjustments to participate fully in our interview process, please let us know.

We understand that everyone has a unique set of skills and experiences and that not everyone will meet all of the requirements listed above. If you’re passionate about self-driving cars and think you have what it takes to make a positive impact on the world, we encourage you to apply.
At Wayve we're committed to creating a diverse, fair and respectful culture that is inclusive of everyone based on their unique skills and perspectives, and regardless of sex, race, religion or belief, ethnic or national origin, disability, age, citizenship, marital, domestic or civil partnership status, sexual orientation, gender identity, veteran status, pregnancy or related condition  (including breastfeeding) or any other basis as protected by applicable law.  

For more information visit Careers at Wayve. 

To learn more about what drives us, visit Values at Wayve 

For US candidates only, please visit E-Verify Notice and Participation and Right to Work

DISCLAIMER: We will not ask about marriage or pregnancy, care responsibilities or disabilities in any of our job adverts or interviews. However, we do look to capture information about care responsibilities, and disabilities among other diversity information as part of an optional DEI Monitoring form to help us identify areas of improvement in our hiring process and ensure that the process is inclusive and non-discriminatory.



Wayve Mountain View, California, USA Office

709 N Shoreline Blvd, Mountain View, California, United States, 94043

Wayve Sunnyvale, California, USA Office

605 W. California Ave, Sunnyvale, United States, 94086

Similar Jobs

3 Days Ago
In-Office
Sunnyvale, CA, USA
Entry level
Entry level
Artificial Intelligence • Transportation
Build scalable performance tooling for AI model training and inference across cloud and embedded hardware. Responsibilities include profiling workloads, conducting roofline analysis, identifying layer- and operator-level bottlenecks, predicting latency, memory, utilization, and compute costs, and monitoring regressions. The role requires close collaboration with model, compiler, runtime, platform, and hardware teams to establish performance targets and translate measurements into actionable recommendations.
Top Skills: PythonPyTorch
8 Minutes Ago
Easy Apply
Hybrid
San Francisco, CA, USA
Easy Apply
207K-253K Annually
Senior level
207K-253K Annually
Senior level
Artificial Intelligence • Cloud • Software
Negotiate customer, vendor, partnership, reseller, and procurement agreements. Advise go-to-market teams, manage commercial deal pipelines, develop contract templates and playbooks, and improve legal workflows using AI tools. Support regional expansion and potentially product, privacy, security, EMEA, and APAC initiatives while managing high-volume work in a fast-paced technology environment.
Top Skills: Ai ToolsCloud-Based ServicesIaasPaasSaaS
48 Minutes Ago
In-Office
San Jose, CA, USA
56K-56K Hourly
Internship
56K-56K Hourly
Internship
Artificial Intelligence • Hardware • Information Technology • Machine Learning
MBA intern supporting CDBU marketing initiatives through market and opportunity assessments, customer-facing solution narratives, vertical playbooks, and go-to-market recommendations. The role analyzes market trends, customer segments, workloads, ecosystems, and whitespace opportunities; translates technical capabilities into business value; uses AI tools for research and content development; and leads an internship project through executive presentation and handoff.
Top Skills: AICloud ComputingCpu/Gpu PlatformsData Center InfrastructureGenerative AiMemory TechnologiesStorage Solutions

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account