Avride Logo

Avride

Lead AI Infrastructure Engineer

Posted 11 Days Ago
Be an Early Applicant
In-Office
Austin, TX
Senior level
In-Office
Austin, TX
Senior level
Lead ML infrastructure engineer owning GPU inference performance for onboard real-time and high-throughput offboard systems. Improve and maintain C++ inference framework, optimize GPU execution, collaborate with applied ML teams, and expand ownership across ML pipelines and distributed infrastructure.
The summary above was generated by AI
About the Team

Our team is at the core of Avride's self-driving stack. We build the base infrastructure layer that powers all autopilot code. It includes a C++ framework for implementing autonomy components, execution graph building and optimization systems, as well as runtimes that execute those graphs, both onboard and in simulation.

The vast part of the execution graph is implemented as a chain of neural network operations. The onboard mode relies on stable latencies of the inference of those networks, while in simulation we also optimize throughput at scale.

About the role

We’re looking for a software engineer with a leadership mindset and deep ML infrastructure experience. You will decide and influence the ML infrastructure layer across the company. The biggest challenge we’re facing at the moment is the effectiveness of GPU inference - both for onboard applications with near real-time guarantees and for offboard cases that target high throughput and deterministic execution. It is the first priority within this role.

What you'll do
  • At first, you will take on the GPU inference framework, focusing on performance
  • Later, the role assumes responsibility and ownership for broader ML infrastructure scattered across ML pipelines
  • Close collaboration with the applied ML team responsible for defining the neural model's architecture
What you'll need
  • Experience with PyTorch
  • Understanding of how GPUs work
  • Experience in diagnosing and resolving performance issues
  • Strong record of building infrastructure including distributed systems
  • 5+ years of experience with C++
  • Programming experience in multi-threaded environments - multiple processes, threads, timers, and interrupts

Candidates are required to be authorized to work in the U.S. The employer is not offering relocation, sponsorship, and remote work options are not available.

Avride is an equal opportunity employer and committed to providing reasonable accommodations to qualified applicants and employees with disabilities to ensure they have equal access to employment opportunities. Avride complies with the Americans with Disabilities Act (ADA), if you need a reasonable accommodation to assist with the application or hiring process, or to perform the essential functions of a job, please email [email protected].

Similar Jobs

19 Days Ago
Remote or Hybrid
United States
150K-170K Annually
Senior level
150K-170K Annually
Senior level
Information Technology • Database • Consulting
Lead design and delivery of agentic AI systems and LLMOps on AWS to generate, validate, and deploy infrastructure-as-code (Terraform). Build multi-agent applications, RAG pipelines, and AIOps for cloud operations; integrate AI into CI/CD and ITSM. Set technical direction, establish guardrails/policy-as-code, curate reusable Terraform modules, and mentor engineering teams.
Top Skills: AiopsAmazon Bedrock AgentsArtifactoryAWSCi/CdGitItsmJenkinsLangchainLlmopsRagSonarqubeTerraform
11 Days Ago
Hybrid
2 Locations
209K-286K Annually
Senior level
209K-286K Annually
Senior level
Fintech • Machine Learning • Payments • Software • Financial Services
Lead design, development, deployment, and support of scalable GenAI platform components (foundation model training, LLM inference, vector search, guardrails, evaluation, observability). Optimize training and inference performance, contribute to technical vision and roadmap, and partner with cross-functional teams to deliver responsible, production-grade AI systems.
Top Skills: Aws UltraclustersC#C++Go (Golang)HuggingfaceJavaLarge Language ModelsLlm InferenceNemo GuardrailsPythonPyTorchScalaSimilarity SearchVectordbs
21 Minutes Ago
Remote or Hybrid
55K-75K Annually
Junior
55K-75K Annually
Junior
Artificial Intelligence • Fintech • Insurance • Marketing Tech • Software • Analytics
Handle inbound and warm sales leads remotely, consult customers on property & casualty insurance needs, convert leads to policyholders, complete paid training and licensing, work scheduled shifts including weekends, and meet remote workspace and wired internet requirements.
Top Skills: Dsl)FiberPcWired High-Speed Internet (Cable

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account