SambaNova Systems Logo

SambaNova Systems

Software Engineering Director - Inference Platform

Reposted One Month Ago
Be an Early Applicant
In-Office
San Jose, CA, USA
245K-325K Annually
Senior level
In-Office
San Jose, CA, USA
245K-325K Annually
Senior level
The role involves leading a software engineering team focused on AI infrastructure and enterprise software, ensuring technical quality and team development while also contributing as a hands-on engineer.
The summary above was generated by AI

SambaNova is a leader in next-generation AI infrastructure, delivering a full-stack inference platform for customers worldwide. At the core of SambaNova's technology is the RDU (Reconfigurable Dataflow Unit) — a chip built on a dataflow architecture rather than the traditional GPU model. Its decode performance is especially strong for agentic workloads like multi-turn agents, code generation, and long-running applications. RDUs are packaged into SambaRack, rack-scale hardware that lets customers deploy state-of-the-art models with better performance, greater energy efficiency, and faster time to value.

About the team

The SambaStack team builds the inference serving platform enterprises and service providers use to host foundation models on RDU hardware. The platform is cloud-agnostic, and the team owns deployment, orchestration, scaling, and reliability for production inference workloads.

About the Role

The SambaStack team is seeking an Engineering Director to lead a software engineering team while remaining a hands-on technical contributor. You'll own the growth and effectiveness of your team, the technical quality of their output, and the roadmap for the inference serving platform. You'll continue to design systems and write code on the platform's hardest problems.

 
Responsibilities

People & Team Leadership

  • Lead, grow, and inspire a software team, fostering a culture of technical excellence, collaboration, and continuous improvement
  • Own hiring for your team: define roles, assess candidates, and build a diverse, high-performing engineering organization
  • Drive the career development of your engineers through regular 1:1s, meaningful feedback, and clear growth plans
  • Set clear goals and expectations, track team progress, and remove blockers to keep the team executing effectively
  • Partner with Product, ML, QA, and DevOps leadership to align team priorities with broader organizational objectives

Technical Contribution

  • Contribute directly as a software engineer: designing systems, writing code, and leading implementation on critical platform initiatives
  • Participate in and lead design reviews, providing authoritative technical guidance on major features and cross-cutting changes
  • Establish and uphold engineering standards, best practices, and coding quality across the team
  • Proactively identify and drive the resolution of systemic risks: performance bottlenecks, reliability gaps, and scaling constraints
  • Stay current on emerging technologies and AI inference trends to inform the team's technical direction
Required Qualifications
  • 12+ years of software engineering experience, including 2+ years leading and managing engineering teams
  • Demonstrated ability to manage, develop, and hire strong engineers while remaining a hands-on technical contributor
  • Strong technical depth in backend systems built with Python, Go, or Rust
  • Experience with AI inference serving, orchestration, and the infrastructure challenges of LLM workloads at production scale
  • Proven ability to drive engineering excellence across an engineering team
  • Excellent communication skills: able to work effectively with engineers, cross-functional partners, and senior leadership
  • Experience running iterative engineering processes in a fast-paced environment
Preferred Qualifications
  • Experience with Kubernetes ecosystem tooling, including Operators and Helm Charts
  • Background in ML application optimization and productionization
  • Experience building or managing teams that deliver cloud-native platforms for enterprise customers
  • Familiarity with AI coding assistants and the developer tooling ecosystem
  • Experience at a high-growth AI infrastructure company

Base Salary Range:

Base Pay Range
$245,000$325,000 USD

Submission Guidelines
Please note that in order to be considered an applicant for any position at SambaNova Systems, you must submit an application form for each position for which you believe you are qualified. 

EEO Policy
SambaNova Systems is an Equal Opportunity/Affirmative Action Employer. All qualified applicants will receive consideration for employment without regard basis of age (40 and over), color, disability, gender identity, genetic information, marital status, military or veteran status, national origin/ancestry, race, religion, creed, sex (including pregnancy, childbirth, breastfeeding), sexual orientation, and any other applicable status protected by federal, state, or local laws.

Benefits Summary for US-Based, Full-Time Employment Positions
SambaNova offers a competitive total rewards package, including the base salary, plus equity and benefits. We cover 95% premium coverage for employee medical insurance, and 77% premium coverage for dependents and offer a Health Savings Account (HSA) with employer contribution. We also offer Dental, Vision, Short/Long term Disability, Basic Life, Voluntary Life, and AD&D insurance plans in addition to Flexible Spending Account (FSA) options like Health Care, Limited Purpose, and Dependent Care. Our library of well-being benefits available to you and your dependents includes a full subscription to Headspace, Gympass+ membership with access to physical gyms, One Medical membership, counseling services with an Employee Assistance Program, and much more.

HQ

SambaNova Systems Palo Alto, California, USA Office

Our Palo Alto office is in a tech complex known for incubating research facilities and borders the Bay Trail along the Don Edwards Wildlife Refuge. Only a 5-minute walk, our employees often fly out of PAO for lunch with colleagues or enjoy happy hour at the nearby Palo Alto Country Club.

Similar Jobs

23 Minutes Ago
In-Office
144K-217K Annually
Senior level
144K-217K Annually
Senior level
Aerospace • Artificial Intelligence • Machine Learning • Robotics • Software
Deploys, integrates, configures, troubleshoots, and supports airborne vision and detection systems across laboratory, production, customer, and field environments. Performs hardware and software integration, wiring and electro-mechanical work, system testing, configuration management, documentation, customer training, and field-service logistics. Coordinates with engineering, production, quality, program, and customer teams while supporting deployments and demonstrations. Requires approximately 50% travel and ability to obtain a Secret clearance.
Top Skills: CanDigital MultimetersErp SystemsEthernetIpc/Whma-A-620LinuxMaintenance-Tracking SystemsNetwork AnalyzersNetworkingOscilloscopesPower SuppliesProduct Lifecycle Management SystemsSerial InterfacesWindows
45 Minutes Ago
Hybrid
Livermore, CA, USA
15-24 Hourly
Entry level
15-24 Hourly
Entry level
eCommerce • Fashion • Retail • Sales • Wearables • Design
Temporary retail associate serving as a stylist and trusted product advisor. Responsibilities include greeting customers, understanding their needs, providing styling recommendations, driving sales through product knowledge and storytelling, completing POS transactions, maintaining stockroom organization, and supporting omni/virtual selling. Requires flexible scheduling, physical ability to handle merchandise, strong communication skills, teamwork, and prior retail experience.
Top Skills: Omnichannel SellingPoint-Of-Sale (Pos) SystemsSocial Media
58 Minutes Ago
Hybrid
San Francisco, CA, USA
255K-300K Annually
Entry level
255K-300K Annually
Entry level
Artificial Intelligence • Productivity • Software
Build and operate Notion’s model layer by integrating frontier models, improving inference reliability, implementing failover and observability, and developing model-level capabilities. The role requires production experience with LLM APIs, evaluation-driven development, and strong judgment around latency, cost, reliability, and output quality. The engineer will also collaborate across AI teams to unblock adoption and ensure fixes are delivered.
Top Skills: Ai Inference SystemsLarge Language ModelsLlm ApisModel EvaluationObservabilityStreamingTool Calling

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account