SkyPilot Logo

SkyPilot

MTS, Engineering Manager

Posted 47 Minutes Ago
Be an Early Applicant
In-Office
San Mateo, CA, USA
Senior level
In-Office
San Mateo, CA, USA
Senior level
Lead and scale SkyPilot’s engineering team while remaining technically involved. Recruit and develop distributed systems and AI infrastructure engineers, establish engineering practices, guide architecture and reliability standards, shape the product roadmap with founders using customer feedback, sequence execution across projects, and serve as an escalation point for major incidents. The role requires experience operating large-scale distributed systems or cloud platforms, managing engineering teams, and familiarity with Kubernetes or Slurm and GPU workloads.
The summary above was generated by AI
About SkyPilot

SkyPilot accelerates the world's most ambitious AI teams. SkyPilot turns fragmented AI compute across clusters into one optimized, highly available and easy-to-use pool: a single AI supercomputer.

SkyPilot (10k+ GitHub stars, 18M+ downloads) manages the GPU fleets of 100s of companies, from Fortune 500s to top AI-natives like Abridge, Applied Compute, Mistral, Unconventional AI, H Company, and Nubank, with usage growing exponentially. Born in the UC Berkeley lab behind Spark and Databricks, our growing team includes top-tier talent from Databricks, Google, Berkeley, MIT, CMU, and Cornell. To date, SkyPilot has raised over $20M in seed funding from top investors (incl. Lux, Coatue, Amplify) and operators (incl. Ali Ghodsi, Jeff Dean, Guillermo Rauch, Amjad Masad, Clem Delangue, Aaron Levie).

The role

We are hiring our very first Engineering Manager to grow and scale our engineering team. You will report directly to our CTO and work closely with the founders.

We are looking for someone who is hands-on, hungry for impact, and takes pride in building high-performing teams. The ideal candidate understands the output of a manager is the sum of the outputs of all of their reports.

You'll lead the engineering team, stay involved in technical work, and shape product directions alongside the founders: reviewing customer call recordings and notes, aggregating requests across accounts, and deciding what the team builds next. You'll have direct exposure to our largest customers.

What you'll do

Build and lead a high-performance engineering team
  • Recruit, coach, and grow a team of distributed systems and AI infrastructure engineers. 
  • Set the engineering culture and practices needed as the team scales, driving high performance and ownership.
  • Give fast, direct feedback.

Drive technical directions
  • Participate in design reviews, code reviews, and architecture decisions. 
  • Own reliability, security, and operational standards for a control plane that frontier AI teams depend on
  • Guide the team as the platform expands into more use cases.

Shape product directions for the platform
  • Work with the CTO and the founders to set priorities. 
  • Review customer call recordings and notes, aggregate feature requests and bugs across customers, and turn them into a prioritized roadmap, weighing retention of existing customers and growth of new accounts.

Deliver
  • Translate the roadmap into milestones, sequence work across parallel projects, track execution, and serve as the escalation point for major incidents.

What we're looking for

  • 8+ years of software engineering experience with 3+ years managing engineers (teams of 10+), ideally at a fast-growing infrastructure or developer-tools company.
  • Experience building and operating large-scale distributed systems or cloud platforms in production, and still able to contribute technically.
  • Hands-on with Kubernetes or Slurm, and familiar with GPU workloads. Experience with schedulers, orchestration, or AI infrastructure is a plus.
  • Comfortable acting as the product manager for your area: synthesizing customer input, running an intake and triage process, and making prioritization calls with incomplete information.
  • Strong product intuition and care for both developer experience and enterprise needs, and able to make and communicate hard tradeoffs.
  • Experience as an early engineering leader at a startup, prior PM experience, or building enterprise platforms (multi-tenancy, RBAC, HA control planes) is a strong plus.

What we offer

  • Competitive compensation and equity
  • Comprehensive medical, dental, vision coverage for you and your dependents
  • The chance to work with some of the best minds in AI infrastructure and distributed systems, with significant autonomy and ownership.
  • A front-row seat at the latest open-source infra startup from Berkeley (lineage: Databricks).
  • Gourmet lunch & dinner for the team to do their best work

Location: San Mateo, CA.

Similar Jobs

9 Minutes Ago
Hybrid
San Jose, CA, USA
224K-387K Annually
Expert/Leader
224K-387K Annually
Expert/Leader
AdTech • eCommerce • Information Technology • Software • Travel • Generative AI
Leads product strategy, roadmap, execution, and team development for CRM and customer communications products across email, push, SMS, in-product, and transactional channels. Oversees customer data, marketing automation, personalization, journey orchestration, experimentation, analytics, integrations, consent, and platform architecture. Partners with Engineering, Data, Marketing, Legal, Privacy, and Operations while making build-versus-buy decisions and applying AI/ML solutions to improve engagement and business outcomes.
Top Skills: Ai/MlAPIsCustomer Data PlatformsData ModelsExperimentation AnalyticsMarketing AutomationPlatform ArchitectureSalesforce Data CloudSalesforce Marketing Cloud
31 Minutes Ago
Remote or Hybrid
California, USA
118K-194K Annually
Senior level
118K-194K Annually
Senior level
Gaming • Information Technology • Mobile • Software • Esports
Lead the DevOps team for WWE 2K by defining technical strategy and building reliable development and production infrastructure. Responsibilities include designing CI/CD pipelines, deployment automation, cloud and containerized environments, developer tooling, observability, documentation, and operational standards. The role also requires cross-team collaboration, technical leadership, and improving workflows across a distributed game development studio.
Top Skills: .NetAWSBashCC#C++Ci/CdGCPGitInfrastructure As CodeJenkinsLinuxNode.jsPerforcePowershellPuppetPythonReactWindows
33 Minutes Ago
In-Office or Remote
7 Locations
136K-245K Annually
Senior level
136K-245K Annually
Senior level
Blockchain • eCommerce • Fintech • Payments • Software • Financial Services • Cryptocurrency
Lead Block’s Compliance Issues Management Oversight program across Square, Cash App, and Afterpay. Own remediation tracking, change governance, scope determinations, quality assurance, escalation, reporting, and closure validation for compliance issues involving AML/BSA, sanctions, and consumer protection. Build governance standards, procedures, templates, and playbooks; facilitate cross-functional forums; advise stakeholders; and scale the program through AI, automation, and workflow tooling.
Top Skills: AICompliance Issue Management SystemsGovernance Reporting ToolsWorkflow Automation

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account