Baseten Logo

Baseten

Technical Program Manager, Infrastructure

Reposted 26 Days Ago
Hybrid
San Francisco, CA, USA
165K-330K Annually
Mid level
Hybrid
San Francisco, CA, USA
165K-330K Annually
Mid level
Lead and drive complex, cross-team infrastructure programs end-to-end: scope migrations, sequence dependencies, manage risk, establish operating rhythms, track action items, create clear plans from ambiguous goals, surface risks, and communicate progress to engineers and senior leadership to ensure programs complete on time and without surprises.
The summary above was generated by AI

ABOUT BASETEN

Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F, led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products.

THE ROLE

At Baseten, we’re looking for a Technical Program Manager to drive our most complex, cross-cutting infrastructure programs. This role will operate across all domains of AI infrastructure, from the GPUs up to the multi-cluster orchestration layer.

This is an execution-first role. The work is less about owning a single system and more about imposing order on ambiguity: standing up the right structures, driving decisions to closure, and making sure nothing falls through the cracks across dozens of stakeholders. If you take satisfaction in turning a chaotic, half-defined initiative into a predictable, well-governed program, this role is for you.

RESPONSIBILITIES

  • Own complex migrations end to end. Lead large-scale infrastructure migrations across teams and domains. This will involve scoping the work, sequencing dependencies, managing risk, and driving them to completion without surprises.

  • Drive process across infrastructure. Establish and run the operating rhythms that keep programs healthy: planning cadences, status reporting, decision logs, risk reviews, and escalation paths. Make the process light enough that teams adopt it and rigorous enough that it actually works.

  • Help managers build the right structures. Partner with engineering managers and leads to design the team structures, ownership boundaries, and working models a program needs to succeed. Spot gaps in accountability before they become problems.

  • Own follow-ups relentlessly. Be the person who tracks every open item, owns the action register, and closes loops. Drive decisions and unblockers to resolution rather than letting them drift.

  • Create clarity from ambiguity. Take loosely defined goals and turn them into scoped, sequenced, resourced plans with clear owners and measurable milestones.

  • Manage risk and dependencies. Maintain a clear-eyed view of cross-team dependencies and risks; surface them early, drive mitigation, and keep leadership informed with honest, signal-rich updates.

  • Communicate up, down, and across. Translate between deeply technical teams and senior leadership. Run effective program reviews and write the kind of crisp status that lets busy people make fast decisions.

REQUIREMENTS

  • 3+ years of technical program management experience, with a track record of delivering large, cross-team infrastructure or platform programs.

  • Demonstrated ownership of complex migrations or comparable multi-domain initiatives from inception through completion.

  • Strong execution discipline: you build the structures, cadences, and tracking that keep programs on the rails, and you hold the line on follow-through.

  • Comfort operating across multiple infrastructure domains. You don't need to be the deepest expert in any one, but you can engage credibly with engineers on compute, storage, networking, and data.

  • Proven ability to influence without authority, aligning teams and managers who don't report to you around shared plans and commitments.

  • Excellent written and verbal communication; you can make complex, technical programs legible to any audience.

  • A bias toward closure: you're uncomfortable with open loops and you drive things to done.

NICE-TO-HAVE

  • Experience designing team structures or operating models alongside engineering leadership.

  • Experience with AI models, their use, and how they operate at the infrastructure level.

  • Background in or exposure to software engineering, SRE, or infrastructure operations.

  • Familiarity with the tooling that supports program execution at scale (e.g., issue tracking, planning, and reporting systems).

WHAT SUCCESS LOOKS LIKE

  • In your first few months, you'll have taken ownership of one or more in-flight programs, established the operating rhythm they were missing, and built the trust that makes teams want you in the room.

  • Within a year, the programs you run are the ones leadership doesn't worry about, because the structure is clear, the risks are visible early, and the follow-ups always close.

BENEFITS

  • Competitive compensation, including meaningful equity.

  • 100% coverage of medical, dental, and vision insurance for employee and dependents

  • Flexible PTO policy including company wide Winter Break (our offices are closed from Christmas Eve to New Year's Day!)

  • Paid parental leave

  • Fertility and family-building stipend through Carrot

  • Company-facilitated 401(k)

  • Exposure to a variety of ML startups, offering unparalleled learning and networking opportunities.

Apply now to embark on a rewarding journey in shaping the future of AI! If you are a motivated individual with a passion for machine learning and a desire to be part of a collaborative and forward-thinking team, we would love to hear from you.

At Baseten, we are committed to fostering a diverse and inclusive workplace. We provide equal employment opportunities to all employees and applicants without regard to race, color, religion, gender, sexual orientation, gender identity or expression, national origin, age, genetic information, disability, or veteran status.

We are an Equal Opportunity Employer and will consider qualified applicants with criminal histories in a manner consistent with applicable law (by example, the requirements of the San Francisco Fair Chance Ordinance, where applicable).

HQ

Baseten San Francisco, California, USA Office

San Francisco, CA, United States

Similar Jobs

Yesterday
Hybrid
Santa Clara, CA, USA
200K-322K Annually
Senior level
200K-322K Annually
Senior level
Artificial Intelligence • Computer Vision • Hardware • Robotics • Metaverse
Lead large-scale DGX Cloud infrastructure programs, including AI capacity enablement, cluster bring-up, maintenance, deployment, security, compliance, and observability. Gather requirements, build roadmaps, manage engineering deliverables, establish KPIs, mitigate risks, coordinate internal and external partners, and communicate progress to executive leadership. Drive adoption of cloud-native solutions and continuous process improvement across global, cross-functional teams.
Top Skills: AgileAi/Ml InfrastructureApi IntegrationCloud InfrastructureJIRAKubernetesNvidia GpusScrumSmartsheetTerraform
7 Days Ago
In-Office or Remote
Santa Clara, CA, USA
168K-322K Annually
Senior level
168K-322K Annually
Senior level
Artificial Intelligence • Computer Vision • Hardware • Robotics • Metaverse
Lead recurring capacity operations for large-scale AI infrastructure programs. Own accelerator-capacity intake, forecasting, allocation reviews, lifecycle tracking, dependency management, dashboards, data quality, operating cadences, risk and decision records, and executive reporting. Partner with engineering, research, infrastructure, data, finance, operations, and leadership teams to resolve constraints, improve automation, and establish durable mechanisms for clear ownership, decisions, and action closure.
Top Skills: Ai/Ml PlatformsAPIsBatch SchedulingCloud InfrastructureCluster ManagementDashboard PlatformsData CentersDistributed SystemsGpu And Accelerator Capacity PlanningObservability PlatformsWorkflow Automation
13 Days Ago
Hybrid
154K-212K Annually
Senior level
154K-212K Annually
Senior level
Automotive • Cloud • Hardware • Software
Lead complex, cross-functional programs to improve developer productivity, DevOps, infrastructure, and cloud platforms. Define strategy, milestones, and success metrics; coordinate across engineering, cloud, security, and data teams; drive incident management improvements; measure program health with data; and create reusable program practices. Influence without direct authority and communicate technical tradeoffs to engineering and executive audiences.
Top Skills: Ai ToolsBuild SystemsCi/CdCloud InfrastructureCloud ServicesDeveloper ToolsDevOpsDistributed SystemsIncident ResponseObservabilitySre

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account