Thinking Machines Lab Logo

Thinking Machines Lab

Product Manager - Deployment

Posted 6 Days Ago
Hybrid
San Francisco, CA, USA
300K-450K Annually
Entry level
Hybrid
San Francisco, CA, USA
300K-450K Annually
Entry level
Own the deployment strategy and roadmap for moving trained and fine-tuned AI models into reliable production use. Define serving workflows, autoscaling, monitoring, rollback, incident response, SLAs, pricing inputs, and API/SDK experiences. Partner with infrastructure, research, engineering, GTM, and users to prioritize improvements, resolve dependencies, guide launches, analyze performance and cost tradeoffs, and build a scalable deployment platform.
The summary above was generated by AI
About Thinking Machines

The mission of Thinking Machines is to build AI that extends human will and judgment. We are training frontier models with Inkling, developing Tinker to let people make models their own, and crafting interfaces that broaden human-AI communication. We believe the future worth building is human, and we're hiring people who want to build it.

About the Role

As Product Manager for Deployment, you will own how Thinking Machines' models and fine-tuned checkpoints go from training into production use. You will shape the path from a trained model to a served, reliable, cost-effective endpoint — covering inference infrastructure, serving APIs, latency and throughput tradeoffs, scaling behavior, observability, and the workflows researchers and external users rely on to deploy their work with confidence.

This is not a mature MLOps role at an established platform. Deployment at Thinking Machines is still being defined: what "production-ready" means for a fine-tuned model, which serving paths we support, how much control users get over performance and cost tradeoffs, and how we scale reliably as usage grows. You will work from infrastructure capability through to a deployment experience that is fast, predictable, and trustworthy.

The strongest candidate has shipped and operated production ML or infrastructure systems before, ideally as an engineer before becoming a product leader, and can reason from strategy down to autoscaling behavior, latency budgets, rollout safety, and the on-call realities of running models in production.

What You'll Do

  • Own deployment strategy, roadmap, and success metrics for taking models and Tinker-trained checkpoints into production, in close partnership with infrastructure, research, engineering, and GTM

  • Define priority deployment paths and workflows across model serving, autoscaling, versioning, rollback, monitoring, and incident response

  • Work at engineering depth on serving architecture, latency and cost tradeoffs, reliability targets, capacity planning, and API/SDK surfaces for deployment

  • Build direct feedback loops with users deploying models in production, and turn scattered signals into a clear view of what's broken, what's missing, and what to prioritize next

  • Drive ambiguous workstreams end to end: technical scoping, dependency resolution, launch readiness, on-call/escalation design, and post-incident learning

  • Connect deployment decisions to the model and infrastructure roadmap, making visible the tradeoffs between flexibility, reliability, and operational cost

  • Shape SLAs, pricing/packaging inputs for hosted inference, and the operating model for a deployment platform expected to scale quickly

  • Do whatever work makes deployment succeed — reviewing a serving config, joining an incident retro, inspecting latency data, or writing the rollout plan for a new model

Skills and Qualifications

  • Experience owning a production ML serving, infrastructure, or deployment product, with direct involvement in reliability, scaling, or performance decisions

  • Track record working at engineering depth with production systems — comfortable discussing latency, throughput, autoscaling, rollback, or incident response in specifics

  • Experience taking a technical product from early usage through to reliable, scaled production use

Preferred qualifications:

  • Background as an engineer or technical founder before moving into product leadership

  • Experience with ML inference infrastructure specifically (model serving frameworks, GPU scheduling, batching, quantization tradeoffs, or similar)

  • Experience operating in a startup, lab, or new product area where the deployment model and roadmap weren't handed to you

  • Comfortable moving between a strategic narrative and a specific technical detail (an autoscaling policy, an SLA definition, a rollout gate) without losing judgment

  • Experience building trust with technical users through evidence, responsiveness, and follow-through rather than process ownership

Logistics

  • Location: This role is based in San Francisco, CA.

  • Compensation: Depending on background, skills and experience, the expected annual salary range for this position is $300,000 - $450,000 USD.

  • Visa sponsorship: We sponsor visas. While we can't guarantee success for every candidate or role, if you're the right fit, we're committed to working through the visa process together.

  • Benefits: Thinking Machines offers generous health, dental, and vision benefits, unlimited PTO, paid parental leave, and relocation support as needed.

Similar Jobs

3 Hours Ago
Remote or Hybrid
United States
60K-81K Annually
Mid level
60K-81K Annually
Mid level
Cloud • Fintech • Software • Business Intelligence • Consulting • Financial Services
Partner with skilled nursing and long-term care clients to assess and improve revenue cycle operations—admissions, payer authorization, MDS/PDPM reimbursement, billing, collections, and denials. Lead client engagements, recommend process and system improvements using EHR and billing data, manage project execution, deliver training, and support change management to optimize reimbursement and financial performance.
Top Skills: Billing SystemsElectronic Health Record (Ehr)Revenue Cycle Management Systems
3 Hours Ago
Remote or Hybrid
United States
107K-144K Annually
Mid level
107K-144K Annually
Mid level
Cloud • Fintech • Software • Business Intelligence • Consulting • Financial Services
Own and prioritize the Workday/ERP product backlog, translate business needs into epics and user stories, partner with Finance and business leaders, support Agile delivery and releases, manage integrations and reporting, ensure compliance and security, and drive continuous improvement and adoption across Workday modules.
Top Skills: AgileWorkday FinancialsWorkday IntegrationsWorkday PsaWorkday Release ManagementWorkday Reporting
15-22 Hourly
Junior
eCommerce • Fashion • Retail • Sales • Wearables • Design
Supports store sales operations by welcoming customers, operating POS and cash wrap, recommending products, processing shipments, replenishing inventory, maintaining stockroom organization, executing visual merchandising updates, and following loss-prevention and housekeeping standards. Requires flexible availability, strong customer service, fashion awareness, attention to detail, and the ability to lift up to 50 pounds occasionally.
Top Skills: InternetIpadLaptopMobile PosPosSocial MediaWalkie-Talkie

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account