Chai Discovery Logo

Chai Discovery

Software Engineer, Platform & Inference

Reposted One Month Ago
In-Office
San Francisco, CA, USA
Senior level
In-Office
San Francisco, CA, USA
Senior level
We are seeking a Software Engineer focused on building resilient infrastructure systems for AI drug discovery, enhancing developer productivity and ensuring compliance with biopharma standards.
The summary above was generated by AI
About Chai Discovery

Chai builds the design suite for molecules. We train frontier models that learn the underlying foundations of biochemical structure and interaction, so scientists can move faster and pursue targets that other methods cannot reach.

AI is reinventing life sciences the same way it reinvented software engineering, and Chai is at the forefront of this shift. Leading pharmaceutical companies like Eli Lilly, Pfizer, and Novartis are adopting our platform to power their drug discovery programs.

We value diverse perspectives and are ready to find greatness in unexpected places.

About the role

Platform engineers make Chai's models fast, cheap, and reliable at scale, and enable the outer loop that accelerates research: the infrastructure and software abstractions used to train, eval, and understand models.

You'll own the serving stack that turns our frontier models into a product scientists depend on: latency, throughput, GPU efficiency, batching, and autoscaling across a large multi-cloud GPU fleet. You'll also contribute to the work that enables turning raw models into product-ready pipelines, and the experiment and observability tooling that lets a researcher ship faster.

You've built high-performance services that developers love, moved ML systems into production at scale, and can see around corners before they become outages.

You'll work closely with the researchers who train the models, the product engineers who build on them, and the commercial team deploying them to the world's largest pharma companies.

About you

We index on systems judgment, ownership, and the scars that come from having run production infrastructure before. We're looking for engineers who get obsessed with hard problems and don't give up easily. We look for:

  • 4+ years building production systems, with real depth in performance, distributed systems, or ML serving

  • Experience optimizing model inference: GPU utilization, batching, quantization, caching, or kernel-level work

  • A platform mindset: you like building the tools and abstractions that make other engineers and researchers faster

  • End-to-end ownership of 24/7 systems, including observability, alerting, and incident response

  • Experience across both 0-to-1 buildouts and 1-to-n scale-ups, with an always-evolving playbook you bring wherever you go

  • The instinct to treat cost and efficiency as first-class constraints, not afterthoughts

A background in biology is not required. What makes the difference is technical excellence, curiosity about the domain, and grit.

We offer

The opportunity to work at the vanguard of AI research and frontier biology, with world-class people, on a mission that matters. We protect & promote a culture of high velocity and ownership. We compensate our team accordingly.

HQ

Chai Discovery San Francisco, California, USA Office

San Francisco, California, United States

Similar Jobs

One Month Ago
In-Office
Mountain View, CA, USA
194K-352K Annually
Senior level
194K-352K Annually
Senior level
Artificial Intelligence • Automotive • Information Technology • Robotics
Design, build, and maintain ML infrastructure for model lifecycle: training, optimization, validation, deployment, inference, and observability. Own ML inference and compiler platforms, develop scalable ML pipelines, continuous testing/monitoring, and support deployment to diverse hardware for autonomous vehicles.
Top Skills: C++Data Workflow Orchestration PlatformsDistributed SystemsLarge Language ModelsMl CompilerMl Inference PlatformPython
One Month Ago
In-Office
Mountain View, CA, USA
160K-241K Annually
Junior
160K-241K Annually
Junior
Artificial Intelligence • Automotive • Information Technology • Robotics
Build and maintain core ML infrastructure: design and implement ML training, optimization, validation, and deployment pipelines; operate an in-house ML inference and compiler platform; implement testing, monitoring, and observability for model lifecycles; collaborate with cross-functional teams to deploy optimized models to vehicle hardware.
Top Skills: C++Data Workflow OrchestrationDistributed SystemsLarge Language ModelsMl CompilerMl Inference PlatformObservabilityPython
One Month Ago
In-Office
Sunnyvale, CA, USA
Mid level
Mid level
Artificial Intelligence • Hardware • Software • Semiconductor
Build and maintain the inference orchestration layer across datacenter clusters. Design and implement highly available, low-latency systems, define platform direction (Kubernetes CRDs/operators), lead production incident response, optimize performance and capacity, drive observability and security, and collaborate with ML, product, and cloud teams to productionize scalable inference services.
Top Skills: C++CertificatesCi/CdCustom Resource Definitions (Crds)GoGpu-Accelerated WorkloadsKubernetesKubernetes OperatorsMl Inference InfrastructureModel ServingMtlsObservabilityTls

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account