Architects and operates scalable customer-hosted cloud and on-premises deployments. Owns high-severity incident resolution, develops automated deployment pipelines and safeguards, builds observability systems, and creates internal tooling to improve developer velocity, reliability, and security. The role requires production Kubernetes and cloud infrastructure experience, strong Linux, networking, and infrastructure-as-code fundamentals, customer communication, and familiarity with AWS, Terraform, GitOps, IAM, and observability tools.
We’re building the scientific intelligence platform for the physical world, to accelerate breakthroughs in semiconductors, batteries, advanced materials, aerospace, and beyond. These industries represent trillions in spend, but are still trapped in outdated software. We’re changing that by building the intelligence layer to help scientists and engineers move faster and solve core problems -- R&D through manufacturing.
We have a once-in-a-generation opportunity to create a category-defining company at the intersection of AI and the physical sciences. And we’re building a world-class team to do it.
We recently raised a seed round led by Greylock, with participation from Neo, BoxGroup, Liquid 2, and top angels including Jeff Dean + leadership at OpenAI and AMD. Our team (ex-Applied Intuition, Glean, SpaceX, Warp, Jane Street, Verkada) works in-person in San Francisco. Our office is located walking distance from 4th and King Caltrain station.
What you’ll do:
- Architect and own complex customer-hosted cloud deployments that run scalably and reliably across diverse customer environments.
- Triage, debug, and resolve high-severity issues end-to-end, building bulletproof automated safeguards and observability to permanently prevent regressions.
- Design fully automated, repeatable deployment pipelines with deep observability to proactively surface edge cases before they hit production.
- Drive developer velocity and build internal tooling that empower the engineering team to ship fast without sacrificing reliability or security.
We would love to meet you if you have experience:
- Operating production Kubernetes and cloud infrastructure.
- Delivering on-prem or customer-hosted software
- Strong Linux, networking and infrastructure-as-code fundamentals
- Independently owning incidents and communicating with customers
- Working with the following key technologies:
- Public cloud/AWS etc, Managed Kubernetes knowledge, IaC tools like Terraform
- GitOps/Argo CD, Linux, networking, IAM and observability tooling
- On-prem deployment; air-gapped experience is a bonus
Bonus points if you have:
- Worked at an early stage startup, founded a company, or plan to start one someday.
- Air-gapped or restricted-network deployment experience
- Multi-cloud, hybrid-cloud or enterprise security experience
- Experience supporting data/AI infrastructure
- Experience or deep curiosity in science.
Learn More about Altara
- Website
- Video
- TechCrunch
- Follow us on LinkedIn and X
Similar Jobs
Fintech • HR Tech
Build and operate AWS cloud infrastructure using Terraform/OpenTofu and Spacelift. Lead multi-account architecture migrations, develop reusable infrastructure modules, improve reliability and disaster recovery, support compliance, troubleshoot cloud environments, and automate recurring issues. Partner with engineering teams, strengthen observability and incident response, and participate in an internal infrastructure support on-call rotation. The role also drives multi-quarter platform initiatives with significant autonomy.
Top Skills:
Ai-Assisted Development ToolsAlertingAWSAws IamAws Iam Identity CenterAws OrganizationsCloudFormationCrossplaneDnsInfrastructure As CodeInfrastructure-As-Code Policy EnforcementIp Address ManagementLoggingMonitoringOpentofuSpaceliftSsoTerraformVpc
Artificial Intelligence • Cloud • HR Tech • Information Technology • Productivity • Software • Automation
Build and operate core infrastructure for Moveworks' agentic AI platform, including distributed data stores, authentication, event streaming, configuration, rate limiting, observability, and reliability systems. Partner with machine learning, search, product, data, and frontend teams to define infrastructure roadmaps, improve microservice performance and scalability, identify bottlenecks, and deliver interdependent engineering projects.
Top Skills:
A/B TestingAlertingC++Cloud PlatformsContainersDistributed LoggingDistributed SystemsDistributed TracingElasticsearchEvent StreamingFeature FlagsGoIstioJavaKafkaMicroservicesMonitoringOpensearchPythonRestful Apis
Artificial Intelligence • Big Data • Consumer Web • eCommerce
Own the distributed serving and data infrastructure powering globally served pages and governed multi-agent build loops. Design caching, indexing, rebuild, reliability, observability, performance, and invariant-based enforcement systems. Take end-to-end responsibility for correctness, uptime, failure handling, latency budgets, and production incidents without a dedicated operations team. The role requires independently modeling complex systems, proving correctness under load, and evolving infrastructure as AI capabilities change.
Top Skills:
Automated TestingCachingData Serving InfrastructureDistributed SystemsIndexingObservabilityPerformance Engineering
What you need to know about the San Francisco Tech Scene
San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.
Key Facts About San Francisco Tech
- Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
- Major Tech Employers: Google, Apple, Salesforce, Meta
- Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
- Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
- Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
- Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine



.png)