Relace Logo

Relace

Infrastructure Engineer

Reposted One Month Ago
Be an Early Applicant
In-Office
San Francisco, CA, USA
Junior
In-Office
San Francisco, CA, USA
Junior
Design and operate systems for high-performance inference and training infrastructure, focusing on reliability, speed, and cost-efficiency. Collaborate with research and product teams.
The summary above was generated by AI

About Us

Relace is building the models and infrastructure that code agents reach for. We power the fastest model on OpenRouter (10,000 tok/s) and deliver optimized small language models designed for retrieval, application, and core code generation functions.

Our technology supports some of the world’s fastest-moving companies — including Lovable, Figma, and Vercel — as they deploy and scale code generation to hundreds of millions of users. We recently raised our Series A from a16z, and we’re growing quickly.

Our team is made up of mathematicians, physicists, and computer scientists who are deeply passionate about their craft. If you thrive on ambitious technical problems, care about elegant systems design, and want to build the foundation of how code gets written at scale, this is the place for you.

The Role

As an Infrastructure Engineer at Relace, you’ll design and operate the systems that power our high-performance inference and training infrastructure. You’ll work closely with our research and product teams to ensure our models run at scale with reliability, speed, and cost-efficiency. This is a hands-on engineering role where you’ll shape how we build and scale the backbone of modern code generation.

You’ll have the opportunity to:

- Architect and manage the infrastructure powering our ultra-fast inference and training stack.

- Build reliable, efficient systems for deploying and scaling ML workloads globally.

- Work on GPU scheduling, distributed systems, and high-performance cloud deployments.

- Optimize performance and cost across compute, networking, and storage layers.

- Collaborate with world-class engineers to push the limits of what small models can do.

Requirements

2+ years of experience writing high-quality production code

Strong experience with cloud infrastructure (AWS, GCP, Azure, or equivalent)

Experience with data science and systems optimization

Familiarity with ML infrastructure, GPU’s, etc. a plus

Work out of our SF office in FiDi

Similar Jobs

Yesterday
In-Office
Palo Alto, CA, USA
120K-120K Annually
Internship
120K-120K Annually
Internship
Artificial Intelligence • Software
Supports deployment, operation, monitoring, troubleshooting, and optimization of Palantir software across production environments. Automates manual workflows and runbooks, improves infrastructure reliability and scalability, and develops solutions using Foundry and Apollo. Collaborates with engineers and cross-functional teams while owning infrastructure projects with minimal supervision. The internship requires programming proficiency, infrastructure troubleshooting skills, and an active or obtainable U.S. security clearance.
Top Skills: ApolloC++Cloud InfrastructureConfiguration ManagementDistributed SystemsFoundryGoJavaJavaScriptLoad BalancingMonitoringObservability ToolsPythonService LogsStorage SystemsTypescript
2 Days Ago
Remote or Hybrid
USA
120K-180K Annually
Mid level
120K-180K Annually
Mid level
Cloud • Computer Vision • Information Technology • Sales • Security • Cybersecurity
Build and maintain CrowdStrike’s golden-image platform and automated pipelines across Linux, macOS, and Windows. Develop Python automation for image builds, validation, drift detection, and testing; manage infrastructure as code with Ansible, Packer, and Terraform; integrate Jenkins, GitLab CI, and Artifactory; and support image lifecycle operations. The role includes troubleshooting production issues, security hardening, compliance collaboration, documentation, cross-team partnership, and participation in on-call rotations.
Top Skills: AnsibleApple Business ManagerArtifactoryAwx/TowerCis BenchmarksDepFalcon SensorGitlab CiJenkinsKubernetesKvmLinuxmacOSMacos MdmPackerPythonStigTerraformVMwareWindows
8 Days Ago
In-Office
171K-243K Annually
Senior level
171K-243K Annually
Senior level
3D Printing • Aerospace • Hardware • Software • Manufacturing
Build, architect, and maintain Vast’s high-performance computing infrastructure for aerospace engineering analysis. Responsibilities include cluster lifecycle management, patching, performance optimization, reliability, application and user onboarding, and integration with NX/Teamcenter design workflows. The role supports engineering teams using structural, mechanical, dynamic, and computational fluid dynamics analyses, while leading HPC initiatives from concept through production deployment.
Top Skills: AbaqusAnsysAWSAws ParallelclusterAzureGCPHpcLinuxNastranNxPuppetPythonSlurmStar-Ccm+TeamcenterTerraformWave6

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account