Fluidstack Logo

Fluidstack

Compute Engineer, Deployment

Reposted One Month Ago
Be an Early Applicant
Remote or Hybrid
Hiring Remotely in San Francisco, CA, USA
164K-206K Annually
Senior level
Remote or Hybrid
Hiring Remotely in San Francisco, CA, USA
164K-206K Annually
Senior level
Own compute turn-up from facility availability to ready-for-service: qualify racks at scale (BMC/BIOS/firmware, burn-in, node and cluster validation), automate qualification and provisioning workflows, triage hardware failures and drive RMAs, run mostly-remote turn-up with brief on-site pulses, and partner across network, ICT, ops, and hardware teams.
The summary above was generated by AI
About Fluidstack

We exist to make humanity more free. For most of human history, you farmed or you starved. Technology gave people more time for the things they wanted to do, instead of things they had to do. Powerful AI will be the biggest lever for human choice we've ever built - but only if models are aligned with what humanity actually wants. There are groups building AI who don't share these goals. Whoever deploys frontier compute infrastructure fastest will decide whether AI expands human freedom or shrinks it.

We're singularly focused on delivering 10 to 100s of GWs of compute faster than anyone else, rethinking every layer of the stack. We acquire power, design and build data centers, and operate them - with teams spanning hardware and software. Speed and scale are our key differentiators. Come be a part of building civilization-scale infrastructure for AI.


We hire people who care deeply about this problem space. If that is you, please apply!

How We Operate
  • Be a barrel. Full autonomy. Own things end to end, take on scope without being asked, no permission required to operate outside your core role.

  • Insane urgency. We drive everything forward as fast as possible.

  • Reason from first principles. Challenge every assumption. Zero analogy thinking, no egos, the best idea wins.

  • Love of the game. The frontier of AI is the most interesting problem of our time. We put in long hours at high intensity to push the frontier forward.

  • Build something that actually matters. If you're going to spend your time, spend it on something that matters to the world.


The Infrastructure Team

Examples of key problems the team is working on

  • Bring gigawatts of accelerators from first power-on to production. Facility availability to ready-for-service across thousands of racks per site, with a new data hall landing every few weeks.

  • Make rack qualification faster than the fleet grows. Firmware baselines, burn-in, and cluster validation proven on every rack before a customer workload touches it, at a pace that never becomes the critical path.

  • Scale by tooling, not headcount. Deployed megawatts grow severalfold next year while the team stays near-flat, because anything done twice by hand becomes software.


Role Scope
  • Own compute turn-up from facility availability to ready-for-service: the stretch after the network hands off and before customers run workloads.

  • Qualify racks at scale: establish firmware baselines, configure BMC and BIOS, run burn-in, and validate at node and cluster level across hundreds of racks per site on GPU and custom accelerator platforms.

  • Drive qualification through the base-management Kubernetes platform and provisioning stack (discovery, imaging, firmware updates, shared services), burning down qual queues with tooling rather than manual runs.

  • Triage hardware failures found in qualification: isolate to component, drive RMA and vendor escalation, and feed failure patterns back into the qual gates.

  • Run turn-up remotely by default, with on-site pulses of roughly a week per data hall as new halls reach facility availability, plus occasional overlapping-site weeks.

  • Partner with network deployment, ICT, data center operations, and hardware teams during turn-up windows, and support incident response on freshly-live capacity.

  • Ability to travel 30-40% of the time to our Data Centers and Labs, as needed.


What We're Looking For

The below is a starting point. We always make space for exceptional people, so if you don't fit this role exactly, tell us where you would.

  • You've brought up server or GPU fleets at scale, hundreds of nodes or more, and taken them all the way to production.

  • You work deep in Linux and out-of-band management: BMC, IPMI, and Redfish are daily tools for you, not occasional lookups.

  • You've automated hardware workflows in Python or Go rather than clicking through them, and the second time you do anything by hand you turn it into software.

  • You've worked physically in data halls, racking, cabling, and swapping components, and you're just as effective acting as remote hands or directing them.

  • You triage failures methodically across hardware, firmware, and software, isolating the fault to a component before reaching for a fix.

  • You travel for turn-up windows when a new data hall comes online.

  • Bonus: Kubernetes-based bare-metal provisioning. Accelerator platform bringup (NVIDIA, AMD, or custom). Burn-in and stress harness design. DCIM and inventory tooling.

We are committed to pay equity and transparency.

Fluidstack is an Equal Employment Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, national origin, sexual orientation, gender identity, disability and protected veterans’ status, or any other characteristic protected by law. Fluidstack will consider for employment qualified applicants with arrest and conviction records pursuant to applicable law.

You will receive a confirmation email once your application has successfully been accepted. If there is an error with your submission and you did not receive a confirmation email, please email [email protected] with your resume/CV, the role you've applied for, and the date you submitted your application-- someone from our recruiting team will be in touch.

Similar Jobs

13 Minutes Ago
Remote or Hybrid
US
130K-150K Annually
Senior level
130K-150K Annually
Senior level
AdTech • Consumer Web • Digital Media • eCommerce • Marketing Tech
Leads brand positioning and go-to-market narratives for CPG brands. Develops advertiser-facing materials and packages editorial opportunities using audience and marketplace insights. Partners with editorial, sales, strategy, design, and research teams to enable revenue-generating programs. Manages timelines, deliverables, storytelling quality, and process improvements across brand marketing workstreams.
Top Skills: KeynotePowerPoint
14 Minutes Ago
Remote
United States
128K-216K Annually
Senior level
128K-216K Annually
Senior level
Aerospace • Artificial Intelligence • Computer Vision • Software • Analytics • Defense • Big Data Analytics
Build scalable systems and machine learning models that transform satellite imagery into labeled geospatial data and production computer vision products. Responsibilities include developing data and annotation pipelines, tracking lineage and versioning, deploying inference services, evaluating models, supporting cloud infrastructure, and improving labeling workflows. The role requires strong Python, cloud, containerization, CI/CD, infrastructure-as-code, production support, and deep learning expertise, with preferred experience in GCP, Vertex AI, spatial databases, and satellite imagery.
Top Skills: Ci/CdCloud RunCloud SqlCloud StorageComputer VisionDeep LearningDockerFionaGdalGeopandasGeotiffGoogle Cloud PlatformIamInfrastructure As CodeKubeflow PipelinesMachine LearningPostgisPostgresPythonRasterioShapelyVertex AiVertex Ai PipelinesWorkflow Orchestration
15 Minutes Ago
Remote or Hybrid
USA
15-20 Hourly
Entry level
15-20 Hourly
Entry level
eCommerce • Fashion • Retail • Sales • Wearables • Design
Provides personalized styling and product recommendations, builds customer relationships, drives sales, processes POS transactions, and maintains stockroom and sales-floor operations. The role requires strong communication, product knowledge, retail experience, teamwork, flexibility for nights and weekends, and the ability to lift and maneuver merchandise.
Top Skills: Omnichannel/Virtual Selling ToolsPoint-Of-Sale (Pos) SystemsSocial Media

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account