General Compute Logo

General Compute

Founding Platform Engineer

Posted 2 Days Ago
Be an Early Applicant
Hybrid
San Francisco, CA, USA
Entry level
Hybrid
San Francisco, CA, USA
Entry level
Build the foundational inference cloud platform, including its control plane, API, serving layer, routing, model placement, scheduling, reliability, and observability. The role involves scaling a heterogeneous GPU and ASIC fleet, partnering with data center and model deployment teams, and shaping platform architecture as an early individual contributor at a small startup.
The summary above was generated by AI
About us

General Compute is the neocloud for alternative chips.

Inference is fragmenting: purpose-built silicon from SambaNova, Cerebras, Positron, d-Matrix, and others already beats GPUs on decode, and we productionize that hardware — we buy the racks, find the data center space, and run it for our customers. Each piece of hardware runs the workload it's actually built for: prefill stays on GPUs, decode moves to the chip built for it, and today that means generating tokens 5–7× faster than existing GPU-based competitors. Our customers are frontier labs, fast-growing AI application companies, and asset-light clouds.

We closed a $15M seed round in May 2026, and have since closed a $400M debt facility — $100M funded upfront by Upper90, with the balance available for drawdown — collateralized by our inference chips.

About the role

You will build the inference cloud itself — the control plane, API, and serving layer that turn racks into a sellable product. There's no existing platform team to inherit or manage, no legacy system to work around, and no established playbook to follow — just the platform itself to build, with reliability treated as core infrastructure from day one rather than something bolted on after the first outage.The technical problem is also genuinely unsolved elsewhere. The fleet is heterogeneous by design — GPUs for prefill, multiple ASIC vendors for decode — so there's no single-vendor playbook to lean on; you'll be defining how a mixed-hardware inference cloud gets scheduled, routed, and served reliably, in close partnership with the teams standing up the physical fleet.

What you'll do:
  • Build/own the control plane — routing, model placement, scheduling across a mixed ASIC/GPU pool

  • Build the API and serving layer exposing rack capacity as a sellable product

  • Build in reliability and observability from day one

  • Scale the platform ahead of the demand curve

  • Partner closely with data center deployment and model bring-up teams

  • Be a founding technical voice on platform architecture

What we need from you:
  • Strong systems engineering background on cloud control planes/serving infra at scale

  • Comfort being a high-impact IC rather than a manager

  • Track record building reliability from scratch

  • Comfort with hardware heterogeneity/ambiguity

  • Genuine interest in being an early hire at a ~6–7 person company.

Nice-to-haves:
  • LLM-serving infra experience (vLLM, TGI, Ray Serve, etc.)

  • Experience running non-NVIDIA accelerators (TPUs/ASICs) in production.

Similar Jobs

18 Days Ago
Hybrid
160K-190K Annually
Senior level
160K-190K Annually
Senior level
Artificial Intelligence • Healthtech • Sales • Software
Lead the platform engineering for a voice-first AI product serving life sciences field teams. Own backend services, APIs, real-time systems, multi-agent runtime, cloud infrastructure, integrations (CRM, Slack, Teams), observability, security/compliance, and data/auditability. Ship internal and customer-facing tooling and collaborate closely with co-founders and a small engineering team.
Top Skills: APIsCloud InfrastructureMicrosoft TeamsMulti-Agent SystemsObservabilityReal-Time StreamingSalesforceSlackVeevaVoice Ai
One Month Ago
In-Office
Senior level
Senior level
AdTech • Big Data • Marketing Tech • Analytics
Design and build the platform layer, solve challenging infrastructure and systems problems, define and drive the roadmap, and scale systems to handle trillions of data points to support growth of major consumer brands.
Senior level
Financial Services
Manage end-to-end client and employee events across North America, including strategy, registration, speaker logistics, budgets, compliance, vendor coordination, production, sponsorships, and onsite execution. Build project plans, manage event communications and reporting, troubleshoot registration technology, oversee event materials and audiovisual needs, and collaborate with senior stakeholders, sales, compliance, planners, and vendors. The role requires advanced Microsoft Office skills, strong organization, attention to detail, and approximately 15% travel.
Top Skills: Microsoft CopilotExcelMS OfficeMicrosoft PowerpointMicrosoft WordRegistration Management SystemsVirtual Event Platforms

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account