Maximum of 25 job preferences reached.
Top Engineering Jobs in San Francisco, CA
Cloud • Information Technology • Machine Learning
Leads technical due diligence for prospective data center sites, evaluating electrical, mechanical, civil, structural, architectural, utility, zoning, and infrastructure feasibility. Reviews provider proposals, capacity data, conceptual layouts, and building documentation; identifies risks, upgrades, cost and schedule impacts; prepares go/no-go recommendations and technical assessments. Coordinates with utilities, landlords, developers, internal subject matter experts, and design teams, then hands secured sites to Design Managers. Frequent travel is required.
Top Skills:
CdusChilled Water PlantsDashboardsIt/LvMepPower DistributionSubstationsTcs LoopsUtility Infrastructure
Cloud • Information Technology • Machine Learning
Plans, procures, and manages backbone network capacity, including wavelength and dark fiber services. Reviews vendor solutions, forecasts network demand, optimizes circuit utilization, ensures redundancy and diversity, manages telecommunications vendors, maintains technical documentation, and supports network expansion. The role collaborates with engineering, operations, and product teams to resolve capacity issues and incorporate emerging technologies into global backbone planning.
Top Skills:
CienaCiscoCoherent OpticsDark FiberDwdmFiberlocatorGisInfineraIp RoutingIp SwitchingNetwork AutomationNokia
Cloud • Information Technology • Machine Learning
Design and review data center white space infrastructure for high-density computing environments. Coordinate electrical, mechanical, cooling, low-voltage, cabling, and network pathway systems; review drawings, submittals, RFIs, and field conditions; oversee consultants; resolve design and constructability issues; and support construction, commissioning, and deployment. Contribute to design standards, scalability, technical documentation, and mentorship across CoreWeave’s data center portfolio.
Top Skills:
AsanaAshraeAutocadBicsiBluebeamBranch CircuitsBuswaysCable TrayFan WallsGrounding SystemsHot Aisle ContainmentLadder RackLiquid CoolingNfpaOverhead ConveyancePdusRack Power DistributionRemote CdusRevitRppsSmartsheetStructured CablingTia-942Uptime Institute
Cloud • Information Technology • Machine Learning
Build and operate full-stack applications and AI-facing features for CoreWeave’s internal data platform. Responsibilities include developing TypeScript, React, Next.js, and Python services; designing APIs and relational data models; integrating analytical data platforms; deploying on Kubernetes; and owning production reliability, testing, observability, and on-call support. The role also involves collaborating with non-engineers to define processes, shipping LLM-backed applications, and evaluating AI output quality.
Top Skills:
BigQueryCi/CdDelta LakeDockerHelmHudiIcebergKafkaKubernetesLanggraphNext.JsPostgresPythonReactSnowflakeSparkSQLStarrocksTypescript
Cloud • Information Technology • Machine Learning
Build and ship full-stack software for CoreWeave’s internal data infrastructure and operational applications. Develop TypeScript, React, and Next.js frontends; Python services on Kubernetes; SQL queries and data models; and AI-enabled interfaces such as text-to-SQL tools, retrieval systems, and agents. Own scoped features through production, including testing and post-release fixes, while collaborating with senior engineers and internal users.
Top Skills:
BigQueryCi/CdDockerGitHelmHttp ApisJavaScriptKubernetesLlmsNext.JsPythonReactSnowflakeSparkSQLStarrocksTypescript
New
Cut your apply time in half.
Use ourAI Assistantto automatically fill your job applications.
Use For Free
Cloud • Information Technology • Machine Learning
Serves as a technical lead for customers using CoreWeave cloud infrastructure, with emphasis on Kubernetes in high-performance computing environments. Responsibilities include designing and deploying tailored solutions, leading proofs of concept, optimizing AI/ML workloads, conducting technical reviews, advising on product strategy, collaborating with engineering teams, and representing CoreWeave at industry events.
Top Skills:
Artificial IntelligenceAutomationCloud ComputingDistributed SystemsHigh-Performance ComputingInference FrameworksInfinibandKubernetesMachine LearningMulti-CloudNvidia Collective Communications Library (Nccl)Nvidia GpusScriptingSlurm
Cloud • Information Technology • Machine Learning
Build and operate CoreWeave’s multi-cloud internal infrastructure across GCP, Azure, and AWS. Manage cloud foundations, IAM, networking, Terraform infrastructure as code, monitoring, logging, CI/CD, backups, disaster recovery, and production reliability. Troubleshoot complex infrastructure and application issues, support cloud migrations and acquisitions, implement secure architecture patterns, and contribute to incident response and root-cause analysis. Mentor junior engineers and communicate technical decisions across teams.
Top Skills:
AWSBashCi/CdCloud BuildCloud RunCompute EngineDnsGCPGitGithub ActionsGkeIamLinuxAzurePowershellPythonSecret ManagerTerraformVnetsVpcVpnWindowsWorkload Identity Federation
Cloud • Information Technology • Machine Learning
As a Solutions Architect, engage with customers to architect cloud solutions, optimize workloads, and provide technical leadership while collaborating with engineering teams.
Top Skills:
Distributed TrainingInferenceKubernetesMachine Learning OperationsNetworking EngineeringSlurm Workload Manager
Cloud • Information Technology • Machine Learning
As a Solutions Architect at CoreWeave, you'll lead customer engagement, optimize cloud infrastructure solutions, and collaborate closely with engineering teams on product development and enhancements.
Top Skills:
Distributed TrainingInferenceKubernetesMachine Learning OperationsNetworking EngineeringNvidia GpusSlurm Workload Manager
Cloud • Information Technology • Machine Learning
Lead and expand the SRE/Production Engineering team to ensure platform reliability, performance, and operational efficiency. Drive SRE strategy, automation-first practices (IaC, Terraform, Kubernetes), incident management, postmortems, SLO/SLA frameworks, on-call strategy, and system hardening. Collaborate with engineering, product, hardware, and security teams to build scalable, resilient, self-healing AI-focused cloud infrastructure.
Top Skills:
Ai InfrastructureBare MetalChaos EngineeringDpusGpu-Accelerated WorkloadsHpc ClustersInfrastructure-As-CodeKubernetesService MeshSlo/Sla FrameworksTerraform
Let Your Resume Do The Work
Upload your resume to be matched with jobs you're a great fit for.
Success! We'll use this to further personalize your experience.
All Filters
Total selected ()
No Results
No Results


