CoreWeave Logo

CoreWeave

Senior Manager, Technical Support Engineering - Cloud

Reposted One Month Ago
In-Office
2 Locations
198K-264K Annually
Senior level
In-Office
2 Locations
198K-264K Annually
Senior level
Lead and grow an infrastructure support team to maintain and optimize physical data center and GPU-based compute environments. Own incident and escalation management, improve support processes, manage ticket workflows and 24/7 shift logistics, and collaborate with product and engineering to ensure reliable hardware delivery and client communication.
The summary above was generated by AI
CoreWeave is The Essential Cloud for AI™. Built for pioneers by pioneers, CoreWeave delivers a platform of technology, tools, and teams that enables innovators to build and scale AI with confidence. Trusted by leading AI labs, startups, and global enterprises, CoreWeave combines superior infrastructure performance with deep technical expertise to accelerate breakthroughs and turn compute into capability. Founded in 2017, CoreWeave became a publicly traded company (Nasdaq: CRWV) in March 2025. Learn more at www.coreweave.com.

CoreWeave is The Essential Cloud for AI™. Built for pioneers by pioneers, CoreWeave delivers a platform of technology, tools, and teams that enables innovators to build and scale AI with confidence. Trusted by leading AI labs, startups, and global enterprises, CoreWeave combines superior infrastructure performance with deep technical expertise to accelerate breakthroughs and turn compute into capability. Founded in 2017, CoreWeave became a publicly traded company (Nasdaq: CRWV) in March 2025. Learn more at www.coreweave.com.

About the Team

The Technical Support Engineering - Infrastructure team sits within CoreWeave's Global Field Organization (GFO) and is the front line for every customer running AI and HPC workloads on our platform. Operating 24/7/365, the team supports the Kubernetes-powered infrastructure behind the AI revolution: GPU compute, high-performance networking and storage, Slurm/HPC clusters, and the large-scale, mission-critical training workloads that run on them.

We operate on a Direct-to-Expert model. Instead of routing customers through generic support tiers, we get the right expert on the problem fast while a single owner stays accountable for the issue end-to-end. That means fewer handoffs, clearer ownership, and a high-touch experience that our customers feel on every ticket. The team is the connective tissue between our customers and CoreWeave's engineering organization. We triage and resolve deep technical issues directly, coordinate with Product Engineering, Specialist Field Engineers, and domain specialists when needed, and feed what we learn in the field straight back into the product roadmap.

About the Role

As the Senior Manager of Technical Support Engineering - Infrastructure, you'll own, scale, and continuously improve CoreWeave's infrastructure support function — a large, globally distributed 24/7/365 organization of skilled engineers who resolve our customers' most complex technical challenges with deep expertise, efficiency, and empathy. High-touch, expert-led support is one of CoreWeave's clearest differentiators, which gives this role prominent visibility and influence across the company.

You'll operate at the department level, setting the technical and operational bar for the entire function and designing the operating systems (coverage/staffing model, Direct-to-Expert partnerships, quality program, metrics) that let support scale reliably. You'll lead with empathy and invest in each person's growth, while building the structure and processes that let the team scale with CoreWeave's hyper-growth without compromising the culture we care deeply about.

In this role, you will:
  • Own the strategy, health, and performance of the entire 24/7/365 infrastructure support function, helping scale coverage and capability across regions and domains as CoreWeave grows.
  • Own talent acquisition and retention by hiring, onboarding, and developing engineers through diligent performance management and coaching tailored to each individual's needs.
  • Stay hands-on: dig into complex, customer-impacting issues alongside your team and serve as a senior technical escalation point, ensuring the highest quality of support.
  • Build and facilitate the enablement and career-development frameworks for onboarding, technical training, and progression paths that raise capability across the whole function and create a repeatable path for career growth.
  • Implement quality assurance measures, including ticket reviews and best-practice playbooks, that raise the bar on resolution speed, accuracy, and consistency.
  • Own the support operating model for the function including  roles and responsibilities, escalation boundaries, coverage/staffing model, and the response and resolution targets (SLOs) that keep every customer's experience consistently excellent as volume scales.
  • Lead customer communication during critical incidents and resolve conflicts with clarity, composure, and empathy.
  • Track and report on KPIs focused on team performance and customer satisfaction, and own the strategic planning for the team's growth and scalability.
  • Own the cross-functional interface between support and Product Engineering, Specialist Field Engineers, and domain teams,  building the alignment mechanisms and feedback loops that make the Direct-to-Expert model work at scale for escalations, live incidents, and customer needs.
  • Champion the voice of the customer, turning recurring support patterns into product, tooling, and process improvements.
  • Help set the multi-quarter vision and operating plan for infrastructure support, and represent the function in company-level planning, headcount, and prioritization discussions.
Who you are:
  • You have 5+ years of people-leadership experience and 8+ years total in technical support/operations, including running a 24/7 support function at scale in a cloud operations environment.
  • You have a strong background in Linux, containerization technologies, and Kubernetes, and you understand virtualization and cloud computing concepts. Experience at a hyperscaler or cloud infrastructure provider is a strong plus.
  • You lead with empathy and aren't afraid to get your hands dirty. You do the work alongside your direct reports and model the standard you set.
  • You're energized by leadership excellence and talent development: diligent performance management, coaching, and growing each engineer according to their individual needs.
  • You've built enablement and quality programs that scaled across a function or multiple teams — and can show the measurable improvement they drove.
  • You've designed the operating model for a support org, including coverage/staffing model, escalation boundaries, SLOs, and evolved it as the org scaled.
  • You're a calm, clear communicator with customers and executives during critical incidents, and you resolve conflicts effectively across teams and organizational boundaries.
  • You think in systems and multi-quarter plans.  You've defined KPI/SLO frameworks, reported to senior leadership, and owned capacity and growth planning for a function, not just a single team.
  • You've led a globally distributed team across time zones.
  • Leadership & Communication: proven ability to lead through senior talent and set direction for a function, with executive-level communication skills.
  • Strategic & Operational Planning: you build operating systems, plans, and metrics that scale a function beyond what any one person can hold.
  • Problem-Solving & Adaptability: robust problem-solving skills and adaptability at organizational scale, in a fast-paced, hyper-growth environment.
  • Program Management: experience with program-management tools and methodologies.
Bonus points if you have:
  • Experience supporting AI/ML, HPC, or GPU-accelerated workloads at scale.
  • Hands-on Kubernetes operations experience (CKA certification a plus).
  • Familiarity with Slurm/SUNK, RDMA networking, distributed storage, and observability tooling such as Grafana.
  • You have some experience with infrastructure as it relates to Data Center Operations.

Wondering if you're a good fit? We believe in investing in our people, and value candidates who can bring their own diversified experiences to our teams – even if you aren't a 100% skill or experience match. Here are a few qualities we've found compatible with our team. If some of this describes you, we'd love to talk.

  • You're an expert in what it takes to be an excellent leader and foster an environment in which people are excited and inspired to participate.
  • You love to dive into problems, test for solutions, and enjoy engaging with customers.
  • You're excited and curious about AI.

Why CoreWeave?

At CoreWeave, we work hard, have fun, and move fast! We're in an exciting stage of hyper-growth that you will not want to miss out on. We're not afraid of a little chaos, and we're constantly learning. Our team cares deeply about how we build our product and how we work together, which is represented through our core values:

  • Be Curious at Your Core
  • Act Like an Owner
  • Empower Employees
  • Deliver Best-in-Class Client Experiences
  • Achieve More Together

We support and encourage an entrepreneurial outlook and independent thinking. We foster an environment that encourages collaboration and provides the opportunity to develop innovative solutions to complex problems. As we get set for take off, the growth opportunities within the organization are constantly expanding. You will be surrounded by some of the best talent in the industry, who will want to learn from you, too. Come join us!

The base salary range for this role is $198,000 to $264,000. The starting salary will be determined based on job-related knowledge, skills, experience, and market location. We strive for both market alignment and internal equity when determining compensation. In addition to base salary, our total rewards package includes a discretionary bonus, equity awards, and a comprehensive benefits program (all based on eligibility).


What We Offer

The range we’ve posted represents the typical compensation range for this role. To determine actual compensation, we review the market rate for each candidate which can include a variety of factors. These include qualifications, experience, interview performance, and location.

In addition to a competitive salary, we offer a variety of benefits to support your needs. The benefits below reflect our US-based offerings for full-time employees; for roles in other locations, benefits vary and are shared during the hiring process. These include:

  • Medical, dental, and vision insurance - 100% paid for by CoreWeave
  • Company-paid Life Insurance 
  • Voluntary supplemental life insurance 
  • Short and long-term disability insurance 
  • Flexible Spending Account
  • Health Savings Account
  • Tuition Reimbursement 
  • Ability to Participate in Employee Stock Purchase Program (ESPP)
  • Mental Wellness Benefits through Spring Health 
  • Family-Forming support provided by Carrot
  • Paid Parental Leave 
  • Flexible, full-service childcare support with Kinside
  • 401(k) with a generous employer match
  • Flexible PTO
  • Catered lunch each day in our office and data center locations
  • A casual work environment
  • A work culture focused on innovative disruption

California Applicants

California Consumer Privacy Act 

Equal Opportunity & Accommodations

CoreWeave is an equal opportunity employer, committed to fostering an inclusive and supportive workplace. All qualified applicants and candidates will receive consideration for employment without regard to race, color, religion, sex, disability, age, sexual orientation, gender identity, national origin, veteran status, or genetic information.

As part of this commitment and consistent with the Americans with Disabilities Act (ADA), CoreWeave will ensure that qualified applicants and candidates with disabilities are provided reasonable accommodations for the hiring process, unless such accommodation would cause an undue hardship. If reasonable accommodation is needed, please contact: [email protected].

Export Control Compliance

This position requires access to export controlled information. To conform to U.S. Government export regulations applicable to that information, applicant must either be (A) a U.S. person, defined as a (i) U.S. citizen or national, (ii) U.S. lawful permanent resident (green card holder), (iii) refugee under 8 U.S.C. § 1157, or (iv) asylee under 8 U.S.C. § 1158, (B) eligible to access the export controlled information without a required export authorization, or (C) eligible and reasonably likely to obtain the required export authorization from the applicable U.S. government agency.  CoreWeave may, for legitimate business reasons, decline to pursue any export licensing process.

CoreWeave Sunnyvale, California, USA Office

CoreWeave Sunnyvale, CA Office

Sunnyvale, California, United States

Similar Jobs at CoreWeave

2 Days Ago
In-Office
Sunnyvale, CA, USA
157K-210K Annually
Senior level
157K-210K Annually
Senior level
Cloud • Information Technology • Machine Learning
Lead technical programs delivering GPU and CPU compute fleets from infrastructure handover through provisioning, validation, operational acceptance, and steady-state operations. Translate customer and business requirements into delivery plans, manage dependencies and blockers, assess capacity gaps, drive prioritization, establish readiness criteria, and coordinate cross-functional engineering and operations teams. Define delivery metrics, improve forecasting, and identify tooling, automation, and process improvements to increase provisioning throughput and operational readiness.
Top Skills: Computer NetworkingCpuData Center InfrastructureGpuHardware/Software IntegrationInfrastructure AutomationServer ProvisioningSQL
2 Days Ago
In-Office
Sunnyvale, CA, USA
157K-210K Annually
Senior level
157K-210K Annually
Senior level
Cloud • Information Technology • Machine Learning
Own and operationalize CoreWeave’s third-party risk management and external data-sharing programs. Responsibilities include maintaining data inventories, defining review and approval controls, monitoring data flows, conducting third-party security risk assessments, evaluating privacy and security controls, tracking remediation, reporting risk metrics, and advising technical and executive stakeholders. The role also contributes to TPRM strategy, privacy compliance, audit support, and cross-functional coordination with engineering, procurement, IT, and Legal.
Top Skills: Ccpa/CpraDpasGdprIso 27001Iso 27701Nist CsfPythonSccsSoc 2SQL
2 Days Ago
In-Office
Sunnyvale, CA, USA
157K-210K Annually
Junior
157K-210K Annually
Junior
Cloud • Information Technology • Machine Learning
Manage and mature CoreWeave’s third-party risk management program across the vendor lifecycle. Responsibilities include conducting security risk assessments, evaluating controls and compliance evidence, tracking findings through remediation, performing vendor reassessments, maintaining vendor inventories and risk registers, and improving questionnaires, workflows, and documentation. The role coordinates with engineering, procurement, IT, Legal, vendors, and business stakeholders while supporting security, privacy, audit, and compliance objectives.
Top Skills: Cloud TechnologiesCsps/HyperscalersGrc PlatformsSecurity Control FrameworksTprm Platforms

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account