xAI Logo

xAI

Software Engineer - Kernels/CUDA (C++)

Reposted One Month Ago
Be an Early Applicant
In-Office
Palo Alto, CA, USA
180K-440K Annually
Senior level
In-Office
Palo Alto, CA, USA
180K-440K Annually
Senior level
Design, build, and optimize massive GPU clusters and low-level CUDA kernels for AI training and inference. Work on Linux kernel internals, virtualization (KVM, Firecracker), custom orchestration, profiling, and infrastructure-as-code to maximize performance and scalability for production AI workloads.
The summary above was generated by AI

SpaceXAI’s mission is to create AI systems that can accurately understand the universe and aid humanity in its pursuit of knowledge. Our team is small, highly motivated, and focused on engineering excellence. This organization is for individuals who appreciate challenging themselves and thrive on curiosity. We operate with a flat organizational structure. All employees are expected to be hands-on and to contribute directly to the company’s mission. Leadership is given to those who show initiative and consistently deliver excellence. Work ethic and strong prioritization skills are important. All employees are expected to have strong communication skills. They should be able to concisely and accurately share knowledge with their teammates.

ABOUT THE ROLE:

We are building one of the world’s largest AI supercomputers from the ground up. As part of the Compute Infrastructure team, you will own both the raw GPU supercomputer and the platform layer that runs on top of it. You will work across the full stack — from low-level GPU kernel optimizations and Linux kernel internals to massive-scale orchestration and virtualization — to make training and inference at SpaceXAI as fast, reliable, and scalable as possible.

This is a broad, high-impact role that combines hardcore supercompute and compute infrastructure work. Your contributions will directly accelerate Grok’s training speed and overall AI progress.

RESPONSIBILITIES:

  • Design, build, and optimize massive GPU clusters for extreme-scale training and inference workloads
  • Develop and tune low-level CUDA kernels (GeMM, Attention, etc.), using CUTLASS, Tensor Cores, and Nsight for maximum performance
  • Profile, debug, and eliminate bottlenecks across GPU memory hierarchy, networking fabric, filesystems, and multi-GPU operation
  • Collaborate closely with AI research teams to deliver production-grade performance and scalability
PREFERRED SKILLS AND EXPERIENCE:
  • Deep low-level systems programming (C/C++/PTX/SASS)
  • Strong experience with large-scale GPU clusters or distributed compute infrastructure at production scale
  • Hands-on work with GPU kernel optimization (CUTLASS, custom kernels, Nsight profiling)
  • Track record of building or running high-performance infrastructure for AI workloads (training or inference platforms)
  • Ability to reason from first principles and optimize for both memory-bound and compute-bound scenarios
COMPENSATION AND BENEFITS:

$180,000 - $440,000 USD

Base salary is just one part of our total rewards package at SpaceXAI, which also includes equity, comprehensive medical, vision, and dental coverage, access to a 401(k) retirement plan, short & long-term disability insurance, life insurance, and various other discounts and perks.

SpaceXAI is an equal opportunity employer. For details on data processing, view our Recruitment Privacy Notice.

HQ

xAI Palo Alto, California, USA Office

1450 Page Mill Road, Palo Alto, CA, United States

xAI San Francisco, California, USA Office

3180 18th St., San Francisco, CA, United States

Similar Jobs

43 Minutes Ago
In-Office
144K-217K Annually
Senior level
144K-217K Annually
Senior level
Aerospace • Artificial Intelligence • Machine Learning • Robotics • Software
Deploys, integrates, configures, troubleshoots, and supports airborne vision and detection systems across laboratory, production, customer, and field environments. Performs hardware and software integration, wiring and electro-mechanical work, system testing, configuration management, documentation, customer training, and field-service logistics. Coordinates with engineering, production, quality, program, and customer teams while supporting deployments and demonstrations. Requires approximately 50% travel and ability to obtain a Secret clearance.
Top Skills: CanDigital MultimetersErp SystemsEthernetIpc/Whma-A-620LinuxMaintenance-Tracking SystemsNetwork AnalyzersNetworkingOscilloscopesPower SuppliesProduct Lifecycle Management SystemsSerial InterfacesWindows
An Hour Ago
Hybrid
Livermore, CA, USA
15-24 Hourly
Entry level
15-24 Hourly
Entry level
eCommerce • Fashion • Retail • Sales • Wearables • Design
Temporary retail associate serving as a stylist and trusted product advisor. Responsibilities include greeting customers, understanding their needs, providing styling recommendations, driving sales through product knowledge and storytelling, completing POS transactions, maintaining stockroom organization, and supporting omni/virtual selling. Requires flexible scheduling, physical ability to handle merchandise, strong communication skills, teamwork, and prior retail experience.
Top Skills: Omnichannel SellingPoint-Of-Sale (Pos) SystemsSocial Media
An Hour Ago
Hybrid
San Francisco, CA, USA
255K-300K Annually
Entry level
255K-300K Annually
Entry level
Artificial Intelligence • Productivity • Software
Build and operate Notion’s model layer by integrating frontier models, improving inference reliability, implementing failover and observability, and developing model-level capabilities. The role requires production experience with LLM APIs, evaluation-driven development, and strong judgment around latency, cost, reliability, and output quality. The engineer will also collaborate across AI teams to unblock adoption and ensure fixes are delivered.
Top Skills: Ai Inference SystemsLarge Language ModelsLlm ApisModel EvaluationObservabilityStreamingTool Calling

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account