xAI Logo

xAI

Software Engineer, Data Platform

Reposted One Month Ago
Be an Early Applicant
In-Office
Palo Alto, CA, USA
180K-440K Annually
Senior level
In-Office
Palo Alto, CA, USA
180K-440K Annually
Senior level
Build, operate, and scale distributed data infrastructure (Kafka, Spark, Flink, Trino, HDFS) to enable real-time ML pipelines and analytics at petabyte scale. Design high-throughput, low-latency ingestion and transport, optimize performance, debug distributed systems, and collaborate with ML and product teams to ensure reliable, production-grade data movement and compute.
The summary above was generated by AI

SpaceXAI’s mission is to create AI systems that can accurately understand the universe and aid humanity in its pursuit of knowledge. Our team is small, highly motivated, and focused on engineering excellence. This organization is for individuals who appreciate challenging themselves and thrive on curiosity. We operate with a flat organizational structure. All employees are expected to be hands-on and to contribute directly to the company’s mission. Leadership is given to those who show initiative and consistently deliver excellence. Work ethic and strong prioritization skills are important. All employees are expected to have strong communication skills. They should be able to concisely and accurately share knowledge with their teammates.

ABOUT THE ROLE:

The Data Platform team builds and operates the infrastructure responsible for all large-scale data transport and processing across the company. We own and manage core systems including Apache Kafka, HDFS, Spark, Flink, and Trino, enabling real-time ML pipelines, feed ranking, experimentation, analytics, and observability at petabyte scale. Our team deals with latency-critical workloads, high-throughput streaming, and distributed compute systems that require fault tolerance, performance, and absolute reliability.

As a software engineer on the Data Platform team, you will design, build, and operate the distributed systems powering SpaceXAI's data movement and compute. You will take ownership of infrastructure components that process trillions of events daily, driving the scalability, performance, and reliability of the systems that power product and ML workloads across the company.

RESPONSIBILITIES:
  • Design and implement high-throughput, low-latency data ingestion and transport systems.
  • Scale and optimize multi-tenant Kafka infrastructure supporting real-time workloads.
  • Extend and tune Spark, Flink, and Trino for demanding production pipelines.
  • Build interfaces, APIs, and pipelines enabling teams to query, process, and move data at petabyte scale.
  • Debug and optimize distributed systems, with a focus on reliability and performance under load.
  • Collaborate with ML, product, and infrastructure teams to unblock critical data workflows.
BASIC QUALIFICATIONS:
  • Proven expertise in distributed systems, stream processing, or large-scale data platforms.
  • Proficiency in  Rust, Go, Scala  or similar systems languages.
  • Hands-on experience with  Kafka, Flink, Spark, Trino, or Hadoop*in production.
  • Strong debugging, profiling, and performance optimization skills.
  • Track record of shipping and maintaining critical infrastructure.
  • Comfortable working in fast-moving, high-stakes environments with minimal guardrails.
COMPENSATION AND BENEFITS:

$180,000 - $440,000 USD

Base salary is just one part of our total rewards package at SpaceXAI, which also includes equity, comprehensive medical, vision, and dental coverage, access to a 401(k) retirement plan, short & long-term disability insurance, life insurance, and various other discounts and perks.

SpaceXAI is an equal opportunity employer. For details on data processing, view our Recruitment Privacy Notice.

HQ

xAI Palo Alto, California, USA Office

1450 Page Mill Road, Palo Alto, CA, United States

xAI San Francisco, California, USA Office

3180 18th St., San Francisco, CA, United States

Similar Jobs

Yesterday
Hybrid
Santa Clara, CA, USA
240K-420K Annually
Expert/Leader
240K-420K Annually
Expert/Leader
Artificial Intelligence • Cloud • HR Tech • Information Technology • Productivity • Software • Automation
Architect and engineer RaptorDB deployment, operations, backup, restore, replication, failover, and resilience across hybrid infrastructure. Lead technical direction, cross-functional architecture, cloud-native adoption, performance optimization, automated testing, deployment pipelines, incident response, and operational documentation. Partner with database development, infrastructure, systems engineering, security, and support teams to operate highly available databases at scale.
Top Skills: Cloud-Native PostgresqlDatabase OperatorsGoKubernetesLinuxPythonRaptordb
3 Days Ago
Hybrid
San Francisco, CA, USA
281K-334K Annually
Expert/Leader
281K-334K Annually
Expert/Leader
Artificial Intelligence • Productivity • Software
Lead Notion’s Data Foundations team by setting multi-quarter platform strategy, developing engineers and technical leaders, and guiding complex initiatives across data lakes, streaming, distributed compute, governance, and reliability. Drive enterprise-ready capabilities including data residency, access controls, lifecycle management, and key management. Improve Kafka, Debezium, S3/Iceberg, Spark, EMR, and Athena scalability, reliability, usability, and cost efficiency while partnering across engineering, AI, search, infrastructure, and security teams.
Top Skills: Amazon AthenaAmazon EmrAmazon S3Apache IcebergSparkDebeziumKafka
6 Days Ago
Hybrid
Santa Clara, CA, USA
176K-308K Annually
Senior level
176K-308K Annually
Senior level
Artificial Intelligence • Cloud • HR Tech • Information Technology • Productivity • Software • Automation
Develop highly scalable Java software for ServiceNow’s platform persistence systems. Manage data lifecycles, storage, retrieval, security, performance, and recovery across relational and distributed systems. Collaborate cross-functionally on data initiatives, telemetry, visualization, and Kubernetes deployments while applying concurrency, parallel programming, and AI-assisted development practices.
Top Skills: Amazon S3Apache IcebergJavaJavaScriptKubernetesRelational DatabasesServicenow Platform

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account