RevolutionParts Logo

RevolutionParts

Staff Data Engineer

Posted 21 Days Ago
Remote or Hybrid
Hiring Remotely in United States
170K-185K Annually
Expert/Leader
Remote or Hybrid
Hiring Remotely in United States
170K-185K Annually
Expert/Leader
Own the 2-3 year architecture and reliability of high-volume catalog, pricing, and inventory ingestion. Lead migrations from legacy batch to modern systems, define data SLAs, build observability and orchestration, mentor senior engineers, and resolve cross-team, high-severity data problems.
The summary above was generated by AI
RevolutionParts is not just a pioneering force in the automotive eCommerce realm; we're actively seeking passionate and talented individuals to join our squad of Revolutionaries (yes, that's what we call ourselves!). As leaders in providing streamlined, user-friendly solutions, we empower automotive brands to maximize online sales. Our commitment to technology, top-notch customer service, and a profound understanding of the automotive market sets us apart. If you're ready to revolutionize the eCommerce space for automotive parts and accessories, consider joining our dynamic team of Revolutionaries.

The Role
Most data engineering roles hand you a Jira board. This one hands you a whiteboard and asks what should be on it.
RevolutionParts powers parts and accessories commerce for thousands of automotive dealers and OEMs across North America. The data behind all of it (catalog, pricing, inventory) moves through a high-volume ingestion system that has scaled with the business. It was the right architecture for where we were. It isn’t the right architecture for where we’re going.

We need someone who can keep this system reliable today while making it obsolete on a timeline they define.
The target architecture doesn't exist yet. The technical bar for this domain gets set by whoever takes this role. If that's an uncomfortable amount of open space, this probably isn't the right fit. If it sounds like the kind of problem worth leaving your current job for, read on.

Responsibilities
Strategic Leadership & Architectural Ownership
  • You are the technical authority for data ingestion at RevolutionParts. You lead through expertise, not authority.
  • Own the 2-3 year architectural vision for data ingestion. That means the destination, the migration sequence, the tradeoffs at each stage, and the criteria that determine when the current system has earned its retirement.
  • Set the engineering standards that govern how every team builds on and interfaces with core data infrastructure: schema design, data contracts, query optimization, observability. What you establish here becomes the organization’s baseline.
  • Shape technical strategy across Product, BI, Platform Engineering, and Executive Leadership. Not as an advisor. As the person who drives alignment, cuts through ambiguity, and owns the outcomes of complex multi-quarter initiatives from discovery through delivery.
  • Take ownership of the highest-severity, most ambiguous problems in the data domain: the ones that cross team boundaries, have no clear owner, and have already resisted resolution.

Execution & Operational Excellence

  • Hold ultimate accountability for the architecture and production performance of our catalog, pricing, and inventory ingestion systems, with the technical depth to make decisions no one else in the organization is positioned to make.
  • Define the reliability bar for data across the organization. Build the monitoring, alerting, and validation frameworks that turn data quality from a best-effort into a contractual commitment with clear SLAs and owners.
  • Make final, binding technical debt decisions for the ingestion domain, weighing immediate stability against long-term architectural health. Document the reasoning with enough clarity that it survives organizational change 18 months from now.
  • Elevate the technical ceiling of the data engineering organization through direct mentorship of Senior Engineers on distributed systems, high-volume database performance, and data modeling at scale. Your impact here compounds beyond your own output.

Requirements
10+ years in data or software engineering, at least 3 at Staff level or equivalent owning architectural decisions on high-volume production systems.

Python and Spark/PySpark at petabyte scale — production systems, not notebooks. You tune Spark from first principles: partition strategy, join optimization, dynamic allocation, skew diagnosis.
  • Designed and operated distributed job execution systems: dynamic compute provisioning, variable workload profiles, job isolation, and resource contention at scale.
  • Deep experience with message queue architectures in production: fan-out patterns, poison pill handling, dead letter queues, consumer lag at scale.
  • Built observability into systems that had none — monitoring, alerting, lineage, and pipeline health designed in from the ground up, not dashboards bolted on afterward.

Built pipeline orchestration infrastructure, not just DAGs. You have strong opinions about operability because you've inherited systems that weren't.

  • You set the engineering quality bar. Reliable, efficient, documented, testable, maintainable — and you hold the team to the same standard.

Led a migration from legacy batch infrastructure (custom schedulers, daemon-based systems, cron pipelines) to modern architecture without taking down production. We'll go deep on this in the interview.
  • Deep AWS in production: EKS, EC2 fleet management, SQS, RDS. Operated at scale, not just deployed into it.
  • Kubernetes in production — workload behavior, compute right-sizing for variable job profiles, failure modes under load.
  • Streaming in production: Kafka, Flink, Kinesis, or Redpanda. You've made the batch-vs-streaming call in both directions and can defend either.
  • Cloud data warehouse architecture — Snowflake, BigQuery, or Databricks. Clustering, partitioning, cost management, mixed analytical and operational workloads.
  • You use AI coding tools daily and have shipped production work because of it.
  • You write architecture docs engineers trust and can brief a VP on the same decision. Both matter at this level.
  • BS or MS in Computer Science, Engineering, or equivalent.

AI Fluency & Modern Tooling
At RevolutionParts, we expect team members to actively use modern tools — including AI-powered systems — to improve decision-making, productivity, and quality of work.
This includes:
  • Using AI tools responsibly to accelerate research, analysis, documentation, and problem-solving
  • Exercising strong judgment around data privacy, accuracy, and ethical use
  • Continuously learning and adapting as AI capabilities evolve
Proven examples of using AI to improve outcomes in prior roles is expected.

RevolutionParts is proud to provide all full-time Revolutionaries with a comprehensive employment package including competitive compensation, career development, benefits, 401K match, parental leave, and many more valuable perks. You can learn more about our core-value driven culture at our career page.

RevolutionParts is an Equal Opportunity Employer; we value diversity. We do not discriminate on the basis of race, religion, color, national origin, gender, gender orientation, gender identity or expression, sexual identity, sexual orientation, age, marital status, family status, genetic information, veteran status, or disability status.

Please Note: You will only receive correspondence through the GEM ATS or from a @revolutionparts.com email address. If you are receiving communication through any other platform or domain, it may be fraudulent, and we urge you to ignore the communication.
Compensation
The base pay range for this role is $170,000 – $185,000 per year.

Similar Jobs

3 Days Ago
Remote or Hybrid
286K-392K Annually
Senior level
286K-392K Annually
Senior level
Fintech • Machine Learning • Payments • Software • Financial Services
Leads technical strategy for enterprise data pipelines and data-sharing platforms. Builds scalable, resilient, high-performance systems using AWS, lakehouse architecture, Kafka, Flink, Spark, Snowflake, and Databricks. Develops code, drives engineering excellence and technology adoption, influences enterprise stakeholders, advises on platform capabilities, mentors engineers, and recruits technical talent. The role requires deep data engineering and architecture expertise, hands-on leadership, and innovation across internal and external data environments.
Top Skills: AWSDatabricksFlinkJavaKafkaLakehousePythonScalaSnowflakeSparkSQL
7 Days Ago
Easy Apply
Remote
U.S.
Easy Apply
187K-255K Annually
Senior level
187K-255K Annually
Senior level
Artificial Intelligence • Enterprise Web • Software • Design • Generative AI
Designs and operates batch, streaming, and real-time data platforms and pipelines using Spark, Kafka, Iceberg, Airflow, and cloud infrastructure. Owns data lake evolution, data quality, observability, reliability, governance, event instrumentation, schema management, and privacy controls for PII, retention, deletion, and residency. Leads complex initiatives end-to-end, troubleshoots production issues, mentors engineers, and develops AI agent harnesses that enforce data engineering standards and quality gates.
Top Skills: AirflowAmazon EmrApache IcebergChange Data Capture (Cdc)Ci/CdDruidEksInfrastructure As CodeKafkaKubernetesMwaaSparkSpark Structured StreamingSQL
15 Days Ago
Easy Apply
Remote
USA
Easy Apply
207K-244K Annually
Senior level
207K-244K Annually
Senior level
Artificial Intelligence • Blockchain • Fintech • Financial Services • Cryptocurrency • NFT • Web3
Architect, build, and operate low-latency market data systems, including feed handlers, normalization, distribution, and exchange connectivity. Design high-throughput infrastructure for real-time trading data, improve reliability and performance, participate in observability, on-call, and incident response, and deliver well-tested features end to end. Partner with engineering, product, and cross-functional stakeholders while leading a lean, high-impact team.
Top Skills: AeronC++FixGenerative AiItch/OuchJavaMulticastSbe

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account