Mytra Logo

Mytra

Senior Data Engineer

Posted 4 Days Ago
In-Office
Brisbane, CA, USA
180K-200K Annually
Senior level
In-Office
Brisbane, CA, USA
180K-200K Annually
Senior level
Design and operate a scalable data platform for high-volume robotics telemetry and control-plane data. Responsibilities include data capture, storage tiers, ingestion, transformation engines, durable processing, schema and ID-key management, retention and egress policies, replication, CI/CD, monitoring, and alerting. The role supports cloud, on-premises, edge, and outage-resilient deployments while enabling downstream analytics and machine learning.
The summary above was generated by AI

About Mytra:

We’re creating an entirely new way to solve the most ubiquitous problem in industry - moving and storing material. We’re applying robotics and distributed software to create a new class of product for this $1T market. We’re focused on the supply chain industry first. The industry is in a massive bind with the continued growth of e-commerce, sharp rise in costs, and supply chain disruptions. What has been a “sleepy” industry for decades is now at the epicenter of sustaining the global economy.

About the role

Mytra is looking for a Senior Data Software Engineer to build the data platform that the rest of our software reads from.

Every robot and station puts out about 2 MB/s of telemetry and uploads a 60 MB recording every couple of minutes, roughly 6 to 8 GB per edge component per hour. We're designing for 10,000+ storage cells a site, across many sites. Our control plane records every job, fault, and operator action. We require a system that unifies, aggregates, and transforms this data; as well as unlocking the power of our data for downstream insights.

This includes the services that pull data off the wire, the stores it lands in, the ID keying that lets you join across sources, and the pipeline machinery that turns raw capture into the tables products and agents read. Your services will run inside customer facilities, cloud facilities, be durable for multi-day network outages, and enforces each customer's rules for data egress. As much of the services are not built, you will also have the opportunity to architect your solution from the group up.

What you’ll do:
  • Design the data capture lanes, storage tiers, and pipeline layout
  • Re-design the ingestion service with something that scales
  • Build the transform machinery. You own the engine and what it guarantees; the team writes the transforms that run on it.
  • Build the durable path for what the system does and has to remember
  • Build the ID keying and crosswalk tables that connect a stored fact to its site, robot, hardware revision, calibration, layout, and pallet
  • Own retention, cost, and egress as policy instead of plumbing. Retention policies per record and the site-to-cloud replication path.
  • Run what you build. These services run in our cloud and inside customer sites, so you own their CI/CD, monitoring, and alerting.
Ideal Candidates
  • 5+ years building and running production backend or data infrastructure
  • You’ve written production code in Go, Python, Rust, Java or C++
  • You've worked with a durable message or streaming system (NATS/JetStream, Kafka, Pulsar, or similar) and know how delivery guarantees, backpressure, and replay actually behave
  • You've used a columnar store, relational databases, and No-SQL databases in production and you have a strong understanding of how trade offs work in these technologies when used in data solutions
  • You've shipped services and kept them running: CI/CD, monitoring, alerting, and being on the hook when they break
  • You've designed schemas or data contracts that had to survive change
  • BS/MS in Computer Science, or the equivalent from having done the work
  • You've built and run data services yourself; whether ingestion, CDC, or event pipelines, rather than writing jobs on top of someone else's platform
  • You know how columnar stores like ClickHouse and object-storage lakes like Parquet or Iceberg work, and have opinions about which belongs where. Schema governance, failure modes, and cost are design inputs for you, not afterthoughts.
  • Uncertainty isn’t scary, and you’re willing to realize mistakes and pivot
  • You can convey complex ideas and technical concepts to a diverse audience
  • Openness, inquisitiveness, and constant ambition are important values
  • You enjoy collaborative working and taking time to help others
  • You seek and value ownership, accountability, and taking on responsibility
Nice to have:
  • Experience with robotics, IoT, or industrial telemetry at volume, including formats like MCAP
  • GCP (GCS, BigQuery/BigLake), Kubernetes, and data systems run on-prem, at the edge, or air-gapped, where you can't lean on a cloud service being reachable
  • Prep the data store for ML or RL training to read directly
  • Prometheus and Grafana experience as a consumer of the data pipeline
A Final Note:

If this role excites you, we encourage you to apply — even if you don’t check every box. At Mytra, we build bots, but we hire humans. We value thoughtful challengers who communicate with clarity and respect, and we’re lifelong learners committed to getting better every day. We move fast, learn faster, and show up for each other under pressure. The best outcomes come from diverse perspectives working together to build something meaningful and built to last

Benefits Include:
  • Competitive compensation and equity grants at a high-growth company backed by top-tier VCs
  • Fully subsidized health coverage, including medical (baseline plans), dental, and vision for employees and dependents
  • 401(k) plans and employer-subsidized life insurance
  • Fully subsidized lunch and snacks at HQ—we eat and share stories together at the “long” table
  • Generous PTO and company-paid holidays, including one week over the winter break so everyone can recharge together
  • Voluntary pet insurance, Voluntary Life Insurance, Accident, Critical Illness, and Hospital Indemnity
  • Fully subsidized tax advisory services and education to help you understand your equity
  • Commuter benefits, including a up to $150 monthly commuter benefit
  • Lively, modern combined office and lab space where we rapidly iterate through design, build, and test phases
  • Fully equipped onsite gym and showers at headquarters
Pay Notice

Compensation for this position will be set based on the candidate's experience level and location. The posted salary band is applicable only to the San Francisco Bay Area and is subject to change.

HQ

Mytra Brisbane, California, USA Office

427 Valley Dr, Brisbane, California, United States, 94005

Similar Jobs

2 Days Ago
Remote or Hybrid
United States
Senior level
Senior level
AdTech • Consumer Web • Digital Media • eCommerce • Marketing Tech • SEO
Build and support scalable data platforms and pipelines, leading migrations such as Snowflake to BigQuery and transitioning reporting to Looker. Responsibilities include data architecture, ETL/ELT, API and marketing integrations, data quality, production troubleshooting, warehousing, performance optimization, and platform modernization. The role partners with analytics and business teams, owns projects through production, documents solutions, and provides technical guidance.
Top Skills: AWSAzureBigQueryConfluenceDraw.IoGCPGitJIRAKafkaLookerLucidchartMiroModeNotionPower BIPythonSnowflakeSparkSQLTableauTalend
2 Days Ago
Easy Apply
Hybrid
San Francisco, CA, USA
Easy Apply
198K-293K Annually
Senior level
198K-293K Annually
Senior level
Fintech • Payments • Financial Services
Design, develop, deploy, and operate scalable ELT pipelines and data architectures for developer experience products. Integrate and validate diverse data sources, improve pipeline and query performance, establish data governance and quality practices, and build reliable data sources for proactive recommendations. Collaborate with data scientists, analysts, engineers, product managers, and customers while mentoring teams and promoting scalable data engineering standards.
Top Skills: AirflowElt PipelinesGitHadoopKafkaPysparkPytestPythonSparkSQL
4 Days Ago
Remote or Hybrid
16 Locations
110K-200K Annually
Senior level
110K-200K Annually
Senior level
Information Technology • Productivity • Software • Infrastructure as a Service (IaaS)
Design and scale lakehouse architecture, including Bronze, Silver, and Gold data layers. Build reliable streaming and batch pipelines with Kafka, Spark, and Airflow; manage Iceberg, Delta, and Hudi table formats; optimize Trino, Starburst, and Databricks queries; and improve platform reliability, data quality, and performance. Collaborate with data scientists, analysts, and product teams to deliver scalable data platform solutions.
Top Skills: Apache AirflowApache HudiApache IcebergApache KafkaSparkDatabricksDelta LakeLakehouse ArchitectureMedallion ArchitecturePythonSQLStarburstTrino

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account