Mill Logo

Mill

Senior Data Engineer — Recommendations & Data Platform

Posted 18 Days Ago
Hybrid
San Bruno, CA, USA
185K-210K Annually
Senior level
Hybrid
San Bruno, CA, USA
185K-210K Annually
Senior level
Design, build, and maintain scalable data pipelines and a customer-facing recommendation engine (including LLM logic). Manage data warehouse, data quality monitoring, CI/CD for pipelines, and partner with product, analytics, engineering, and marketing to ensure consistent metrics and observability.
The summary above was generated by AI

Mill is a waste prevention technology company reimagining what it means to eliminate waste, starting with food. We build smart systems and infrastructure for homes, businesses, and municipalities that transform food scraps from landfill-bound waste into valuable resources, including chicken feed. Tens of thousands of Mill’s residential food recyclers are already helping households divert millions of pounds of food scraps every year, paving the way for our upcoming launch of Mill Commercial—the industry’s first end-to-end solution for managing, understanding, and preventing food waste in commercial environments (e.g. grocery, restaurants, food services). At Mill, we are passionate about building easy-to-use, beautifully designed technologies that keep food in the food system and out of landfills.

As a Data Engineer at Mill, you'll touch systems end-to-end — from raw ingestion to the recommendation a customer sees in the app to managing the data warehouse. You'll architect a warehouse model one week and tune recommendation logic the next. You'll partner closely with product, engineering, data analytics, and marketing teams.

What You'll Do
  • Build and operate the customer-facing recommendation engine that turns food waste data into actionable recommendations — purchasing suggestions, anomaly explanations, operational nudges — including LLM-based logic where useful
  • Aggregate and unify disparate data streams to guarantee the precision, reliability, and protection of our analytical assets.
  • Co-create and drive Mill's self-serve analytics strategy with analysts and data scientists — defining golden datasets, setting SLAs for freshness and reliability, and establishing the governance model that keeps self-serve trustworthy at scale
  • Own data quality monitoring against those SLAs — build alerting, validation frameworks, and observability tooling so golden dataset issues get caught before they become business problems
  • Collaborate with software engineers to instrument new product features and ensure event data flows cleanly into the warehouse, meeting the governance and quality bar required for it to become part of a golden dataset
  • Design, build, and maintain scalable data pipelines across Mill's product and operational systems
  • Define and maintain the metrics, table endorsements, and business logic that analysts and stakeholders rely on — so everyone across the company is working from the same numbers
What We're Looking For
  • Have designed, built, or operated a recommendation system in production — one that combines multiple data sources into a single customer-facing output — not just contributed data to someone else's model
  • Experience building recommendation or personalization logic using LLMs (prompt-based scoring, retrieval-augmented generation, agent-based reasoning) in a live product, not just a prototype
  • Have built and operated data pipelines in production using Python, SQL, and tools like dbt, Airflow, Fivetran, or similar against a cloud data warehouse (e.g., Snowflake, BigQuery, Redshift) — including handling failures, backfills, and schema changes after launch
  • Have brought CI/CD discipline to data pipelines or product logic (automated testing, staged rollout, rollback), with a track record of measuring whether a change actually improved outcomes, not just shipping it
  • 5 years of experience operating data engineering systems in production
  • A bias toward clarity and action
Nice to Have
  • Exposure to distributed systems concepts (partitioning, consistency, fault tolerance)
  • Hands-on experience with infrastructure as code (Terraform, Pulumi) in a cloud environment
  • Experience with Hex, Mixpanel, Tableau, or similar BI/analytics tools
  • Familiarity with data contract or data mesh patterns
  • Experience with event tracking or product analytics

The estimated base salary range for this position is $185k to $210k, which does not include the value of benefits or a potential equity grant. A wide range of factors are considered in making compensation decisions, including but not limited to skill sets, market conditions, experience and training, licensure and certifications, and business and organizational needs.

HQ

Mill San Bruno, California, USA Office

1099 Sneath Lane, San Bruno, CA, United States, 94066

Similar Jobs

3 Days Ago
Remote or Hybrid
United States
190K-240K Annually
Senior level
190K-240K Annually
Senior level
Big Data • Cloud • Productivity • Software • Database • Analytics • Automation
As a Senior Data Engineer at Jellyfish, you'll build and maintain data pipelines, optimize orchestration, automate CI/CD processes, and enhance data integration while ensuring high performance and reliability.
Top Skills: AirflowBigQueryDagsterDatabricksDbtPrefectPysparkPythonRedisSnowflakeSQLTerraform
Yesterday
Hybrid
San Francisco, CA, USA
170K-220K Annually
Senior level
170K-220K Annually
Senior level
Artificial Intelligence • Big Data • Machine Learning • Software
The Sr. Data Engineer will build and manage data pipelines for large datasets, collaborate with teams, and leverage cloud technologies for data processing.
Top Skills: AirflowApache BeamSparkBigQueryCassandraDataflowGCPJavaKotlinPostgresPythonScylla
5 Days Ago
In-Office or Remote
United States
Junior
Junior
Artificial Intelligence • Cloud • Information Technology • Machine Learning • Software • Sports • Analytics
The Data Engineer I will build and maintain data pipelines, contribute to platform code, and integrate data sources while growing alongside senior engineers in a production-focused role.
Top Skills: AirflowAWSBigQueryDatabricksDbtDelta LakeIcebergPythonRedshiftSnowflakeSparkSQLTerraformTrino

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account