Top Data Engineer Jobs in San Francisco Bay Area, CA

4 Days AgoSaved
Hybrid
San Francisco Bay Area, CA
157K-245K Annually
Mid level
157K-245K Annually
Mid level
Artificial Intelligence • Information Technology • Machine Learning • Natural Language Processing • Productivity • Software • Generative AI
Build and own scalable finance and revenue data pipelines, models, and datasets supporting ARR, NRR, bookings, billing, and revenue attribution. Ensure data quality, reliability, monitoring, reconciliation, and self-service access while partnering with Finance, Revenue Operations, Analytics, Product, and Engineering.
Top Skills: Apache AirflowSparkCi/CdClaude CodeCodexDatabricksDatabricks WorkflowsDbtDelta LakeGitSnowflakeSQL
5 Days AgoSaved
Hybrid
San Francisco Bay Area, CA
209K-286K Annually
Senior level
209K-286K Annually
Senior level
Fintech • Machine Learning • Payments • Software • Financial Services
Design, build, and lead scalable cloud data platforms, pipelines, and applications. The role architects data engineering patterns, evaluates platforms such as Snowflake and Databricks, delivers distributed and streaming workloads, and ensures operational resilience. Responsibilities include collaborating with product and engineering teams, influencing technical decisions, mentoring engineers, communicating data concepts, and leading large-scale initiatives independently.
Top Skills: AgileAirflowAWSCassandraDagsterDatabricksDynamoDBEmrGlueGCPJavaAzureMongoDBMonte CarloNosql DatabasesPythonRedshiftRelational DatabasesScalaSnowflakeSparkSplunkSQL
4 Days AgoSaved
Remote or Hybrid
San Francisco Bay Area, CA
140K-215K Annually
Senior level
140K-215K Annually
Senior level
Cloud • Computer Vision • Information Technology • Sales • Security • Cybersecurity
Architects, deploys, and operates scalable data platform infrastructure supporting analytics. Responsibilities include administering Airflow and Superset, managing Terraform infrastructure, building CI/CD pipelines, enforcing security and compliance, developing DBT data models, cataloging assets in OpenMetadata, writing Python automation, and mentoring engineers. The role requires extensive infrastructure and data engineering experience across cloud platforms, containers, data security, and workflow orchestration.
Top Skills: Apache AirflowApache SupersetAWSAzureCi/CdData ModelingDbtDockerGCPIamKubernetesMachine Learning PipelinesOciOpenmetadataPythonSQLTerraform
Reposted 14 Days AgoSaved
Hybrid
San Francisco Bay Area, CA
124K-280K Annually
Senior level
124K-280K Annually
Senior level
Artificial Intelligence • Professional Services • Business Intelligence • Consulting • Cybersecurity • Generative AI
Lead design and implementation of enterprise data architecture and cloud-based data solutions. Manage large data engineering projects, develop data models and pipelines, ensure governance and security compliance, advise clients strategically, and build high-performing inclusive teams.
Top Skills: AWSAzureDatabricksGCPSnowflake
Reposted 17 Days AgoSaved
Hybrid
San Francisco Bay Area, CA
157K-245K Annually
Mid level
157K-245K Annually
Mid level
Artificial Intelligence • Information Technology • Machine Learning • Natural Language Processing • Productivity • Software • Generative AI
Design, build, and operate large-scale ETL pipelines, data lakes, and data platforms processing billions of daily events. Own data quality, freshness, reliability, monitoring, and observability for foundational datasets. Develop self-service ETL frameworks and tooling, contribute to architecture and technical strategy, and partner with product, engineering, machine learning, data science, and leadership teams to deliver scalable data solutions.
Top Skills: Amazon RedshiftApache AirflowApache FlinkApache HudiApache IcebergApache KafkaSparkBigQueryCi/CdDagsterDatabricksDbtDelta LakeGitPythonSnowflakeSpark Structured StreamingSQLTerraform
20 Days AgoSaved
Easy Apply
Hybrid
San Francisco Bay Area, CA
Easy Apply
50-50 Annually
Internship
50-50 Annually
Internship
Artificial Intelligence • Blockchain • Fintech • Financial Services • Cryptocurrency • NFT • Web3
Build and maintain data pipelines and scalable cloud infrastructure supporting analytics and machine learning. Integrate internal and external data sources, ensure data integrity, contribute to analytical investigations, document findings, and apply data governance, security, and performance optimization practices while collaborating with senior engineers and cross-functional teams.
Top Skills: APIsAWSAzureCloud-Native DesignDistributed SystemsGCPGenerative AiLarge Language Models (Llms)MicroservicesPythonRelational DatabasesSQL
19 Days AgoSaved
Remote or Hybrid
San Francisco Bay Area, CA
125K-159K Annually
Mid level
125K-159K Annually
Mid level
AdTech • Automotive • Big Data • Consumer Web
Administer and enhance Edmunds’ Databricks data platform and AWS infrastructure. Build and maintain ETL pipelines, infrastructure-as-code tooling using Terraform or CDK, and operational dashboards, alerts, and reports. Collaborate with business, engineering, analytics, security, and infrastructure teams to support data platform users and AI solutions. Evaluate new technologies, troubleshoot platform issues, and improve operational, cost, and security visibility.
Top Skills: SparkAWSAws CdkDatabricksInfrastructure As Code (Iac)PythonScalaSQLTerraform
6 Days AgoSaved
Hybrid
San Francisco Bay Area, CA
165K-270K Annually
Senior level
165K-270K Annually
Senior level
Security • Software • Cybersecurity • Automation
Build and scale data products supporting analytics, engineering services, and AI/ML models. Responsibilities include optimizing queries and data pipelines, developing data models and governance processes, administering ETL/ELT and reverse ETL orchestration, and creating documentation and training. The role also involves mentoring, interviewing, and hiring data engineers while helping define the organization’s data engineering strategy.
Top Skills: AirbyteAirflowCensusChange Data Capture (Cdc)DbtEltETLFivetranHightouchLookerMl/Ai PipelinesModePythonReverse EtlSigmaSnowflakeSQLStitch
7 Days AgoSaved
Hybrid
San Francisco Bay Area, CA
281K-334K Annually
Expert/Leader
281K-334K Annually
Expert/Leader
Artificial Intelligence • Productivity • Software
Lead Notion’s Data Foundations team by setting multi-quarter platform strategy, developing engineers and technical leaders, and guiding complex initiatives across data lakes, streaming, distributed compute, governance, and reliability. Drive enterprise-ready capabilities including data residency, access controls, lifecycle management, and key management. Improve Kafka, Debezium, S3/Iceberg, Spark, EMR, and Athena scalability, reliability, usability, and cost efficiency while partnering across engineering, AI, search, infrastructure, and security teams.
Top Skills: Amazon AthenaAmazon EmrAmazon S3Apache IcebergSparkDebeziumKafka
One Month AgoSaved
Hybrid
San Francisco Bay Area, CA
153K-207K Annually
Mid level
153K-207K Annually
Mid level
Cloud • Healthtech • Social Impact • Software • Biotech
Build and operate production-grade ELT pipelines from Benchling, Salesforce, and third-party systems into Snowflake using dbt. Own ingestion, transformation, warehouse, BI, governance, monitoring, testing, access controls, privacy, cost, and performance. Partner with AI engineering to provide trusted data for internal AI applications and agentic workflows, contribute to warehouse and metrics-layer architecture, and translate cross-functional stakeholder needs into scalable data solutions.
Top Skills: AirflowAWSDbtPythonSalesforceSnowflakeSQL
4 Days AgoSaved
Hybrid
San Francisco Bay Area, CA
300K-350K Annually
Mid level
300K-350K Annually
Mid level
Artificial Intelligence • Machine Learning • Software • Generative AI
Build and own marketing data, attribution, tracking, audience sync, and conversion infrastructure. Integrate product, CRM, warehouse, and advertising platforms; send offline conversions and LTV signals to ad platforms; maintain event schemas and dashboards; debug data discrepancies; implement privacy-safe tracking; and partner cross-functionally with growth, product, marketing, and data teams.
Top Skills: AdjustAmplitudeAppsflyerAttBigQueryBranchCmpConsent ModeEltETLGCPGoogle AdsGtmLinkedin AdsMeta AdsMetabasePrivacy SandboxPythonSkanSQLStripeTiktok Ads
4 Days AgoSaved
Hybrid
San Francisco Bay Area, CA
180K-230K Annually
Senior level
180K-230K Annually
Senior level
Fintech • Software
Own and scale Collective’s data platform by building reliable batch and event-driven pipelines into BigQuery, modeling data with dbt, and enforcing quality, observability, security, and cost standards. Partner with engineers, analysts, and business stakeholders to deliver trustworthy datasets, metrics, reporting, and AI-enabled warehouse access. The role includes production ownership, incident response, governance, infrastructure practices, and support for analytics and LLM-based tools.
Top Skills: AirflowAmplitudeBigQueryCi/CdClaude CodeCloud ComposerDagsterDatadogDbtFivetranGitGoogle Cloud PlatformInfrastructure As CodeKafkaLlm ToolsMetabaseMetric StoresPub/SubPythonSemantic LayersSQLTerraform
New

Cut your apply time in half.

Use ourAI Assistantto automatically fill your job applications.

Use For Free
Application Tracker Preview
4 Days AgoSaved
In-Office
San Francisco Bay Area, CA
72K-92K Hourly
Mid level
72K-92K Hourly
Mid level
Information Technology
Design, optimize, and maintain scalable data architectures, pipelines, and processing solutions. Ensure data quality, perform SQL and Python analysis, build dashboards and visualizations, and support marketing and product analytics. Lead tooling development, requirements gathering, QA, stakeholder communications, and cross-functional initiatives while providing technical guidance and managing team decisions.
Top Skills: Data PipelinesData ProcessingETLPythonSQL
10 Days AgoSaved
Easy Apply
Hybrid
San Francisco Bay Area, CA
Easy Apply
186K-232K Annually
Senior level
186K-232K Annually
Senior level
Artificial Intelligence • Big Data • Healthtech • Biotech • Pharmaceutical
Build and operate reliable data systems and products supporting clinical operations, drug evaluation, business development, analytics, machine learning, and AI agents. Responsibilities include designing pipelines, canonical data models, data contracts, warehouse transformations, quality and observability practices, governance for regulated data, and incident response. The role partners closely with Product Engineering and Data Science, uses AI-assisted development, contributes to architecture, and mentors other engineers.
Top Skills: DagsterDbtDockerGitLlmsOpentofuPythonSnowflakeSQLTerraform
Junior
Energy • Utilities
Build and maintain production data pipelines, develop dbt models in Snowflake, create financial and budget data products, support platform reporting, migrate legacy publishing workflows, and resolve data quality issues. The role also supports department onboarding, documentation, testing, code review, and collaboration with analysts and senior engineers.
Top Skills: DbtPythonSnowflakeSQLTerraform
5 Days AgoSaved
Hybrid
San Francisco Bay Area, CA
46-46 Annually
Internship
46-46 Annually
Internship
Productivity • Professional Services • Software • Design
Build, test, document, and optimize data pipelines and foundational datasets for Figma’s data warehouse. Interns will support data quality, scalability, analytics, experimentation, and cross-functional data needs while owning a scoped project from definition through implementation, testing, documentation, and results sharing.
Top Skills: Data PipelinesData WarehousesETLPythonSQL
24 Days AgoSaved
Remote or Hybrid
San Francisco Bay Area, CA
160K-260K Annually
Expert/Leader
160K-260K Annually
Expert/Leader
Artificial Intelligence • Cloud • Payments • Software • Business Intelligence • Generative AI • Automation
Define and govern enterprise-scale data architecture across batch, streaming, warehouse, lakehouse, transactional, and AI use cases. Establish standards for data quality, lineage, access, cataloging, governance, observability, and SLAs. Architect AI-enabled workflows, resolve complex architecture issues, influence roadmaps, and mentor engineers through hands-on technical leadership. The role requires 15+ years of software, data engineering, or architecture experience and expertise in large-scale data platforms and modeling.
Top Skills: AIBatch ProcessingBigQueryData CatalogsData WarehousesDbtFeature StoresGCPLakehousesOlapOltpStreaming ArchitecturesVector Stores
Reposted 8 Days AgoSaved
In-Office
San Francisco Bay Area, CA
Junior
Junior
Real Estate • PropTech
Build and operate HomeLight’s data engineering systems, including scalable pipelines, ETL workflows, data warehouses, databases, schemas, and APIs. Analyze and visualize data, develop performance metrics, optimize data release processes, and provide reliable data for algorithms and internal teams. The role also involves applying LLMs and prompt engineering, designing service-oriented architectures, and building robust, highly available systems.
Top Skills: AirflowAPIsAWSChatgptClaude CodeCodexCursorDjangoElasticsearchETLLlmsPrompt EngineeringPythonRuby On Rails
5 Minutes AgoSaved
In-Office or Remote
San Francisco Bay Area, CA
Senior level
Senior level
Digital Media • Fintech • Information Technology • Machine Learning • Financial Services • Cybersecurity • Automation
Leads the design and implementation of scalable data architectures, ETL pipelines, data warehouses, migrations, and cloud data solutions. Provides technical leadership through system analysis, requirements gathering, design documentation, code and architecture reviews, SDLC management, production deployments, and mentoring junior engineers. Develops integrations using multiple programming languages, ETL platforms, cloud services, APIs, and CI/CD tools.
Top Skills: Amazon AthenaAmazon Ec2Amazon EmrAmazon RdsAmazon RedshiftAmazon S3SparkAWSAws GlueAws IamAws Lake FormationAws LambdaBitbucketCC++GitGoogle Cloud Platform (Gcp)Ibm DatastageJavaJavaScriptJenkinsAzureOracle Data Integrator (Odi)Pl/SqlPythonRest ApisTalendTeradataUnix
Reposted 4 Hours AgoSaved
Remote
San Francisco Bay Area, CA
800K-1M Annually
Mid level
800K-1M Annually
Mid level
Information Technology • Professional Services • Consulting
Build and maintain large-scale data processing pipelines using Apache Spark (PySpark/Scala) and Python. Design ETL workflows, implement distributed data processing, use version control (Git), and work with cloud platforms (GCP/AWS/Azure).
Top Skills: SparkAWSAzureETLGCPGitPysparkPythonScala
10 Days AgoSaved
In-Office
San Francisco Bay Area, CA
150K-220K Annually
Entry level
150K-220K Annually
Entry level
Artificial Intelligence • Software • Analytics • Financial Services
Build and operate data pipelines that integrate customer financial data from ERPs, warehouses, and payment systems. Develop AI agents that automate ERP implementations, partner with implementation teams to understand and transform customer data, and contribute to shared AI platform capabilities. The role requires strong full-stack engineering fundamentals, autonomous execution, and comfort working in an ambiguous, fast-moving environment.
Top Skills: Next.JsPythonReact
10 Days AgoSaved
In-Office
San Francisco Bay Area, CA
180K-220K Annually
Senior level
180K-220K Annually
Senior level
Software
Build and scale batch and streaming data pipelines using Databricks, Spark, Kafka, AWS, Python, SQL, dbt, and Fivetran. Develop ETL/ELT and reverse ETL workflows, curated datasets, dashboards, data quality controls, and data activation integrations with Salesforce, HubSpot, and Braze. Partner with Finance, Marketing, Product, and Engineering on reporting, forecasting, segmentation, and AI/ML applications, including RAG, embeddings, and agent-based systems.
Top Skills: Amazon AthenaAmazon EmrAmazon S3Apache AirflowApache KafkaSparkAWSAws GlueBrazeDatabricksDbtEmbeddingsFivetranHubspotLlmsMcpPysparkPythonRagSalesforceSQLVector Databases
Reposted 11 Days AgoSaved
In-Office
San Francisco Bay Area, CA
145K-215K Annually
Senior level
145K-215K Annually
Senior level
Artificial Intelligence • Healthtech • Telehealth
Build and maintain data pipelines to ingest, normalize, and deliver messy healthcare data (FHIR, HL7, CCDAs, PDFs) into structured models for AI and clinical use. Implement ETL, real-time webhooks, document extraction, PII sanitization, storage on AWS, and monitor data quality while ensuring compliance with healthcare interoperability standards.
Top Skills: AWSCcdasClaudeCursorDynamoDBEhrFhirHieHl7JSONPdfPostgresPythonS3WebhooksXML
11 Days AgoSaved
In-Office
San Francisco Bay Area, CA
144K-153K Annually
Mid level
144K-153K Annually
Mid level
News + Entertainment • Sports
Build and optimize scalable ELT/ETL pipelines, data models, schemas, and monitoring systems in BigQuery. Ensure data quality, governance, compliance, and pipeline reliability while collaborating with data scientists, analysts, engineers, and business partners. Develop generative AI data products and document engineering standards. The role requires cloud data experience, SQL, programming, orchestration tools, and strong communication skills.
Top Skills: Apache AirflowBigQueryCi/CdDbtEltETLGenerative AiGoogle Cloud PlatformPythonSQL
Reposted 2 Days AgoSaved
Remote
San Francisco Bay Area, CA
Mid level
Mid level
Fintech • Software • Analytics • Financial Services
Design, build, and maintain scalable data pipelines, data lakes, and databases; ingest and map customer financial datasets; monitor pipeline reliability; translate business requirements into data flows and analytical insights to ensure high-quality, usable data.
Top Skills: APIsBigQueryBigtableData LakesDatabricksDbtFivetranGithub ActionsMariadbMongoDBMySQLNoSQLPostgresRedshiftSnowflakeSQL
All Filters
JobType
New Jobs
Job Category
Experience
Industry
Company Name
Company Size

Sign up now Access later

Create Free Account