AHEAD Logo

AHEAD

Data Engineer

Posted 11 Days Ago
Remote
Hiring Remotely in United States
150K-180K Annually
Mid level
Remote
Hiring Remotely in United States
150K-180K Annually
Mid level
Build and operate cloud data capabilities, including ingestion pipelines, transformations, data models, curated data products, and data-quality controls. Use Snowflake, dbt, SQL, and Python to deliver governed data for analytics, applications, automation, and AI workflows. Implement testing, CI/CD, documentation, lineage, access controls, monitoring, incident resolution, and performance optimization. Collaborate with engineering, analytics, governance, security, and business teams in an Agile, product-oriented environment with active AI-assisted development.
The summary above was generated by AI

The Data Engineer, Data Platform will build and operate the data capabilities that help AHEAD teams access trusted, usable, and well-managed information. This role will develop ingestion pipelines, transformations, data models, and curated data products in the modern cloud data platform, with an emphasis on Snowflake and dbt. 

The role will support data coming from enterprise applications and services, including Salesforce, Hatch, NetSuite, Signal, and approved APIs. The Data Engineer will help make data available for analytics, applications, automation, and AI-enabled workflows through consistent engineering patterns,documented definitions, appropriate access controls, and dependable operational practices. Active use of AI throughout the software development lifecycle is a core expectation of this role, including AI-assisted code generation, automated testing, documentation, troubleshooting, and review with appropriate human validation. 

Working under the Director, Data Platform and alongside the Data Governance Lead, this role will contribute to a product-oriented engineering team. The role will partner with data consumers and other engineering teams to understand requirements, deliver useful platform capabilities, and improve the speed and consistency of data delivery. 

Duties/Responsibilities

  • Build, maintain, and improve batch and low-latency data ingestion pipelines from enterprise systems, APIs, and other approved sources. 
  • Follow the AI SDLC by actively using approved AI coding tools and agents to generate, refactor, explain, and review code; validate generated output through engineering judgment, testing, and peer review. 
  • Use AI to generate and improve unit, integration, data-quality, and regression tests, then verify that automated tests accurately validate the intended behavior. 
  • Use AI-assisted workflows to create and maintain technical documentation, data-product documentation, runbooks, lineage notes, and change summaries as part of delivery. 
  • Build toward coordinated multi-agent delivery patterns that can divide and accelerate discovery, implementation, testing, documentation, and operational support while preserving human accountability. 
  • Develop SQL and Python solutions that collect, validate, transform, and publish data for downstream consumption. 
  • Use Snowflake and dbt to implement reliable transformations, reusable models, curated datasets, and data products across raw, common, and curated layers. 
  • Translate business and technical requirements into source mappings, data models, acceptance criteria, and maintainable engineering solutions. 
  • Partner with analytics, application, AI, Integration Platform, and business teams to make data available through governed and documented access patterns. 
  • Apply data quality checks for completeness, freshness, uniqueness, consistency, referential integrity, and other relevant quality dimensions. 
  • Add metadata, documentation, lineage, ownership, and usage guidance to data products so consumers can find and understand the data they use. 
  • Implement secure access patterns in partnership with Data Governance and Security teams, including role-based access, classification tags, masking, and row- or column-level controls when appropriate. 
  • Build automated tests and deployment processes that support consistent delivery through development, quality assurance, and production environments. 
  • Monitor pipeline health, data freshness, processing performance, and failures; troubleshoot issues and participate in incident resolution. 
  • Optimize Snowflake workloads, queries, transformations, and storage patterns for performance, reliability, and cost discipline. 
  • Support the curation and publication of cross-system data needed for shared business context, entity-aware access, reporting, automation, and AI use cases. 
  • Work with the Integration Platform and semantic-layer capabilities, including Horizon, to support consistent business meaning and reusable data access. 
  • Participate in backlog refinement, estimation, code review, technical documentation, and iterative delivery within an Agile engineering team. 
  • Identify opportunities to simplify delivery, reduce duplicate work, improve platform standards, and strengthen the reliability of data engineering practices. 

Education and Experience

    • Bachelor’s degree in computer science, information systems, engineering, mathematics, or a related field, or equivalent experience. 

    • 3 or more years of experience in data engineering, software engineering, analytics engineering, or a related technical role. 

    • Professional experience writing production-quality SQL and Python. 

    • Experience building or supporting data pipelines, transformations, and data models in a cloud data environment. 

    • Experience with Snowflake, dbt, or comparable cloud data warehouse and transformation technologies. 

    • Understanding of data modeling, ELT/ETL patterns, pipeline orchestration, APIs, and source-system integration. 

    • Experience with software engineering practices including source control, code review, automated testing, and CI/CD. 

    • Demonstrated active use of AI-assisted software development tools for code generation, test creation, documentation, debugging, or review. 

    • Ability to follow an AI SDLC and identify practical opportunities for multiple cooperating agents to improve delivery speed, consistency, and coverage. 

    • Understanding of data quality, metadata, lineage, access control, privacy, and secure handling of enterprise data. 

    • Ability to investigate data issues, communicate findings clearly, and work through ambiguity with teammates and stakeholders. 

    • Ability to collaborate effectively with engineers, analysts, product owners, governance partners, security teams, and business stakeholders. 

Preferred

    • Experience with Azure services, serverless functions, cloud storage, or other cloud-native data engineering capabilities. 

    • Experience with REST or GraphQL APIs and data ingestion from enterprise applications such as Salesforce, Hatch, NetSuite, or similar systems. 

    • Familiarity with orchestration, event-driven processing, observability, data catalogs, lineage tooling, or data quality platforms. 

    • Experience supporting semantic models, MCP-based access, or other governed interfaces for analytics, applications, automation, or AI workflows. 

    • Experience working with master data, reference data, entity resolution, or shared business definitions across multiple systems. 

    • Experience operating data products with documented ownership, access expectations, quality measures, and support procedures. 

    • Experience using AI agents or agentic workflows to support software delivery, data engineering, testing, documentation, or platform operations. 

    • Curiosity about emerging data platform technologies and a practical approach to adopting them. 

     

Physical Requirements

     
  • Ability to safely and successfully perform the essential job functions consistent with the ADA, FMLA, and other federal, state, and local standards, including meeting qualitative and/or quantitative productivity standards. 
    • Ability to maintain regular, punctual attendance consistent with the ADA, FMLA, and other federal, state, and local standards. 

    • Primarily office and computer-based work with standard engineering and collaboration expectations for an enterprise technology role. 

AHEAD San Francisco, California, USA Office

2000 Crow Canyon Place Suite 250, San Francisco, United States, 94583

Similar Jobs

Yesterday
Remote or Hybrid
US
105K-150K Annually
Senior level
105K-150K Annually
Senior level
Information Technology
Designs, develops, and supports analytics data solutions for clients, including Snowflake and dbt ingestion and transformation pipelines, data marts, validation, monitoring, governance, and documentation. Manages multiple consulting projects, budgets, timelines, quality assurance, troubleshooting, client communication, and implementation planning. Requires strong data engineering experience, regulated-industry knowledge, stakeholder communication, and willingness to travel.
Top Skills: AWSDbtSnowflakeSQL
Yesterday
In-Office or Remote
124K-207K Annually
Senior level
124K-207K Annually
Senior level
Artificial Intelligence • Healthtech • Machine Learning • Natural Language Processing • Biotech • Pharmaceutical
Build and operate production data pipelines supporting analytics, AI, and agentic workflows. Responsibilities include implementing canonical data models, maintaining Databricks or Snowflake platforms, monitoring reliability, responding to incidents, validating healthcare data mappings, improving performance and cost, and documenting architecture. Requires strong SQL and Python skills, cloud data platform experience, ETL/ELT orchestration expertise, and healthcare or pharmaceutical data experience.
Top Skills: Ai/Ml WorkflowsDatabricksEltETLHedisOmopPythonSnowflakeSQL
2 Days Ago
Remote or Hybrid
USA
125K-159K Annually
Mid level
125K-159K Annually
Mid level
AdTech • Automotive • Big Data • Consumer Web
Administer and enhance Edmunds’ Databricks data platform and AWS infrastructure. Build and maintain ETL pipelines, infrastructure-as-code tooling using Terraform or CDK, and operational dashboards, alerts, and reports. Collaborate with business, engineering, analytics, security, and infrastructure teams to support data platform users and AI solutions. Evaluate new technologies, troubleshoot platform issues, and improve operational, cost, and security visibility.
Top Skills: SparkAWSAws CdkDatabricksInfrastructure As Code (Iac)PythonScalaSQLTerraform

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account