Worldly Logo

Worldly

Senior Data Engineer

Posted 15 Days Ago
Remote
Hiring Remotely in United States
135K-165K Annually
Senior level
Remote
Hiring Remotely in United States
135K-165K Annually
Senior level
Own and evolve Worldly’s production data lake, Postgres warehouse, CDC and streaming pipelines, dbt transformations, orchestration, BI access controls, and graph integrations. Maintain reliable, governed analytics infrastructure using Iceberg, Trino, Dagster, Terraform, and AWS. Support production generative AI workflows involving embeddings, similarity search, extraction, and classification. Responsibilities include data quality, monitoring, security, incident response, disaster recovery, schema governance, and collaboration with data science and analytics stakeholders.
The summary above was generated by AI
Senior Data Engineer

Location: Remote - US

About Worldly
Worldly is the world's most comprehensive impact intelligence platform — delivering real data to businesses on impacts within their supply chain. Worldly is trusted by 40,000 global brands, retailers, and manufacturers to provide the single source of ESG intelligence they need to accelerate business and industry transformation.

Through strategic and meaningful customer relationships, Worldly provides key insights into supplier performance, product impact, trends analysis, and compliance. When a company wants to change how business is done, we enable that systemic shift.

Backed by a dedicated global team of individuals aligned by values, Worldly proudly operates as a public benefit corporation with backing from mission-aligned investors. Want to learn more? Read our story.

About the Opportunity

Worldly is hiring a hands-on data engineer with a passion for sustainability to join our dynamic team. You will take on a primary role in building, operating, maintaining, and evolving the systems that support our internal analytics and power our customer-facing analytics platforms.

In this role, you will:

  • Own our production data lake — a CDC-fed, medallion (bronze/silver/gold) lake on Apache Iceberg, queried through Trino — as well as our independent Postgres data warehouse.

  • Play a core role in integrating our graph database with the warehouse, lake, and primary databases, including migrating legacy pipelines onto a single, governed integration layer.

  • Apply generative AI and embeddings-based techniques already running in production (e.g., entity matching, document extraction) to keep expanding what our data can do.

What You'll Do

Data Warehouse & Data Lake

  • Operate and evolve our Postgres data warehouse (schema, performance, access controls) and build analytics-ready datasets from it.

  • Own the lake end-to-end: CDC ingestion (source database → message bus → streaming writer → Iceberg bronze/silver/gold), the Trino query layer, and the table catalog.

  • Own pipeline health — latency SLAs, schema-drift detection, source reconciliation — and the GitOps/Terraform infrastructure and backup/DR posture underneath it.

  • Bring consistent schemas, documented lineage, and clear ownership to a data estate that's grown quickly through acquisitions.

Orchestration, Transformation & Reporting

  • Maintain and evolve our Dagster-orchestrated dbt pipelines: sensor-triggered and scheduled builds, data quality tests, branch-based versioning for curated releases.

  • Operate our BI/reporting layer, including per-user, policy-based data access enforced at the query engine (not just the dashboard), and consistent metric definitions across dashboards.

Graph Integration

  • Own pipelines integrating the graph database with the warehouse, lake, and primary databases, working within canonical data models and a single write path.

  • Partner with data science to migrate legacy direct-to-graph services onto that shared path, and to evolve relational structures into graph-native models.

GenAI/NLP Enablement

  • Support and extend production genAI workflows — embeddings/similarity search, LLM-based extraction and classification — and keep our data infrastructure "AI-ready."

We'd Like to See

  • 5+ years in data engineering, analytics engineering, or data platform engineering.

  • Advanced SQL and relational database experience (Postgres, MongoDB).

  • Hands-on graph database experience in production, including integrating graph models with warehouses and lakes — core, not peripheral, to this role.

  • Experience with open table formats and medallion lake architectures (e.g., Apache Iceberg) and distributed SQL engines (e.g., Trino, Presto).

  • Experience with streaming/CDC pipelines (e.g., Kafka or Pulsar, Debezium, Flink or similar).

  • Strong Python skills for pipelines, automation, and operational tooling.

  • Experience with dbt orchestrated by a modern scheduler (e.g., Dagster, Airflow).

  • AWS infrastructure experience (EKS, VPC, IAM, S3) via infrastructure-as-code (e.g., Terraform), plus CI/CD, GitOps (e.g., ArgoCD), and Docker.

  • Experience with analytics data modeling, metric definitions, and automated monitoring/data-quality controls.

  • Experience operating production data systems: incident triage, root-cause analysis, runbooks, reliability improvements.

  • Comfortable working with cross-functional/analytics stakeholders (Jira/Confluence, Agile).

  • Familiarity with data security practices (PII protection, encryption, access management).

It Helps If You Have

  • Experience with BI tooling supporting per-user, policy-based data access (e.g., Superset with SSO impersonation).

  • Experience with policy-based access control (e.g., OPA) and identity platforms (e.g., Keycloak).

  • Familiarity with multi-region data residency (e.g., separate EU/China data handling).

What We Can Offer You

  • Medical, Dental, and Vision Insurance are offered through multiple PPO options. Worldly covers 90% of employee premiums and 60% of spouse/dependent premiums.

  • Company-sponsored 401k with up to 4% match for US employees.

  • Incentive Stock Options.

  • 100% Parental Paid Leave.

  • Unlimited PTO.

  • 12 paid company holidays.

Life at Worldly
Our team is motivated to transform the way products are made. By helping our customers succeed in a new era of sustainable production, we can build technology that makes a difference on a planetary level.

Our team represents over 15 countries and brings unique experiences from technology to farming to the table. Surround yourself with kind, enthusiastic, and dedicated people who put collaboration and growth at the center of our shared goals.

Benefits and Perks

  • Earn a competitive salary and performance-based bonuses. Get healthcare, retirement matching, and equity for US employees.

  • Use the office stipend to get the supplies you need—combat Zoom fatigue with no-meeting Fridays.

  • Flexible time off. Take the time you need to recharge. Our culture encourages team members to explore and rest to be their best selves.

  • We're remote, not lonely. Join the culture committee, coffee chats, or a variety of other interest groups.

Travel Notice
Roles at Worldly may require occasional travel to support business needs, including but not limited to team collaboration, customer engagement, or company events.

Equity Statement
We believe reflecting the diversity of those we strive to serve is essential. True innovation happens when everyone has room at the table, including the tools, resources, and opportunity to excel. We're dedicated to building a culturally and experientially diverse team that leads and works with empathy and respect.

Compensation Overview
Annual Base Rate (USD): $135,000 - $165,000
10% Annual Bonus
Incentive Stock Option Package
Work-From-Home Stipends

*Final compensation figures will be determined based on a wide variety of factors, including experience and location. These factors will be evaluated and considered by Worldly throughout the entirety of this process.

HQ

Worldly Kensington, California, USA Office

Kensington, California, United States

Similar Jobs

17 Hours Ago
In-Office or Remote
7 Locations
168K-297K Annually
Senior level
168K-297K Annually
Senior level
Blockchain • eCommerce • Fintech • Payments • Software • Financial Services • Cryptocurrency
Lead data modeling and ownership of critical pipelines supporting Block’s product data foundation. Build reliable, governed, and monitored datasets, data quality and lineage systems, experimentation infrastructure, and AI-assisted automation. Partner with product and engineering teams to translate business needs into end-to-end data solutions, participate in on-call support, and maintain pipeline SLAs.
Top Skills: AirflowDatabricksDbtGitOmniPrefectPythonSnowflakeSQLTerraform
Yesterday
Remote or Hybrid
7 Locations
168K-297K Annually
Senior level
168K-297K Annually
Senior level
Blockchain • Fintech • Mobile • Payments • Software • Financial Services
Build and optimize data models, pipelines, monitoring, lineage, and data quality systems supporting Block’s product data foundation. Own critical data engineering solutions across their full lifecycle, ensure pipeline reliability through on-call support, and translate business and product needs into automated data solutions. Develop AI-assisted workflows and agents for data quality, issue resolution, and operational efficiency while partnering with engineering, product, data science, and machine learning teams.
Top Skills: AirflowDatabricksDbtGitOmniPrefectPythonSnowflakeSQLTerraform
2 Days Ago
Remote or Hybrid
United States
109K-183K Annually
Senior level
109K-183K Annually
Senior level
Artificial Intelligence • Cloud • Sales • Security • Software • Cybersecurity • Data Privacy
Design, build, and operate scalable batch and streaming data pipelines, lakehouse and warehouse models, and data services. Own datasets end to end, improve Snowflake, Iceberg, Spark, and Flink performance and reliability, implement governance and observability, and support graph-serving data models. Partner with product and engineering teams, participate in on-call, review code and designs, and mentor junior engineers.
Top Skills: AirflowApache CassandraApache FlinkApache IcebergSparkAWSAzureClaude CodeCloudFormationCursorDatadogDbtGithub CopilotGoogle Cloud PlatformGrafanaJavaKafkaKubernetesMlopsOpensearchPrometheusPythonScalaSnowflakeSQLTerraform

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account