Postscript Logo

Postscript

Staff Platform Engineer

Posted 22 Days Ago
Remote
Hiring Remotely in USA
190K-230K Annually
Senior level
Remote
Hiring Remotely in USA
190K-230K Annually
Senior level
Design, build, and operate a next-generation CDP/messaging platform at scale. Define infra topology, scaling, blast-radius limits, and safety guardrails for AI agents. Own reliability, capacity, cost, security boundaries, incident leadership, and author self-verifying runbooks and checks so automation can operate safely.
The summary above was generated by AI

Postscript is the AI messaging platform trusted by 20,000+ Shopify brands — including Brooklinen, Ruggable, True Classic, and Dr. Squatch. 

With a mission to make SMS your number one revenue driving channel, we built the best-in-class SMS marketing platform and launched revenue-driving AI features not available on any other platform. And we're just getting started. 

Backed by Greylock and Y Combinator, fully remote since 2018.

Job DescriptionStaff Platform Engineer (Infrastructure)

We are seeking an experienced Staff Platform Engineer to help build and operate Acorn, Postscript's next-generation CDP and messaging platform: ~52 Rust/Python services on EKS, built to run at 500K+ events/sec for 20K+ merchants.

What makes this role different: the platform is agent-operated. Deploys, migrations, drift reconciliation, and infra standup are automated and encoded as self-verifying runbooks that AI agents execute. We are not hiring someone to run those procedures by hand. We are hiring the person who designs the system, writes the guardrails that let agents execute it safely, and owns the failures no runbook covers yet.

Primary duties
  • System & Platform Design: Own infra topology, scaling and blast-radius boundaries, the build/deploy graph, database ownership boundaries, and cost ceilings — the decisions agents execute but can't make.
  • Guardrail Engineering: Build the validation gates, idempotency checks, drift detectors, and runbooks that keep both humans and agents from doing damage. Your deliverable is the safety system, not the deploy.
  • Escalation Tier: Diagnose the novel failures — consumer lag, DLQs, materialized-view chains, consent ordering. Write the runbook the first time; agents handle it after.
  • Reliability, Capacity & Cost: Own the 500K events/sec target: load testing, Karpenter/HPA tuning, Database sizing, SLOs and alerting.
  • Security & Secrets Boundary: Own IAM/IRSA, secret rotation, the auth chain, and what the agent/MCP surface is allowed to do.
  • Agent-Fleet Operability: Keep skills, runbooks, and agent context in sync with reality. A stale runbook is a wrong actor, executed confidently.
  • Incident Leadership: Lead incidents and capture every fix as a new runbook plus regression test.
What We’ll Love About You
  • Experience: 5+ years operating production Kubernetes and cloud infrastructure (AWS) at scale.
  • IaC Discipline: Strong Terraform and GitOps/Kustomize practice, with a bias for making operations idempotent, automated, and self-verifying.
  • Incident Judgment: Real incident-response experience on distributed-systems failures — databases, streaming, consumer lag, data integrity.
  • Data Systems Fluency: You know how Postgres, a columnar store, and a log/streaming bus behave under load.
  • Agent Fluency: Comfort working alongside AI agents — writing the guardrails, runbooks, and checks that let automation operate safely, and knowing where a human must stay in the loop.
  • Instincts: You reach for "encode this as a check" before "I'll remember to do this." You'd rather delete a manual step than document it. You treat a wrong runbook as a production bug.
What You’ll Love About Us
  • Salary range of USD $190,000 to $230,000 base plus significant equity (we do not have geo based salaries)
  • High growth startup - plenty of room for you to directly impact the company and grow your career!
  • Work from home (or wherever)
  • Fun - We’re passionate and enjoy what we do
  • Competitive compensation and opportunity for equity
  • Flexible paid time off

For information about how we use your personal data, please see our U.S. Job Applicant Privacy Notice

You are welcome here. Postscript is an ever-evolving place of equal employment for talented individuals.

Similar Jobs

2 Days Ago
Remote or Hybrid
District of Columbia, USA
160K-246K Annually
Senior level
160K-246K Annually
Senior level
Automotive • Big Data • Information Technology • Robotics • Software • Transportation • Manufacturing
Build and maintain a full-stack platform that connects AI models to business processes and third-party platforms. Productionize prototypes into reliable, testable agentic and automation systems, design complex workflow orchestration, integrate APIs and data systems, implement observability and resilience, and evangelize best practices while acting as technical escalation for deployed systems.
Top Skills: Azure Ai FoundryDatabricksGleanGooglePostgresPythonReactRest ApisViteWorkflow Orchestration
8 Days Ago
Remote
United States
220K-260K Annually
Expert/Leader
220K-260K Annually
Expert/Leader
Fintech • Information Technology • Software
Lead design and evolution of SentiLink's data platform for fraud detection. Architect scalable batch and streaming pipelines, improve reliability, performance, and observability, set engineering standards, mentor engineers, drive cross-team technical initiatives, participate in production support/on-call, and evaluate build-vs-buy and emerging technologies.
Top Skills: AWSDockerEksEmrFlinkGlueGoHadoopKafkaKubernetesLambdaOpensearchPostgresPythonRedshiftS3SnsSparkSqs
14 Days Ago
Easy Apply
Remote
USA
Easy Apply
218K-257K Annually
Senior level
218K-257K Annually
Senior level
Artificial Intelligence • Blockchain • Fintech • Financial Services • Cryptocurrency • NFT • Web3
Lead end-to-end ML systems for identity verification, including document authenticity, face match, liveness and deepfake detection. Build GNN-based identity graphs, behavioral/device-intelligence models, vendor benchmarking/evaluation, and production enforcement. Mentor engineers and align ML architecture with Product, Compliance, Risk, and Security stakeholders.
Top Skills: Computer Vision/BiometricsGenerative AiGraph Neural Networks (Gnns)LlmsModel Serving InfrastructureNlpPythonPyTorchSequence ModelsTensorFlow

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account