Decagon Logo

Decagon

Staff Software Engineer, Infrastructure

Reposted 10 Days Ago
Be an Early Applicant
Hybrid
San Francisco, CA, USA
300K-300K Annually
Senior level
Hybrid
San Francisco, CA, USA
300K-300K Annually
Senior level
Design and operate production infrastructure for high-scale systems, implementing services, optimizing performance, and improving developer tools.
The summary above was generated by AI

About Decagon

Decagon is the leading conversational AI platform empowering every brand to deliver concierge customer experiences.

Our technology enables industry-defining enterprises like Avis Budget Group, Block’s Cash App and Square, Chime, Oura Health, and Hunter Douglas to deploy AI agents that power personalized, deeply satisfying interactions across voice, chat, email, SMS, and every other channel.

We’re building a future where customer experiences are being redefined from support tickets and hold music to faster resolutions, richer conversations, and deeper relationships. We’re proud to be backed by world-class investors who share that vision, including a16z, Accel, Bain Capital Ventures, Coatue, and Index Ventures, along with many others.

We’re an in-office company, driven by a shared commitment to excellence and velocity. Our values — Just Get It Done, Invent What Customers Want, Winner’s Mindset, and The Polymath Principle — shape how we work and grow as a team.

About the Team

The Infrastructure team builds and operates the foundations that power Decagon: networking, data, ML serving, developer platform, and real‑time voice. We partner closely with product, data, and ML to deliver high‑scale, low‑latency systems with clear SLOs and great developer ergonomics.

We organize around five focus areas:

  • Core Infra: The foundational cloud stack—networking, compute, storage, security, and infrastructure‑as‑code—to ensure reliability, scale, and cost efficiency.

  • Data Infra: Streaming/batch data platforms powering analytics/BI and customer‑facing telemetry, including for customer‑managed and on‑prem environments.

  • ML Infra: GPU and model‑serving platforms for LLM inference with multi‑provider routing and support for on‑prem/air‑gapped deployments.

  • Platform (DevEx): CI/CD, paved paths, and core services that make shipping fast, safe, and consistent across teams.

Our mission is to deliver magical support experiences — AI agents working alongside humans to resolve issues quickly and accurately.

About the Role

We’re hiring a Senior Infrastructure Engineer to design, build, and operate production infrastructure for high‑scale, low‑latency systems. You’ll own critical services end‑to‑end, improve reliability and performance, and create paved‑paths that let every Decagon engineer ship confidently.

In this role, you will

  • Design and implement critical infrastructure services with strong SLOs, clear runbooks, and actionable telemetry.

  • Partner with research and product teams to architect solutions, set up prototypes, evaluate performance, and scale new features.

  • Tune service latencies: optimize networking paths, apply smart caching/queuing, and tune CPU/memory/I/O for tight p95/p99s.

  • Evolve CI/CD, golden paths, and self‑service tooling to improve developer velocity and safety.

  • Support various deployment architectures for customers with robust observability and upgrade paths.

  • Lead infrastructure‑as‑code (Terraform) and GitOps practices; reduce drift with reusable modules and policy‑as‑code.

  • Participate in on‑call and drive down toil through automation and elimination of recurring issues.

Your background looks something like this

  • 8+ years building and operating production infrastructure at scale.

  • Depth in at least one area across Core/Data/AI‑ML/Platform/Voice, with curiosity to learn the rest.

  • Proven track record meeting high availability and low latency targets (owning SLOs, p95/p99, and load testing).

  • Excellent observability chops (OpenTelemetry, Prometheus/Grafana, Datadog) and incident response (PagerDuty, SLO/error budgets).

  • Clear written communication and the ability to turn ambiguous requirements into simple, reliable designs.

Even better if you have

  • Experience being an early backend/platform/infrastructure engineer at another company

  • Strong Kubernetes experience (GKE/EKS/AKS) and experience across multiple cloud providers (GCP, AWS, and Azure)

  • Experience with customer‑managed deployments


Compensation

$200K – $400K + Offers Equity
This range reflects the expected compensation for this role. Compensation within the range is determined based on experience, skills, and the scope of responsibilities, with flexibility for candidates who demonstrate exceptional impact.
In addition to base salary, we offer competitive equity. Final compensation may vary based on location within the United States.

Benefits

We proudly offer the following benefits for our full-time employees:

  • Take what you need vacation policy (subject to local requirements; UK employees receive 25 days of statutory leave)

  • Medical, Dental, and Vision benefits for you and your family

  • Life Insurance and Disability Benefits

  • Retirement Plan (e.g., 401K, pension)

  • Parental Leave

  • Fertility and family building benefits through Carrot

  • Daily lunches and snacks in the office to keep you at your best

These benefits are described in more detail in Decagon’s policies, may vary by location, and can change at any time according to applicable compensation and benefits plans.

HQ

Decagon San Francisco, California, USA Office

2261 Market St, 5378, San Francisco, California, United States, 94114

Similar Jobs

13 Days Ago
In-Office or Remote
2 Locations
188K-275K Annually
Senior level
188K-275K Annually
Senior level
Cloud • Information Technology • Machine Learning
The Staff Software Engineer will design and implement the backend architecture for Marimo's molab, focusing on high availability, low latency, and system stability, while utilizing CoreWeave's infrastructure.
Top Skills: Cloud InfrastructureDistributed SystemsGpu Resource AllocationKubernetesPython
4 Days Ago
In-Office
San Francisco, CA, USA
190K-250K Annually
Senior level
190K-250K Annually
Senior level
Healthtech
Lead a small team building data and ML infrastructure for large-scale medical imaging ML. Architect and implement distributed compute platforms (training and inference), cloud-data systems, and production deployments. Mentor engineers, write high-performance Python code (including cross-language bindings), use frameworks like Ray and Kubernetes, apply infrastructure-as-code, and collaborate with researchers to enable scalable, secure ML workflows and monitoring.
Top Skills: Apache IcebergAWSAzureC++CdkGCPKubernetesPythonRayTerraform
4 Days Ago
Hybrid
San Mateo, CA, USA
230K-275K Annually
Senior level
230K-275K Annually
Senior level
Artificial Intelligence • Hardware • Robotics • Software
Build and operate Skydio's cloud infrastructure and Kubernetes fleet, improve continuous delivery, collaborate across teams to add platform capabilities, enhance security controls, and drive cost-saving initiatives. Hands-on role spanning infrastructure and product code changes.
Top Skills: Ci/CdCloud PlatformsContinuous DeliveryContinuous DeploymentDevOpsGoKubernetesPython

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account