Infinity Constellation Logo

Infinity Constellation

Staff Software Engineer - Supernal

Reposted 18 Hours Ago
Remote
Hiring Remotely in USA
10-10 Annually
Senior level
Remote
Hiring Remotely in USA
10-10 Annually
Senior level
This role involves owning and evolving the backend platform for AI employees, mentoring engineers, optimizing system performance, and driving architectural decisions at Supernal.
The summary above was generated by AI
Staff Software Engineer
About Supernal

Supernal helps small-to-medium businesses hire their first AI employee. Our AI teammates are built using intelligent, agentic workflows deployed on a proprietary platform. We deliver working, value-generating AI Employees—not tools—that handle real business processes alongside human teams.

The Role

We're looking for a Staff/Principal Software Engineer to own and evolve the core platform that powers our AI employees. This is a technical leadership position responsible for the systems that enable our agents to scale reliably: the Django backend, distributed task infrastructure, event-driven architecture, Kubernetes deployments, and observability stack.

You'll work across the full system—from database query optimization to Helm chart tuning to designing new platform abstractions. You'll be a force multiplier for the engineering team, driving architectural decisions, eliminating scaling bottlenecks, and establishing patterns that make the platform more robust and developer-friendly.

This role reports to the Director of Engineering and involves significant autonomy in shaping technical direction.

What You'll Own
  • Drive platform architecture decisions and align the team on scalable patterns and long-term maintainability

  • Review a high volume of code, design docs, and architectural proposals for scalability, reliability, security, and operability

  • Be a technical mentor and force multiplier: unblock engineers, raise the bar on production readiness, and establish platform best practices

  • Own and evolve the core backend platform (Django/DRF/ASGI) performance and correctness

  • Scale async execution across Celery + Dramatiq + Temporal/Cortex; implement resilient workflow patterns (retries, circuit breakers, graceful degradation)

  • Optimize PostgreSQL/pgvector (query tuning, connection pooling) and caching strategies

  • Maintain and improve Kubernetes deployment infrastructure (GKE, Helm, Terraform/OpenTofu) and CI/CD + rollout strategies. Own KEDA autoscaling policies and resource allocation across worker pools.

  • Own reliability of RabbitMQ, Redis, and PostgreSQL infrastructure; lead incident response and post-mortems

  • Extend OpenTelemetry + Datadog instrumentation, dashboards, alerts, and SLOs; profile and reduce latency/memory bottlenecks

What We're Looking ForRequired
  • 10+ years building and operating production backend systems at scale

  • Deep expertise in Python (Django preferred) and relational databases (PostgreSQL)

  • Hands-on experience with Kubernetes, Helm, and cloud infrastructure (GCP preferred)

  • Strong background in distributed systems: message queues, event sourcing, workflow orchestration

  • Production experience with async task systems (Celery, Dramatiq, or similar)

  • Track record of debugging complex production issues across multiple services

  • Ability to work autonomously and drive technical initiatives without close supervision

  • Clear technical communication—able to explain tradeoffs and build consensus

Preferred
  • Experience with Temporal or similar workflow engines

  • Background in LLM infrastructure, RAG systems, or AI/ML platforms

  • Familiarity with OpenTelemetry, Datadog, or similar observability stacks

  • Experience with KEDA or other Kubernetes autoscaling solutions

  • Contributions to multi-tenant SaaS platform architecture

  • History of improving developer experience and platform abstractions

What Success Looks Like
  • Platform services maintain high availability with predictable performance under load

  • Scaling bottlenecks are identified and resolved proactively

  • New features ship faster because platform primitives are well-designed and documented

  • Incidents are rare, quickly detected, and thoroughly addressed

  • Engineers across the team adopt platform patterns and best practices

  • Technical debt is systematically identified and paid down

  • You're a trusted technical voice in architectural discussions

Compensation & Logistics
  • Compensation: Competitive salary commensurate with experience (Staff/Principal level)

  • Location: Remote

  • Type: Full-time

  • Requirements: Overlap with Americas timezones for collaboration; reliable high-speed internet

Similar Jobs

A Minute Ago
In-Office or Remote
104K-145K Annually
Senior level
104K-145K Annually
Senior level
Artificial Intelligence • Fintech • Information Technology • Logistics • Payments • Business Intelligence • Generative AI
Lead RevOps data modeling, ETL/DBT development on Snowflake, SQL-based analytics, and Tableau reporting. Optimize and maintain RevOps systems (Salesforce, Clari, Outreach, Gong/Chorus), ensure data quality, apply advanced analytics/AI concepts for forecasting, and partner cross-functionally to deliver actionable revenue insights.
Top Skills: Ai/MlChorusClariDbtGongOutreachRevenue IntelligenceSalesforceSnowflakeSQLTableau
A Minute Ago
In-Office or Remote
150K-170K Annually
Senior level
150K-170K Annually
Senior level
Artificial Intelligence • Fintech • Information Technology • Logistics • Payments • Business Intelligence • Generative AI
Drive account expansion within assigned enterprise customers by maintaining executive sponsorship, developing account plans, identifying and closing cross-sell opportunities, and partnering with Customer Value Managers and solution teams to deliver Coupa's spend management solutions and measurable business value.
Top Skills: CoupaCoupa AiCoupa PayCoupa Supply ChainTotal Spend Management Platform
9 Minutes Ago
In-Office or Remote
215K-358K Annually
Senior level
215K-358K Annually
Senior level
Artificial Intelligence • Healthtech • Machine Learning • Natural Language Processing • Biotech • Pharmaceutical
Lead strategy and investment for Pfizer Oncology patient solutions platforms and co-pay/patient support programs. Transform programs from reactive to AI-enabled predictive support, set outcome-based KPIs, manage the US & Global Patient Experience team, own budgets and vendor selection, partner cross-functionally, and represent Pfizer externally on patient support innovation.
Top Skills: AIConversational AiPredictive Analytics

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account