Istari Digital Logo

Istari Digital

Sr. Cloud Infrastructure Engineer

Reposted 18 Days Ago
Remote
Hiring Remotely in US
135K-220K Annually
Senior level
Remote
Hiring Remotely in US
135K-220K Annually
Senior level
Design, implement, and operate scalable AWS and Kubernetes infrastructure using Terraform and automation. Improve CI/CD, observability, security posture, and developer experience; troubleshoot production incidents, maintain runbooks, and support customer-managed deployments.
The summary above was generated by AI

We are seeking a Senior Cloud Infrastructure Engineer to join our Engineering team. This role is central to building, operating, and scaling the cloud infrastructure that powers our platform, and to making our software easy to deploy both in our cloud-hosted environments and in customer-managed environments.

This engineer will work closely with platform and product engineering to improve the reliability, automation, and operational maturity of our AWS and Kubernetes environments. Some of our environments serve regulated, security-sensitive customers; while deep security specialization is not required, this engineer should be comfortable operating within an established, security-conscious baseline.

The ideal candidate combines strong hands-on infrastructure expertise with a practical, execution-focused mindset and a bias toward automation.

Core Responsibilities

  • Design, implement, and maintain scalable, reliable infrastructure in AWS

  • Operate and improve Kubernetes-based environments, including production workloads

  • Build and maintain infrastructure as code using Terraform

  • Improve CI/CD pipelines, deployment workflows, and release automation in partnership with engineering teams

  • Build and maintain the packaging and reference architectures customers use to install our software in their own environments

  • Strengthen observability across the platform, including monitoring, logging, alerting, and actionable dashboards

  • Improve developer experience through tooling, environment automation, and self-service infrastructure

  • Operate within and preserve the established security and compliance posture of our environments

  • Monitor, troubleshoot, and resolve complex infrastructure issues with clear and timely communication

  • Participate in incident response and post-incident analysis

  • Develop and maintain documentation, runbooks, and technical standards

  • Identify opportunities to improve cost efficiency, performance, and resilience across environments

Required Qualifications

  • Minimum of 5 years of experience in DevOps, Infrastructure Engineering, Platform Engineering, or Site Reliability Engineering

  • Strong hands-on experience with AWS in production environments

  • Proven experience operating Kubernetes in production

  • Strong experience with Terraform and infrastructure-as-code practices

  • Proven experience building or improving CI/CD pipelines and deployment automation

  • Solid understanding of cloud networking, IAM, secrets management, and operational controls

  • Experience with monitoring, logging, and observability tooling

  • Scripting proficiency in Python, Go, Bash, or similar

  • Excellent troubleshooting and problem-solving skills in complex production environments

  • Strong communication skills with the ability to explain technical concepts to both technical and non-technical stakeholders

  • Must live/work in the U.S.

  •  Preferred Qualifications
  • Experience operating stateful workloads on Kubernetes, such as databases or message queues, including persistent storage and backup/recovery

  • Experience with GitOps-based deployment workflows

  • Experience packaging software for customer-managed or self-hosted deployment (e.g., Helm charts)

  • Familiarity with compliance or security frameworks such as FedRAMP, NIST, SOC 2, or similar

  • Experience with PostgreSQL, cloud storage platforms, and production networking patterns

  • Experience with configuration management tools such as Ansible

  • Experience with additional cloud platforms such as Azure or GCP

  • Experience with service mesh or advanced Kubernetes networking

  • Experience supporting customer-facing or mission-critical production infrastructure

  • Top Secret Security Clearance

Similar Jobs

21 Days Ago
Remote or Hybrid
UT, USA
125K-180K Annually
Senior level
125K-180K Annually
Senior level
Cloud • Computer Vision • Information Technology • Sales • Security • Cybersecurity
Lead enterprise cloud architecture, governance, and infrastructure enablement. Define reusable IaC patterns (Terraform), cloud networking (Transit Gateway, multi-account VPCs), and migration/data center exit strategies. Build architecture review processes, drive adoption of cloud-native practices (Kubernetes, CI/CD), and translate technical decisions for stakeholders while coordinating security and compliance.
Top Skills: Ai/LlmArgocdAWSAws IamAws Network FirewallCi/CdGitlab Ci/CdGitopsHelmKubernetesOctopus DeployTerraformTerraform ModulesTransit GatewayVpc
11 Days Ago
Remote
US
168K-194K Annually
Senior level
168K-194K Annually
Senior level
Big Data • eCommerce
Design, build, and operate shared cloud infrastructure across AWS, Kubernetes, Terraform, Databricks, and Cloudflare. Lead SRE and DevOps initiatives involving reliability, observability, CI/CD, incident response, disaster recovery, cost optimization, and developer self-service. Partner with application and data engineering teams to improve workload operations, deployment safety, and infrastructure scalability. Participate in on-call rotations and contribute to technical standards, architecture, documentation, and sustainable 24/7 operations.
Top Skills: AnthropicAWSAws CloudformationCi/CdCloudflareDatabricksDatadogInfrastructure As CodeKubernetesOpenaiTerraform
17 Days Ago
Remote
United States
Senior level
Senior level
Cloud • Information Technology • Cybersecurity
Leads Azure, AWS, and hybrid cloud deployment, migration, modernization, networking, security, automation, governance, backup, and disaster recovery projects. Designs highly available cloud architectures, implements Infrastructure as Code, supports virtual desktop environments, assesses customer readiness, and develops migration roadmaps. Provides technical leadership, mentoring, escalation support, architecture review participation, and customer advisory services.
Top Skills: Active DirectoryAd CsAmazon EbsAmazon Ec2Amazon EfsAmazon FsxAmazon S3Amazon Web Services (Aws)Amazon WorkspacesArm TemplatesAws Application Migration Service (Mgn)Aws BackupAws CliAws CloudformationAws ConfigAws Direct ConnectAws GuarddutyAws InspectorAws Migration HubAws Security HubAws Transit GatewayAws VpcAzure Active Directory Domain ServicesAzure Application GatewayAzure BackupAzure CliAzure ExpressrouteAzure Load BalancerAzure MigrateAzure MonitorAzure Site RecoveryAzure SqlAzure StorageAzure Traffic ManagerAzure Virtual DesktopAzure Virtual MachinesAzure Virtual NetworkAzure Vpn GatewayBicepCiscoDhcpDnsHypervisorsLog AnalyticsAzureNerdio For MspPalo AltoPowerPointPowershellPythonSQL ServerTerraformVisio

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account