Worth Logo

Worth

Senior DevOps Engineer, Infrastructure & Reliability

Posted 14 Hours Ago
In-Office or Remote
Hiring Remotely in Miami, FL
Senior level
In-Office or Remote
Hiring Remotely in Miami, FL
Senior level
Build and operate scalable, reliable infrastructure using Terraform, Kubernetes, AWS, and CI/CD automation. Responsibilities include managing EKS platforms, improving deployment pipelines, securing networking and IAM, enhancing observability, optimizing cloud costs, implementing disaster recovery, reducing operational toil, and modernizing legacy infrastructure. The role also leads incident response, reliability initiatives, platform adoption, and cross-team infrastructure projects.
The summary above was generated by AI

Worth AI, a leader in the computer software industry, is looking for a Senior DevOps Engineer to join our Infrastructure team with a singular mission: to make our systems faster, more reliable, and more resilient while making life dramatically easier for engineers shipping software. 

This is a hands-on build role. You will spend most of your time writing Terraform, tuning Kubernetes workloads, automating things that are currently manual, and shipping infrastructure changes to production. You'll join a small platform team with an established roadmap and existing patterns, and a strong voice in how the work gets built.

  • Implement scalable Infrastructure-as-Code patterns using tools like Terraform to standardize cloud provisioning and reduce configuration drift.
  • Own and evolve our Kubernetes platform (EKS or self-managed), ensuring workloads are secure, scalable, and resilient by default.
  • Optimize CI/CD pipelines to improve deployment frequency, reduce lead time, and increase confidence in releases.
  • Design and enforce secure networking, IAM, and secrets management strategies across environments.
  • Improve observability by refining metrics, logs, and tracing using tools like DataDog, ensuring actionable insight into system health.
  • Optimize cloud cost efficiency through rightsizing, autoscaling strategies, and architectural improvements.
  • Implement disaster recovery planning, backup strategies, and multi-region resilience initiatives.
  • Refactor brittle or manually managed infrastructure into automated, testable, and reproducible systems.
  • Introduce new infrastructure tooling or architectural shifts and drive adoption through documentation, workshops, and hands-on support.
  • Partner with engineering teams to eliminate friction in CI/CD, deployments, and cloud environments.
  • Communicate technical trade-offs clearly across engineering and product stakeholders, balancing speed with safety.
Technology Stack
  • Cloud & Infrastructure: AWS (EKS, RDS, MSK, S3, Lambda, IAM, VPC)
    Containerization & Orchestration: Kubernetes, ArgoCD
    Infrastructure-as-Code: Terraform
    CI/CD: GitHub Actions
    Monitoring & Observability: DataDog
    Data & Messaging: PostgreSQL, Kafka, Redis
    Languages (as needed): Bash, Python, TypeScript, JavaScript

Requirements
  • 8+ years in DevOps, SRE, or infrastructure engineering.
  • Proven experience designing and operating production Kubernetes environments at scale.
  • Deep hands-on expertise with AWS infrastructure and cloud networking.
  • Strong experience building and maintaining Terraform modules across large cloud environments.
  • Demonstrated ownership of CI/CD systems and measurable improvement of DORA metrics.
  • Experience leading incident response processes and driving meaningful postmortem outcomes.
  • Strong understanding of distributed systems, event-driven architectures (Kafka), and database performance (PostgreSQL).
  • Proven ability to modernize legacy infrastructure and eliminate manual operational toil.
  • Track record of taking a scoped infrastructure project from an ambiguous starting point to production without needing daily direction.
  • Demonstrated ability to build trust across teams while raising the reliability bar.
Success Metrics
  • System Reliability: Maintain or exceed defined SLO/SLA targets with reduced incident frequency and duration.
  • Infrastructure Stability: Reduce production incidents caused by misconfiguration, manual processes, or infrastructure drift.
  • Operational Efficiency: Increase the percentage of infrastructure managed through code and automation.
  • Cost Optimization: Improve cloud cost efficiency without sacrificing reliability or performance.
Bonus Points (Nice to Have)
  • Experience coding applications
  • Experience operating high-throughput Kafka clusters (MSK or self-managed).
  • Strong background in database performance tuning (PostgreSQL, Redis).
  • Experience implementing autoscaling strategies for high-traffic systems.
  • Familiarity with service mesh technologies.
  • Experience building internal developer platforms (IDP).
  • Background in security best practices (zero-trust networking, policy-as-code).
  • Experience with multi-region or globally distributed systems.
  • Experience introducing platform-wide reliability frameworks (SLOs, error budgets, chaos testing).

All Remote Hires will be required to travel to Orlando, Florida at least twice per year for Town Halls and team collaboration, in addition to orientation in Orlando.


Benefits
  • Health Care Plan (Medical, Dental & Vision)
  • Retirement Plan (401k)
  • Life Insurance
  • Flexible Paid Time Off
  • 9 paid Holidays
  • Family Leave
  • Remote
  • Hybrid work (for Orlando Associates)
  • Free Food & Snacks (Orlando)
  • Wellness Resources

Similar Jobs at Worth

12 Hours Ago
In-Office or Remote
Senior level
Senior level
Artificial Intelligence • Fintech • Software • Financial Services
Own enterprise new-logo acquisition and expansion while personally closing complex, high-value deals and carrying an individual quota. Hire, coach, and develop Enterprise AEs and SDRs; establish forecasting, pipeline, conversion, and sales operating cadences. Partner with Marketing, Product, Customer Success, Partnerships, Legal, and Finance, while presenting performance and strategic recommendations to executives and the board.
Top Skills: ChallengerChorusCommand Of The MessageForce ManagementGongHubspotLinkedin Sales NavigatorMeddpiccOutreachSalesloftZoominfo
12 Hours Ago
In-Office or Remote
Mid level
Mid level
Artificial Intelligence • Fintech • Software • Financial Services
Own vulnerability management across endpoints, servers, and cloud infrastructure; prioritize remediation and track SLA performance. Harden AWS environments, improve identity management, investigate cloud alerts, and support application security through SAST, DAST, dependency scanning, secure code reviews, and threat modeling. Resolve security tickets, document incidents and remediation, and lead security initiatives while collaborating with engineering, IT, and compliance.
Top Skills: AWSAws ConfigBashCloudtrailDastGuarddutyIamJIRALinearOwasp Top 10PythonQualysS3SastScaSecurity HubSnykSoc 2TenableVpcWiz
12 Hours Ago
In-Office or Remote
Entry level
Entry level
Artificial Intelligence • Fintech • Software • Financial Services
Generates and qualifies new business opportunities through research, cold calls, email, and social outreach. The role conducts needs analysis, communicates product value, schedules sales meetings and demos, maintains CRM records, and collaborates with sales and marketing teams. It also requires staying informed about AI, industry trends, and competitors while meeting lead-generation targets.
Top Skills: Artificial IntelligenceCrm SoftwareHubspot

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account