Focal Systems Logo

Focal Systems

Senior Site Reliability Engineer

Sorry, this job was removed at 08:11 p.m. (PST) on Wednesday, May 28, 2025
In-Office
San Francisco, CA
In-Office
San Francisco, CA

Similar Jobs

3 Days Ago
Remote or Hybrid
San Diego, CA, USA
127K-215K Annually
Senior level
127K-215K Annually
Senior level
Artificial Intelligence • Cloud • HR Tech • Information Technology • Productivity • Software • Automation
The Senior Site Reliability Engineer will maintain and develop the reliability of ServiceNow's cloud infrastructure, ensuring high performance and scalability while providing 24/7 support for government clients.
Top Skills: AWSAzureJavaScriptLinuxMariadbMySQLPostgresPython
3 Days Ago
Remote or Hybrid
San Diego, CA, USA
127K-215K Annually
Senior level
127K-215K Annually
Senior level
Artificial Intelligence • Cloud • HR Tech • Information Technology • Productivity • Software • Automation
As a Senior Site Reliability Engineer, you will maintain cloud infrastructure reliability, automate tasks, and drive technical resolutions across the technology stack, focusing on improving system design and operations.
Top Skills: AWSAzureJavaScriptLinuxMariadbMySQLPostgresPython
5 Days Ago
In-Office
Costa Mesa, CA, USA
154K-231K Annually
Senior level
154K-231K Annually
Senior level
Aerospace • Artificial Intelligence • Hardware • Robotics • Security • Software • Defense
The Senior Site Reliability Engineer will manage infrastructure, improve CI/CD pipelines, and advocate for reliability practices while supporting various teams across the company.
Top Skills: AnsibleAWSAzureBashCloudFormationDockerGoGoogle Cloud PlatformHelmKubernetesPowershellPuppetPythonRustSccmTeamcenterTerraform

Location: San Francisco - hybrid (1-2 days per week) 
Salary: $170-190k + stock 


Company Description

Focal Systems is the industry leader in retail AI solutions. We are a Silicon Valley based startup that has more than doubled in size every year since inception. We are a Deep Learning first company. Our mission is to automate and optimize brick and mortar retail using deep learning computer vision. Focal Systems has been deployed at scale with the top retailers in the world. We are looking for smart, creative and passionate people who want to help build a great and enduring company and deploy Deep Learning to the world! 


Mission of the role:
To enable us to scale from 200k to 1 million cameras


Job Summary

As a Sr. DevOps/Site Reliability Engineer (SRE) at our company, you will play a pivotal role in ensuring the smooth operation and continuous improvement of our infrastructure, deployment processes, and overall system reliability.



Responsibilities

  • Set up and manage blue/green and canary deployments to ensure smooth launches without downtime.
  • Operate multiple large GCP Kubernetes clusters and fine tune for reliability vs cost
  • Manage the various distributed services of the company, ensuring to always provide graceful updates, comprehensive test coverage, tracking of logs, and 99.9% uptime
  • Work with Backend, Frontend and Deep Learning teams and write infrastructure automation code for their needs
  • Identify scalability bottlenecks through load testing and plan infrastructure architecture
  • Create tools to provide transparency/ease of access into the company's rich datasets stored across varying geographic locations and data formats
  • Design, build, and manage a robust Continuous Integration and Continuous Deployment (CI/CD) pipeline.
  • Lead uptime improvement processes including: postmortem review, on-call setup.

Requirements 

  • Solid experience in an infrastructure or Site Reliability Engineer (SRE) role 
  • Cloud Services experience with Google Cloud Platform (GCP)
  • Hands-on experience with containerization (Docker) and orchestration platforms (Kubernetes) required
  • Experience in cloud cost management
  • Great understanding of SQL, networking, distributed systems, operating systems (debian) and software engineering practices
  • Experience with messaging systems
  • Terraform or other Infrastructure as Code automation solution
  • Operating Relational SQL databases and Redis at terabyte scale. 
  • Proven experience with setting up monitoring/alerting and reliability engineering
  • Scriptings skills in Python
  • Must be comfortable with 12-hour on call rotations

Nice to have experience:

  • GitOps 
  • Setting up automation for complex load testing scenarios
  • Tuning Deep Learning pipelines with Python, Pytorch and Multiprocessing
  • Backend programming with Python

Why Focal Systems

Strong Values and Mission - We are a tightly-knit team with an ambitious mission and a strong set of core values, which define our approach to business and have successfully guided us since inception.
Exceptional Team - We are a team of hard-working, fun-loving professionals from some of the most eminent universities, research labs, and tech companies of our time. We pride ourselves on recruiting exceptional individuals to help us redefine the state-of-the-art.
Outstanding Partners - We work with 10+ of the largest retailers in the world and have a world-class roster of investors, advisors and partners to support & advise us in our endeavors.


Benefits

We care deeply about the health, happiness, and wellbeing of all of our employees. We offer:

  • Competitive Salary & Attractive Stock
  • Paid Time Off 
  • Quarterly Team Retreats
  • Education grants

HQ

Focal Systems Burlingame, California, USA Office

1300 Old Bayshore Hwy, 255, Burlingame, CA, United States, 94010

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account