As a founding member of the Observability team, you'll build and integrate observability solutions, ensuring system reliability and performance while mentoring others and driving automation.
ABOUT RETOOL
Nearly every company in the world runs on custom software for critical operations like tracking performance metrics, handling support workflows, building admin dashboards, and countless processes you might never have thought of. But most companies don't have the resources to properly invest in these tools, leading to a lot of old, clunky internal software, or worse, teams still stuck in manual and spreadsheet workflows.
AI has changed who gets to build software. The definition of "developer" now includes analysts, operators, and domain experts creating solutions directly—and the tools they reach for are multiplying by the week. That's both an opportunity and a challenge: as more people build with more AI tools, the risk of shipping ungoverned software into production grows just as fast.
At Retool, we're building the platform that makes all of it safe to ship. Build with any AI tool you want, then deploy into one place that connects to your real business data, enforces enterprise policies automatically, and lets teams create once and reuse everywhere with shared, trusted components. The cost of building software has collapsed. The cost of governing it hasn't—and that's the problem we solve.
Developers and domain experts have already automated over 100 million hours of work on our platform, freeing them to focus on creative problem-solving and strategic work that drives real business value. The people closest to the problem can now build the software to solve it, safely, and within enterprise guardrails.
Let's build the future together.
WHY WE'RE LOOKING FOR YOU:
Retool is in an exciting hyper-growth phase and we need infrastructure engineers to tackle our rapid scaling challenges. These scaling challenges are unique both in scope and in technical complexity as we scale both the company and the product.
As a member of the Cloud Platform team, you will build and automate foundational systems for reliability, security, performance, developer velocity, and scale. Whether you’re interested in wide breadth or deep depth, there are both opportunities for both. You will work cross functionally, supporting the engineering, product, design, customer success, support, and sales functions.
IN THIS ROLE, YOU'LL:
- Scale Retool’s core cloud platform for high availability and performance globally
- Work collaboratively with the rest of the engineering team to deliver infrastructure for core and emerging products
- Evolve our backend architecture/infrastructure for both cloud and on-premise deployments
- Work with the team to set and prioritize our roadmap to maximize customer impact
- Define and automate developer process/workflow
- Support our systems in production
- Develop new data solutions
- Build monitoring and observability for production systems
THE SKILLSET YOU'll BRING:
- You have a track record of delivering engineering projects and process improvements
- You have experience scaling cloud infrastructure
- You enjoy building and productionizing developer productivity tools, frameworks, and other aspects of platform engineering
- You have experience with inner workings of Linux, containers (Docker, containerd), and container orchestration technologies (e.g. docker-compose, Kubernetes)
- You have a track record of building productive, collaborative relationships, both within an engineering org and across the broader company
- You enjoy the ambiguity and high-ownership culture of early-stage startups
- You are pragmatic, solution-oriented, and scrappy
- You enjoy working collaboratively with a broad range of job functions and roles
- You have experience with our tech stack: Node, Postgres, Docker, Kubernetes
- You have experience scaling relational databases (Postgres preferred)
- You have good knowledge of cloud, on-prem, traffic routing, service architecture in multi-region setup.
For candidates based in the United States, the pay range(s) for this role is listed below and represents base salary range for non-commissionable roles or on-target earnings (OTE) for commissionable roles. This salary range may be inclusive of several career levels at Retool and will be narrowed during the interview process based on a number of factors such as (but not limited to), scope and responsibilities, the candidate’s experience and qualifications, and location.
Additional compensation in the form(s) of equity and/or commission are dependent on the position offered. Retool provides a comprehensive benefit plan, including medical, dental, vision, and 401(k). Pay and benefits are subject to change at any time, consistent with the terms of any applicable compensation or benefit plans.
The base pay range for this role is $163,710 – $306,000 per year.
Retool offers generous benefits to all employees and hybrid work location. For more information, please visit the benefits and perks section of our careers page!
Retool is currently set up to employ all roles in the US and specific roles in the UK. To find roles that can be employed in the UK, please refer to our careers page and review the indicated locations.
Retool San Francisco, California, USA Office
Retool's headquarters is in San Francisco's Mission District within walking distance to the 16th St. Bart station, great coffee shops, and SF institutions like Dandelion Chocolate and Tartine Manufacturing. We have dedicated parking, secure bike storage, and 24/7 onsite security.
Similar Jobs
Financial Services
Lead design, build, and operate scalable, secure cloud platform capabilities. Develop Java/Spring Boot services, implement IaC with Terraform, run Kubernetes in HA, deliver CI/CD, apply SRE practices, use observability/APM tooling, and prototype AI/agentic workflows. Mentor engineers and enforce reliability, security, and compliance.
Top Skills:
Agentic AiAi-Assisted Development ToolsAlbAmazon EksApi GatewayAWSAws LambdaBlazemeterDatadogDynatraceEcsElbGitJavaJava 17JmeterJulesJunitKubernetesMockitoPolicy-As-CodeRedisRest ApisRoute 53Secrets ManagementSpinnakerSplunkSpring BootTerraform
Cloud • Computer Vision • Information Technology • Sales • Security • Cybersecurity
Build and operate a hyper-scale data lake and analytics platform: develop Java microservices and Spark/Scala/Flink pipelines, own a new graph database, design ultra-high-scale data platforms, optimize performance and latency, enable efficient querying for analytics and ML, and collaborate on feature design and stability across teams.
Top Skills:
Apache FlinkSparkAWSCassandraClickhouseDruidDynamoDBGraph DatabaseIcebergJavaKotlinKubernetesMySQLPinotPostgresPythonRayScala
Financial Services
Lead design and delivery of scalable, secure AWS-based platform solutions. Develop and maintain production code and automation, champion AI-assisted engineering practices, produce architecture documentation, analyze large datasets for resiliency and performance, implement CI/CD and observability, and mentor junior engineers.
Top Skills:
Ai-Assisted Development ToolsAppdynamicsAWSAws S3CloudwatchData LakesDatadogDynatraceEc2EcsEksElasticsearchEmrEvent BusGrafanaJavaJenkinsNoSQLPythonSpinnakerSplunkSQL
What you need to know about the San Francisco Tech Scene
San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.
Key Facts About San Francisco Tech
- Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
- Major Tech Employers: Google, Apple, Salesforce, Meta
- Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
- Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
- Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
- Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine


