Ford Motor Company Logo

Ford Motor Company

Director of Cloud SRE

Posted 5 Days Ago
In-Office or Remote
Hiring Remotely in United States
142K-268K Annually
Expert/Leader
In-Office or Remote
Hiring Remotely in United States
142K-268K Annually
Expert/Leader
Leads SRE engineering leaders and engineers while defining enterprise observability, reliability, and platform strategy across GCP, on-premise, manufacturing, distribution, and campus environments. Oversees vendor-agnostic tooling, OpenTelemetry integrations, CI/CD observability, SRE maturity models, and Agentic AI initiatives. Drives adoption of SRE practices, develops technical roadmaps, partners with senior leadership and operational teams, and maintains hands-on architectural and technical credibility.
The summary above was generated by AI

We are looking for a Director of Cloud SRE to lead a team of engineering leaders and engineers responsible for federating core SRE principles across a global, hybrid technology organization. This role drives the strategy, architecture, and roadmap for our internal observability and reliability tooling — spanning telemetry standards, developer pipeline integration, and application team SRE maturity — and works in close partnership with peer SRE leaders to extend that strategy consistently across cloud-native and on-premise environments alike.

This is a builder's role as much as a leader's role. You will guide a team that develops internally-owned tooling built intentionally to remain vendor-agnostic so the organization is never architecturally locked to a single observability provider and drives Agentic AI deeper into our SRE ecosystem. While our primary application runtime is GCP, this leader must be equally comfortable partnering across the SRE organization to extend reliability and observability standards into data center, manufacturing, distribution, and global campus environments — meeting engineering and operations teams where they are, not just where the platform lives.

The ideal candidate blends technical depth with organizational fluency: someone who can sit in an architecture review and a roadmap planning session with equal credibility, who has personally built and operated production systems, and who can partner effectively across a broader SRE leadership team to advance a long-term, unified observability strategy.

Responsibilities
  • Partner with fellow SRE leaders to define and drive a multi-year, holistic strategy for unified observability and SRE platform offerings, spanning cloud-native (GCP) and on-premise (data center, manufacturing, distribution, campus) environments.

  • Lead, develop, and grow a team of engineering managers/leads and individual contributor engineers, building organizational depth in SRE practice and platform engineering.

  • Drive the roadmap for internally-built observability tooling, ensuring architecture remains vendor-agnostic and portable across telemetry backends (OpenTelemetry-first design, current integration with Dynatrace) with a focus on Agentic AI platforms to simplify correlation data.

  • Help federate core SRE principles — SLIs/SLOs, error budgets, incident management, toil reduction, capacity and reliability engineering — across application and platform teams enterprise-wide, working alongside peer SRE leaders rather than centralizing reliability as a bottleneck.

  • Partner with developer experience and platform engineering teams to embed observability and reliability tooling directly into CI/CD pipelines and source repositories, shifting reliability left in the development lifecycle.

  • Contribute to an SRE maturity model, providing application teams a clear, staged path to deepen their own reliability practice with SRE org support and self-service tooling.

  • Build cross-domain relationships with manufacturing, plant, and OT engineering leadership, in partnership with other SRE leaders, to extend reliability and observability discipline into environments with materially different constraints (legacy protocols, air-gapped or constrained networks, safety-critical operations).

  • Represent SRE platform direction to senior technology leadership, including architecture governance bodies, and act as an escalation point for major reliability and observability initiatives within your team's scope.

  • Contribute to vendor relationship and technology decisions related to observability tooling, balancing build-vs-buy tradeoffs against long-term platform and cost strategy.

  • Ensure the team maintains hands-on technical currency — reviewing designs, contributing to architecture decisions, and staying credible as a technical leader.

Qualifications
  • Bachelor's degree in Computer Science, Engineering, or related field, or equivalent practical experience.

  • 10+ years of experience in Site Reliability Engineering, platform engineering, or infrastructure engineering

  • 4+ years in a people leadership role managing engineering leaders and/or engineers.

  • Demonstrated experience building and operating observability platforms at scale, with hands-on depth in OpenTelemetry and at least one enterprise observability platform (Dynatrace, Datadog, New Relic, Splunk, or similar).

  • Proven track record designing and delivering internally-built developer tooling, including integration with CI/CD pipelines, source control platforms, and developer workflows.

  • Strong working knowledge of public cloud architecture (GCP strongly preferred; AWS/Azure acceptable) and demonstrated ability to extend reliability practices into hybrid or on-premise environments.

  • Deep understanding of core SRE principles: SLIs/SLOs, error budgets, incident management and postmortem practice, toil reduction, capacity planning, and reliability-by-design.

  • Experience operating in environments with heterogeneous infrastructure — cloud, data center, and OT/manufacturing or industrial environments a strong plus.

  • Demonstrated ability to build a long-term technical strategy and translate it into an executable roadmap, balancing tactical remediation against multi-year platform investment.

  • Strong executive communication skills — able to represent technical strategy to senior leadership and align cross-functional stakeholders around a unified direction.

  • Experience with infrastructure-as-code (Terraform or equivalent) and modern software delivery practices (agile/PI planning experience a plus).

Preferred:

  • Experience building or scaling an SRE function within a large, matrixed enterprise.

  • Familiarity with emerging AI/agentic observability standards (OTel GenAI semantic conventions) and their application to platform tooling.

  • Experience establishing SRE maturity models or capability frameworks used to guide staged team adoption.

 

This role requires up to 10% travel.

 

You may not check every box, or your experience may look a little different from what we've outlined, but if you think you can bring value to Ford Motor Company, we encourage you to apply!
As an established global company, we offer the benefit of choice. You can choose what your Ford future will look like: will your story span the globe, or keep you close to home? Will your career be a deep dive into what you love, or a series of new teams and new skills? Will you be a leader, a changemaker, a technical expert, a culture builder…or all of the above? No matter what you choose, we offer a work life that works for you, including:
• Immediate medical, dental, vision and prescription drug coverage
• Flexible family care days, paid parental leave, new parent ramp-up programs, subsidized back-up child care and more
• Family building benefits including adoption and surrogacy expense reimbursement, fertility treatments, and more
• Vehicle discount program for employees and family members and management leases
• Tuition assistance
• Established and active employee resource groups
• Paid time off for individual and team community service 
• A generous schedule of paid holidays, including the week between Christmas and New Year’s Day 
• Paid time off and the option to purchase additional vacation time. 
 

This position is leadership level 5 and ranges from $141,700-$268,300.     
Final determination of salary grade will be based on candidate's skills and experience, and base salary will be set within the applicable range according to job scope, responsibility and competitive market value.
For more information on salary and benefits, click here: https://fordcareers.co/LL5

Visa sponsorship is NOT available for this position.

 

Candidates for positions with Ford Motor Company must be legally authorized to work in the United States. Verification of employment eligibility will be required at the time of hire.

We are an Equal Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, religion, color, age, sex, national origin, sexual orientation, gender identity, disability status or protected veteran status. In the United States, if you need a reasonable accommodation for the online application process due to a disability, please call 1-888-336-0660.

#LI-Remote

    #LI-DS2  

Ford Motor Company Palo Alto, California, USA Office

3200 Hillview, Palo Alto, CA, United States, 94304

Similar Jobs

4 Minutes Ago
Easy Apply
Remote or Hybrid
Easy Apply
196K-219K Annually
Senior level
196K-219K Annually
Senior level
AdTech • Enterprise Web • Information Technology • Machine Learning • Marketing Tech • Sales
As a Staff Data Scientist, you will lead complex marketplace problem-solving, architect ML systems, mentor teams, and drive data science strategy in partnership with engineering and product management.
Top Skills: AirflowAWSAzureGCPKubeflowPythonPyTorchSQLTensorFlowTfx
8 Minutes Ago
Easy Apply
Remote
United States
Easy Apply
195K-280K Annually
Senior level
195K-280K Annually
Senior level
Big Data • Fintech • Mobile • Payments • Financial Services
Build, deploy, and maintain AI-powered agents, APIs, and applications for People operations on Snowflake. Turn messy business requirements into production systems, integrate with Workday/Notion/case tools, ensure data governance and security, design multi-model LLM reliability controls, and operate full lifecycle including CI/CD, monitoring, and incident fixes while partnering with non-technical stakeholders.
Top Skills: APIsCase Management ToolsCi/CdContainerizationData GovernanceDbtDocker (Containers)GitGitLlmsMonitoringNotionPythonQuicksilverRbacsSecrets ManagementSnowflakeSnowpark Container ServicesWorkday
9 Minutes Ago
Remote or Hybrid
7 Locations
264K-395K Annually
Expert/Leader
264K-395K Annually
Expert/Leader
eCommerce • Fintech • Hardware • Payments • Software • Financial Services
Lead technical direction and architecture for Neighborhoods mobile experiences across Cash App and Square. Drive cross-organizational mobile initiatives, stay hands-on in iOS development, partner with backend/web/Android teams, mentor engineers, and apply AI-assisted workflows to accelerate feature development and quality.
Top Skills: BazelClaude CodeCombineCursorGooseObjective-CProtocol BuffersSwiftSwift ConcurrencySwiftuiUikitWorkflow

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account