Jobs and Careers
FO

Director of Cloud SRE

Ford Motor Company
United States, United StatesRemotefull_timeVerifiedPosted 20 Aug 2026
💰 $268,300/yr($141,700/yr$268,300/yr)

About the role

We are looking for a Director of Cloud SRE to lead a team of engineering leaders and engineers responsible for federating core SRE principles across a global, hybrid technology organization. This role drives the strategy, architecture, and roadmap for our internal observability and reliability tooling — spanning telemetry standards, developer pipeline integration, and application team SRE maturity — and works in close partnership with peer SRE leaders to extend that strategy consistently across cloud-native and on-premise environments alike.

This is a builder's role as much as a leader's role. You will guide a team that develops internally-owned tooling built intentionally to remain vendor-agnostic so the organization is never architecturally locked to a single observability provider and drives Agentic AI deeper into our SRE ecosystem. While our primary application runtime is GCP, this leader must be equally comfortable partnering across the SRE organization to extend reliability and observability standards into data center, manufacturing, distribution, and global campus environments — meeting engineering and operations teams where they are, not just where the platform lives.

The ideal candidate blends technical depth with organizational fluency: someone who can sit in an architecture review and a roadmap planning session with equal credibility, who has personally built and operated production systems, and who can partner effectively across a broader SRE leadership team to advance a long-term, unified observability strategy.

  • Partner with fellow SRE leaders to define and drive a multi-year, holistic strategy for unified observability and SRE platform offerings, spanning cloud-native (GCP) and on-premise (data center, manufacturing, distribution, campus) environments.

  • Lead, develop, and grow a team of engineering managers/leads and individual contributor engineers, building organizational depth in SRE practice and platform engineering.

  • Drive the roadmap for internally-built observability tooling, ensuring architecture remains vendor-agnostic and portable across telemetry backends (OpenTelemetry-first design, current integration with Dynatrace) with a focus on Agentic AI platforms to simplify correlation data.

  • Help federate core SRE principles — SLIs/SLOs, error budgets, incident management, toil reduction, capacity and reliability engineering — across application and platform teams enterprise-wide, working alongside peer SRE leaders rather than centralizing reliability as a bottleneck.

  • Partner with developer experience and platform engineering teams to embed observability and reliability tooling directly into CI/CD pipelines and source repositories, shifting reliability left in the development lifecycle.

  • Contribute to an SRE maturity model, providing application teams a clear, staged path to deepen their own reliability practice with SRE org support and self-service tooling.

  • Build cross-domain relationships with manufacturing, plant, and OT engineering leadership, in partnership with other SRE leaders, to extend reliability and observability discipline into environments with materially different constraints (legacy protocols, air-gapped or constrained networks, safety-critical operations).

  • Represent SRE platform direction to senior technology leadership, including architecture governance bodies, and act as an escalation point for major reliability and observability initiatives within your team's scope.

  • Contribute to vendor relationship and technology decisions related to observability tooling, balancing build-vs-buy tradeoffs against long-term platform and cost strategy.

  • Ensure the team maintains hands-on technical currency — reviewing designs, contributing to architecture decisions, and staying credible as a technical leader.

  • Bachelor's degree in Computer Science, Engineering, or related field, or equivalent practical experience.

  • 10+ years of experience in Site Reliability Engineering, platform engineering, or infrastructure engineering

  • 4+ years in a people leadership role managing engineering leaders and/or engineers.

  • Demonstrated experience building and operating observability platforms at scale, with hands-on depth in OpenTelemetry and at least one enterprise observability platform (Dynatrace, Datadog, New Relic, Splunk, or similar).

  • Proven track record

Apply for this role

Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.

Apply Now →Generate Application Kit

Free account required — sign up in 30s

Company

Ford Motor Company

View company profile →