Senior Software Engineer
Red HatAbout the role
Red Hat is seeking a Senior Software Engineer to join the CI/CD and delivery engineering team within Red Hat OpenShift Service on AWS (ROSA). ROSA is Red Hat's fully managed Kubernetes platform on AWS, built on a large-scale multi-tenant architecture where Red Hat operates shared control plane infrastructure while customers securely run workloads in their own AWS accounts.
This team owns the systems and practices that define how ROSA software moves from commit to production — across global AWS regions, safely and at speed. You will design, build, and operate the CI/CD pipelines, testing infrastructure, release automation, and delivery tooling that the broader ROSA engineering organization depends on every day. The stack includes Prow, ci-operator, Argo CD, Terraform, Konflux, and app-interface/qontract, with a strong emphasis on pipeline reliability, progressive delivery, automated quality gates, and infrastructure-as-code validation.. There is no separate operations team: what you build, you run — including operating across a multi-region fleet of management clusters and service clusters via OCM APIs.
This is not a role where CI/CD is a side responsibility. It is the job. You will own the end-to-end delivery pipeline for a complex, multi-region managed Kubernetes service — designing systems that catch problems before they reach production, enable safe and frequent releases, and give engineers fast, reliable feedback on every change. Success requires deep expertise in CI/CD architecture, build and release engineering, automated testing strategy, and production delivery, combined with strong Kubernetes and AWS fluency and the judgment to make sound trade-offs in a high-stakes production environment.
As Red Hat continues evolving toward a more AI-enabled software development lifecycle, you will also help define how AI-powered tooling and automation are integrated into CI/CD — from intelligent test selection and AI-assisted code review to agentic pipeline automation and LLM-powered release validation. Our engineering culture values strong ownership, technical depth, open collaboration, and continuous improvement.
What you will do
Own and evolve ROSA's CI/CD platform. Design, build, and operate the CI/CD pipelines and delivery infrastructure that ROSA engineering depends on to ship software safely and frequently. Own the full pipeline lifecycle — from source integration and build automation through testing, artifact management, promotion, and production deployment. Ensure pipelines are fast, reliable, observable, and secure. You are accountable for the health and reliability of the delivery path the same way a platform engineer is accountable for uptime.
Design and implement automated testing and quality gates. Build testing infrastructure and automated quality gates that catch defects, regressions, security issues, and configuration drift before they reach production. Define and enforce testing strategies across unit, integration, end-to-end, and infrastructure validation stages. Ensure that test results are fast, trustworthy, and actionable — flaky tests and slow feedback loops are problems you own and fix.
Drive progressive delivery and release engineering. Implement and operate progressive delivery strategies — canary deployments, sector-based staged rollouts, and progressive promotion through environment gates — that allow ROSA to release changes safely across global AWS regions. Build release automation that is repeatable, auditable, and resilient. Own the tooling and processes that determine how and when software reaches production.
Build and maintain infrastructure-as-code and configuration validation. Ensure that Terraform, Kubernetes manifests, and other infrastructure-as-code artifacts are validated, linted, and tested as part of every pipeline run. Integrate policy-as-code and security scanning into the delivery pipeline so that infrastructure changes meet compliance and security requirements before they are applied..
Leverage AI to accelerate CI/CD and delivery. Explore and integrate AI-powered tooling into the CI/CD lifecycle — intelligent test selection, automated root cause analysis for pipeline failures, and agentic CI triage. Help shape how the broader engineering team adopts AI-assisted delivery practices.
Operate what you build with a high reliability bar. The CI/CD platform is production infrastructure. Own its reliability with clear SLOs, monitoring, alerting, and incident response. When a pipeline is broken or a release is bl
Apply for this role
Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.
Apply Now →Generate Application KitFree account required — sign up in 30s