Jobs and Careers
SE

Senior DevOps/SRE Engineer

SEI
United Statesfull_timeVerifiedPosted 13 Aug 2026
💰 $170,000/yr($140,000/yr$170,000/yr)

About the role

 

We are looking for a Senior Site Reliability Engineer to work as part of a lean, productfocused engineering organization. This role is about building and operating reliable cloudbased systems by writing code, automating infrastructure and delivery workflows, and reducing friction for developers and users. You will work closely with product and application engineers to design, deploy, and operate systems with clear ownership and practical engineering judgment. We expect you to use modern tooling, including AIassisted tools where appropriate, to speed up automation, troubleshooting, and operations while remaining accountable for correctness, security, and reliability. This role favors simple, effective solutions, hands��on ownership, and continuous improvement within small Agile teams. 

 

What you will do: 

  • Design, build, and operate cloud infrastructure for critical production and nonproduction applications with reliability and simplicity as primary goals 

  • Architect and evolve multi-account AWS foundations (organizations, accounts, IAM boundaries, guardrails, and environment separation) to enable secure, scalable delivery 

  • Design and operate cloud networking architecture (VPCs, routing, segmentation, ingress/egress, connectivity patterns) to support reliability, security, and compliance requirements 

  • Treat reliability, security, and compliance as firstclass design concerns throughout the system lifecycle 

  • Build tooling and automation that reduces errors, shortens recovery time, and improves daytoday operations 

  • Implement monitoring, logging, and alerting that make system behavior observable and actionable 

  • Use AIassisted tools to accelerate infrastructure delivery, automation, troubleshooting, and rootcause analysis, applying engineering judgment to validate outcomes 

  • Implement reliability guardrails for releases (progressive delivery, safe rollbacks, change risk controls) and provide production support during deployments. 

  • Participate in incident response, perform root cause analysis, and drive durable improvements that prevent recurrence 

  • Work closely with application engineers to coown system design, operation, and continuous improvement 

  • Maintain clear, lightweight documentation that supports shared ownership and effective oncall operations 

 

 What we n

Apply for this role

Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.

Apply Now →Generate Application Kit

Free account required — sign up in 30s

Company

SEI

View company profile →