Sr. Software Reliability Engineer
AbbottAbout the role
JOB DESCRIPTION:
Working at Abbott
At Abbott, you can do work that matters, grow, and learn, care for yourself and your family, be your true self, and live a full life. You’ll also have access to:
Career development with an international company where you can grow the career you dream of.
Employees can qualify for free medical coverage in our Health Investment Plan (HIP) PPO medical plan in the next calendar year.
An excellent retirement savings plan with a high employer contribution.
Tuition reimbursement, the Freedom 2 Save student debt program, and FreeU education benefit - an affordable and convenient path to getting a bachelor’s degree.
A company recognized as a great place to work in dozens of countries worldwide and named one of the most admired companies in the world by Fortune.
A company that is recognized as one of the best big companies to work for as well as the best place to work for diversity, working mothers, female executives, and scientists.
The Opportunity
We’re looking for a strong Senior Site Reliability Engineer (SRE) who’s ready to roll up their sleeves and make a real impact. This is a hands-on role focused on designing, implementing, and optimizing scalable, reliable infrastructure that powers our life-critical medical device platforms.
If you thrive in complex environments, enjoy solving infrastructure challenges, and are passionate about building systems that are both resilient and compliant with healthcare regulations—this is the role for you.
As a Senior SRE, you’ll work closely with engineering, QA, cybersecurity, and regulatory teams to ensure our systems meet the highest standards of availability, performance, and compliance. You’ll also play a key role in on-call production support, helping monitor systems, respond to incidents, and drive continuous improvements in reliability and observability
What You’ll Work On
- System Reliability & Performance: Design and maintain fault-tolerant infrastructure for medical device platforms across cloud and on-prem environments.
- Automation & Tooling: Develop and maintain tools for deployment, monitoring, and incident response.
- Incident Management: Lead root cause analysis and postmortems, driving continuous improvement.
- Compliance & Security: Ensure systems meet regulatory standards (e.g., HIPAA, FDA CFR Part 11, ISO 13485).
- Observability: Implement and optimize monitoring and alerting systems using tools like Prometheus, Grafana, ELK, or Datadog.
- Collaboration:
- Partner with software engineers to embed reliability into application design.
- Work closely with QA and validation teams to support software verification and validation processes.
- Collaborate with cybersecurity teams to ensure secure infrastructure and data protection.
- Participate in cross-functional planning and sprint ceremonies to align reliability goals with product development.
Required Qualifications:
- Bachelor’s degree in computer science, Engineering, or related field.
- 5+ years of experience in SRE, DevOps, or Infrastructure Engineering.
- Strong experience with cloud platforms (AWS, Azure, or GCP).
- Proficiency in infrastructure-as-code tools (Terraform, Ansible, etc.).
- Expertise in container orchestration (Kubernetes, Docker).
- Solid understanding of CI/CD pipelines and automation frameworks.
Preferred Qualifications
- Master’s degree in computer science, Engineering, or related field.
- Experience in regulated industries (medical devices, pharma, healthcare).
- Knowledge of FDA software validation processes.
- Experience with real-time monitoring and alerting systems.
- Certifications in cloud technologies or SRE practices (e.g., AWS Certified DevOps Engineer, Google SRE).
- Familiarity with compliance frameworks relevant to medical devices.
- Excellent communication and collaboration skills.
Apply for this role
Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.
Apply Now →Generate Application KitFree account required — sign up in 30s