Jobs and Careers
DU

System Analyst - Site Reliability Engineer II

Duke University
United Statesfull_timeVerifiedPosted 8 Jan 2025

About the role

At Duke Health, we're driven by a commitment to compassionate care that changes the lives of patients, their loved ones, and the greater community. No matter where your talents lie, join us and discover how we can advance health together.

Occupational Summary
The DHTS Systems Analyst-Site Reliability Engineer (SRE) is responsible for designing, implementing, and maintaining large-scale distributed systems with a focus on reliability, scalability, and performance.
The SRE collaborates with development teams to ensure that applications and services are designed and operated to meet reliability targets and scale efficiently. This role involves working with Kubernetes for
on-premises environments and Azure Kubernetes Service (AKS) for cloud-based solutions.

 

Essential Tasks/Responsibilities
Level 2 (DHTS System Analyst 2)
• Participate in on-call rotations to respond to system alerts and incidents.
• Assist in troubleshooting and resolving system issues and outages across both on-premises and cloud environments.
• Collaborate with development teams to improve system reliability and efficiency across onpremises and cloud infrastructures.
• Independently design and implement monitoring solutions for complex systems in OpenShift and AKS environments.
• Lead incident response efforts and coordinate with multiple teams during outages, considering the nuances of both on-premises and cloud infrastructures.
• Develop and implement automation solutions to improve system reliability and efficiency across OpenShift and AKS platforms.
• Conduct thorough root cause analysis for incidents and propose long-term solutions that align with the organization's hybrid infrastructure strategy.
• Contribute to the design and implementation of disaster recovery and business continuity plans, leveraging both on-premises and cloud resources.
• Mentor junior team members and provide technical guidance on OpenShift and AKS best practices.
• Participate in the evaluation and implementation of new technologies and tools that complement OpenShift and AKS environments.
�� Collaborate with development teams to define and implement SLIs, SLOs, and SLAs across both platforms.
• Contribute to the development of architectural improvements to enhance system reliability and scalability in a hybrid infrastructure model.
Required Qualifications at this Level

Education
Bachelor's degree in a related field is preferred, or equivalent work experience.
Experience
• Level 2 (DHTS System Analyst 2): Minimum 5 years of software development experience and/or
IT solutions engineering.
Required Skills and Knowledge
Level 2 (DHTS System Analyst 2)
• Familiarity with project management and Agile/SCRUM methodologies
• Proficiency in at least one programming language (e.g., Python, Go, Java)
• Familiarity with version control systems (e.g., Git)
• Familiarity with CI/CD technologies like GitLab CI or GitHub Actions
• Basic understanding of server administration (preferably Linux)
• Understanding of networking topologies, firewall rules, and certificate management
• Ability to analyze customer requirements and translate into effective solutions
• Critical thinking and problem-solving skills
• Strong customer service orientation
• Strong experience with Application Development Lifecycle, with a DevOps focus
• Proficiency in script writing (e.g., Ansible Playbooks, Helm Charts)
• Extensive experience with containerization and orchestration technologies (Docker, Kubernetes)
• Strong experience with CI/CD technologies and practices
• Advanced knowledge of server administration (preferably Linux)
• Solid understanding of networking topologies, firewall rules, and certificate management
• Proven ability to analyze complex customer requirements and translate into effective solutions
• Advanced troubleshooting and root cause analysis skills
• Strong project management skills, including Agile/SCRUM experience
• Experience with cloud platforms (AWS, Azure, GCP) and services (SaaS, IaaS, PaaS, FaaS)
• Knowledge of Enterprise Architecture best practices
• Familiarity with AI and ML concepts
Desired Skills
• Red Hat OpenShift certifications
• Azure DevOps and Infrastructure certifications
• CKA (Certified Kubernetes Administrator) or CKAD (Certified Kubernetes Application Developer) certifications
• Experience with multi-cloud environments
• Knowledge of FHIR APIs and healthcare-specific technologies
• Excellent time management, organizational, and task prioritization skills
• Strong presentation skills
• Ability to communicate effectively with non-technical staff and members of interdisciplinary teams
• Ability to interact well and ef

Apply for this role

Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.

Apply Now →Generate Application Kit

Free account required — sign up in 30s

Company

Duke University

View company profile →