Jobs and Careers
RO

Senior Site Reliability Engineer (SRE)

Roche
Sant Cugat del Vallès, Spainfull_timeVerifiedPosted 12 Jun 2025

About the role

<p>At Roche you can show up as yourself, embraced for the unique qualities you bring. Our culture encourages personal expression, open dialogue, and genuine connections,  where you are valued, accepted and respected for who you are, allowing you to thrive both personally and professionally. This is how we aim to prevent, stop and cure diseases and ensure everyone has access to healthcare today and for generations to come. Join Roche, where every voice matters.</p><h3></h3><p></p><h3>The Position</h3><p><span>The role requires the candidate to be available for on-call duty service, responding promptly to urgent issues and emergencies outside of regular working hours, ensuring that critical situations are addressed in a timely and effective manner</span></p><p></p><p><b>Who We Are</b> </p><p><span>At Roche, we are passionate about transforming patients’ lives, and we are bold in both decision and action - we believe that good business means a better world. That is why we come to work every single day. We commit ourselves to scientific rigor, unassailable ethics, and access to medical innovations for all. We do this today to build a better tomorrow. </span></p><p></p><p><span>Roche is strongly committed to a diverse and inclusive workplace. We strive to build teams that represent a range of backgrounds, perspectives, and skills. Embracing diversity enables us to create a great place to work and to innovate for patients.</span></p><p></p><p><span>Roche is building a global site reliability engineering (SRE) team that will support commercial and internal solutions. This team will have the mindset of building and creating engineering solutions to solve a broad spectrum of problems.</span></p><p><span> </span></p><p><b>Step into the Future of IT Infrastructure with Roche!</b></p><p><span>As a seasoned Site Reliability Engineer (SRE) at Roche, you'll leverage your deep software engineering expertise to propel our IT infrastructure to new heights of robustness, scalability, and reliability. This isn't just a role—it's an invitation to shape the backbone of critical infrastructures and drive our technological innovations forward.</span></p><p></p><p><b>Your Mission</b></p><p><span>Design and maintain cutting-edge tools, scripts, and frameworks that automate repetitive tasks, streamline software deployment, and manage expansive systems with unparalleled efficiency.</span></p><p></p><p><span>Partner closely with forward-thinking development teams to architect and implement high-performance solutions that elevate system efficiency, optimize resource utilization, and enhance deployment processes for superior uptime and user satisfaction.</span></p><p></p><p><b>Your Impact</b></p><p><span>Lead the charge in incident management and response. Detect system anomalies, troubleshoot swiftly, and conduct thorough root cause analyses to prevent recurring issues.</span></p><p></p><p><span>Champion continuous improvement by refining monitoring and alerting mechanisms, conducting insightful post-incident reviews, and embedding best practices in software lifecycle management. Your strategic foresight and meticulous planning will ensure our systems are not only reliable but also superlatively performant.</span></p><p></p><p><span>By joining our elite team, you will play a pivotal role in delivering seamless experiences to our end-users, exceeding business and customer demands, and solidifying Roche's reputation as a leader in IT innovation.</span></p><p></p><p><b>Your Core Responsibilities</b></p><ul><li><p><b>Reliability Mastery:</b><span><b> </b>Proactively monitor and maintain system reliability using advanced tools like DataDog, VictorOps, ELK, Grafana, and Prometheus. Become a key player in ensuring system stability and performance.</span></p></li><li><p><b>Uptime Guardian:</b><span> Ensure optimal uptime and performance by swiftly identifying issues and responding to alerts with precision.</span></p></li><li><p><b>Technical Troubleshooter</b>:<span> Basic understanding of Architecture and designs to<span> </span> deep dive into complex technical issues, troubleshoot, investigate, and resolve them. Collaborate seamlessly with engineering teams to enable timely and effective resolutions.</span></p></li><li><p><b>Service Excellence:</b><span> Maintain and consistently achieve defined SLAs, SLIs, and SLOs, ensuring service levels are consistently met or exceeded.</span></p></li><li><p><b>Automation Innovator:</b><span> Develop and deploy automation scripts (using Python or other scripting languages) to streamline operations, enhance system efficiencies, and reduce manual tasks.</span></p></li><li><p><b>Cloud Steward</b>:<span> Manage and maintain robust infrastructure across AWS and Azure environments, implementing best practices to ensure peak performance, reliability  of cloud-based applications. Drive cost optimization through best practice implementation and continuous vigilance. </span></p></li><li><p><b>Cross-functional Colla

Apply for this role

Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.

Apply Now →Generate Application Kit

Free account required — sign up in 30s

Company

Roche

View company profile →