Jobs and Careers
RI

Senior Site Reliability Engineer

Rithum
Ireland - Remote, IrelandRemotefull_timeVerifiedPosted 27 Aug 2025

About the role

Rithum™ is the world’s most trusted commerce network, accelerating how brands, suppliers, and retailers work together to deliver seamless e-commerce experiences. We provide an unmatched platform for brands and retailers, enabling them to accelerate growth, optimise operations across channels, scale product offerings and enhance margins.

Today, more than 40,000 companies trust Rithum to grow their business across hundreds of channels, representing over $50 billion in annual GMV. Using our commerce, marketing, and delivery solutions, our customers create optimised consumer shopping journeys from beginning to end.

 

Overview

As a Senior Site Reliability Engineer in our Platform Engineering Organization, you help to build and run large-scale, distributed, fault-tolerant systems. In this role, you are involved in the complete lifecycle of our products from inception to operation, ensuring they are reliable, performant and meet appropriate uptime and availability targets. You design and maintain resilient systems, implement robust observability through metrics, logging, and tracing, and build automation that improves deployment, monitoring, and incident response workflows.  This includes leveraging AI/ML for intelligent alerting, anomaly detection, and predictive incident response to enhance system reliability and scalability.  Working with others in the organization, you help develop and influence operational tooling, best practices, and standards that empower the engineering organization and help ensure Rithum's effective and efficient operations. As a Senior Engineer, you operate independently, self-prioritizing work, design and lead projects from start to completion, engaging with stakeholders for successful delivery. You mentor and assist less experienced people on the team and coach them to help improve their skills. 

 

Responsibilities

  • Collaborate with developers, Client Support, and cross-functional teams to build production automation, analysis tools, and improving reliability and performance.
  • Design, implement, and maintain robust application monitoring and observability systems for a distributed, highly available, and scalable software stack leveraging AI/ML to detect anomalies and asset with incidents.
  • Analyse and resolve problems in legacy environments while designing and implementing modern, scalable solutions from the ground up.
  • Participate in the rotating on-call schedule, ensuring that user emergencies, platform alerts, and support requests are addressed.
  • Drives automation and operational efficiency.

 

Qualifications 

Minimum Qualifications  

  • 3+ years' experience working as an SRE, DevOps Engineer or related
  • Experience with logging and monitoring systems like CloudWatch, Grafana or Prometheus
  • Experience with AWS foundations, including compute, storage, and security
  • Good AWS knowledge including application design, migration support, cost planning, capacity allocation, and application resiliency
  • Expertise in creating multi-region cloud systems with a solid disaster recovery plan
  • Experience with both high-level and scripting languages like Python, Bash or Typescript
  • Experience troubleshooting and debugging complex, distributed applications
  • IaC experience automating infrastructure with CDK, Terraform or Ansible
  • Experience with continuous deployment pipelines and containerization like EKS or ECS
  • Strong understanding of software engineering fundamentals, including object-oriented design, modular architecture, and maintainable coding practices.

Preferred Qualifications 

  • You have a bachelor's degree, or higher, in Computer Science or related field; or equivalent practical experience demonstrating strong software engineering fundamentals.
  • Experience working in a highly collaborative environment with both platform and product teams,
  • Excellent collaboration and communication skills, consistently learning new technologies and helps foster an environment of continuous improvement and innovation.
  • Client satisfaction focus.

 

Travel Required

Up to 10%

 

Other Duties

Please note this job description is not designed to cover or contain a comprehensive listing of activities, duties or responsibilities required of the employee for this job. Duties, responsibilities, and activities may change at any time with or without notice.

 

What it’s like to work at Rithum 

When you join Rithum, you can expect to work with smart risk-takers, courageous collaborators, and curious minds.

As part of the Rithum team, you are valued, supported, and included. Guided by a

Apply for this role

Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.

Apply Now →Generate Application Kit

Free account required — sign up in 30s

Company

Rithum

View company profile →