Jobs and Careers
DO

Senior Platform Engineer

DoiT International
Remote Ireland, IrelandRemotefull_timeVerifiedPosted 25 Mar 2025

About the role

Location

Our Senior Platform Engineer will be an integral part of our Engineering teams in EMEA. This role is based remotely in EMEA in one of our legal entities: UK, Ireland, Israel, Estonia, or Spain. The role is also available to contractors in other East Europe locations and Portugal.

Who We Are
DoiT works alongside cloud-driven organizations to optimize your cloud use so you can focus on business growth and innovation. We deliver consulting, support, and training to solve cloud challenges, a FinOps-certified platform to navigate and automate spend, and access to volume pricing and billing support for a streamlined procurement process.

Our global team of cloud experts have decades of experience in the analytics, optimization, and governance of cloud architecture, as well as specializations in Kubernetes, GenAI, and more. An award-winning strategic partner of AWS, Google Cloud, and Microsoft Azure, DoiT works alongside more than 3,800 customers worldwide.

The Opportunity

As a Platform Engineer, you will be responsible for building and maintaining the foundational infrastructure that empowers our development teams. This is an Individual Contributor role requiring hands-on work with AWS & GCP, Kubernetes, and Terraform. You will contribute to the design, implementation, and automation of our platform, ensuring scalability, reliability, and security.

Responsibilities

  • Function as an individual contributor within the team: actively collaborating with peers through thorough code reviews, providing constructive support and mentorship, and contributing to a unified technical direction for the platform. This role also requires collaboration with individuals in feature teams, providing them with support and working with them to facilitate the adoption of developed platform features.
  • Architect, Design, and Implement Infrastructure as Code (IaC) using Terraform: You will be responsible for the comprehensive lifecycle management of our infrastructure through Terraform. This involves designing modular and reusable Terraform configurations, managing state effectively, implementing robust testing strategies, and ensuring that our infrastructure is consistently provisioned and managed in a predictable and repeatable manner.
  • Deploy, Manage, and Optimize Kubernetes Clusters on AWS (EKS) and GCP (GKE): You will take ownership of the deployment, configuration, and ongoing maintenance of our Kubernetes clusters on AWS Elastic Kubernetes Service (EKS) and GCP Google Kubernetes Engine (GKE). This includes managing node groups, configuring network policies, implementing service meshes, handling cluster upgrades, and ensuring high availability and fault tolerance. You will also be responsible for monitoring cluster health, performance, and resource utilization, and proactively addressing any issues that arise.
  • Develop and Maintain Sophisticated CI/CD Pipelines for Platform Components: You will design, implement, and maintain robust Continuous Integration/Continuous Deployment (CI/CD) pipelines specifically tailored for our platform components. This involves integrating various tools like Argo CD or Atlantis, automating build processes, implementing comprehensive testing strategies, and ensuring seamless deployment of platform updates. You will also focus on optimizing pipeline performance and reducing deployment times.
  • Diagnose, Troubleshoot, and Resolve Platform-Related Issues: You will be the primary point of contact for diagnosing and resolving platform-related issues, including performance bottlenecks, scalability challenges, and security vulnerabilities. This involves utilizing advanced troubleshooting techniques, analyzing logs and metrics, and collaborating with development teams to identify and resolve root causes. You will also contribute to creating comprehensive incident response plans and post-mortem analyses.
  • Drive Automation Initiatives to Streamline Operational Tasks and Enhance System Reliability: You will champion automation initiatives to eliminate manual operational tasks, reduce human error, and improve overall system reliability. This involves developing scripts, tools, and workflows to automate tasks such as infrastructure provisioning, configuration management, and monitoring. You will also proactively identify opportunities for automation and drive continuous improvement in our operational processes.
  • Act as a Strategic Partner to Development Teams, Understanding and Addressing Their Infrastructure Needs: You will foster strong relationships with development teams, acting as a trusted advisor and strategic partner. You will actively engage with them to understand their infrastructure requirements, provide

Apply for this role

Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.

Apply Now →Generate Application Kit

Free account required — sign up in 30s

Company

DoiT International

View company profile →