Jobs and Careers
PR
Senior Site Reliability Engineer (SRE)
PragmatikePortugalRemotefull_timeVerifiedPosted 2 Dec 2025
About the role
<p><strong>Job Description</strong></p><p><strong></strong><strong>Location:</strong> Fully remote EU timezone (CET ±2h)<br/><strong>Start date:</strong><span> ASAP<br/></span><strong>Languages:</strong><span> Fluent English is mandatory<br/></span><strong>Industry:</strong><span> Cloud Computing </span></p>
<p>We are hiring at Pragmatike to expand our team and drive the growth of our internal projects.</p>
<p>Our focus is on developing cutting-edge solutions in Cloud Computing, while fostering a culture of collaboration and innovation. Joining us means being part of a passionate team where your ideas and skills directly contribute to shaping tomorrows technologies.</p>
<p>If you're excited about working on ambitious projects in a dynamic and flexible environment, we'd love to hear from you!</p>
<p><strong>Responsabilities:</strong></p>
<ul>
<li>
<p>Operate and maintain Linux-based infrastructure (Debian/Ubuntu).</p>
</li>
<li>
<p>Deploy, manage, and scale Kubernetes clusters across bare-metal, virtualized, and on-prem environments.</p>
</li>
<li>
<p>Oversee full cluster lifecycle: upgrades, node pools, networking, storage, and security hardening.</p>
</li>
<li>
<p>Implement automation for provisioning and operations using Ansible, Bash/Python, and GitOps workflows.</p>
</li>
<li>
<p>Design and maintain networking architecture including VLANs, L2/L3 routing, VPNs, and multi-site connectivity.</p>
</li>
<li>
<p>Build automated deployment workflows (PXE boot, Preseed, cloud-init).</p>
</li>
<li>
<p>Deploy and maintain observability stacks (Prometheus/Grafana, Loki, ELK, Graylog).</p>
</li>
<li>
<p>Lead incident response activities, define SLOs/SLIs, and optimize alerting and monitoring pipelines.</p>
</li>
<li>
<p>Manage virtualization and orchestration layers (OpenStack, Proxmox, VMware).</p></li></ul>
<p><strong>Requirements:</strong>Expert-level, hands-on experience operating Kubernetes in production environments.</p>
<ul><li>Strong proficiency with Linux systems administration (Debian/Ubuntu).</li><li>Solid understanding of networking fundamentals (VLANs, routing, VPNs).</li><li>Experience building and maintaining automation workflows (Ansible, Bash/Python, Git-based).</li><li>Experience with observability stacks such as Prometheus, Grafana, ELK, Loki, or Graylog.</li><li>Background with virtualization technologies (OpenStack, Proxmox, VMware).</li><li>Strong understanding of distributed systems and container orchestration.</li><li>Ability to work autonomously in a fast-paced, engineering-driven environment.<strong></strong></li></ul>
<p><strong>Nice To Have:</strong></p>
<ul><li>Experience with service mesh (Istio, Linkerd) or advanced CNI implementations.<p></p>
</li><li>
<p>Knowledge of Cloudflare APIs, DNS automation, or tunnel configurations.</p>
</li><li>
<p>Experience with GPU infrastructure, node preparation, or resource scheduling.</p>
</li><li>
<p>Familiarity with security best practices (RBAC, firewalls, network policies).</p>
</li><li>
<p>Exposure to IT asset management or license tracking workflows.</p></li></ul>
<p><strong>Why Join Us:</strong></p>
<ul><li>100% remote work with flexible hours</li><li>High-impact role with autonomy and ownership</li><li>Collaborative and international engineering team</li><li>Cutting-edge tech stack with strong focus on reliability and automation.</li></ul>
Apply for this role
Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.
Apply Now →Generate Application KitFree account required — sign up in 30s