Jobs and Careers
KO
Senior Site Reliability Engineer
Kontakt.ioremote in Poland, PolandRemotefull_timeVerifiedPosted 4 Feb 2025
About the role
<b><a href="http://kontakt.io/" rel="noopener noreferrer">Kontakt.io</a> is building the platform that care operations run on.</b><br/>We reduce waste, cut costs, and improve revenue by improving throughput, asset utilization and staff productivity. Our platform uses AI, RTLS, and EHR data to enable self-learning agents to automate workflows, adapt in real-time, and orchestrate all of care delivery operations.<br/>Easy to deploy and scale, it gives a clear picture of spaces, equipment, and people, eliminating inefficiencies and enhancing the patient experience. With measurable 10X ROI and over 20+ use cases, <a href="http://kontakt.io/" rel="noopener noreferrer">Kontakt.io</a> is the go-to platform for better and faster care delivery operations.<br/>As a <b>Site Reliability Engineer (SRE)</b> at <a href="http://Kontakt.io" rel="noopener noreferrer">Kontakt.io</a>, you will be responsible for ensuring the <b>scalability, availability, and security</b> of our cloud-based <b>AI-driven healthcare platform</b>. You will collaborate with software, data, and infrastructure teams to build <b>highly resilient</b> and <b>automated systems</b>, allowing hospitals and care facilities to <b>operate seamlessly</b> and <b>without downtime</b>.Your expertise in <b>cloud infrastructure, automation, monitoring, and performance optimization</b> will directly impact how healthcare organizations leverage real-time data to enhance patient care and operational efficiency.<br/>If you are passionate about <b>highly available systems, automation, and making an impact in healthcare</b>, join <b>Kontakt.io</b> and help us <b>build the future of smart care operations</b>!
<h3>Key Responsibilities:</h3>
<ul>
<li>Design and maintain <b>highly available, fault-tolerant, and scalable</b> cloud infrastructure.</li><li>Implement <b>SLOs, SLIs, and SLAs</b> to track system reliability and optimize uptime.</li><li>Participate in <b>24/7 on-call rotation</b></li><li>Oversee production platform deployments</li><li>Monitor <b>latency, traffic, errors, and system health</b> using modern observability tools.</li><li>Conduct <b>root cause analysis (RCA) and post-mortems</b> to continuously improve system resilience.</li><li>Automate <b>infrastructure provisioning</b> using <b>Terraform, Ansible, or Pulumi</b>.</li><li>Implement <b>CI/CD pipelines</b> to ensure seamless and safe deployments.</li><li>Enable <b>self-healing mechanisms</b> using Kubernetes operators, auto-scaling, and fault detection.</li><li>Ensure compliance with <b>HIPAA, GDPR, and other healthcare data regulations</b>.</li><li>Define and execute <b>disaster recovery (DR) and business continuity plans</b>.</li><li>Manage and optimize <b>AWS </b>environments for cost-efficiency and performance.</li><li>Deploy and manage <b>observability tools </b>and build real-time <b>alerting and response</b> frameworks</li><li>Establish <b>best practices for logging, debugging, and performance monitoring</b>.</li><li>Improve <b>incident response automation</b> through runbooks, AI-based anomaly detection, and predictive analytics.</li></ul>
<h3>What You Bring</h3>
<ul>
<li><b>3+ years</b> of experience as an <b>SRE</b></li><li>Strong expertise in <b>Kubernetes, Docker, and container orchestration</b>.</li><li>Experience managing <b>cloud-native environments (AWS)</b>.</li><li>Experience with <b>event-driven architectures, Kafka, or real-time data streaming</b>.</li><li>Knowledge of <b>machine learning infrastructure.</b></li><li>Previous experience in <b>healthcare, compliance (HIPAA), and highly regulated environments</b>.</li><li>Proficiency in <b>Infrastructure as Code (IaC)</b> using Terraform.</li><li>Deep knowledge of <b>networking, DNS, load balancing, and security best practices</b>.</li><li>Experience with <b>CI/CD pipelines</b> (Jenkins, CI, or ArgoCD).</li><li>Hands-on experience with <b>monitoring and logging tools</b> (Prometheus, Grafana, ELK, OpenTelemetry).</li><li>Strong programming skills in <b>Python, Golang, or Bash</b> for automation.</li><li>Knowledge of <b>machine learning infrastructure.</b></li><li>Previous experience in <b>healthcare, compliance (HIPAA), and highly regulated environments</b>.</li></ul>
<h3>We offer:</h3>
<ul>
<li>Work on a <b>mission-driven platform</b> that improves healthcare operations and patient outcomes.</li><li>B2B contract or an employment agreement</li><li>Competitive salary and <b>stock option plan</b></li><li>Collaborate with <b>top engineers, data scientists, and AI experts</b>.</li><li>Flexible <b>remote or hybrid work</b> options (office in Krakow)</li><li>Collaborative and self-organized environment</li><li>private medical care, cafeteria system</li></ul><b>We Make Things Easy</b><u>Easy to Use.</u><span> Simplicity is harder than complexity. Each of our apps focuses on a single user and a specific problem. We create solutions for everyone to help them get things done.</span><u>Easy to Buy.</u><span> We simplify pricing with a single, per-be
Apply for this role
Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.
Apply Now →Generate Application KitFree account required — sign up in 30s