Jobs and Careers
ZE
Senior Site Reliability Engineer
ZenitechHungaryfull_timeVerifiedPosted 14 Aug 2025
About the role
<h2>The Role</h2><p dir="ltr">We are looking for a versatile <strong>Site Reliability Engineer (SRE) </strong>with <strong>strong software engineering skills in </strong><strong>Java and Javascript/Frontend</strong> to help us ensure the performance, scalability, and resilience of critical systems for our Client. As part of a multidisciplinary team, you’ll bridge the gap between development and operations, applying your programming expertise to automation, tooling, monitoring, and infrastructure reliability.<br/><br/></p><p dir="ltr">In this role, you'll not only maintain systems uptime but also<strong> write and review production-quality code</strong>, develop internal tooling, and contribute to front-end/backend reliability. Your contributions will directly impact services like the lottery website, mobile APIs, CMS, and internal developer experience.</p><p dir="ltr"><br/></p><h2>What you will do<br/></h2><p></p><ul>
<li><strong>Website:</strong> Ensure reliable performance of the public-facing website and its cloud infrastructure (AWS).</li>
<li><strong>Mobile Apps: </strong>Support the backend and monitoring needs for Android and iOS native apps.</li>
<li><strong>Backend-for-Frontend (BFF) APIs: </strong>Maintain and optimise APIs using Java and support API-level observability.</li>
<li><strong>Geo Location Services: </strong>Maintain the live geolocation service and its integration with the BFF layer.</li>
<li><strong>CMS (Magnolia): </strong>Support CMS services and automate related infrastructure and deployment tasks.</li>
<li><strong>Internal Tools: </strong>Build and enhance operational tools using Java or JavaScript (React) to improve team efficiency.</li>
<li><strong>Front-End Observability:</strong> Contribute to front-end reliability by instrumenting user-facing apps with monitoring (e.g., AppD RUM).</li>
<li><strong>Non-Production Environments: </strong>Manage and enhance dev/test/staging environments to ensure parity with production.</li>
<li><strong>On-Call Support</strong>: Participate in a rotating on-call schedule to respond to P1/P2 incidents and restore services swiftly.</li>
</ul><h2>Requirements</h2><p><strong>Software Engineering (Java + Frontend)</strong></p><ul>
<li>Strong programming experience in <strong>Java 21,</strong> including <strong>Spring Boot</strong>, RESTful APIs, and integration testing.</li>
<li>Experience developing or maintaining <strong>React </strong>applications (React 17+), including component libraries and API integration.</li>
<li>Competency in writing <strong>automation scripts</strong> and tools using <strong>Bash, Python, </strong>or <strong>Node.js.</strong></li>
<li>Understanding of frontend observability techniques, including <strong>RUM</strong>, browser-based metrics, and user experience monitoring.</li>
<li>Experience with testing frameworks (JUnit, Cypress, Jest) and code quality practices (linting, formatting, CI gates).</li>
</ul><p><strong>Cloud Architecture & DevOps</strong></p><ul>
<li>Proficiency with<strong> Infrastructure as Code (IaC)</strong>, particularly <strong>Terraform</strong>.</li>
<li>Hands-on with <strong>AWS services</strong> (ECS, Lambda, EC2, S3, CloudFront, CloudWatch).</li>
<li>Familiar with container orchestration tools such as <strong>Docker</strong>, <strong>ECS</strong>, or <strong>Kubernetes</strong>.</li>
<li>CI/CD pipeline experience using <strong>GitHub </strong><strong>Actions</strong>, <strong>Jenkins</strong>, or similar.</li>
</ul><p><strong>Monitoring, Observability, and Incident Management</strong></p><ul>
<li>Set up and manage monitoring dashboards using observability tools such as <strong>Prometheus, Grafana, Splunk Observability, AppDynamics, Otel, </strong>or <strong>New Relic.</strong></li>
<li>Ability to query logs using <strong>ELK </strong><strong>stack</strong>, <strong>Splunk</strong>, <strong>Logz.io</strong>, or <strong>Cloudwatch</strong>.</li>
<li>Lead incident response, perform root cause analysis, and collaborate on post-incident improvements.</li>
</ul><p><strong>Security, Compliance, and Automation</strong></p><ul>
<li>Apply <strong>DevSecOps</strong> practices across services and environments.</li>
<li>Automate repetitive tasks and operational toil to support team velocity.</li>
<li>Contribute to performance tuning of Java-based services and React apps.</li>
</ul><p><strong>Collaboration & Communication</strong></p><ul>
<li>Participate in code reviews across backend and frontend repos.</li>
<li>Collaborate with developers, QA engineers, DevOps, and product stakeholders in agile rituals.</li>
<li>Contribute to architectural discussions and planning with a reliability-first mindset.</li>
</ul><p><strong>Soft Skills:</strong></p><ul>
<li>Proactivity</li>
<li>Curiosity</li>
<li>Influencing skills to bring process improvements to Development teams</li>
</ul><p><strong>Nice to Have</strong></p><ul>
<li>Familiarity with <strong>Magnolia CMS </strong>or similar platforms.</li>
<li>Experience su
Apply for this role
Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.
Apply Now →Generate Application KitFree account required — sign up in 30s