Jobs and Careers
EX

Middle/Senior Site Reliability Engineer

Exadel
Hungary, Poland, Hungaryfull_timeVerifiedPosted 5 Oct 2023

About the role

We're looking for a talented Middle/Senior Site Reliability Engineer who will be embedded within the product development team and manage those applications’ overall reliability and availability.

The customer is looking for a person who understands what SR Engineering is and how to apply it while working with the team, so they can help to onboard the team onto these practices. Taking ownership of these responsibilities and providing guidance for the team is also crucial. You should have a passion for troubleshooting, getting to the root cause of any identified issue, resolving it, and owning the lifecycle of that feedback within the application teams.

About the Customer:
The leading provider of vehicle lifecycle solutions, with headquarters in Chicago, enables the companies that build, insure, and replace vehicles to power the next generation of transportation. Its platform delivers advanced mobile, artificial intelligence, and car technologies. It connects a network of 350+ insurance companies, 24,000+ repair facilities, hundreds of parts suppliers, and dozens of third-party data and service providers. The customer's collective solutions enhance productivity and help clients deliver better experiences for end consumers.

About the Project:
The project is a highly unique and customizable application that calculates vehicle valuations for the insurance industry. This application was completely re-engineered in 2015 from Mainframes to a Java platform. The new application has been built using the latest design patterns and can be highly customized to fit insurance carriers' requirements without major application changes. The customer is currently adding several critical enhancements that are now possible given the new platform.

Project Tech Stack:
Java 8, Java 11
Spring MVC, Spring Boot
Docker & Kubernetes
Weblogic
Oracle DB
Query & AngularJS
Hibernate
PostGresDB
Vue.Js

Key Areas of Focus:

  • Reducing Technical Debt and Toil
  • Observability/System Monitoring
  • Incident Response throughout SDLC
  • Problem Management
  • Supporting the Product Development team

Project delivery process:
Daily meetings, weekly statuses, plannings, and reviews depending on the product team's needs, between 3 pm - 6/7 pm CETCentral European Time.

Requirements:

  • Past enterprise-level experience in DevOps, Software, Infrastructure, or Site Reliability Engineering with the ability to demonstrate an understanding of high-level technical briefs, talks, and ideas
  • Expertise intense leading teams in troubleshooting, issue resolution, or escalations
  • Ability to document solutions, SRE architectural patterns, and best practices to ensure that teams have guidance as needed
  • Proven ability to dig through metrics, logs, and available sources to triage and resolve an incident at any time
  • Proficiency in the full software delivery lifecycle
  • Understanding of Microservices and APIs
  • Experience and interest in working in an Agile environment
  • Versed in system management, monitoring, and analysis in order to identify opportunities to improve service health, manageability, and reliability improvements
  • Eager to problem solve and troubleshoot issues that may arise day to day
  • Сommunication and interpersonal skills

Nice To Have:

  • Experience functioning as an SRE in maintaining the reliability of the applications and infrastructure
  • Proficient in infrastructure as code practices
  • Knowledge of building CI/CD pipelines from scratch
  • Able to troubleshoot complicated, cross-platform issues by handling OS, Networking, Database, and applications in cloud-based environments.

Responsibilities:

  • Help build an SRE culture by sharing best practices, approaches, documentation, and code with other engineering teams across the organization
  • Document tribal knowledge as you acquire it over time by creating runbooks/playbooks and ensuring critical system information is readily available to those who need it through dashboards
  • Configuring and maintaining the monitoring tooling as it relates to the target application
  • Monitor application/infrastructure and take steps to improve overall system software performance, availability, and reliability by incorporating changes through defined feedback loops within the software delivery lifecycle
  • Apply automation to any tasks/parts of the system that are performed manually
  • CollaborateWork closely with software developers and testers to ensure the product is responding correctly to non-functional requirements such as security, performanc

Apply for this role

Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.

Apply Now →Generate Application Kit

Free account required — sign up in 30s

Company

Exadel

View company profile →