Senior Site & Reliability Engineer
ExadelAbout the role
We are looking for an experienced Senior Site&Reliability Engineer to join our team of professionals!
Site Reliability Engineers create a bridge between development and operations by applying a software engineering mindset to system administration topics. They split their time between operations/on-call duties and developing systems and software that help increase site reliability and performance. SRE automates redundancy, and they automate manual tasks that can turn into programmatic tasks to keep the stack up and running. Site reliability engineers are able to oversee software and the performance of the entire technology stack.
Work at Exadel - Who We Are:
Since 1998, Exadel has been engineering its own software products and custom software for clients of all sizes. Headquartered in Walnut Creek, California, Exadel currently has 2700+ employees in development centers across America, Europe, and Asia. Our people drive Exadel’s success, and they are at the core of our values, so Exadel is a people-first cultured company.
About the Customer:
The customer is one of the largest international clothing-retail companies known for its fast-fashion clothing. The business concept is to offer fashion and quality at the best price in a sustainable way.
Project Team:
When you join our team, you'll be immersed in a culture where teammates always help each other achieve better results. We believe that together we are greater and can find brilliant solutions by sharing ideas.
Requirements:
- Strong coding skills using Python
- Good understanding of different operating systems, such as Windows/Linux. Ability to perform scripting using PowerShell/Bash
- Advanced knowledge in DevOps: working on CI/CD (with Jenkins), using Docker and Kubernetes
- Hands-on experience and knowledge of version control and monitoring tools (APM or General Monitoring)
- Deep understanding of various databases (Spark/Hadoop/Azure Data Factory).
- Understanding of Cloud Native applications relatable to Azure/GCP (Computing/Solutions)
- Ability to streamline Automated Incident process with SLO/SLI, understanding of the response systems
- Functional knowledge of Architecture and Design principles
- Software-centric, open-minded and curious mindset, strong problem-solving skills
Nice to have:
- Knowledge of Data Manipulation, ETL
- Familiarity with Agile Architecture Delivery
English level:
Intermediate+
Responsibilities:
- General systems uptimes
- Systems performance
- Latency
- Incident and outage management
- Systems and application monitoring
- Change management
- Capacity planning
Advantages of Working with Exadel:
- You can build your expertise with our Client Engagement team, who provide assistance with existing and potential projects
- You can join any Exadel Community or create your own to communicate with like-minded colleagues
- You can participate in continuing education as a mentor or speaker. You will not only be emotionally but also financially rewarded for mentoring
- You can take part in internal and external meetups as a speaker or listener. We support you in broadening your horizons and encourage knowledge sharing for all of our employees
- You can learn English with the support of native speakers
- You can take part in cultural, sporting, charity, and entertainment events
- Working at Exadel means always upgrading your skills and proficiency, so we provide plenty of opportunities for professional development. If you’re looking for a challenge that will lead you to the next level of your career, you’ve found the right place
- We work hard to ensure honest and open relations between employees and leadership so our offices are friendly environments
Apply for this role
Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.
Apply Now →Generate Application KitFree account required — sign up in 30s