Jobs and Careers
TR
Senior SRE (b2b, Remote)
TripleTenSpainRemotefull_timeVerifiedPosted 23 Jan 2025
About the role
<h3>Description</h3>
<p><a href="https://tripleten.com/business/" rel="noopener noreferrer" target="_blank">TripleTen for Business</a> empowers companies to achieve their business goals by bridging talent gaps in Data Science, AI for professionals, Python Development, and Management.</p><p>Our transformative approach includes tailored training programs, informed by comprehensive pre-training assessments, ensuring precise alignment with client needs. With expert-led content and personalized mentoring, we help employees excel and achieve new levels of proficiency.</p><p>We are looking for a <strong>Senior Site Reliability Engineer. </strong>In this role, you will take ownership of <strong>ensuring service high availability</strong>*, documenting infrastructure details, and empowering developers through training and guidance on working with it*</p>
<h3>What you will do</h3>
<ul><li>Develop infrastructure, and write solutions to simplify operations.</li><li>Build processes to achieve and maintain 99.99% uptime, and improve the exercise process.</li><li>Develop automation and service reliability, plan resources, and reduce ops in development.</li><li>Build infrastructure and monitoring, help developers solve infrastructure problems, train developers to solve problems independently, and improve the observability of infrastructure, monitoring, schedules, and alerts.</li></ul><p><br/></p>
<h3>Requirements</h3>
<ul><li>2+ years of Site Reliability Experience.</li><li><strong>Experience working with Prometheus - must have.</strong></li><li>Experience working with Kubernetes, GitLab CI, and Ansible.</li><li>Experience working with Unix systems (we have Ubuntu) and the console.</li><li>Understanding the basics of TCP/IP to build networks, how web services work, REST API, and gRPC.</li><li>Experience performing diagnostics, including interpreting the output of Ps, Top, Strace, Perf, and TCPDump.</li><li>Understanding of how user applications interact with the operating system, including familiarity with system calls, processes, and threads.</li><li>Willingness to build high-load systems and understanding of how to do that.</li><li>Understanding of fault tolerance and service scaling.</li><li>High degree of emotional intelligence, ability to find common ground with colleagues and work as part of a team.</li><li><strong>Must be professionally fluent in English </strong></li></ul><p><strong>Nice to have:</strong></p><ul><li>Experience working with AWS and Terraform.</li><li>Experience programming in Python / Golang or desire to learn how.</li></ul>
<h3>What we can offer you</h3>
<ul><li><strong>Full-time remote collaboration with a convenient schedule.</strong> Professional freedom, where we trust your experience instead of wasting each other's time and effort micromanaging;</li><li><strong>A diverse and tight-knit team.</strong> Our teammates are spread out across Serbia, the US, Israel, Georgia, Armenia, Latin America, and more. They’ve worked at all of big techs, ed-techs, design agencies, and cultural institutions;</li><li><strong>Comfortable digital workspace.</strong> We use Miro, Notion, Google Workspace, Jira, etc.— to make working together process seamless.</li></ul><p><br/></p>
Apply for this role
Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.
Apply Now →Generate Application KitFree account required — sign up in 30s