Jobs and Careers
TA
Site Reliability Engineer - Observability Team - Suresnes on site
TalendFranceRemotefull_timeVerifiedPosted 11 Mar 2023
About the role
WHO WE ARE:
Talend, a leader in data integration and data governance, is changing the way the world makes decisions. In order to compete and win, IT and business leaders need data that they can trust and understand instantly. Talend Data Fabric is the only platform that seamlessly combines an extensive range of data integration and governance capabilities to actively manage the health of corporate information. This unified approach is unique and essential to delivering complete, clean, and uncompromised data in real-time to all employees. It has made it possible to create innovations like the Talend Trust Score™, an industry-first assessment that instantly quantifies the reliability of any data set. Over 7,250 customers have chosen Talend to run their businesses on healthy data. Talend is recognized as a leader in its field by leading analyst firms and industry media. We pride ourselves in our values of Passion, Agility, Team Spirit, and Integrity. Every one of our 1,400 employees brings a certain je ne sais quoi that makes Talend special.
Our SRE team of 40 engineers is divided into 4 squads: - Observability: Monitoring, alerting, tracing,- Runtime: Kubernetes stacks evolution, Ingress controller and performance,- Backends: Backend infrastructure, data persistence, - Security: Security, stability, and scalability.
Talend is looking for a Site Reliability Engineer - Observability to join our growing team in France (Nantes office or Remote), team of 40 members. In this role, you will join the SRE squad (7 members) where you’ll be responsible for the monitoring, alerting, tracing of our Talend Cloud service. You’ll act as key player to propagate the Observability practices among the other teams. You’ll get to work hands-on with plenty of exciting technology and scale challenges as we grow to support millions of transactions across hundreds of servers in our Talend Cloud environment. We are seeking candidates with expertise on both development and system administration.
Responsibilities· Ensure high reliability and availability of the Talend Cloud platform· Develop new infrastructure features based on Cloud Technology in a Public Cloud environment· Define and evangelize cloud-related optimizations and best practices to improve reliability, scalability and performance· Develop effective tooling, alerts, and response to both identify and address reliability risks· Work with fellow operations engineers and development teams on complex problems, and make decisions and recommendations about alerting/monitoring after analyzing possible courses of action· Be responsible for troubleshooting cloud infrastructure, systems, network, and application stacks· Manage and deploy internal and external monitoring solutions· Perform on-call duty as part of a team maintaining the availability and performance of our cloud infrastructure as well as the various internal services and systems that these core services depend on. Mandatory Skills· Bachelor's in computer science or a relevant field.· Experience with IaaC / configuration mgmt. / systems automation tools at scale (e.g., Terraform);· Practical experience in Cloud engineering / architecture (at least one from AWS or Azure) and microservices domain (Docker, Kubernetes...).· Practical experience with Agile methodologies Optional Skills· Two years or more of practical experience in managing Cloud infrastructure· One year working with the following observabilities technologies :o Prometheuso ElasticSearcho Grafanao Thanoso Tempo and Opentelemetry tooling· Knowledge and practical experience in one or more of the following technologies would be a good fit:o AKS, EKSo CI/CD (Github Action)o scripting / automation (bash, Python, GoLang)· One year working with GitOps deployment tooling (FluxCD, Sealed Secret)Willing to work in pair programming mode with team members
#LI-MM1 #LI-remote
AND NOW, A LITTLE ABOUT US:
Talend has received some pretty impressive accolades along the way:
- 7,250+ global customers rely on Talend for their data health- Named a Leader for Data Integration Tools by Gartner (for the 7th year in a row)- Named a Leader for Data Quality Solutions by Gartner (for the 5th year in a row)- Named a Leader in The Forrester Wave™: Enterprise Data Fabric (for the 2nd time in a row)- Top 100 best Leadership teams, compensation and benefits 2022 by Comparably.com- Ranked in the DBTA “100 Companies that Matter Most in Data” We are passionate about helping companies become more data driven; and, if we can be honest, we are all geeks at heart who pride ourselves on the vibrant company culture that we have built.
As a global employer, Talend believes our success depends on diversity, inclusion and mutual respect among our team members. We want to look like our customers, and we recr
Talend, a leader in data integration and data governance, is changing the way the world makes decisions. In order to compete and win, IT and business leaders need data that they can trust and understand instantly. Talend Data Fabric is the only platform that seamlessly combines an extensive range of data integration and governance capabilities to actively manage the health of corporate information. This unified approach is unique and essential to delivering complete, clean, and uncompromised data in real-time to all employees. It has made it possible to create innovations like the Talend Trust Score™, an industry-first assessment that instantly quantifies the reliability of any data set. Over 7,250 customers have chosen Talend to run their businesses on healthy data. Talend is recognized as a leader in its field by leading analyst firms and industry media. We pride ourselves in our values of Passion, Agility, Team Spirit, and Integrity. Every one of our 1,400 employees brings a certain je ne sais quoi that makes Talend special.
Our SRE team of 40 engineers is divided into 4 squads: - Observability: Monitoring, alerting, tracing,- Runtime: Kubernetes stacks evolution, Ingress controller and performance,- Backends: Backend infrastructure, data persistence, - Security: Security, stability, and scalability.
Talend is looking for a Site Reliability Engineer - Observability to join our growing team in France (Nantes office or Remote), team of 40 members. In this role, you will join the SRE squad (7 members) where you’ll be responsible for the monitoring, alerting, tracing of our Talend Cloud service. You’ll act as key player to propagate the Observability practices among the other teams. You’ll get to work hands-on with plenty of exciting technology and scale challenges as we grow to support millions of transactions across hundreds of servers in our Talend Cloud environment. We are seeking candidates with expertise on both development and system administration.
Responsibilities· Ensure high reliability and availability of the Talend Cloud platform· Develop new infrastructure features based on Cloud Technology in a Public Cloud environment· Define and evangelize cloud-related optimizations and best practices to improve reliability, scalability and performance· Develop effective tooling, alerts, and response to both identify and address reliability risks· Work with fellow operations engineers and development teams on complex problems, and make decisions and recommendations about alerting/monitoring after analyzing possible courses of action· Be responsible for troubleshooting cloud infrastructure, systems, network, and application stacks· Manage and deploy internal and external monitoring solutions· Perform on-call duty as part of a team maintaining the availability and performance of our cloud infrastructure as well as the various internal services and systems that these core services depend on. Mandatory Skills· Bachelor's in computer science or a relevant field.· Experience with IaaC / configuration mgmt. / systems automation tools at scale (e.g., Terraform);· Practical experience in Cloud engineering / architecture (at least one from AWS or Azure) and microservices domain (Docker, Kubernetes...).· Practical experience with Agile methodologies Optional Skills· Two years or more of practical experience in managing Cloud infrastructure· One year working with the following observabilities technologies :o Prometheuso ElasticSearcho Grafanao Thanoso Tempo and Opentelemetry tooling· Knowledge and practical experience in one or more of the following technologies would be a good fit:o AKS, EKSo CI/CD (Github Action)o scripting / automation (bash, Python, GoLang)· One year working with GitOps deployment tooling (FluxCD, Sealed Secret)Willing to work in pair programming mode with team members
#LI-MM1 #LI-remote
AND NOW, A LITTLE ABOUT US:
Talend has received some pretty impressive accolades along the way:
- 7,250+ global customers rely on Talend for their data health- Named a Leader for Data Integration Tools by Gartner (for the 7th year in a row)- Named a Leader for Data Quality Solutions by Gartner (for the 5th year in a row)- Named a Leader in The Forrester Wave™: Enterprise Data Fabric (for the 2nd time in a row)- Top 100 best Leadership teams, compensation and benefits 2022 by Comparably.com- Ranked in the DBTA “100 Companies that Matter Most in Data” We are passionate about helping companies become more data driven; and, if we can be honest, we are all geeks at heart who pride ourselves on the vibrant company culture that we have built.
As a global employer, Talend believes our success depends on diversity, inclusion and mutual respect among our team members. We want to look like our customers, and we recr
Apply for this role
Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.
Apply Now →Generate Application KitFree account required — sign up in 30s