Sr. Site Reliability Engineer
ForcepointAbout the role
Who is Forcepoint?
Forcepoint simplifies security for global businesses and governments. Forcepoint’s all-in-one, truly cloud-native platform makes it easy to adopt Zero Trust and prevent the theft or loss of sensitive data and intellectual property no matter where people are working. 20+ years in business. 2.7k employees. 150 countries. 11k+ customers. 300+ patents. If our mission excites you, you’re in the right place; we want you to bring your own energy to help us create a safer world. All we’re missing is you!
Forcepoint is seeking a Senior Site Reliability Engineer to join our Site Reliability Engineering Team. The SRE role will focus standardising key Site Reliability Engineering principles across Forcepoint products, and help maintain world-class reliability of our services for our customers.
The SRE role actively targets risk to service availability for customers by partnering with Engineering and Operations teams leveraging modern observability tooling and service restoration methodologies focused on automation and infrastructure as code where possible.
The ideal candidate will have a broad background spanning both applications and infrastructure. They will have direct experience with multiple coding language, core SRE practices & design methodologies.
Job Description:
Monitor, measure and improve the reliability, availability and scalability of Forcepoint products and infrastructure
Partner with Engineering to perform Operations Readiness of our products, ensuring that the products meet architecture & observability design requirements
Lead the New Product Introduction Process (NPI) for SRE organisation, ensuring SRE team is able to successfully provide operational support to the product
Embed across multiple Engineering teams, participate in early stage design discussions, and ensure high-availability and scalability criteria is considered in product design
Identify manual routine operational practices and build robust automation capabilities using code and modern tools
Collaborate with Product Developers and business stakeholders to gather requirements for enabling and improving performance monitoring for applications and services
Engage in Incident response and participate in post-mortem analysis to investigate root cause and capture contributing factors for remediation
Perform analytics on previous incidents and trend/usage patterns to better predict issues and take proactive actions
Design and build custom tools as needed to support process optimization, challenging the status-quo and improving operational efficiency
Participate in 24*7 rotational shifts & On-Call for handling production operation issues
Engage in service capacity planning and demand forecasting, software performance analysis and system tuning
Create meaningful dashboards/reports for application telemetry and infrastructure health for pro-actively identifying performance constraints and bottlenecks
Partner with Engineering to review product architecture, and recommend design enhancements to improve scalability and availability of our products
Requirements:
7+ years of experience with a strong understanding of cloud-based architecture and operations. Hands-on experience with Amazon Web Services is preferred.
Experience in administration/build/management of Linux systems
Foundational understanding of Infrastructure and Platform Technology stacks
Strong understanding of Networking concepts and theories, such as different protocols (TCP/IP, UDP, routing protocols, etc), VLAN configuration, DNS, OSI layers, and load balancing
Understanding of security architecture and certificate management
Working knowledge of Infrastructure and Application monitoring platforms such as Grafana Cloud, Xymon, LibreNMS etc.
Working knowledge of Incident Response and Alerting platforms such as PagerDuty, Opsgenie, XMatters
Apply for this role
Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.
Apply Now →Generate Application KitFree account required — sign up in 30s