Jobs and Careers
LU
Senior Site Reliability Engineer
Lumin DigitalRemote- United States, United StatesRemotefull_timeVerifiedPosted 28 Aug 2023
About the role
Our Site Reliability Engineers (SRE) are good developers with an operations mindset. They enjoy reducing or completely eliminating manual tasks, are excellent problem solvers, and know automation is the key to operating a large-scale system.
SREs make sure that our application is highly available and Service Level Objectives (SLO) are met. SREs work closely with our Software Engineers (SWE) using their interest in operations and development skills to ensure new features follow SRE best practices and are supportable.
ESSENTIAL FUNCTIONS:-CI/CD.Monitor and resolve issues in all environments. Ensure SLO and uptime are met.-Ensure SRE concerns are addressed from the time a feature is designed through its deployment to production.-Work on the SRE scrum team-Engage in capacity planning and demand forecasting, anticipating performance bottlenecks and scaling the environment as needed.-Change management.-Uptime and SLO reporting.
KNOWLEDGE, SKILLS & ABILITIES: -Cultural fit. -Humility. -Strong sense of ownership, customer service, and integrity. -Willing to walk in the mud.-Commitment to continually improving yourself.-Operational expertise with a desire to eliminate manual tasks:DevOps approach - Automation and resilient systems are key.Monitoring and Alerting - Monitor the right things. Alert appropriately:Self heal.Involve people when needed.Log tickets when no immediate action is required.-Remain calm in trying circumstances.-Exceptional full stack and environment troubleshooting skills.-Expert-level knowledge of at least one configuration management system (Chef, Ansible, Puppet, etc.).-Understanding of standard networking protocols and components such as: HTTP, DNS, TCP/IP, ICMP, the OSI Model, Subnetting and Load Balancing.-Security mindset. -Data cannot and will not be compromised.-Driven to ensure that being on-call is boring.-Exceptional written and verbal communication skills.-Past history working on an agile scrum team.-Expert hosting in the Cloud. -AWS preferred, but Google Cloud and Azure are also of interest.-Experience with a microservice architecture running in containers (Docker or other containerization technology).-Experience with Terraform and Kubernetes-Understand CI / CD and ability to architect the workflow.-Willing to participate in a 24x7 on-call rotation.
DESIRED SKILLS:-2+ years of experience as a software engineer. C#, Angular, JavaScript preferred.-AWS Certification preferred but not essential: SysOps and/or Solutions Architect ideal.-Experience with Amazon RDS, EKS, CloudWatch, etc.-Experience with Docker tooling and ecosystem.
Education:Bachelor’s degree or higher in Computer Science, or equivalent experience.
SREs make sure that our application is highly available and Service Level Objectives (SLO) are met. SREs work closely with our Software Engineers (SWE) using their interest in operations and development skills to ensure new features follow SRE best practices and are supportable.
ESSENTIAL FUNCTIONS:-CI/CD.Monitor and resolve issues in all environments. Ensure SLO and uptime are met.-Ensure SRE concerns are addressed from the time a feature is designed through its deployment to production.-Work on the SRE scrum team-Engage in capacity planning and demand forecasting, anticipating performance bottlenecks and scaling the environment as needed.-Change management.-Uptime and SLO reporting.
KNOWLEDGE, SKILLS & ABILITIES: -Cultural fit. -Humility. -Strong sense of ownership, customer service, and integrity. -Willing to walk in the mud.-Commitment to continually improving yourself.-Operational expertise with a desire to eliminate manual tasks:DevOps approach - Automation and resilient systems are key.Monitoring and Alerting - Monitor the right things. Alert appropriately:Self heal.Involve people when needed.Log tickets when no immediate action is required.-Remain calm in trying circumstances.-Exceptional full stack and environment troubleshooting skills.-Expert-level knowledge of at least one configuration management system (Chef, Ansible, Puppet, etc.).-Understanding of standard networking protocols and components such as: HTTP, DNS, TCP/IP, ICMP, the OSI Model, Subnetting and Load Balancing.-Security mindset. -Data cannot and will not be compromised.-Driven to ensure that being on-call is boring.-Exceptional written and verbal communication skills.-Past history working on an agile scrum team.-Expert hosting in the Cloud. -AWS preferred, but Google Cloud and Azure are also of interest.-Experience with a microservice architecture running in containers (Docker or other containerization technology).-Experience with Terraform and Kubernetes-Understand CI / CD and ability to architect the workflow.-Willing to participate in a 24x7 on-call rotation.
DESIRED SKILLS:-2+ years of experience as a software engineer. C#, Angular, JavaScript preferred.-AWS Certification preferred but not essential: SysOps and/or Solutions Architect ideal.-Experience with Amazon RDS, EKS, CloudWatch, etc.-Experience with Docker tooling and ecosystem.
Education:Bachelor’s degree or higher in Computer Science, or equivalent experience.
Apply for this role
Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.
Apply Now →Generate Application KitFree account required — sign up in 30s