Senior Platform & Infrastructure Developer - Observability
Arctic WolfAbout the role
Arctic Wolf, with its unicorn valuation, is the leader in security operations in an exciting and fast-growing industry—cybersecurity. We have won countless awards for our excellence in security operations and remain dedicated to providing an industry-leading customer and employee experience.
Our mission is simple: End Cyber Risk. We’re looking for a Senior Platform & Infrastructure Developer - Observability to be part of making this/that happen.
About the Role
We continue to expand our highly talented Infrastructure teams and are seeking a Senior Platform & Infrastructure Developer - Observability with production operations experience to join us. In this role, you’ll work with the Observability and Platform Engineering teams to design, develop, and maintain solutions to help other R&D teams monitor the behavior and performance of their workloads, reduce likelihood and impact of incidents, and better troubleshoot issues through logging, tracing and appropriate alerting.
Candidates with an operations (DevOps/SysOps/TechOps) background who have supported infrastructure at scale are also encouraged to apply for this role. If you are a firm believer in Infrastructure as Code, continuous deployment/delivery practices, and helping teams understand how their services behave in real-world scenarios, then you might be a great fit!
Technical Responsibilities
System design, configuration, integration, deployment, and operations of Observability systems and tools. These systems include collection of metrics/logs/events from many backend services deployed across multiple AWS accounts and regions and consumed by multiple teams.
Working with engineering teams to enable them to support their services from development to production.
Ensure our Observability platform exceeds goals for availability, capacity, efficiency, scalability, and performance as well as meeting our internal SLOs.
Build the next generation of observability integrating with Istio.
Write libraries and APIs that provide a simple, unified interface to other developers when they use our monitoring, logging, and event processing systems.
Enhance the existing alerting capabilities with Slack, Jira and PagerDuty
Helping build a continuous deployment system guided by metrics and data.
Bring anomaly detection into the observability stack.
Participate in 24x7 on-call rotation after at least 6 months of employment; maximum commitment one week per month.
What You Know
Strong with Python or Go
Cloud of choice, preference for AWS - Lambda, CloudWatch, IAM, EC2, ECS, S3
Solid understanding of Kubernetes
Prometheus, PromQL, Thanos, AlertManager, Grafana, etc
Strong knowledge of standard monitoring protocols/frameworks - Prometheus/Influx line format, SNMP, JMX, etc
Elastic stack, syslog, CloudWatch Logs
Comfortable working with git, Github, and common CI/CD approaches
IAC tooling like CloudFormation or Terraform
How You Do Things
Excited to use your expertise and be prescriptive about the right way forward.
Able to work well alongside SRE, platform and development teams.
Able to work independently and know when to reach out for support.
Passionate about automation - we do everything-as-code.
Other interesting things we’d cheer about:
Java
Some familiarity with open Observability initiatives (e.g., Open Tracing, Open Census, Open Metrics)
Knowledge of Kafka
Familiar with monitoring/observability in GCP and Azure
AWS Certifications
Comfortable with SQL
About Arctic Wolf
At Arctic Wolf we’re cultivating a collaborative and productive work environment that welcomes a diversity of backgrounds, cultures, and ideas to make our teams even stronger as we grow globally. We’ve been named one of the 50 Most Innovative Companies in the world for 2022 (Fast Company)—and the 2nd Most Innovative Security Company. This is in addition to consecutive awards from Top Workplace USA (2021, 2022), Best Places to Work - USA (2021, 2022) and Great Place to Work - Canada (2021, 2022).
Our Values
Arctic Wolf recognizes that success comes from delighting our customers, so we work together to ensure that happens every day. We believe in diversity and inclusion, and truly value the unique qualities and unique perspectives all employees bring to the organization. And we appreciate that—by protecting people’s and organizations’ sensitive data and seeking to end cyber risk— we get to work in an industry t
Apply for this role
Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.
Apply Now →Generate Application KitFree account required — sign up in 30s