Staff Site Reliability Engineer-TS/SCI Clearance with Polygraph
Northrop GrummanAbout the role
Description
At Northrop Grumman, our employees have incredible opportunities to work on revolutionary systems that impact people's lives around the world today, and for generations to come. Our pioneering and inventive spirit has enabled us to be at the forefront of many technological advancements in our nation's history - from the first flight across the Atlantic Ocean, to stealth bombers, to landing on the moon. We look for people who have bold new ideas, courage and a pioneering spirit to join forces to invent the future, and have fun along the way. Our culture thrives on intellectual curiosity, cognitive diversity and bringing your whole self to work — and we have an insatiable drive to do what others think is impossible. Our employees are not only part of history, they're making history.Today’s dynamic global security threats require solutions both big and small – solutions living within the Northrop Grumman Microelectronics Center (NGMC). Boasting state-of-the-art design capabilities, multiple processing nodes, electrical testing, environmental and QCI screening, and failure analysis, the NGMC is a leader in designing, fabricating, packaging, and delivering discriminating microelectronics to the military, aerospace, and commercial markets. For more than 70 years, we have been offering a wide range of trusted foundry and semiconductor services that deliver high performing and reliable microelectronics. Our wide breadth of technologies and capabilities allows us to provide our customers with unique “More than Moore” solutions. Microelectronics | Northrop Grumman
One of our most challenging new fields is Microelectronics Design and Applications (MDA), which combines the unique properties of superconductivity and quantum mechanics to develop radical new energy-efficient computing systems. MDA is seeking a Staff Site Reliability Engineer with demonstrated ability to support systems that enable development of new technologies in support of our innovative MDA teams working on emerging supercomputing technologies.
What You'll Get To Do:
As a Staff Site Reliability Engineer, you will have an opportunity to be part of the DevOps Platform Engineering team. You will help improve the performance and reliability of solutions and services used by many software teams across the organization. This includes building and maintaining systems in our platform using automated methods such as IaC and CaC. This role will be hands-on keyboard in a highly collaborative environment with challenging problems to solve.
Roles and Responsibilities:
Support the MDA DevOps initiative with any program specific responsibilities allocated to the role to include:
Design and manage platform architecture in air-gapped environments
Deploy and maintain Kubernetes-based clusters to support containerized applications with scalability, reliability, and security (i.e. K8s, OpenShift)
Develop and manage infrastructure-as-code (IaC) solutions to automate and standardize platform tool deployments (e.g. Terraform, Ansible, Fluentd)
Build and optimize CI/CD pipeline templates in GitLab to streamline application deployment and test workflows across various environments
Deploy and maintain robust monitoring, alerting, and observability tools (e.g. Prometheus, Grafana, ELK) to enhance performance, reliability, and visibility
Automate incident management processes, including root cause analysis and self-healing mechanisms, to improve platform stability
Ensure compliance with security best practices throughout the platform and processes
Collaborate with development, Information Assurance, and IT teams to align infrastructure with platform requirements
Stay up to date on emerging technologies and trends to incorporate relevant advancements into the DevOps Platform
Applicants must have an active TS/SCI Clearance with Polygraph in order to be considered. Must be a US Citizen.
This position will serve 100% onsite in Linthicum / Annapolis Junction, MD.
Basic Qualifications for Staff Site Reliability Engineer:
Bachelor's degree in a STEM discipline with 12+ years of relative experience; Master's degree in a STEM discipline with 10+ years of relative experience; PhD and 7+ years of relative experience.
Proficient with Linux operating systems and command line
Apply for this role
Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.
Apply Now →Generate Application KitFree account required — sign up in 30s