Senior Infrastructure & Platform Engineer
Stoke SpaceAbout the role
At Stoke, we believe a thriving space economy will enable a vibrant, sustainable, and equitable future here on Earth. That is why we’re building Nova, our fully and rapidly reusable launch vehicle. Designed for daily flight, Nova tackles the core challenges of space transportation by reducing cost, increasing availability, and improving reliability. By radically lowering launch costs and increasing flight cadence, we’re helping create a truly scalable space industry.
Our team is mission-driven, collaborative, and empowered to take ownership of their work. If you want to work alongside some of the most dedicated and talented people on Earth, we’d love to have you join us.
Description
Reusable launch systems are the key to seamlessly connecting Earth and space. Just as our rocket systems are designed to be reliable, automated, and efficient, our infrastructure must embody these same principles to enable our engineering teams to move fast while maintaining the highest standards of security and compliance.
We are looking for a Senior Infrastructure & Platform Engineer to own and evolve the foundational infrastructure that powers Stoke’s engineering operations. You will be responsible for AWS GovCloud and commercial cloud architecture, Infrastructure as Code development, GitHub Enterprise Server operations, and the platform engineering systems that enable our teams to build rockets. This role requires deep technical expertise in AWS, networking, security compliance (ITAR/FedRAMP), and automation, combined with a passion for building reliable, self-service infrastructure that scales with our mission.
You will work closely with engineering teams across Stoke to understand their infrastructure needs, design and implement robust solutions using Pulumi and TypeScript, and build the tools and automation that make infrastructure operations seamless. This is a high-impact role where your work directly enables rocket development, test operations, and mission-critical systems.
You must be ready to stay focused, move quickly, self-direct, and learn on the fly.
Responsibilities
- Design, develop, and maintain Pulumi projects across multiple AWS accounts using TypeScript, implementing best practices for modularity, testing, and deployment automation
- Own the administration, scaling, and reliability of our self-hosted GitHub Enterprise Server instance and custom ephemeral runner system built on AWS Spot Fleet
- Design and implement AWS architectures across GovCloud and commercial regions, including VPC design, Transit Gateway networking, VPN connectivity, and cross-account access patterns
- Implement and maintain infrastructure controls for ITAR and FedRAMP compliance, including IAM policies, KMS encryption, CloudTrail audit logging, VPC security, and network segmentation
- Build self-service tools and automation for internal developers, including CI/CD integrations, developer portal infrastructure, and workflow automation systems
- Develop and maintain CI/CD pipelines, including 100+ GitHub Actions workflows; implement OIDC authentication for secure cloud deployments; optimize build and deployment pipelines
- Design and implement multi-region network architectures, including Transit Gateway peering, site-to-site VPNs, routing policies, NACLs, and security group strategies
- Operate container platforms across Docker, ECS/Fargate, and EKS, including image management and runtime security
- Implement comprehensive monitoring and alerting (CloudWatch, Datadog), perform cost analysis and optimization, and establish operational excellence practices
- Troubleshoot infrastructure issues across the stack, respond to security events, and implement post-incident improvements to prevent recurrence
- Produce clear technical documentation, runbooks, and architectural decision records; mentor team members on infrastructure best practices
Qualifications
- Bachelor’s or Master’s degree in Computer Science, Software Engineering, or a related technical field, or equivalent practical experience
- 5–8 years of experience in infrastructure engineering, platform engineering, DevOps, or site reliability engineering roles
- Proven track record of designing and implementing production AWS infrastructure at scale
- Experience working with security and compliance requirements (ITAR, FedRAMP, SOC 2, or similar frameworks)
- Strong proficiency in Infrastructure as Code using Pulumi (TypeScript preferred)
- Deep experience with AWS GovCloud and core services, including EC2, VPC, IAM, KMS, S3, Lambda, RDS, ECS/Fargate, CloudWatch, and CloudTrail
- Strong understanding of VPC design, subnets, routing tables, Transit Gateway, VPNs, security groups, NACLs, and network security principles
Apply for this role
Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.
Apply Now →Generate Application KitFree account required — sign up in 30s