About the role
<div class="content-intro"><p class="p1">Gradial helps marketers and creatives move from idea to execution faster. Our platform turns intent into action, automating website updates, design system migrations, and ongoing content optimization while preserving brand integrity across every touchpoint.</p> <p class="p1">Backed by leading investors, we’re building software that adapts to the user, not the other way around. We move with urgency, operate with ownership, and solve hard problems from first principles. If you want to do ambitious work, take real responsibility, and help define the future of AI-native content operations, you’ll do your best work here.</p></div><p><strong>The Role&nbsp;</strong></p> <p>As a <strong>Senior</strong> <strong>Infrastructure&nbsp;Engineer</strong> at Gradial, you will architect and evolve the systems that power our AI-driven content operations platform. This role is ideal for someone who thrives in startup-to-scale up environments and brings a deep understanding of how to make infrastructure reliable, secure, and scalable.</p> <p>You’ll play a critical role in building on our core systems, supporting rapid product iteration and ensuring the platform is built for growth. If you’ve owned infrastructure in production, guided system evolution, and want to shape the future AI, we’d love to meet you.</p> <p><strong>What You'll Own</strong></p> <ul> <li>Design and maintain scalable, secure, and resilient infrastructure to support Gradial’s AI platform.</li> <li>Lead Kubernetes cluster management, CI/CD pipelines, observability tooling, and infrastructure-as-code efforts.</li> <li>Anticipate scaling needs and proactively evolve infrastructure architecture to support growth and reliability.</li> <li>Take full ownership of real-time, compute-intensive services: designing, deploying and maintaining to meet high performance standards with minimal oversight.</li> <li>Establish and enforce best practices for system reliability, performance monitoring, and disaster recovery.</li> <li>Evaluate and implement infrastructure automation tools to improve deployment velocity and reduce operational burden.</li> <li>Act as a strategic voice on infrastructure investment, technical debt management, and long-term scalability planning.</li> </ul> <p><strong>What We're Looking For</strong></p> <ul> <li>5+ years of experience in DevOps, SRE or platform engineering roles.</li> <li>Proven track record designing and operating large-scale, production-grade infrastructure.</li> <li>Deep expertise in Kubernetes, cloud-native architecture, and container orchestration.</li> <li>Proficiency with infrastructure-as