Jobs and Careers
NS
Senior Infrastructure Engineer (OpenStack Ironic Specialist)
NscaleEMEA; Germany; Netherlands; Norway; Poland; Spain; UK, Germanyfull_timeVerifiedPosted 24 Jul 2026
About the role
<h2>About Nscale</h2>
<p>Nscale is the GPU cloud engineered for AI. We provide cost-effective, high-performance infrastructure for AI start-ups and large enterprise customers. Nscale enables AI-focused companies to achieve superior results by reducing the complexity of AI development. Our GPU cloud bolsters technical capabilities and directly supports strategic business outcomes, including cost management, rapid innovation, and environmental responsibility.</p>
<p>We thrive on a culture of relentless innovation, ownership, and accountability, where every team member takes pride in their work and drives it with excellence and urgency. As an Nscaler, you’ll build trust through openness and transparency, where everyone is inspired to do their best work. If you join our team, you’ll be contributing to building the technology that powers the future.</p>
<h2>About the Role</h2>
<p>We’re hiring an <strong><strong>Infrastructure Engineer (OpenStack Ironic Specialist)</strong></strong> to design, operate, and continuously improve the bare metal provisioning platforms that underpin Nscale’s infrastructure.</p>
<p>This role sits within the <strong><strong>Infrastructure Engineering team</strong></strong>, which is responsible for the design, implementation, operation, and ongoing improvement of the infrastructure stack supporting both internal and customer-facing services. You’ll work closely with <strong><strong>network, compute, data centre, support, and pre-sales teams</strong></strong>, while also serving as a specialist escalation point for advanced provisioning and hardware issues.</p>
<p>This is a high-impact role focused on <strong><strong>OpenStack Ironic</strong></strong>, automated hardware lifecycle management, and the reliable operation of large-scale physical infrastructure. You’ll also help connect Nscale to the broader <strong><strong>upstream OpenStack community</strong></strong>, ensuring our bare metal platforms evolve in line with real operational needs and industry direction.</p>
<h2>What you'll be doing</h2>
<p><strong><strong>Bare Metal Provisioning & Lifecycle Management</strong></strong></p>
<ul>
<li value="1"><strong><strong>Design</strong></strong> scalable and resilient bare metal provisioning platforms with a strong focus on <strong><strong>OpenStack Ironic</strong></strong>.</li>
<li value="2"><strong><strong>Own</strong></strong> the full lifecycle of physical infrastructure, including discovery, enrolment, provisioning, cleaning, deprovisioning, and hardware state management.</li>
<li value="3"><strong><strong>Build</strong></strong> and maintain provisioning workflows for a wide range of hardware profiles, including <strong><strong>GPU-enabled</strong></strong> and high-performance server platforms.</li>
<li value="4"><strong><strong>Support</strong></strong> platform upgrades, lifecycle management, and operational improvements across <strong><strong>Ironic</strong></strong> and its dependencies.</li>
</ul>
<p><strong><strong>Automation & Platform Integration</strong></strong></p>
<ul>
<li value="1"><strong><strong>Manage</strong></strong> and improve integrations between <strong><strong>Ironic</strong></strong> and related OpenStack services such as <strong><strong>Nova, Neutron, Glance, Keystone, and Placement</strong></strong>.</li>
<li value="2"><strong><strong>Drive</strong></strong> automation for hardware onboarding, firmware and BIOS configuration, deployment workflows, validation, and recovery.</li>
<li value="3"><strong><strong>Implement</strong></strong> infrastructure automation using <strong><strong>infrastructure-as-code</strong></strong> and configuration management approaches.</li>
<li value="4"><strong><strong>Ensure</strong></strong> provisioning platforms and operational processes align with security, compliance, and operational standards.</li>
</ul>
<p><strong><strong>Troubleshooting, Reliability & Operational Support</strong></strong></p>
<ul>
<li value="1"><strong><strong>Troubleshoot</strong></strong> complex issues across provisioning pipelines, <strong><strong>PXE/iPXE</strong></strong>, BMC interfaces, out-of-band management, image deployment, network boot, and hardware compatibility.</li>
<li value="2"><strong><strong>Act</strong></strong> as a <strong><strong>3rd/4th line escalation point</strong></strong> for advanced bare metal and provisioning incidents.</li>
<li value="3"><strong><strong>Perform</strong></strong> root cause analysis and implement long-term fixes to improve platform reliability and repeatability.</li>
<li value="4"><strong><strong>Participate</strong></strong> in on-call rotations and incident response activities for critical infrastructure services.</li>
</ul>
<p><strong><strong>Cross-Functional Collaboration & Community Engagement</strong></strong></p>
<ul>
<li value="1"><strong><strong>Collaborate</strong></strong> with network, compute, data centre, and support teams to deliver reliable physical infrastructure services
Apply for this role
Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.
Apply Now →Generate Application KitFree account required — sign up in 30s