Senior Hardware engineer (R&D / GPU / AI)
NebiusAbout the role
<div class="content-intro"><p><strong>About Nebius:</strong></p> <p>Nebius is leading a new era in cloud infrastructure for the global AI economy. We are building a full-stack AI cloud platform that supports developers and enterprises from data and model training through to production deployment, without the cost and complexity of building large in-house AI/ML infrastructure.</p> <p>Built by engineers, for engineers. From large-scale GPU orchestration to inference optimization, we own the hard problems across compute, storage, networking and applied AI.</p> <p>Listed on Nasdaq (NBIS) and headquartered in Amsterdam, we have a global footprint with R&D hubs across Europe, the UK, North America and Israel. Our team of 1,500+ includes hundreds of engineers with deep expertise across hardware, software and AI R&D.</p></div><p><strong>The role</strong></p> <p>Nebius is looking for a System Engineer (Servers Hardware R&D Team) to support our expanding North American operations. This position requires occasional on-site presence in our Data Center locations as needed.</p> <p><strong>Your responsibilities will include:</strong></p> <ul> <li>Participate in the design, deployment, and maintenance of high-performance cloud systems optimized for AI workloads.</li> <li>Arrange and perform hardware R&D tests and experiments on-site in data center environments.</li> <li>Troubleshoot and resolve complex system issues related to GPUs, networking (InfiniBand, NVLink), PCIe, and server infrastructure.</li> <li>Conduct deep investigations into hardware, software, and networking issues to ensure optimal system performance and reliability.</li> <li>Develop and execute test plans and methodologies for advanced GPU, InfiniBand, and compute systems to benchmark and validate performance.</li> <li>Collaborate closely with cross-functional engineering and operations teams to improve system performance and reliability.</li> <li>Monitor system performance and continuously fine-tune configurations for maximum efficiency.</li> </ul> <p><strong>What we expect you to have:</strong></p> <ul> <li>Strong knowledge of modern server architecture, particularly in high-performance, GPU-based environments.</li> <li>Hands-on experience with GPUs, networking, NVLink, and PCIe technologies.</li> <li>Proficiency in Linux systems, with experience using Python and Bash for automation and tooling.</li> <li>Demonstrated ability to troubleshoot complex hardware, software, and networking issues.</li> <li>Experience with deep problem investigation, root cause analysis, and performance optimization in cloud or high-performance computing environments.<
Apply for this role
Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.
Apply Now →Generate Application KitFree account required — sign up in 30s