Senior Compute Hardware Architect, Enterprise Products
NVIDIAAbout the role
Join the NVIDIA Enterprise Products team as a Senior Compute Hardware Architect, where your passion and expertise in compute hardware, networking, storage, and cloud-native software will be pivotal. We are on the lookout for a multifaceted professional with a profound understanding of distributed systems, datacenter architecture, and large network design. As a key member of our team, you will engage in a collaborative, multi-disciplinary approach to craft scalable datacenter implementations for enterprise-grade AI systems.
Your role involves translating high-level goals into detailed specifications for platforms and datacenter architectures, leading to the development of robust implementations. Collaborating with other specialists in networking, software, and storage domains, you will be developing, validating, and profiling reference cluster designs specifically tailored for enterprise datacenter environments. At the core of our mission is the building and validation of on-prem cloud-ready solutions that seamlessly interoperate across various cloud service providers (CSPs), enabling the realization of hybrid enterprise AI solutions. If you are ready to embark on a journey that combines innovation, collaboration, and groundbreaking technology, then this opportunity is tailored for you.
What you’ll be doing:
Own the creation of scalable datacenter solutions for enterprise AI/ML systems
Craft detailed requirements for pioneering infrastructure patterns and datacenter wide architectures
Create and validate cluster designs, optimizing them for enterprise facilities
Collaborate closely with other experts in networking, software, and storage to drive innovation
Lead multi-disciplinary projects, addressing high-level goals and complex challenges
Engineer on-premises cloud-native solutions that flawlessly integrate with diverse cloud providers
Assume a pivotal role for the compute and hardware architecture domain, driving expertise and excellence
Showcase a multidisciplinary understanding of datacenter systems, networking, operating systems, I/O technology, and accelerators
Conduct TCO analysis, optimizing datacenter efficiency for cost-effectiveness
What we need to see:
Bachelor's degree or equivalent experience
12+ years hardware or infrastructure architecture experience
Cluster Design Proficiency: Expertise in architecting cluster designs for on-prem cloud-native platforms
Architectural Design Skills: Domain proficiency in computer hardware emphasizing scalability, portability, security and resilience
Platform Evaluation Expertise: Ability to evaluate diverse platforms, capturing differences and conducting research on their behavior under varied workloads
Cloud-Native Knowledge: Possess a deep understanding of cloud-native architecture concepts and practices, especially for high availability, scalability, resilience, performance, and security in the compute domain
Hardware Patterns Mastery: Understand and apply computer hardware patterns at a chassis, rack, and cluster level in designing reference architectures
Effective Communication: Talent in presenting technical concepts optimally through strong written and oral skills to both technical and non-technical audiences
Technical Leadership in Cluster Design: Leadership in designing clusters with a technical emphasis on compute and server setup, power infrastructure, effective cooling, thermal regulation, and cabling.
System-Level Thinking: Demonstrate a comprehensive grasp of system design, bus topologies, and accelerators to improve the overall quality of cluster reference designs.
Ways to stand out from the crowd:
Certifications in Leading Infrastructure Platforms: Hold relevant certifications such as Red Hat Certified Architect, IBM Certified Solution Architect – Cloud Computing Infrastructure v3, and VMware Certifications (Data Center Virtualization, Cloud Management and Automation, Network Virtualization)
Wider Infrastructure Expertise: Demonstrate wider experience beyond Compute in Storage, Networking, Infrastructure, Platform Sizing, and Infrastructure Cost Reduction for a broad approach to system architecture
TCO Analysis Skills: Exhibit skills in Total Cost of Ownership (TCO) analysis specifically focused on datacenter architectures to optimize resource utilization and efficiency
With competitive salaries and a generous benefits package, we are widely considered to be one of the technology world’s most desirable employers! NVIDIA is widely considered to be one of the technology world’s most desirable employers. We have some of the most
Apply for this role
Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.
Apply Now →Generate Application KitFree account required — sign up in 30s