EP

Principal Solutions Architect (Req#1048)

ePlus Technology, inc.
San Ramon, USAfull_timePosted 10 Jul 2026

About the role

<hr> <p style="text-align: center;"><span style="font-family: arial, helvetica, sans-serif; font-size: 10pt;"><strong>Overview</strong></span></p> <hr> <p>We are seeking an elite Solutions Architect to lead the end-to-end design, sizing, and deployment of NVIDIA AI Factory-aligned infrastructure. In this highly technical, customer-facing role you will translate complex AI and machine learning workload requirements into fully engineered infrastructure solutions spanning colocation facilities, GPU compute, high-performance networking, parallel storage, and the complete NVIDIA AI software stack.</p> <p>You will serve as a trusted technical advisor to enterprise and hyperscale customers, partnering with sales, product, and engineering teams to win and deliver transformational AI infrastructure programs. Your expertise will directly shape how organizations build and operate production AI Factories capable of training frontier models, running large-scale inference fleets, and accelerating data science pipelines at scale.</p> <hr> <p style="text-align: center;"><span style="font-family: arial, helvetica, sans-serif; font-size: 10pt;"><strong>Your Impact</strong></span></p> <hr> <h2>Solution Design & Architecture</h2> <ul> <li>Lead discovery workshops to capture AI/ML workload requirements, including model training scale, inference SLAs, data pipeline throughput, and multi-tenancy needs.</li> <li>Architect full-stack AI Factory solutions aligned to NVIDIA reference architectures, integrating colocation, GPU compute, networking, storage, and software layers.</li> <li>Develop detailed Bills of Materials (BOMs), rack elevation diagrams, network topology drawings, and power/cooling budgets for customer proposals.</li> <li>Define GPU cluster architectures using NVIDIA DGX, HGX, and MGX systems with B200, B300, and GB300 Blackwell SXM and NVLink-Switch configurations.</li> <li>Design RTX PRO 6000 Blackwell Server Edition deployments for inference-optimized and enterprise AI workloads.</li> <li>Conduct workload sizing and TCO/ROI modeling to validate infrastructure dimensioning for training, finetuning, and inference at scale.</li> </ul> <h2>Colocation & Facility Planning</h2> <ul> <li>Specify colocation requirements including critical power load (MW-scale), UPS and generator configurations, and PUE targets.</li> <li>Design high-density GPU deployments utilizing air-cooled, direct liquid cooling (DLC), and rear-door heat exchanger configurations.</li> <li>Define meet-me room (MMR) and cross-connect requirements; specify carrier-neutral telecom diversity strategies.</li> <li>Engage colocation providers and data center operators to valid

Apply for this role

Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.

Generate Application Kit

Free account required — sign up in 30s

Company

ePlus Technology, inc.

View all open roles →