Principal Engineer, AI Serving Framework Architect (Software)
Samsung SemiconductorAbout the role
<div class="content-intro"><p><span style="font-size: 12pt;"><strong>Please Note:</strong></span></p> <p>To provide the best candidate experience amidst our high application volumes, each candidate is limited to 10 applications across all open jobs within a 6-month period. </p> <p><span style="font-size: 12pt;"><strong>Advancing the World’s Technology Together</strong></span></p> <p>Our technology solutions power the tools you use every day--including smartphones, electric vehicles, hyperscale data centers, IoT devices, and so much more. Here, you’ll have an opportunity to be part of a global leader whose innovative designs are pushing the boundaries of what’s possible and powering the future. </p> <p>We believe innovation and growth are driven by an inclusive culture and a diverse workforce. We’re dedicated to empowering people to be their true selves. Together, we’re building a better tomorrow for our employees, customers, partners, and communities.</p></div><p><strong>Job Title: Principal engineer, AI Serving Framework Architect (Software)</strong></p> <p>The Architecture Research Lab (ARL) focuses on addressing fundamental system-level bottlenecks in modern AI, particularly in <strong>memory capacity/bandwidth and system-scale communication</strong>. By leveraging Samsung’s world-class memory technologies, ARL explores and defines next-generation AI system architectures that deliver step-function improvements in performance, efficiency, and scalability. <br>We are seeking a <strong>Principal AI System Architect</strong> who will play a key role in <strong>bridging AI workloads, system architecture, and hardware design</strong>. In this role, you will develop system-level performance models, drive architecture-level design decisions, and propose forward-looking AI system architectures that shape Samsung’s long-term AI platform strategy.</p> <p><strong>Location</strong>: Daily onsite presence at our San Jose office in alignment with our Flexible Work policy</p> <p><strong>Job ID</strong>: 42853</p> <p><strong>What You’ll Do </strong></p> <ul> <li>As a Tech Lead, leading research teams in Korea and proposing technical direction</li> <li>Research on dynamic scheduling methodologies for maximizing AI inference performance in multi-rack scale memory-centric systems, comprised of heterogeneous compute-capable memory and hierarchical memory</li> <li>Investigating methods to accelerate search operations in RAG’s vector DB and AI Agent’s knowledge-graph by leveraging compute-capable memory</li> <li>Studying strategies for optimally placing KVCach
Apply for this role
Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.
Apply Now →Generate Application KitFree account required — sign up in 30s