Senior Staff, Data Scientist
Samsung Semiconductor, Inc.About the role
Please Note:
To provide the best candidate experience with our high application volumes, we limit applications to a total of 10 over 6 months.
Advancing the World’s Technology Together
Our technology solutions power the tools you use every day--including smartphones, electric vehicles, hyperscale data centers, IoT devices, and so much more. Here, you’ll have an opportunity to be part of a global leader whose innovative designs are pushing the boundaries of what’s possible and powering the future.
We believe innovation and growth are driven by an inclusive culture and a diverse workforce. We’re dedicated to empowering people to be their true selves. Together, we’re building a better tomorrow for our employees, customers, partners, and communities.
The Customer Quality & Reliability (Q&R) team is accountable to identify any major product quality and reliability issues as early as possible and help resolve them as swiftly as possible. You will be part of an incubation team within this organization working on in-field telemetry intended to transform the Customer Quality Experience for Samsung memory products. Fault Management is the future of quality to minimize system downtime within AI/ML hardware deployments and workloads of the future. We analyze trends and patterns from enormous memory fleet telemetry to bucketize failures and perform virtual root-cause analysis. Telemetry analysis helps us design solutions to proactively avoid system downtime. We conduct research and develop both in-house and collaboratively in the industry with the opportunity to publish our findings through whitepapers and conferences. We are looking for innovative and passionate thinkers who can work in a start-up environment and are excited to shape the future of data centers around the world. Join us in our mission!
What You'll Do
- Partner with Sales/Product/Engineering teams to understand business requirements and gather data coming from different telemetry customers to consolidate into a single data pipeline.
- Develop forecasting and statistical machine learning models as per business use case.
- Consolidate codebases, database management, scripting, modeling and optimization, GUI design and deployment.
- Interface with customers to establish the value add of enabling in-field fault management and mitigating systems in order to improve field failure rate of memory subsystems.,
- Propose and develop platform fault management modules for memory subsystems.
- Propose and develop platform RAS (Reliability Availability Serviceability) algorithms for memory fault management.
- Contribute to define industry standards on memory fleet telemetry along with development of sophisticated predictive algorithms to manage hardware faults.
- Stay up-to-date on latest industry trends in the hardware fault management space. Read technical papers, blogs, conference talks as well as publish whitepapers in conferences.
- Drive alignment with multiple stakeholders/teams within US and Korea.
- Drive communication and interaction with internal and external customers regarding project goals and solution.
Location: Hybrid with at least 3 days in office in San Jose, CA office location remainder of time to work remotely
Job ID: 42321
What You Bring
- Bachelors with 15+ years of relevant industry experience, or Masters with 13+ years or PhD with 10+ years in hardware fault management, reliability, data center fleet management or related technical field preferred.
- Expertise in data mining, ML codebases, ML task development or automation, modeling, forecasting using statistics, analyzing trends and patterns, GUI design and deployment.
- Knowledge of platform memory subsystem, platform RAS (Reliability Availability Serviceability).
- Linux kernel commit experience.
- Familiarity with data center operating system and platform concepts (x86, ARM).
- Project management with the ability to write, edit, clarify and maintain consolidated status in real-time.
- Knowledge of platform memory subsystem from bare metal to OS level transactions.
- Innovative and creative, you proactively explore new ideas and adapt quickly to change.
- Excellent communication and interpersonal skills.
- Ability to work independently and as part of a team.
- You’re inclusive, adapting your style to the situation and diverse global norms of our people.
- An avid learner, you approach challenges with curiosity and resilience, seeking data to help build understanding.
- You’re collaborative, building relationships, humbly offering support and openly welcoming approaches.
- Innovative and creative,
Apply for this role
Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.
Apply Now →Generate Application KitFree account required — sign up in 30s
Similar roles
Senior Technical Project Manager
Fiserv
Senior Advisor - AF-PLM Customer Engagement (PEO-Level Liaison)
Sabel Systems Technology Solutions, LLC
Senior Clinical Specialist, Coronary - Queens/Brooklyn
Abbott
$133,300/yr