AWS Cloud Platform Support Engineer
Western DigitalAbout the role
Company Description
At Western Digital, our vision is to power global innovation and push the boundaries of technology to make what you thought was once impossible, possible.
At our core, Western Digital is a company of problem solvers. People achieve extraordinary things given the right technology. For decades, we’ve been doing just that. Our technology helped people put a man on the moon.
We are a key partner to some of the largest and highest growth organizations in the world. From energizing the most competitive gaming platforms, to enabling systems to make cities safer and cars smarter and more connected, to powering the data centers behind many of the world’s biggest companies and public cloud, Western Digital is fueling a brighter, smarter future.
Binge-watch any shows, use social media or shop online lately? You’ll find Western Digital supporting the storage infrastructure behind many of these platforms. And, that flash memory card that captures and preserves your most precious moments? That’s us, too.
We offer an expansive portfolio of technologies, storage devices and platforms for business and consumers alike. Our data-centric solutions are comprised of the Western Digital®, G-Technology™, SanDisk® and WD® brands.
Today’s exceptional challenges require your unique skills. It’s You & Western Digital. Together, we’re the next BIG thing in data.
Job Description
ESSENTIAL DUTIES AND RESPONSIBILITIES:
- Responsible for managing, and supporting cloud solutions in AWS such as S3, EMR, Presto, Aurora and Redshift.
- Responsible in maintaining overall health of Big Data Edge platform including troubleshooting storage, network, and necessary compute.
- Responsible in building healthy DevOps platform with policies and procedures around implementing and supporting our global hybrid cloud container strategies using Kubernetes deployed in Google Cloud Platform (GCP) and Amazon Web Services (AWS).
- Expertise in modern application concepts such as 12-factor application development using infrastructure as a code discipline.
- Experience in setting up continuous integration of source code pipelines using Bitbucket, Jenkins, Terraform, Ansible, etc., is preferred.
- Ability to build continuous deployments using Docker, Artifactory, Spinnaker, Argo CD, etc., is required with a strong advocate of DevOps principles.
- Passionate about managing and supporting modern software-as-a- service (SaaS) design principles using preferred cloud service provides or on-premises solutions.
This position requires partnering with various Western Digital manufacturing, engineering, and IT teams in understanding factory-critical workloads and designing solutions. Big data platform, BDP team provides self-service data and application platforms to enable machine learning (ML) capabilities to engineering and data science community. The ideal candidate should be passionate about working with various cloud native technologies to handle various Service Level Agreements (SLA). Candidate should be versatile to experiment with fail-fast approach to adopt to new technologies and natural troubleshooting capabilities. Communication with internal customers, external vendors and co-workers in a clear and professional manner is expected.
Qualifications
REQUIRED
- Bachelor of Engineering (or) Master of Engineering in Computer Science, Information Technology, Computer Information Systems (or) relevant working experience in IT field
- 8+ years of experience in handling enterprise level production applications, infrastructure for storage, memory, network, compute, and virtualization.
- Minimum 4+ years of architecting, managing, and supporting AWS EMR, EKS, Presto and Redshift cloud solutions for manufacturing environment.
- Hands-on Python and Unix shell scripting is required with a strong advocate of DevOps principles in setting up continuous integration of source code pipelines using Bitbucket, Jenkins, Terraform, Ansible and continuous deployment pipelines using Artifactory and Spinnaker.
SKILLS
- Strong troubleshooting skills in finding root cause of an incident with respect to compute, network, storage resource bottlenecks and proactively identifying corrective actions of existing production workloads is required.
- Experience in using monitoring/alerting observability tools such as Splunk, Prometheus & Grafana
- Candidate should have experience in handling on-call alerts using PagerDuty and responding to ServiceNow incidents.
- Candidate should have an appetite to learn new relevant technologies to support solutions.
Additional Information
Western Digital is committed to providing equal opportunities to all applicants and employees and will not discriminate based on the
Apply for this role
Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.
Apply Now →Generate Application KitFree account required — sign up in 30s