About our group:
As a part of the Global Infrastructure Services team, our goal is to provide reliable compute, storage and network services to all Seagate factories and design center operations. Our motto is to architect and transform our compute/storage/network systems with standard consistent architecture globally to achieve stability, service agility, simplified support, and ease of management by any Administrators across the globe. Seagate is the leader in enterprise and cloud storage solutions, 40% of the world’s data is stored on Seagate devices.
About the role - you will:
Seagate's IT Globals Infrstructure Services (GIS) organization is looking for a Linux DevOps Engineer to optimize and continuously improve our business processes and projects. The Engineer will assist with the IT Infrastructure team with the Kubernetes cluster operator role to automate the support activities, and analyze and provide key insights to optimize the microservices environment for efficiency and better ROI.
Duties will include but not limited to:
- Architect, deploy, and manage highly available Kubernetes clusters on Rancher or other Kubernetes management platforms.
- Design and implementation of CI/CD pipelines on GitLab platform, ensuring automated build, test, and deployment processes.
- Understand the Kubernetes cluster operator role, collaborate closely with development teams to containerize applications and optimize their performance for Kubernetes orchestration.
- Implement comprehensive monitoring, logging, and alerting solutions for Kubernetes clusters using tools such as Prometheus, Grafana, Zabbix, and OpenSearch.
- Ensure the compute and GPU resources are effectively utilized, security and compliance of Kubernetes clusters and services through proper configuration and adherence to best practices and policies
- Inventory the Kubernetes microservices, analyze its distribution and usage, trend its weekly/monthly/yearly growth from compute/storage perspective, and find means to optimize this environment for efficiency.
- Stay abreast of emerging technologies and industry trends in cloud computing, DevOps practices, containers, and Kubernetes
- Write and maintain scripts and tools (mostly Python & Bash based scripting) to streamline processes and ensure operational efficiency.
About you:
- At least 5+ years in a Reliability Engineering, DevOps or IT infrastructure focused role.
- Candidate with lesser years of experience or education may be considered for a junior grade.
- Automate infrastructure provisioning and management using tools such as Terraform, Ansible, or similar technologies.
- Thorough understanding of CI/CD concepts and experience with CI/CD tools GitLab, Harbor, Jenkins.
- Good understanding on Automation and configuration management tools such as Ansible, Terraform, PowerShell.
- Familiarity with cloud platforms such as GCP and Azure.
- Proficiency in a high-level language like Python and Shell Scripting.
- Produce documentation for tasks completed, according to established standards.
- Demonstrate good judgment in selecting methods and techniques for obtaining solutions.
- Thorough knowledge of TCP/IP, Linux operating system, file systems, disk/storage technologies and storage protocols.
- Professional and positive communication skills.
- Ability to proactively identify and resolve issues.
- Must be comfortable with ambiguity and be a resourceful, savvy professional that is committed to finding solutions and making an impact.
- Certification in Kubernetes Administration will be a plus.
Your experience includes: