63415P-Software Engineer Staff
Juniper NetworksAbout the role
Software Engineer, Staff
Location: Sunnyvale, CA
We are seeking an experienced passionate engineering leader with AIML - DevOps Platform Engineering Functional Manager skills with over 12 years of experience to technically lead our DevOps team. This role is crucial in ensuring the reliability, scalability, and performance of our platform services. The ideal candidate will have a strong background in software development, systems architecture, and operational excellence, combined with exceptional leadership and strategic planning skills.
Responsibilities
- Manage the DevOps function for AIML platform engineering, including:
- Lead and mentor a team of DevOps engineers and AI/ML specialists, fostering a culture of collaboration, innovation, and continuous improvement.
- Oversee the recruitment, training, and professional development of team members.
- Set performance goals, conduct regular evaluations, and provide constructive feedback.
- Develop and execute a strategic roadmap for platform engineering and DevOps initiatives, particularly those involving AI/ML, aligned with business objectives.
- Collaborate with cross-functional teams, including software development, AI/ML research, QA, and IT operations, to align DevOps strategies with organizational goals.
- Manage budgets and forecast for AI/ML Ops platform tools and resources.
- Work closely with product management and stakeholders to understand requirements and deliver robust AI/ML solutions.
- Foster strong communication channels between development, operations, AI/ML teams, and other departments.
- Approving projects based on feasibility, resource constraints, and alignment with business goals
- Provide regular updates on project status, risks, and mitigation strategies to senior leadership.
- Oversee the work of other DevOps and platform engineering professionals by:
- Setting priorities for AIML Ops platform engineering tasks and projects
- Allocating resources for platform development, infrastructure management, and AI/ML deployments
- Establishing and maintaining best practices for AIML Ops platform engineering
- Design, implement, and maintain highly available and scalable platform infrastructure for AI/ML applications.
- Ensure the stability and performance of all platform services, with a focus on AI/ML workloads.
- Drive the adoption of infrastructure-as-code practices and tools.
- Identify areas for process improvement and implement solutions to enhance efficiency and reduce downtime.
- Drive the automation of routine tasks to minimize manual intervention and reduce operational overhead.
- Collaborate with data scientists and AI/ML engineers to integrate AI/ML models into production environments.
- Ensure robust monitoring, logging, and alerting for AI/ML model performance and reliability.
- Drive the automation of AI/ML model deployment and retraining processes.
Required Skills and Qualifications:
- Bachelor’s degree in Computer Science, Information Technology, or a related field; Master’s degree preferred.
- 12+ years of experience in software development, systems engineering, or DevOps, with at least 2 years in a leading and managing teams.
- Proven experience managing large-scale, complex platform infrastructure and DevOps teams, with a balanced focus on AI/ML and software development.
- Strong expertise in cloud platforms (AWS, Azure, GCP), containerization (Docker, Kubernetes), and automation tools (Terraform, Ansible, Jenkins).
- Deep understanding of CI/CD pipelines, monitoring, logging, and alerting frameworks.
- Extensive experience with AI/ML model deployment, monitoring, and optimization.
- Proficiency in one or more programming languages (e.g., Go, Python, Shell).
- Excellent problem-solving skills and a track record of implementing innovative solutions.
- Ability to influence and motivate others without direct supervisory authority
- Exceptional communication, leadership, and interpersonal skills.
Preferred Qualifications:
- Relevant certifications in cloud platforms, DevOps tools, or AI/ML technologies.
- Experience with AI/ML frameworks (TensorFlow, PyTorch, Scikit-learn) and data engineering tools.
- Familiarity with microservices architecture and serverless computing.
- Strong understanding of security best practices and compliance standards in AI/ML and software development environments.
Minimum Salary: $159,200.00
Maximum Salary:$228,850.00
The pay range for this position is expected to be between $159,200.00 and $228,850.00/year; however, the base pay offered may vary depending on multiple individualized
Apply for this role
Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.
Apply Now →Generate Application KitFree account required — sign up in 30s