Senior DevOps Engineer
iTradeNetworkAbout the role
About iTradeNetwork
At iTradeNetwork, we provide advanced supply chain software and insights tailored to the food & beverage industry. Our mission is clear and ambitious: To feed the world. From the start, we’ve been dedicated to tackling the most pressing challenges within food and beverage supply chains, delivering innovative solutions and expert support that make a measurable impact.
Our cutting-edge technology helps businesses streamline complex procurement and fulfillment processes, minimize food waste, optimize inventory, manage compliance risk, and scale profitably. We’re proud to serve an elite customer base, including 13 of the top 25 North American grocers, 8 of the top 10 foodservice distributors, and 8 of the top 10 global food and beverage manufacturers.
JOB SUMMARY
We are seeking a highly skilled Sr. DevOps Engineer to join our operations and security engineering team and help design, build and operate reliable, scalable, and secure cloud and on-prem infrastructure. This role is DevOps-first, with a strong emphasis on automation, CI/CD and platform reliability, while working closely with Security (SecOps) to ensure security is embedded naturally into every stage of the SDLC and operations workflows.
You will partner with application engineering, security and operations teams to enable high-velocity software delivery, resilient systems, and well-architectured infrastructure across cloud and on-premise environments. You may also serve as a technical advisor on infrastructure, networking, and security design, helping teams make pragmatic, scalable decisions.
Key Responsibilities:
Core DevOps - Automation, CI/CD & Infrastructure as Code (Primary Focus)
- Build and optimize CI/CD pipelines using Bitbucket, GitHub Actions, GitLab CI, Jenkins, Cloud Build or cloud-native tooling.
- Automate cloud infrastructure provisioning using Terraform, Helm, Kubernetes Operators, or Crossplane.
- Develop automation for operational activities using Python, Go, or Bash, promoting reusable and modular code.
Cloud & Container Platform Engineering
- Design, secure, and operate Google Cloud Platform (GCP) environments (Compute Engine, GKE, Cloud SQL, Cloud Storage, VPC, IAM, Pub/Sub, Cloud Build, Cloud Run).
- Manage multi-environment Kubernetes clusters across dev, QA, staging, and production.
- Implement and maintain service mesh (Istio/Linkerd), API gateways, ingress controllers, and workload identity protections.
- Build internal platform tooling aligned with platform engineering best practices, self-service deployments, golden paths, reusable blueprints.
Security Integration & Governance
- Integrate security-as-code into CI/CD using tools such as Snyk, Trivy, GitLab/GitHub Advanced Security, Checkov, OPA, or Aqua.
- Implement supply chain security controls including SBOM generation, dependency scanning, secrets scanning, signed artifacts, and CI/CD hardening.
- Lead threat modeling, secure design reviews, automated code scanning, and periodic penetration testing.
- Manage and improve Zero Trust security controls, identity and access policies, secrets management (Vault, Google Secrets Manager), and workload identity.
Data Center Operations, Security & Compliance
- Administer and harden Linux and Windows server environments, covering OS provisioning, patching
- Performance tuning, certificate management, and vulnerability remediation.
- Implement and maintain backup and disaster recovery solutions, including backup policies, replication, restore testing, and RPO/RTO compliance.
- Deploy and enforce Zero Trust security architectures within the data center, including network segmentation, identity-based access controls, MFA integration, and secure application exposure.
- Manage email infrastructure and external vendors, handling technical escalations, SLA adherence, capacity planning, licensing, and security compliance.
Observability, Monitoring & Reliability
- Implement full-stack observability solutions using Prometheus, Grafana, Elastic, OpenTelemetry, Datadog.
- Improve logging, tracing, metrics, SLO/SLI dashboards, and alerting playbooks.
- Perform performance tuning, root-cause analysis, and increase system resilience via automation, chaos testing, and runbook improvements.
AI-Enabled DevOps
- Leverage AI/ML-assisted operations such as anomaly detection, predictive scaling, log intelligence, and automated remediation platforms.
- Integrate AI tools to improve CI/CD testing, code scanning, and vulnerability management workflows.
Advisory
Apply for this role
Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.
Apply Now →Generate Application KitFree account required — sign up in 30s