Senior DevOps Engineer
Ingram Barge CompanyAbout the role
Company Description
Ingram Barge is a quality marine transporter on America’s inland waterways since 1946, starting out as a small, family-owned business and growing into what we are today: the largest dry cargo carrier and one of the top chemical carriers on the river. As the leading carrier, we operate a fleet of approximately 140 towboats and 5,000 barges. We're committed to being the best at whatever we do. We're continuously growing, adapting, and responding, and we're successful because of the outstanding hard work and creative energy of our associates.
Job Description
Ingram Barge is seeking a Senior DevOps Engineer to join our dynamic DevSecOps team. This person will work alongside our Systems Architect, Application Development Architect, and Security Engineer and focuses on operationalizing our cloud-native infrastructure, enhancing CI/CD pipelines, ensuring system reliability and resilience, and providing 24x7 operational support.
What you will be doing:
Pipeline & Automation
- Designing and implementing advanced CI/CD pipeline features using GitLab
- Developing and maintaining Terraform modules for infrastructure provisioning
- Creating and optimizing Ansible playbooks for configuration management and deployment automation
- Integrating security scanning and compliance checks into deployment pipelines
Container & Kubernetes Operations
- Building, configuring, and maintaining Azure Kubernetes Service (AKS) clusters
- Developing and optimizing Helm charts for application deployments
- Implementing and managing GitOps workflows
- Monitoring and troubleshooting containerized applications and cluster performance
Infrastructure & Reliability
- Implementing Infrastructure as Code best practices using Terraform and Ansible
- Designing and executing disaster recovery procedures and business continuity plans
- Performing system patching, upgrades, and maintenance activities
- Establishing and maintaining comprehensive monitoring, alerting, and observability solutions using Prometheus and Grafana
Cost Optimization & Resource Management
- Monitoring and analyzing Azure cloud spending patterns and resource utilization
- Implementing cost optimization strategies including right-sizing, reserved instances, and auto-scaling policies
- Developing dashboards and reports for cost tracking and forecasting
- Collaborating with teams to optimize resource allocation and eliminating waste
Monitoring & Observability
- Designing and implementing comprehensive monitoring solutions using Prometheus for metrics collection
- Building and maintaining Grafana dashboards for infrastructure, application, and business metrics
- Configuring intelligent alerting rules and escalation procedures
- Establishing SLIs, SLOs, and error budgets for critical services
24x7 Support & Incident Response
- Participating in on-call rotation for 24x7 production support
- Leading Tier 3 incident response efforts for production outages and system issues
- Performing root cause analysis and implementing preventive measures
- Collaborating with development teams on performance optimization and troubleshooting
- Maintaining runbooks and documentation for operational procedures
Qualifications
Knowledge, Skills, and Abilities:
Technical Expertise (5+ years)
- Strong experience with Kubernetes (AKS preferred) and container orchestration
- Proficiency in Infrastructure as Code: Terraform and Ansible
- Advanced GitLab CI/CD pipeline development and optimization
- Experience with GitOps methodologies and leading toolsets like Helm, Flux and/or ArgoCD
- Python scripting for automation and pipeline tasks
- Azure cloud services and networking concepts
Monitoring & Cost Management
- Hands-on experience with Prometheus for metrics collection and alerting
- Proficiency in Grafana for dashboard creation and data visualization
- Experience with Azure Cost Management tools and FinOps practices
- Knowledge of resource optimization techniques and auto-scaling strategies
- Understanding of cloud pricing models and cost allocation methods
DevOps & SRE Practices
- Incident management and post-mortem processes
- 2
Apply for this role
Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.
Apply Now →Generate Application KitFree account required — sign up in 30s