Public Cloud Observability Engineer - Vice President
CitiAbout the role
About the Opportunity
Are you a seasoned technologist with a passion for building cutting-edge enterprise products and a hands-on approach to engineering? Join Citi's Cloud Technology Services (CTS) team and be part of our commitment to transform Citi technology leveraging game-changing Cloud capabilities to drive agility, efficiency, and innovation.
We're providing our businesses with a competitive edge by leveraging public cloud scale and enabling new infrastructure economics. As the Cloud Engineer – Public Cloud Observability (Inventory Product Team) - VP you will play a pivotal role in shaping and executing our public cloud strategy.
You will be part of a team that continues to deliver big! From building a cloud based High Performance Compute (HPC) platform to run huge risk calculations, to enabling Citi to leverage GenAI at scale, all the way to enabling payments solutions, this team is at the forefront of innovation.
What You’ll Do
Technical Expertise: hands-on technical contribution within a product team that focused on the Public Cloud Foundation , supporting Citi's secure and enterprise-scale adoption of Public Cloud.
Collaborative Development: contribute to a team of cloud engineers and full-stack software developers, building and deploying solutions that advance the public cloud strategy.
Automation: Identify and develop automation initiatives to improve processes related to public cloud services consumption, enhancing client satisfaction and delivering business value.
Cross-Functional Partnership: collaborate with teams across Citi's technology landscape to ensure alignment between public cloud initiatives and broader business goals.
Engineering Excellence: contribute to defining and measuring success criteria for service availability and reliability within the specific product domain.
Compliance Advocacy: ensure adherence to relevant standards, policies, and regulations, prioritizing the protection of Citi's reputation, clients, and assets.
Who You Are
You are a talented cloud engineer with a proven track record of hands-on infrastructure automation development, a deep expertise in public cloud, and a passion for engineering best practices. You have:
-- Cloud Engineering Expertise: A deep understanding of public cloud services adoption at scale. Expert-level understanding of AWS/GCP Observability and Inventory across:
++ Google Cloud Observability – Cloud Asset Inventory, Cloud Monitoring, Cloud Functions
++ AWS Observability tools – AWS Config, CloudWatch, AWS Lambda
++ Experience with Python to automate API integrations and data workflows
-- Infrastructure as Code (IaC) Hands On Expertise: demonstrable experience with the following:
++ Programming Languages: Python and Go.
++ CI/CD: Terraform, Harness, Tekton, Jenkins, etc.
++ Testing Automation: Terratest, Cucumber, PytestBD, AWS Fault Injection Simulator (FIS), Chaos Mesh, etc
++ Experience deploying and operating infrastructure on at least one major cloud platform (AWS, GCP)
- Agile and DevOps Mindset: familiarity with Agile Development, DevOps, and SRE practices.
- Adaptability: demonstrated ability to quickly learn new technologies and adapt to changing project requirements
- Strategic Thinking: experience evaluating complex requirements and rationalizing them into a consistent service offering.
- Excellent communication: the ability to effectively communicate technical concepts to both technical and non-technical audiences.
- Exceptional teamwork and collaboration: with a proven ability to effectively contribute within a cross-functional team environment.
Qualifications:
6+ years of experience in a dedicated Observability, Monitoring, SRE, or DevOps role with a strong focus on building and managing cloud environments.
Proven expertise with at least one major cloud provider (AWS or GCP preferred).
Deep understanding of monitoring concepts, metrics collection, log aggregation, and distributed tracing.
Extensive experience with architecting and implementing observability platforms and tools (e.g., Prometheus, OpenTelemetry, Fluentbit, OpAMP).
Proficiency in scripting and automation (e.g., Python, Go).
Experience with Infrastructure as Code (IaC) tools like Terraform or CloudFormation.
Strong understanding of containerization technologies (Docker, Kubernetes) and their observability challenges.
Excellent problem-solving skills and the ability to diagnose complex technical issues across distributed systems.
Strong co
Apply for this role
Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.
Apply Now →Generate Application KitFree account required — sign up in 30s