Software Architect - Observability Platform
SalesforceAbout the role
To get the best candidate experience, please consider applying for a maximum of 3 roles within 12 months to ensure you are not duplicating efforts.
Job Category
Software EngineeringJob Details
About Salesforce
Salesforce is the #1 AI CRM, where humans with agents drive customer success together. Here, ambition meets action. Tech meets trust. And innovation isn’t a buzzword — it’s a way of life. The world of work as we know it is changing and we're looking for Trailblazers who are passionate about bettering business and the world through AI, driving innovation, and keeping Salesforce's core values at the heart of it all.
Ready to level-up your career at the company leading workforce transformation in the agentic era? You’re in the right place! Agentforce is the future of AI, and you are the future of Salesforce.
Salesforce and Google Cloud have embarked on a groundbreaking partnership worth $2.5 billion to revolutionize customer relationship management (CRM) through advanced artificial intelligence (AI). By integrating Google's Gemini AI models into Salesforce's Agentforce platform, we're enabling businesses to harness multi-modal AI capabilities—processing images, audio, and video—to deliver unparalleled customer experiences. Join our team of talented engineers and help us advance the integration of Salesforce applications on Google Cloud Platform (GCP). You will have the unique opportunity to work at the forefront of IDP, AI and cloud computing and contribute to enabling a full suite of Salesforce applications on Google Cloud. You will get an opportunity to build a platform on GCP to enable agentic solutions on Salesforce.
Our Public Cloud engineering teams are responsible for innovating and maintaining a large scale distributed systems engineering platform that ships hundreds of features to production for tens of millions of users across all industries every day. Our users count on our platform to be highly reliable, lightning fast, supremely secure, and to preserve all of their customizations and integrations every time we ship. You will need deep experience of delivering large scale systems, and experience of owning service at very high scale in efficient and automated way.
Responsibilities
Drive the vision of enabling a full suite of Salesforce applications on Google Cloud in collaboration with teams across geographies
Architecting and implementing large scale Observability platforms
Drive execution and delivery by collaborating with cross functional teams, architects, product owners and engineers
Be technically involved and work closely with product owner, other architects and engineering teams to build the best in class services and systems at scale
Present architecture vision to Executives, Architecture councils and to the community beyond Salesforce.
Provide technical guidance, career development, and mentoring to team members
Encourage ideas, have your point of view and lead with influence
Make critical decisions, innovate aggressively while managing risk and ensure the success of the product
Maintaining and fostering our culture by interviewing and hiring only the most qualified individuals with an eye towards diversity
Contributing to development tasks such as coding and feature verifications to assist teams with release commitments, to gain understanding of the deeply technical product as well as to keep your technical acumen sharp
Determine root-cause for all production level incidents and write corresponding high-quality RCA reports
Experience/Skills required
Masters / Bachelors degree required in Computer Science, Software Engineering, or Equivalent Experience
Experience creating solution architecture of large-scale Observability platforms
Experience architecting and designing systems for fault tolerance, scale-out approaches, and stability.
Experience with contributions to open source projects is desired.
Experience in one (and preferably more) of the following languages: Java, Python, or Go
Experience working with AIOps or AI SRE related tooling to automate Incident resolutions
Define service level objectives (SLOs) and service level indicators (SLIs) to represent and measure service quality
Should have experience using and running monitoring systems like Prometheus, Splunk, Grafana, Zipkin, OTEL or similar systems.
Identify areas to improve service resiliency through techniques such as chaos engineering, performance/load testing, etc
Working knowledge of Kubernetes, servic
Apply for this role
Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.
Apply Now →Generate Application KitFree account required — sign up in 30s