Director, Software Engineering - Core Application Production Engineering
SalesforceAbout the role
To get the best candidate experience, please consider applying for a maximum of 3 roles within 12 months to ensure you are not duplicating efforts.
Job Category
Software EngineeringJob Details
About Salesforce
We’re Salesforce, the Customer Company, inspiring the future of business with AI+ Data +CRM. Leading with our core values, we help companies across every industry blaze new trails and connect with customers in a whole new way. And, we empower you to be a Trailblazer, too — driving your performance and career growth, charting new paths, and improving the state of the world. If you believe in business as the greatest platform for change and in companies doing well and doing good – you’ve come to the right place.
At Salesforce, we work on solving the hard, long term engineering problems that underlies all of our products. Focus is on building powerful and simple to use frameworks, services and software components that will be used by our products to support the exponential growth of our business while we deliver value to our end customers.
Our team owns mission critical services like the Core application server and async processing that supports the Salesforce platform across multiple substrates including AWS among others. These tier0 distributed services support billions of transactions at B2C scale and are resilient, highly available and scalable. Architected on CNA principles, our services are built to leverage OSS platforms like K8s, Service mesh and Spinnaker for continuous deployment.
Our goal is to innovate at scale and leverage AI for self configuration, self detection and self healing. As service owners, we track availability with observability, detect and forecast anomalies, use alert correlation for incident causation and prioritize automation with AIOps to reduce operational pain and toil along with innovation.
We are looking for passionate engineers to join our engineering team, who love to own modern large scale services in production and drive our charter forward. If you’re fired up about software performance, solving complex problems, automating everything, and working with great engineers, this is the job for you!
Your responsibilities include:
Responsibilities
- Manage a distributed team that is spread across many countries and time zones
- Mentor team members to help expand their skillset, perform to their full potential and grow in their careers
- Hire diverse talented engineers to expand the team
- Wear multiple hats as either a Scrum Master/Product Owner and an Engineering Manager and work with different stakeholders to align on objectives, priorities, tradeoffs and risks
- Participate in architectural discussions and design reviews
- Develop deeper insights into incidents across Salesforce services and influence the engineering backlog to address incidents proactively
- Foster a culture that promotes diversity and inclusion
- Complete service ownership, right from influencing product architecture to operating service seamlessly in production
- Be on-call for escalations to address complex problems in real-time and keep services operational and highly available
- Analyze and remediate production incidents for the Core Application Server and asynchronous processing platform
- Leverage AIOps platform to continuously improve anomaly detection, automate runbooks and drive our MTTD & MTTR goals
- Collaborate with Systems engineering teams for activities such as providing inputs for OS patching, JDK upgrade and software configuration
- Collaborate with technical writers to create, update and review documentation for users and operators
- Continuously raise standards of engineering excellence by implementing best DevOps practices
Required Skills
- Bachelors Degree in Computer Science or equivalent experience
- 5 years of work experience
- Knowledge of OO programming and concepts and experience coding in Java, C++ or Python
- Ability to debug complex distributed systems to understand system design with an eye for performance and scalability bottlenecks and provide recommendations to optimize code
- In-depth, hands-on experience with Linux, networking, server, and cloud architectures
- Exposure to container related technologies such as Kubernetes, Docker, etc.
- Proficiency with source control, continuous integration, and testing pipelines
Preferred Skills
- Overall 10+ years experience in a production engineering/performance engineering/DevOps/SRE or similar role working on high scale distributed systems with at least 2+ years experience as a Manager
- Experience analyzing heap dumps
- Experience instrumenting code and profiling applications
- Experience evaluating
Apply for this role
Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.
Apply Now →Generate Application KitFree account required — sign up in 30s