Jobs and Careers
TD

Staff Site Reliability Engineer (US)

TD
5900 North Andrews Avenue, United Statesfull_timeVerifiedPosted 22 Oct 2024
💰 $196,000/yr($113,000/yr$196,000/yr)

About the role

Work Location:

Fort Lauderdale, Florida, United States of America

Hours:

40

Pay Details:

$113,000 - $196,000 USD

TD is committed to providing fair and equitable compensation opportunities to all colleagues. Growth opportunities and skill development are defining features of the colleague experience at TD. Our compensation policies and practices have been designed to allow colleagues to progress through the salary range over time as they progress in their role. The base pay actually offered may vary based upon the candidate's skills and experience, job-related knowledge, geographic location, and other specific business and organizational needs. 

As a candidate, you are encouraged to ask compensation related questions and have an open dialogue with your recruiter who can provide you more specific details for this role.

Line of Business:

Technology Solutions

Job Description:

Site reliability engineering, or SRE, is the intersection of software engineering and systems operations. It's about designing, building, and maintaining the systems, services, and application. And it's more than just keeping the lights on, it's proactively identifying potential issues and implementing solutions to prevent them from happening in the first place.

At TD, SRE ensures that our services – both our internally critical and our externally-visible systems – have reliability and uptime to deliver our legendary customer experience.  

In this role, you'll be challenged to think ahead and find ways to prevent the unexpected. This is a thrilling and challenging field that requires a combination of technical skills, creativity and problem-solving. Be on the front line, ensuring that millions of customers can access the information and services they need, when they need them. This is an opportunity to make a real-world impact. Join us for this truly exciting journey.

Responsibilities

  • Identify optimal ways to improve the design and operation of systems to make them more scalable, more reliable, and more efficient. Candidate should be able to implement the required changes.
  • Work within product teams to support TD's business objectives and operational support goals providing domain expertise on strategic Infrastructure as well as Business project related activities.
  • Review technical deliverables throughout the design and development phase to ensure systems adhere to SRE best practices.
  • Lead the definition and implementation of service-level objectives (SLO) for key technical and business driven measures.
  • Influence and partner with key technology and product team members in the design and development of solutions that promote automation, innovation, and the reduction of toil.
  • Guide and educate the technology organization around SRE practices and the need for scalable and resilient production systems.
  • Write / contribute to existing documentation or educational content and adapt content based on product / program updates and user feedback.
  • Actively participate in, or lead design and code reviews with peers and stakeholders to examine system components for resiliency issues.

Depth & Scope:

  • Expert Site Reliability Engineering role with comprehensive expertise in leading-edge theories, engineering practices, extensive coding and scripting
  • Advanced and highly specialized knowledge of applications, systems, networks, innovation models, design activities, best practices, business / organization, Bank standards, and may fulfill a governance role
  • Engineering specialist assigned to work autonomously on high profile, complex and/or high-risk technology initiatives with significant impact to the organization
  • Provides technical leadership / consulting / direction to multiple businesses and product teams, growing capability across the organization
  • Resolves unique and complex problems that have a broad impact on the business
  • Authoritative expert on site reliability issues within area of specialization
  • Understands the journey of an enterprise transformation where there is a hybrid cloud/non-cloud operating model.
  • Drives end/end accountability of products and services across the enterprise through collaboration and transparency
  • Primarily works at the product umbrella, segment, LOB or Product Family level

Education & Experience:

  • University degree in Computer Science or related technical field involving systems engineering or equivalent practical experience.
  • 10+ years of engineering experience (e.g. Software or platform)

Apply for this role

Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.

Apply Now →Generate Application Kit

Free account required — sign up in 30s

Company

TD

View company profile →