Skip to content

Site Reliability Engineer (SRE) - Scotiabank, Toronto

scotiabank jobsScotiabank·Toronto, ONfull time

Posted: July 6, 2026

Apply for this jobExpires: August 22, 2026

Browse more Scotiabank jobs · See other jobs in Toronto

job description

AI Summary

Scotiabank is seeking a Site Reliability Engineer (SRE) in Toronto, Ontario. This technical role focuses on implementing service improvements, automating repetitive tasks, and leading Disaster Recovery exercises. Key requirements include extensive experience in production support, real-time streaming data, Apache Kafka, and CI/CD pipelines. Candidates should possess strong technical skills in Java, SQL, cloud microservices, and scripting.

About the Site Reliability Engineer (SRE) Role

As a Site Reliability Engineer (SRE) at Scotiabank in Toronto, you will join a purpose-driven and inclusive team committed to delivering exceptional results. This technical role involves implementing, measuring, and gathering insights from Operational Level Indicators to identify areas for significant service improvements. Your work will directly impact the availability, performance, resilience, and incident management of critical banking systems. A core aspect of this position is the implementation of jobs as per Runbooks and the automation of repetitive tasks to reduce toil, enhancing overall operational efficiency. You will also lead and perform crucial Disaster Recovery (DR) exercises, engaging various teams to ensure thorough technical validation. Furthermore, participating in technical vulnerability assessments and providing recommendations for remediation and enhancements will be a key responsibility, contributing to the high level of security and reliability expected from Scotiabank's services. This role offers a challenging environment where complex problem-solving and continuous improvement are highly valued, requiring a proactive approach to anticipate and resolve issues before they impact service delivery.

Required Technical Skills and Experience

To succeed as a Site Reliability Engineer (SRE) with Scotiabank, candidates must demonstrate a robust set of technical skills and significant professional experience. We are looking for individuals with at least 7 years of hands-on technical working experience in production support and in-depth troubleshooting of major incidents and problem management. This includes a minimum of 3 years of hands-on working experience with real-time streaming data projects in operations. Specific technical experience includes 2 years with Apache Kafka for event management and 3 years using Splunk and/or Dynatrace for monitoring and alerts, crucial for maintaining a high level of service availability. Furthermore, 5 years of hands-on technical working experience with software build and CI/CD deployment pipelines, utilizing tools such as Jenkins, Gradle/Maven, or Bitbucket, is essential. Candidates must also be able to read Java code for troubleshooting and debugging, possess a strong technical understanding of RESTful Services, and be proficient in creating SQL Queries for relational databases. Technical knowledge of microservices in Cloud environments (GCP and/or Azure) is required, alongside the ability to write scripts using UNIX shell scripting languages and Python. Strong communication and organizational skills are also vital for this role, ensuring effective collaboration within our inclusive work environment. A post-secondary education in Computer Science, Engineering, or Mathematics is expected, providing the foundational technical level required for this challenging position.

categories

about scotiabank

similar jobs