Talent.com
Scotiabank
System Reliability EngineerScotiabank • Toronto, ON, CA
No longer accepting applications
System Reliability Engineer

System Reliability Engineer

Scotiabank • Toronto, ON, CA
9 days ago
Job type
  • Full-time
Job description

Requisition ID: 265177

Join a purpose driven winning team, committed to results, in an inclusive and high-performing culture.

The System Reliability Engineer will be working in a cross functional technology team responsible for the bank's Customer data and will contribute to the overall success of the team ensuring specific individual goals, plans, initiatives are executed / delivered in support of the team’s business strategies and objectives. Ensures all activities conducted are following governing regulations, internal policies, and procedures.

The incumbent should be a self-starter and be able to work independently with little or no supervision. They should have strong communication skills, a client focused mindset, and take accountability and ownership of tasks.

Is this role right for you? In this role, you will:

  • Contributes to a customer focused culture to deepen client relationships and leverage broader Bank relationships, systems, and knowledge.
  • Understand how the Bank’s risk appetite and risk culture should be considered in day-to-day activities and decisions.
  • Actively pursues effective and efficient operations of his/her respective areas in accordance with Scotiabank’s Values, its Code of Conduct and the Global Sales Principles, while ensuring the adequacy, adherence to and effectiveness of day-to-day business controls to meet obligations with respect to operational, compliance, AML/ATF/sanctions and conduct risk.
  • Drive reliability engineering strategy and platform standardization across teams.
  • Introduce chaos engineering practices to test system resilience.
  • Lead major incident management and stakeholder communication.
  • Mentor junior engineers and promote SRE best practices.
  • Ensure system reliability and uptime by designing, implementing, and maintaining highly available and fault-tolerant systems.
  • Monitor production systems using observability tools (e.g., Dynatrace, Grafana, Splunk) to proactively detect and resolve issues.
  • Develop and maintain automation for deployment, scaling, and incident response using scripts and Infrastructure as Code (IaC).
  • Manage incident response processes, including root cause analysis (RCA), postmortems, and continuous improvement of system resilience.
  • Improve observability and logging standards to enhance troubleshooting and system insights.
  • Experience coding in a professional environment, taking requirements from concept to production use.
  • Ability to work collaboratively and communicate clearly and concisely with both technical and non-technical audiences.
  • Troubleshoot production incidents, job failures, and provide support for production applications.
  • Analyze and resolve incident tickets assigned to the group based on severity and priority and identify the root cause for resolution.
  • Ensure incident and change management processes are executed as mandated, partnering with various internal teams.
  • Provide Release Management support including post-release health checks and monitoring of applications and ensure timely communications to upstream and downstream teams.
  • Develop, document and standardize plans and processes for preventive maintenance steps to ensure system stability and availability.
  • Provide after-hours support via an on-call pager on a rotational basis for production incidents, application releases during a maintenance windows and other maintenance activities.
  • Lead or support on-call rotations to maintain 24/7 service reliability and quick issue resolution.
  • Champions a high-performance environment and contributes to an inclusive work environment.

Do you have the skills that will enable you to succeed in this role? We'd love to work with you if you have:

  • 3+ years’ experience in any ETL platform like iWay, Informatica, Talend, DataStage etc.
  • 3+ years of experience in Application Support, Log Monitoring, Debugging and Incident Management process.
  • 3+ years of Unix Shell Scripting and prior experience with Java based applications.
  • Highly analytical and good understanding of databases, technical architecture, and experience in working with SQL.
  • Experience in Application Support, Log Monitoring, Debugging and Incident Management process
  • Undergraduate Degree in Computer Science, Computer engineering or Technical equivlant
  • Excellent communication skills (verbal/written/presentation).

Location(s): Canada : Ontario : Toronto

Scotiabank is a leading bank in the Americas. Guided by our purpose: "for every future", we help our customers, their families and their communities achieve success through a broad range of advice, products and services, including personal and commercial banking, wealth management and private banking, corporate and investment banking, and capital markets.

At Scotiabank, we value the unique skills and experiences each individual brings to the Bank, and are committed to creating and maintaining an inclusive and accessible environment for everyone. If you require accommodation (including, but not limited to, an accessible interview site, alternate format documents, ASL Interpreter, or Assistive Technology) during the recruitment and selection process, please let our Recruitment team know. If you require technical assistance, please click here. Candidates must apply directly online to be considered for this role. We thank all applicants for their interest in a career at Scotiabank; however, only those candidates who are selected for an interview will be contacted.

Create a job alert for this search

System Reliability Engineer • Toronto, ON, CA

Similar jobs

Impactful Site Reliability Engineer Fostering Reliability and Performance

RootlyToronto
Full-time

Join as an impactful Site Reliability Engineer, shaping the technical future and enhancing system reliability.Tackle rewarding challenges in a collaborative startup atmosphere.As a key player, you’... Show more

 • Promoted

Senior Cloud Reliability Engineer - $157,300 A Year

Autodesk, Inc.East York, Canada
Full-time

This role focuses on ensuring cloud application reliability and performance, designing infrastructure, and optimizing AWS workloads. Show more

 • Promoted

Site Reliability Engineer

DexianToronto, Ontario, Canada
Full-time

Working Location: Toronto, ON (Hybrid 2 days a week in office).The DevOps and Automation is looking for a Site Reliability Engineer with strong expertise in Dynatrace to ensure the reliability, per... Show more

 • Promoted

Lead Site Reliability Engineer At Imanage

iManageToronto, Canada
Full-time

Advance your career as a Lead Site Reliability Engineer at iManage, focused on maintaining and enhancing cloud resilience while enjoying flexible work arrangements.You will play a crucial role in d... Show more

 • Promoted

Remote Senior Site Reliability Engineer Role

ViafouraToronto, ON, CA
Remote
Full-time

Advance your career as a Senior Site Reliability Engineer at Viafoura, specializing in Kubernetes and AWS infrastructure.This remote role positions you to improve our platform's performance and sca... Show more

 • Promoted

Site Reliability Engineer

TELUS DigitalToronto, ON, CA
Full-time

Welcome to TELUS Digital — where innovation drives impact at a global scale.As an award-winning digital product consultancy and the digital division of TELUS, one of Canada’s largest telecommunicat... Show more

 • Promoted

Remote Site Reliability Engineer - Scale Crypto Systems

NewtonToronto, Ontario, Canada
Remote
Full-time

A leading innovative tech company in Toronto is looking for a Site Reliability Engineer.In this pivotal role, you will enhance the reliability and resilience of critical services, manage incidents,... Show more

 • Promoted

Hybrid Systems Engineer: Growth & Benefits

Manion, Wilkins & Associates Ltd.Toronto, ON, CA
Full-time

A Canadian benefits administration firm is looking for an Associate System Engineer to join their team in Toronto, Ontario, on a hybrid basis.The role focuses on providing technical support, managi... Show more

 • Promoted

Site Reliability Engineer

CapgeminiToronto, Ontario, Canada
Full-time

Talent Acquisition Business Partner – Strategic Business Unit at Capgemini America Inc.Choosing Capgemini means choosing a company where you will be empowered to shape your career in the way you’d ... Show more

 • Promoted

System Firmware Technical Engineer (1 yr contract)

AMDMarkham, ON, CA
Full-time

WHAT YOU DO AT AMD CHANGES EVERYTHING.At AMD, our mission is to build great products that accelerate next-generation computing experiences—from AI and data centers, to PCs, gaming and embedded syst... Show more

 • Promoted

Hybrid Hardware Reliability Engineer - Edge & Systems

MicromartToronto
Full-time

A leading technology company based in Toronto is seeking a Hardware Systems Reliability Engineer to ensure the reliability of hardware and software systems within their autonomous retail platform.T... Show more

 • Promoted

Senior Site Reliability Engineer

Morningstar Credit Ratings, LLCToronto, Canada
Full-time

About the Team Investment Services is Morningstar’s internal product group focused on building and maintaining the platforms that power our global data operations.We enable the Managed Investment D... Show more

 • Promoted

Site Reliability Engineer - C$102,700 - C$137,000 A Year

McCain FoodsEast York, Canada
Full-time

Seeking a Site Reliability Engineer to ensure software system reliability and availability by designing resilient architectures, automating infrastructure, and optimizing performance in Azure cloud. Show more

 • Promoted

System Engineer

Matchtech North AmericaToronto, ON, CA
Permanent

Matchtech is working with a key client in Toronto to support the recruitment of a number of Requirements Engineers.These roles are open to various grades/levels and offering competitive renumeratio... Show more

 • Promoted

IBM Site Reliability Engineering Expert

LeadingtalentMarkham, Ontario, Canada
Full-time

Step into a career as a Site Reliability Engineer at IBM, focused on enhancing system reliability and performance.Engage directly with production systems and optimize customer experience.In this ro... Show more

 • Promoted

Site Reliability Engineer, Inference Infrastructure

CohereToronto, Ontario, Canada
Full-time

Cohere is the leading security-first enterprise AI company.We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems.We’re training ... Show more

 • Promoted

Senior System Engineer - C$100,000 - C$130,000 A Year

Morson Talent (Canada & USA)East York, Canada
Full-time

Experienced Senior Systems Engineer sought to lead avionics product development, collaborating with cross-functional teams and ensuring on-time delivery.Role involves system design, requirements ma... Show more

 • Promoted

Senior Distributed Systems Engineer — Scale & Reliability - C$180,000 - C$230,000 A Year

A leading technology companyEast York, Canada
Full-time

Seeking a Senior Distributed Systems Engineer to work on scalable applications for customer post-purchase experience, requiring 7+ years of software engineering experience and cloud platform knowle... Show more

 • Promoted

Senior Site Reliability Engineer Ii - Remote, Scale-Focused - C$183,000 - C$203,000 A Year - Remote

Leading Grocery Delivery ServiceNorth York, Canada
Remote
Full-time

Seeking a Senior Site Reliability Engineer to ensure platform performance, establish incident management, and oversee scalable infrastructure strategies.Requires programming and incident management... Show more

 • Promoted

Senior Site Reliability Engineer- Remote

ClickHouseToronto, ON, CA
Remote
Full-time

Senior Site Reliability Engineer- Remote.Recognized on the 2025 Forbes Cloud 100 list, ClickHouse is one of the most innovative and fast-growing private cloud companies.With more than 3,000 custome... Show more