Talent.com
Scotiabank
System Reliability EngineerScotiabank • Toronto, ON, Canada
System Reliability Engineer

System Reliability Engineer

Scotiabank • Toronto, ON, Canada
24 days ago
Job type
  • Full-time
Job description

Select how often (in days) to receive an alert:

Requisition ID: 265177

Join a purpose driven winning team, committed to results, in an inclusive and high-performing culture.

The System Reliability Engineer will be working in a cross functional technology team responsible for the bank's Customer data and will contribute to the overall success of the team ensuring specific individual goals, plans, initiatives are executed / delivered in support of the team’s business strategies and objectives. Ensures all activities conducted are following governing regulations, internal policies, and procedures.

The incumbent should be a self-starter and be able to work independently with little or no supervision. They should have strong communication skills, a client focused mindset, and take accountability and ownership of tasks.

Is this role right for you? In this role, you will:

  • Contributes to a customer focused culture to deepen client relationships and leverage broader Bank relationships, systems, and knowledge.
  • Understand how the Bank’s risk appetite and risk culture should be considered in day-to-day activities and decisions.
  • Actively pursues effective and efficient operations of his/her respective areas in accordance with Scotiabank’s Values, its Code of Conduct and the Global Sales Principles, while ensuring the adequacy, adherence to and effectiveness of day-to-day business controls to meet obligations with respect to operational, compliance, AML/ATF/sanctions and conductrisk.
  • Drive reliability engineering strategy and platform standardization across teams.
  • Introduce chaos engineering practices to test system resilience.
  • Lead major incident management and stakeholder communication.
  • Mentor junior engineers and promote SRE best practices.
  • Ensure system reliability and uptime by designing, implementing, and maintaining highly available and fault-tolerant systems.
  • Monitor production systems using observability tools (e.g., Dynatrace, Grafana, Splunk) to proactively detect and resolve issues.
  • Develop and maintain automation for deployment, scaling, and incident response using scripts and Infrastructure as Code (IaC).
  • Manage incident response processes, including root cause analysis (RCA), postmortems, and continuous improvement of system resilience.
  • Improve observability and logging standards to enhance troubleshooting and system insights.
  • Experience coding in a professional environment, taking requirements from concept to production use.
  • Ability to work collaboratively and communicate clearly and concisely with both technical and non-technical audiences.
  • Troubleshoot production incidents, job failures, and provide support for production applications.
  • Analyze and resolve incident tickets assigned to the group based on severity and priority and identify the root cause for resolution.
  • Ensure incident and change management processes are executed as mandated, partnering with various internal teams.
  • Provide Release Management support including post-release health checks and monitoring of applications and ensure timely communications to upstream and downstream teams.
  • Develop, document and standardize plans and processes for preventive maintenance steps to ensure system stability and availability.
  • Provide after-hours support via an on-call pager on a rotational basis for production incidents, application releases during a maintenance windows and other maintenance activities.
  • Lead or support on-call rotations to maintain 24/7 service reliability and quick issue resolution.
  • Champions a high-performance environment and contributes to an inclusive workenvironment.

Do you have the skills that will enable you to succeed in this role? We'd love to work with you if you have:

  • 3+ years’ experience in any ETL platform like iWay, Informatica, Talend, DataStage etc.
  • 3+ years of experience in Application Support, Log Monitoring, Debugging and Incident Management process.
  • 3+ years of Unix Shell Scripting and prior experience with Java based applications.
  • Highly analytical and good understanding of databases, technical architecture, and experience in working with SQL.
  • Experience in Application Support, Log Monitoring, Debugging and Incident Management process
  • Undergraduate Degree in Computer Science, Computer engineering or Technical equivlant

Location(s): Canada : Ontario : Toronto

Scotiabank is a leading bank in the Americas. Guided by our purpose: "for every future", we help our customers, their families and their communities achieve success through a broad range of advice, products and services, including personal and commercial banking, wealth management and private banking, corporate and investment banking, and capital markets.

At Scotiabank, we value the unique skills and experiences each individual brings to the Bank, and are committed to creating and maintaining an inclusive and accessible environment for everyone. If you require accommodation (including, but not limited to, an accessible interview site, alternate format documents, ASL Interpreter, or Assistive Technology) during the recruitment and selection process, please let our Recruitment team know. If you require technical assistance, please click here . Candidates must apply directly online to be considered for this role. We thank all applicants for their interest in a career at Scotiabank; however, only those candidates who are selected for an interview will be contacted.

#J-18808-Ljbffr
Create a job alert for this search

System Reliability Engineer • Toronto, ON, Canada

Similar jobs

WOLF Infrastructure Reliability Engineer Role

Wolf Advanced TechnologyAurora, York Region, CA
Full-time

Pursue a full-time Infrastructure Reliability Engineer position at WOLF in Aurora, ON, focusing on secure and scalable aerospace infrastructure.Join a dynamic environment that emphasizes collaborat... Show more

 • Promoted

Systems Engineer

Onico SolutionsRichmond Hill, York Region, CA
Permanent

The Systems Engineer is responsible for defining, designing, integrating, testing, documenting and deploying solutions for assigned projects and tasks.The position will implement and support servic... Show more

 • Promoted

Co-op Engineer- Distributed Agent System

Huawei CanadaMarkham, ON, CA
Full-time

Co‑op Opening for an Engineer (8‑16 months).Our team has an immediate 8‑16‑month opening for a Co‑op Engineer.Established in 2014, the Distributed Scheduling and Data Engine Lab is Huawei Cloud's t... Show more

 • Promoted

Remote Senior Site Reliability Engineer Role

ViafouraToronto, ON, CA
Remote
Full-time

Advance your career as a Senior Site Reliability Engineer at Viafoura, specializing in Kubernetes and AWS infrastructure.This remote role positions you to improve our platform's performance and sca... Show more

 • Promoted

Senior System Engineer

Nexus Systems GroupToronto, ON, CA
Temporary

We're engaging a Senior Systems Engineer for a 6-month contract in Toronto (hybrid).This role is focused on enterprise infrastructure modernization, including Linux migration, Nutanix infrastructur... Show more

 • Promoted

Senior Engineer, System-Level Design Verification

Tenstorrenttoronto, on, Canada
Permanent

Tenstorrent is leading the industry on cutting-edge AI technology, revolutionizing performance expectations, ease of use, and cost efficiency.With AI redefining the computing paradigm, solutions mu... Show more

 • Promoted

Site Reliability Engineer

TELUS DigitalToronto, ON, CA
Full-time

Welcome to TELUS Digital — where innovation drives impact at a global scale.As an award-winning digital product consultancy and the digital division of TELUS, one of Canada’s largest telecommunicat... Show more

 • Promoted

Lead Platform Engineer Enhancing DevOps and System Reliability

Lillio (formerly HiMama)Toronto, ON, CA
Full-time

Transform early childhood education as a Senior Platform Engineer focused on system performance and collaborative tooling.Drive key initiatives for scalable, reliable digital platforms.In this pivo... Show more

 • Promoted

IBM Site Reliability Engineering Expert

LeadingtalentMarkham
Full-time

Step into a career as a Site Reliability Engineer at IBM, focused on enhancing system reliability and performance.Engage directly with production systems and optimize customer experience.In this ro... Show more

 • Promoted

Hybrid Systems Engineer: Growth & Benefits

Manion, Wilkins & Associates Ltd.Toronto, ON, CA
Full-time

A Canadian benefits administration firm is looking for an Associate System Engineer to join their team in Toronto, Ontario, on a hybrid basis.The role focuses on providing technical support, managi... Show more

 • Promoted

System Engineer

Compunnel Inc.Toronto, ON, CA
Full-time

Get notified about new Information Technology Business Analyst jobs in.Information Technology Business Analyst Jobs in United States.IT Business Analyst - Tools Rationalization.IT Healthcare Consul... Show more

 • Promoted

System Firmware Technical Engineer (1 yr contract)

AMDMarkham, ON, CA
Full-time

What you do at AMD changes everything.At AMD, our mission is to build great products that accelerate next‑generation computing experiences—from AI and data centers, to PCs, gaming and embedded syst... Show more

 • Promoted

Co-op Engineer- Distributed Agent System

Huawei Technologies Canada Co., Ltd.Markham, ON, CA
Full-time

Our team has an immediate 8-16 month Co-op opening for an Engineer.Established in 2014, the Distributed Scheduling and Data Engine Lab is Huawei Cloud's technical innovation center in Canada.The la... Show more

 • Promoted

System Administrator

Amphenol Communications SolutionsMarkham, ON, CA
Full-time

Location: Markham, ON • Posted: 7/15/2026 • Location Name: HSIO-Markham • Wage: Depends on Experience.Amphenol Communications Solutions (ACS), a division of Amphenol Corporation, is a world leader ... Show more

 • Promoted

Hybrid Hardware Reliability Engineer - Edge & Systems

MicromartToronto
Full-time

A leading technology company based in Toronto is seeking a Hardware Systems Reliability Engineer to ensure the reliability of hardware and software systems within their autonomous retail platform.T... Show more

 • Promoted

System Reliability Engineer

ScotiabankToronto, Ontario, Canada
Full-time

Select how often (in days) to receive an alert:.Join a purpose driven winning team, committed to results, in an inclusive and high-performing culture.The System Reliability Engineer will be working... Show more

 • Promoted

System Engineer

Matchtech North AmericaToronto, ON, CA
Permanent

Matchtech is working with a key client in Toronto to support the recruitment of a number of Requirements Engineers.These roles are open to various grades/levels and offering competitive renumeratio... Show more

 • Promoted

Senior Site Reliability Engineer- Remote

ClickHouseToronto, ON, CA
Remote
Full-time

Senior Site Reliability Engineer- Remote.Recognized on the 2025 Forbes Cloud 100 list, ClickHouse is one of the most innovative and fast-growing private cloud companies.With more than 3,000 custome... Show more

 • Promoted

Senior Systems Engineer Distributed Systems

UMATRToronto
Full-time

Join a small team as a Senior Systems Engineer specializing in distributed systems and infrastructure in Toronto, Canada.Focus on kernel performance and build essential operational tooling.Candidat... Show more

 • Promoted

Site Reliability Engineer

Momentum Financial Services GroupToronto
Full-time

At Momentum Financial Services Group, we help people move forward by reimagining how money works for those who need it most.With more than 40 years of experience, we’re the team behind Money Mart—C... Show more