Talent.com

Reliability Jobs in Toronto, ON

Create a job alert for this search

Reliability • toronto on

Last updated: 21 hours ago

System Reliability Engineer

ScotiabankToronto, ON, CA
Full-time

Join a purpose driven winning team, committed to results, in an inclusive and high-performing culture.The System Reliability Engineer will be working in a cross functional technology team responsib... Show more

 • New!

Senior Staff Software Eng. Data Platform

0076 eBay CanadaToronto
Full-time

At eBay, we're more than a global ecommerce leader — we’re changing the way the world shops and sells.Our platform empowers millions of buyers and sellers in more than 190 markets around the world.... Show more

Reliability Engineer

Kinross Gold CorporationToronto, ON, CA
Full-time

Location: Downtown Toronto (outside Union Station – TTC & GO accessible).Founded in 1993, Kinross is a Canadian-based senior gold mining company with operations and projects in the United State... Show more

Site Reliability Engineer - SRE

Royal Bank of Canada>TORONTO, Canada
Full-time

We are seeking a Site Resiliency Engineer to be a member of Production Support team and focus on automation, development, implementation and support of Site Reliability Engineering (SRE) solutions.... Show more

Maintenance Repair Analyst

Seaboard Transport GroupNorth York, ON, CA
Full-time

The Maintenance Repair Analyst plays a key role in strengthening fleet reliability, safety and performance for heavy-duty trucks and trailers.By connecting field expertise with fleet leadership, th... Show more

Technical Co-founder (CTO) - AI Legal Front Desk

FutureSightToronto, ON, CA
Remote
Full-time
Quick Apply

Solo practitioners and small law firms still lose work the same way: missed calls, slow callbacks, and intake handled between hearings.Hiring is no answer — payroll, training, and turnover are exac... Show more

Manager, Network Reliability and Resiliency

ServiceNowToronto, Ontario, Canada
CA$125,700.00 yearly
Full-time +1

Due to Government of Canada regulatory requirements, this position requires the successful completion of a Government of Canada Reliability Status screening as a condition ... Show more

Junior RAMS/Systems Assurance Engineer

Egis GroupToronto, Ontario, Canada
Full-time

The RAMS Engineer supports the development and execution of the Reliability, Availability, Maintainability, and Safety (RAMS) program for the LRT project.This role is responsible for conducting RAM... Show more

Site Reliability Engineer

TOTEM Recruteur de talentToronto, ON, CA
Permanent

Schedule: 40 hours/week – 100% remote work.We are looking for an experienced.Working in an AWS and Kubernetes environment, you will help design, automate, monitor, and continuously improve the infr... Show more

Sr. Builder - Mobile (Sr. SDE), Ring

Amazon Development Centre Canada ULC - K03Toronto, Ontario, CAN
Full-time

Ring is redefining how millions of people interact with their homes every single day.As a Senior Builder on this feature team, you'll own and evolve some of the most foundational user experiences i... Show more

Microsoft Power Platform Engineer - Canada Only

Blue MantisToronto, Ontario, CA
Full-time
Quick Apply

You Must Be Located In Canada .The Power Platform Engineer is accountable for designing, implementing, and supporting Blue Mantis’ Power Platform solutions to enable scalable, secure, and enterpris... Show more

Mechanical Engineer – Onshore Reliability

Hudson ManpowerToronto, ON, CA
Full-time

Mechanical Engineer – Onshore Reliability.Bachelor’s Degree in Mechanical Engineering.Oil & Gas / Refinery (Onshore).The Mechanical Engineer – Onshore Reliability will be responsible for improv... Show more

Software Development Engineer, Personnel Resource Manager

Amazon Development Centre Canada ULCToronto, Ontario, CAN
Full-time

The Worblehat team is looking for a Software Development Engineer to join our mission-critical platform that automates operational compliance across Amazon's global fulfillment network.This unique ... Show more

Site Reliability Engineer- TDJP00058343

Randstad CanadaToronto, Ontario, CA
Full-time +2
Quick Apply

Our client, is seeking a talented and proactive Site Reliability Engineer (SRE) / Senior Database Platform Engineer to join their core Data Engineering and Operations team.In this engineering-focus... Show more

Site Reliability Engineer (SRE)

ScotiabankToronto, ON, CA
Full-time

You want to be challenged with complex problem solving taking the learnings forward as continuous improvements.You thrive on supporting critical systems requiring a high level of trust, resilience ... Show more

Executive Director, Reliability & Maintenance Services

ConfidentialToronto, Ontario, CA
Full-time

Executive Director, Reliability & Maintenance Services About the Company Well-regarded organization managing a local airport Industry Airlines/Aviation Type Non Profit Founded 1996 Employees ... Show more

 • Promoted

Reliability Expert - Fully Remote | Upto $120/hr

MercorToronto, Ontario, Canada
CA$80.00 hourly
Remote
Part-time
Quick Apply

Headquartered in San Francisco, our investors include.Incident management / reliability / SRE Evaluator.Evaluate AI-generated artifacts against domain-specific quality rubrics.Identify factual, aes... Show more

Sr. Machine Learning Software Verification Engineer

TalentlabToronto, Ontario, Canada
Full-time

AI Software Test / Validation Engineer.Technology / AI / Semiconductor.Our client is a global technology leader developing next-generation AI and machine learning solutions for on-device applicatio... Show more

Mid-Senior Mining Professionals

Hire Resolve.comToronto, ON, CA
Full-time
Quick Apply

Hire Resolve is assisting mining organizations in hiring experienced mining professionals across Canada.This is a multi-role opportunity spanning several functions within the sector, including mine... Show more

AI Infrastructure Engineer

Palona AIToronto, ON, CA
Full-time
Quick Apply

Palona’s AI agents operate continuously in production, handle real-time guest interactions, integrate with restaurant systems, and face sharp traffic peaks.Infrastructure is therefore part of the p... Show more

People also ask
System Reliability Engineer

System Reliability Engineer

ScotiabankToronto, ON, CA
21 hours ago
Job type
  • Full-time
Job description

Requisition ID: 265177

Join a purpose driven winning team, committed to results, in an inclusive and high-performing culture.

The System Reliability Engineer will be working in a cross functional technology team responsible for the bank's Customer data and will contribute to the overall success of the team ensuring specific individual goals, plans, initiatives are executed / delivered in support of the team’s business strategies and objectives. Ensures all activities conducted are following governing regulations, internal policies, and procedures.

The incumbent should be a self-starter and be able to work independently with little or no supervision. They should have strong communication skills, a client focused mindset, and take accountability and ownership of tasks.

Is this role right for you? In this role, you will:

  • Contributes to a customer focused culture to deepen client relationships and leverage broader Bank relationships, systems, and knowledge.
  • Understand how the Bank’s risk appetite and risk culture should be considered in day-to-day activities and decisions.
  • Actively pursues effective and efficient operations of his/her respective areas in accordance with Scotiabank’s Values, its Code of Conduct and the Global Sales Principles, while ensuring the adequacy, adherence to and effectiveness of day-to-day business controls to meet obligations with respect to operational, compliance, AML/ATF/sanctions and conduct risk.
  • Drive reliability engineering strategy and platform standardization across teams.
  • Introduce chaos engineering practices to test system resilience.
  • Lead major incident management and stakeholder communication.
  • Mentor junior engineers and promote SRE best practices.
  • Ensure system reliability and uptime by designing, implementing, and maintaining highly available and fault-tolerant systems.
  • Monitor production systems using observability tools (e.g., Dynatrace, Grafana, Splunk) to proactively detect and resolve issues.
  • Develop and maintain automation for deployment, scaling, and incident response using scripts and Infrastructure as Code (IaC).
  • Manage incident response processes, including root cause analysis (RCA), postmortems, and continuous improvement of system resilience.
  • Improve observability and logging standards to enhance troubleshooting and system insights.
  • Experience coding in a professional environment, taking requirements from concept to production use.
  • Ability to work collaboratively and communicate clearly and concisely with both technical and non-technical audiences.
  • Troubleshoot production incidents, job failures, and provide support for production applications.
  • Analyze and resolve incident tickets assigned to the group based on severity and priority and identify the root cause for resolution.
  • Ensure incident and change management processes are executed as mandated, partnering with various internal teams.
  • Provide Release Management support including post-release health checks and monitoring of applications and ensure timely communications to upstream and downstream teams.
  • Develop, document and standardize plans and processes for preventive maintenance steps to ensure system stability and availability.
  • Provide after-hours support via an on-call pager on a rotational basis for production incidents, application releases during a maintenance windows and other maintenance activities.
  • Lead or support on-call rotations to maintain 24/7 service reliability and quick issue resolution.
  • Champions a high-performance environment and contributes to an inclusive work environment.

Do you have the skills that will enable you to succeed in this role? We'd love to work with you if you have:

  • 3+ years’ experience in any ETL platform like iWay, Informatica, Talend, DataStage etc.
  • 3+ years of experience in Application Support, Log Monitoring, Debugging and Incident Management process.
  • 3+ years of Unix Shell Scripting and prior experience with Java based applications.
  • Highly analytical and good understanding of databases, technical architecture, and experience in working with SQL.
  • Experience in Application Support, Log Monitoring, Debugging and Incident Management process
  • Undergraduate Degree in Computer Science, Computer engineering or Technical equivlant
  • Excellent communication skills (verbal/written/presentation).

Location(s): Canada : Ontario : Toronto

Scotiabank is a leading bank in the Americas. Guided by our purpose: "for every future", we help our customers, their families and their communities achieve success through a broad range of advice, products and services, including personal and commercial banking, wealth management and private banking, corporate and investment banking, and capital markets.

At Scotiabank, we value the unique skills and experiences each individual brings to the Bank, and are committed to creating and maintaining an inclusive and accessible environment for everyone. If you require accommodation (including, but not limited to, an accessible interview site, alternate format documents, ASL Interpreter, or Assistive Technology) during the recruitment and selection process, please let our Recruitment team know. If you require technical assistance, please click here. Candidates must apply directly online to be considered for this role. We thank all applicants for their interest in a career at Scotiabank; however, only those candidates who are selected for an interview will be contacted.