Talent.com
Royal Bank of Canada>
Lead Site Reliability EngineerRoyal Bank of Canada> • TORONTO, Canada
Lead Site Reliability Engineer

Lead Site Reliability Engineer

Royal Bank of Canada> • TORONTO, Canada
24 days ago
Job type
  • Full-time
Job description

Job Description

We are looking to expand our Digital team at RBC. If you are looking for an exciting, high-growth opportunity with a leading financial institution that is accelerating cloud native area this could be the job for you. Are you looking for a chance to make a difference? Are you someone who embraces change?
We are looking for an individual who exemplifies the attributes of a leader, mentor and decision-maker. We are currently building out our SRE team, with the goal being to provide expertise and tooling to manage the health, security, and availability of their applications in production. We work with other teams to provide guidance throughout the lifecycle of building, deploying, and operating the application.

What Will You Do?

  • Run the production environment by monitoring availability and taking a holistic view of system health

  • Build tools to manage platform infrastructure and applications

  • Debug production issues across services and levels of the stack and provide primary operational support and engineering for multiple large distributed software applications

  • Help adopt and drive the tool creation for application health monitoring and alerting.

  • Improve reliability, quality, and time-to-market of our suite of software solutions

  • Measure and optimize system performance, with an eye toward pushing our capabilities forward, getting ahead of application team needs, and innovating to continually improve

  • Gather and analyze metrics from both operating systems and applications to assist in performance tuning and fault finding.

  • Participate in system design consulting, platform management, and capacity planning.

  • Create sustainable systems and services through automation and uplifts.

  • Balance feature development speed and reliability with well-defined service level objectives.


What Do You Need To Succeed?


Must have:

  • Overall, 4-6 years of support experience in Openshift, Azure & Kubernetes.

  • 4-5 years of experience as an SRE supporting multiple applications and very strong programming skills in Java/SpringBoot/Python.

  • Oracle and SQL database operational experience in the cloud/on-premise and writing/understanding database queries (SQL and/or No-SQL) and Object Oriented design and development

  • Exposure to OCP, and GitHub is desirable

  • Having a good overall understanding of networking-related areas like certificates, load balancers etc.

  • Monitoring using Splunk, Dynatrace, RUM, Grafana & other related tools

  • Experience with the operational aspects of software systems such as monitoring, centralized logging, and alerting.

  • Experience in micro-services, public cloud (Azure preferred) & container technologies and working knowledge of Mainframes & JCL is nice to have


Nice-to-have:

  • Knowledge of public cloud (Microsoft Azure and AWS) and private cloud (OpenShift) platforms and development of applications in multi-cloud, hybrid environments

  • Knowledge of containers and orchestration (e.g: Docker, Kubernetes)

What’s in it for you?


We thrive on the challenge to be our best, progressive thinking to keep growing, and working together to deliver trusted advice to help our clients thrive and communities prosper. We care about each other, reaching our potential, making a difference to our communities, and achieving success that is mutual.
A comprehensive Total Rewards Program including bonuses and flexible benefits, competitive compensation, commissions, and stock where applicable
Leaders who support your development through coaching and managing opportunities
Ability to make a difference and lasting impact
Work in a dynamic, collaborative, progressive, and high-performing team
A world-class training program in financial services
Flexible work/life balance options
Opportunities to do challenging work

#LI

Job Skills

Agile Methodology, Group Problem Solving, IT Systems Integration, Organizational Leadership, Product Services, Software Development Life Cycle (SDLC), System Applications, System Integration Testing (SIT), Systems Software

Additional Job Details

Address:

RBC WATERPARK PLACE, 88 QUEENS QUAY W:TORONTO

City:

Toronto

Country:

Canada

Work hours/week:

37.5

Employment Type:

Full time

Platform:

TECHNOLOGY AND OPERATIONS

Job Type:

Regular

Pay Type:

Salaried

Posted Date:

2026-09-10

Application Deadline:

2026-10-10

Note: Applications will be accepted until 11:59 PM on the day prior to the application deadline date above

Our Employment Opportunities

At RBC, we are guided by living shared values of Client First, Integrity, Collaboration, Respect and Excellence and winning together as One RBC. We believe an inclusive workplace that has diverse perspectives is core to our continued growth as one of the largest and most successful banks in the world. Maintaining a workplace where our employees feel supported to perform at their best, effectively collaborate, drive innovation, and grow professionally helps to bring our Purpose to life and create value for our clients and communities. RBC strives to deliver this through policies and programs intended to foster a workplace based on respect, belonging and opportunity for all.

Join our Talent Community

Stay in-the-know about great career opportunities at RBC. Sign up and get customized info on our latest jobs, career tips and Recruitment events that matter to you.

Expand your limits and create a new future together at RBC. Find out how we use our passion and drive to enhance the well-being of our clients and communities at jobs.rbc.com.

RBC is presently inviting candidates to apply for this existing vacancy. Applying to this posting allows you to express your interest in this current career opportunity at RBC. Qualified applicants may be contacted to review their resume in more detail.

Create a job alert for this search

Lead Site Reliability Engineer • TORONTO, Canada

Similar jobs

Senior Staff Site Reliability Engineer

CerebrasToronto, ON, CA
Full-time

Become a Staff Site Reliability Engineer at Cerebras Systems, revolutionizing AI inference service reliability.Design innovative solutions for operational challenges.This position is crucial for en... Show more

 • Promoted

Site Reliability Engineer

DexianToronto, Ontario, Canada
Full-time

Working Location: Toronto, ON (Hybrid 2 days a week in office).The DevOps and Automation is looking for a Site Reliability Engineer with strong expertise in Dynatrace to ensure the reliability, per... Show more

 • Promoted

Senior Site Reliability Engineer - C$120,000 - C$154,000 A Year

Loblaw Companies LimitedEast York, Canada
Full-time

Seeking a Senior Site Reliability Engineer in Brampton, ON to lead engineering standards, manage production environments, and improve cloud-native solution reliability and performance. Show more

 • Promoted

Senior Site Reliability Engineer

ThinkificToronto, ON, CA
Full-time

Senior Site Reliability Engineer.Senior Site Reliability Engineer.Are you an experienced Site Reliability Engineer looking for a new challenge?.Senior Site Reliability Engineer.Senior Site Reliabil... Show more

 • Promoted

Remote Senior Site Reliability Engineer Role

ViafouraToronto, ON, CA
Remote
Full-time

Advance your career as a Senior Site Reliability Engineer at Viafoura, specializing in Kubernetes and AWS infrastructure.This remote role positions you to improve our platform's performance and sca... Show more

 • Promoted

Senior Site Reliability Engineer - C$130,000 - C$180,000 A Year

Acquird.ioEast York, Canada
Full-time

Senior SRE needed to maintain cloud infrastructure, CI/CD pipelines, and ensure system uptime and reliability.Requires experience with Azure and cloud engineering. Show more

 • Promoted

Staff Site Reliability Engineer - C$124,200 - C$166,700 A Year

Walt Disney Animation StudiosToronto County, Canada
Full-time

Seeking a Staff Site Reliability Engineer with expertise in Linux, software development (Python, Go, Java, Node), CI/CD tools, Git, cloud platforms (AWS, GCP, Azure), and container technologies (Do... Show more

 • Promoted

Site Reliability Engineer - C$110,000 - C$130,000 A Year

Compass DigitalEast York, Canada
Full-time

Join Compass Digital as a Site Reliability Engineer to design, build, and automate cloud-native systems using AWS, Go, and TypeScript in a hybrid work environment across Canada. Show more

 • Promoted

Site Reliability Engineer, Observability - C$110,000 - C$130,000 A Year

PricelineToronto, Canada
Full-time

Role Overview This role is eligible for our hybrid work model: Two days in-office.As a Site Reliability Engineer – Observability, you will play a key part in maturing our observability capabilities... Show more

 • Promoted

Senior Site Reliability Engineer

Morningstar Credit Ratings, LLCToronto, ON, CA
Full-time

Investment Services is Morningstar’s internal product group focused on building and maintaining the platforms that power our global data operations.We enable the Managed Investment Data (MID), Refe... Show more

 • Promoted

Senior Site Reliability Engineer

Sage Recruiting Inc.Toronto, Ontario, Canada
Full-time

This range is provided by Sage Recruiting Inc.Your actual pay will be based on your skills and experience — talk with your recruiter to learn more.Senior Site Reliability Engineer (Founding Role).A... Show more

 • Promoted

Senior Site Reliability Engineer - C$137,200 - C$196,000 A Year

Tubi Media GroupEast York, Canada
Full-time

Seeking a Senior Site Reliability Engineer to design, build, and maintain scalable distributed systems, automate operations with code and CI/CD pipelines, and participate in on-call rotations.Focus... Show more

 • Promoted

Site Reliability Engineer (Intermediate) - C$117,610 - C$158,240 A Year

BitcompleteEast York, Canada
Full-time

Intermediate SRE to build reliable cloud infrastructure, focusing on Kubernetes, Observability, or Developer Experience, collaborating on debugging, migrations, and automation. Show more

 • Promoted

Senior Site Reliability Engineer - C$180,000 - C$200,000 A Year

Sage Recruiting Inc.Toronto County, Canada
Full-time

Seeking a Senior Site Reliability Engineer to build and own the SRE program from scratch for a new fintech platform.Responsibilities include incident response, system design, and automation. Show more

 • Promoted

Site Reliability Engineer

Future Secure AIToronto, Ontario, Canada
Full-time

At Future Secure AI, we're building something genuinely new — and we're looking for people bold enough to build it with us.We work at the frontier of AI, tackling big, real-world problems for globa... Show more

 • Promoted

Site Reliability Engineer

Tangerine BankToronto, Canada
Full-time

Press Tab to Move to Skip to Content Link Select how often (in days) to receive an alert: Tangerine is Canada’s leading direct. Show more

 • Promoted

Senior Site Reliability Engineer

MorningstarToronto, Ontario, Canada
Full-time

Investment Services is Morningstar’s internal product group focused on building and maintaining the platforms that power our global data operations.We enable the Managed Investment Data (MID), Refe... Show more

 • Promoted

IBM Site Reliability Engineering Expert

LeadingtalentMarkham, Ontario, Canada
Full-time

Step into a career as a Site Reliability Engineer at IBM, focused on enhancing system reliability and performance.Engage directly with production systems and optimize customer experience.In this ro... Show more

 • Promoted

Site Reliability Engineer - C$102,700 - C$137,000 A Year

McCain FoodsToronto County, Canada
Full-time

Seeking a Site Reliability Engineer to ensure software system reliability and availability by designing resilient architectures, automating infrastructure, and optimizing performance in Azure cloud. Show more

 • Promoted

Senior Site Reliability Engineer Ii - Remote, Scale-Focused - C$183,000 - C$203,000 A Year - Remote

Leading Grocery Delivery ServiceNorth York, Canada
Remote
Full-time

Seeking a Senior Site Reliability Engineer to ensure platform performance, establish incident management, and oversee scalable infrastructure strategies.Requires programming and incident management... Show more