Talent.com
Royal Bank of Canada>
Staff, Site Reliability Engineer(Global Security)Royal Bank of Canada> • VANCOUVER, Canada
Staff, Site Reliability Engineer(Global Security)

Staff, Site Reliability Engineer(Global Security)

Royal Bank of Canada> • VANCOUVER, Canada
2 days ago
Job type
  • Full-time
Job description

Job Description

What is the opportunity?

We are seeking an experienced and hands-on Staff Site Reliability Engineer (SRE) who is passionate about building scalable systems, automating infrastructure, and improving the reliability of production environments. You will work closely with system support, engineering and infrastructure teams to ensure the availability, performance, and efficiency of our IAM systems and services. This is a technical, execution-focused role with deep engineering work.

What will you do?

  • Serve as the senior-most technical voice for IAM reliability — setting architecture direction and reliability standards, and leading by example through hands-on design, coding, and operations

  • Own service reliability for IAM systems end-to-end by defining and maintaining SLOs, SLIs, and error budgets, and using them to drive prioritization and continuous improvement

  • Design and implement resilient, highly available IAM infrastructure and services — spanning authentication, authorization, identity lifecycle, and privileged access — across multi-region and hybrid-cloud architectures

  • Write and review production-grade code (services, APIs, automation frameworks, and internal tools), applying software engineering rigor — testing, code review, version control, modular design — to reliability work rather than treating it as throwaway scripting

  • Build self-service platforms, reusable modules, and golden-path automation that let application teams provision, deploy, and safely operate IAM services with less hand-holding from the production support team

  • Champion Infrastructure as Code and GitOps practices (Terraform, Ansible, Puppet, Kubernetes/Helm) to eliminate manual toil, enforce consistency, and enable safe, repeatable deployments at scale

  • Own and continuously improve CI/CD pipelines and release engineering for IAM services, embedding reliability, security, and rollback safety directly into the delivery pipeline

  • Lead incident response and on-call operations for high-severity IAM outages and performance degradations, driving root cause analysis, blameless postmortems, and long-term structural remediation

  • Build and evolve observability, monitoring, and alerting pipelines (metrics, logs, traces) using modern tooling to proactively detect, investigate, and resolve availability and security issues before they impact users

  • Develop and test failover strategies and recovery procedures, including chaos engineering exercises, backup validation, and disaster recovery simulations to validate IAM system readiness

  • Orchestrate workload automation, scheduling, and release pipelines across enterprise systems (e.g., Stonebranch, CI/CD platforms) to streamline delivery and reduce manual operational overhead

  • Partner with security, infrastructure, application, and compliance teams to embed IAM into enterprise-wide business continuity and resilience strategy, ensuring alignment with risk and regulatory mandates

  • Mentor and coach engineers on reliability, automation-first thinking, and sound software engineering practice; help raise the bar for how the broader team builds and operates IAM services

What do you need to succeed?


Must Have:

  • 5+ years of experience in Site Reliability Engineering, DevOps or Platform Engineering with demonstrated staff/senior-level technical leadership across SLOs/SLIs, error budgets, and driving continuous reliability improvement at scale

  • Solid software engineering fundamentals — proficient in at least one modern language (Python, Go, Java, or similar), with the ability to design and build production-quality services and tooling, not just automation scripts

  • Proven experience designing, implementing, and operating highly available, fault-tolerant, and scalable systems in production, including hybrid and multi-cloud environments

  • Strong DevOps foundation — experience owning CI/CD pipelines (e.g., Jenkins, GitLab CI, GitHub Actions) and release engineering practices that build reliability into the delivery process itself

  • Platform-engineering mindset — a track record of turning recurring operational work into self-service tooling, reusable modules, or golden-path automation that other engineering teams can adopt independently

  • Experience with containerization and orchestration (Docker, Kubernetes) in production environments

  • Deep knowledge of building and operating monitoring, alerting, and observability platforms (e.g., Prometheus, Grafana, Dynatrace, ELK, Splunk, SIEM) to enable proactive incident detection and response

  • Proven incident management skills — able to lead high-severity incident response, root cause analysis, and postmortem processes, and to drive long-term fixes

  • Proficient in disaster recovery, failover strategies, and resilience testing (e.g., chaos engineering, tabletop exercises) to validate system readiness

  • Solid understanding of cloud platforms (AWS, Azure) and hybrid environments, including experience supporting production workloads at scale

  • Excellent collaboration and communication skills — comfortable working across security, infrastructure, application, and compliance teams, and influencing technical direction without direct authority

Nice to Have:

  • Experience with IAM platforms and security-focused systems (Microsoft Entra, Okta, Idira, SailPoint, Ping, HashiCorp Vault, etc.)

  • Expertise with Infrastructure as Code and configuration management (Terraform, Ansible, Puppet) and scripting (Python, PowerShell, Bash) to reduce manual toil and streamline deployments

  • Familiarity with identity, authentication, authorization, and privileged access management concepts and protocols (OAuth2, OIDC, SAML, LDAP, SCIM)

  • Knowledge of enterprise security architecture and compliance frameworks as they relate to IAM and service resilience

  • Exposure to AIOps or ML-based anomaly detection for proactive reliability management

What’s in it for you?

We thrive on the challenge to be our best, progressive thinking to keep growing, and working together to deliver trusted advice to help our clients thrive and communities prosper. We care about each other, reaching our potential, making a difference to our communities, and achieving success that is mutual.

  • A comprehensive Total Rewards Program including bonuses and flexible benefits, competitive compensation, commissions, and stock where applicable

  • Leaders who support your development through coaching and managing opportunities

  • Ability to make a difference and lasting impact

  • Work in a dynamic, collaborative, progressive, and high-performing team

  • A world-class training program in financial services

  • Opportunities to do challenging work


#LI-POST
#TECHPJ

Job Skills

Agile Working, Application Security, Automation Tools, Business Continuity, Cloud Platform, Cyber Security Management, Decision Making, Enterprise Security Architecture, High Reliability, Hybrid Systems, Identity Access Management (IAM), Information Security Management, Information Technology Security, Infrastructure Penetration Testing, Interpersonal Communication, IT Security Architecture, IT Systems Integration, Microsoft PowerShell, Python (Programming Language), Security Information and Event Management (SIEM), Security Standards, Security Tools, Strategic Thinking, Systems Development Lifecycle (SDLC), Terraform (Software)

Additional Job Details

Address:

16 YORK ST:TORONTO

City:

Toronto

Country:

Canada

Work hours/week:

37.5

Employment Type:

Full time

Platform:

TECHNOLOGY AND OPERATIONS

Job Type:

Regular

Pay Type:

Salaried

Posted Date:

2026-08-25

Application Deadline:

2026-09-08

Note: Applications will be accepted until 11:59 PM on the day prior to the application deadline date above

Our Employment Opportunities

At RBC, we are guided by living shared values of Client First, Integrity, Collaboration, Respect and Excellence and winning together as One RBC. We believe an inclusive workplace that has diverse perspectives is core to our continued growth as one of the largest and most successful banks in the world. Maintaining a workplace where our employees feel supported to perform at their best, effectively collaborate, drive innovation, and grow professionally helps to bring our Purpose to life and create value for our clients and communities. RBC strives to deliver this through policies and programs intended to foster a workplace based on respect, belonging and opportunity for all.

Join our Talent Community

Stay in-the-know about great career opportunities at RBC. Sign up and get customized info on our latest jobs, career tips and Recruitment events that matter to you.

Expand your limits and create a new future together at RBC. Find out how we use our passion and drive to enhance the well-being of our clients and communities at jobs.rbc.com.

RBC is presently inviting candidates to apply for this existing vacancy. Applying to this posting allows you to express your interest in this current career opportunity at RBC. Qualified applicants may be contacted to review their resume in more detail.

Create a job alert for this search

Staff, Site Reliability Engineer(Global Security) • VANCOUVER, Canada

Similar jobs

Senior Site Reliability Engineer - $120,000 - $140,000 A Year

Semios GroupWest End, Canada
Full-time

Senior Site Reliability Engineer to ensure infrastructure scalability, reliability, and performance, focusing on automation, observability, and resilience in a hybrid work environment. Show more

 • Promoted

Staff Software Engineer, Security - C$170,000 - C$250,000 A Year - Remote

Super.ComWest End, Canada
Remote
Full-time

Seeking a Staff-level Security Engineer to join the Security & Privacy team.This role involves acting as a company-wide subject matter expert, managing engineers, influencing strategy, and improvin... Show more

 • Promoted

Staff Site Reliability Engineer, Fabric - C$144,000 - C$200,000 A Year

MongoDBVancouver, Canada
Full-time

Platform Engineering is the department within SRE that is responsible for a range of critical infrastructure and operational functions that support the broader engineering organization.Among these ... Show more

 • Promoted

Site-Based Reliability & Integrity Engineer – Lng Facility - C$115,000 - C$125,000 A Year

Woodfibre Management LtdSquamish, Canada
Full-time

A Canadian LNG project company is seeking a Reliability & Integrity Engineer for their facility in Squamish, BC.This role involves ensuring technical integrity and reliability during construction a... Show more

 • Promoted

Site Reliability Engineer

TELUS DigitalVancouver, Metro Vancouver Regional District, CA
Full-time

Welcome to TELUS Digital — where innovation drives impact at a global scale.As an award-winning digital product consultancy and the digital division of TELUS, one of Canada’s largest telecommunicat... Show more

 • Promoted

REMOTE Protection & Controls Substation Engineer

JobotVancouver, Metro Vancouver Regional District, CA
Remote
Full-time

REMOTE Protection & Controls Substation Engineer.Be among the first 25 applicants.This range is provided by Jobot.Your actual pay will be based on your skills and experience — talk with your recrui... Show more

 • Promoted

Site Reliability Engineer - $86,105 - $128,150 A Year

SmartThingsVancouver, Canada
Full-time

Seeking an SRE/DevOps Engineer to maintain cloud systems, automate infrastructure, optimize workflows, and ensure operational excellence.Collaborate with dev teams and mentor junior engineers. Show more

 • Promoted

Staff Security Engineer: Remote Lead & AppSec Architect

Super.comVancouver, Metro Vancouver Regional District, CA
Remote
Full-time

A technology company in Canada is seeking a Staff Software Engineer specializing in Security to lead and mentor engineers within the Security & Privacy team.The successful candidate will drive appl... Show more

 • Promoted

Senior Site Reliability Engineer - Hybrid & Leadership - $120,000 - $140,000 A Year

A Leading Agricultural Technology CompanyVancouver, Canada
Full-time

A leading agricultural technology company in Vancouver is seeking a Senior Site Reliability Engineer.The role emphasizes leading infrastructure delivery, enhancing product resiliency, and mentoring... Show more

 • Promoted

Security Operations Engineer - $98,400 - $147,600 A Year - Remote

JaneWest End, Canada
Remote
Full-time

Seeking a Security Operations Engineer to manage security services, monitor alerts, triage incidents, and participate in on-call rotations.Focus on reducing toil with AI and automation, and fosteri... Show more

 • Promoted

Security Site Manager- With Relocation Support - C$80,000 - C$90,000 A Year

GuardteckVancouver, Canada
Full-time

Manages security operations for a site, ensuring adherence to standards, client satisfaction, and team development.Requires strong leadership and incident response skills. Show more

 • Promoted

Senior Security Operations Engineer - C$192,000 - C$240,000 A Year

BrexVancouver, Canada
Full-time

Why join usBrex is the AI-powered spend platform.We help companies spend with confidence with integrated corporate cards, banking, and global payments, plus intuitive software for travel and expens... Show more

 • Promoted

Security Site Supervisor

Allied UniversalVancouver, Metro Vancouver Regional District, CA
Full-time

Company Overview: We are North America's leading security and facility services provider with approximately 300,000 service personnel.At Allied Universal(R), we pride ourselves on fostering a promo... Show more

 • Promoted

Staff Site Reliability Engineer - C$124,200 - C$166,700 A Year

Walt Disney Animation StudiosWest End, Canada
Full-time

Seeking a Staff Site Reliability Engineer with expertise in Linux, software development, CI tools, Git, cloud hosting, and container computing to optimize service deployments and improve system ava... Show more

 • Promoted

Senior Site Reliability Engineer, Developer Platform - C$108,200 - C$143,430 A Year

Rivian and Volkswagen Group TechnologiesVancouver, Canada
Full-time

About UsRivian and Volkswagen Group Technologies is a joint venture between two industry leaders with a clear vision for automotive's next chapter.From operating systems to zonal controllers to... Show more

 • Promoted

Senior Security Engineer (Global Security)

RBCVancouver, British Columbia, Canada
Full-time

Shape the future of application security at RBC! Join the Application Security Group and develop innovative solutions to streamline processes, boost efficiency, and unlock new possibilities.We’re s... Show more

 • Promoted

Site Reliability Engineer

ArbitrumVancouver, British Columbia, Canada
Full-time

LayerZero The Future is Omnichain.Founded in 2021, LayerZero’s vision is to create a community of cross-chain developers, building dApps that are no longer constrained by individual blockchain capa... Show more

 • Promoted

Remote Software Engineer: Nuclear Safety-Critical Systems

TRXVancouver, Metro Vancouver Regional District, CA
Remote
Full-time

A leading technology firm is seeking a Software Developer to design and maintain software for nuclear systems in Canada.The ideal candidate has a Bachelor’s degree in Computer Science or a related ... Show more

 • Promoted

Site Reliability Engineer Ii - C$100,000 - C$139,500 A Year

Electronic ArtsVancouver, Canada
Full-time

The Site Reliability Engineer will design, build, and optimize systems.They will develop features and improve the platform for hosting EA's games and create automation tools. Show more

 • Promoted

Site Reliability Engineer - $120,000 - $200,000 A Year

LayerZero LabsVancouver, Canada
Full-time

The SRE will build and maintain blockchain node infrastructure, ensuring reliability and performance, and automating incident detection. Show more