Talent.com
Rootly
Senior Site Reliability EngineerRootly • Toronto, ON, Canada
Senior Site Reliability Engineer

Senior Site Reliability Engineer

Rootly • Toronto, ON, Canada
30+ days ago
Job type
  • Full-time
Job description

About Rootly

At Rootly, we are on a mission to be the go‑to way companies respond when things go wrong, helping every organization be more reliable. We do this by building an industry‑leading incident management platform that allows companies around the world consistently and quickly resolve incidents. We are not simply transforming an industry, we are carving an entirely new +$B segment ourselves and need incredible talent to achieve this ambitious goal together.

Customers love Rootly. Some of the fastest growing companies around the world such as NVIDIA, Figma, Canva, Tripadvisor, Squarespace and more rely on Rootly to power their critical incident management process. They obsess over our delightful enterprise‑ready platform and unique partnership model. See why our customers have reviewed us 5 stars on G2.

Investors love Rootly. We are backed by some of the most respected funds in the world from Y Combinator to operators like the CTO of Dropbox and GitHub. We'd be happy to disclose our entire funding and profitability picture live during the interview. As a culture we relentlessly put transparency first. We conduct monthly financial reviews as a team so everyone has a pulse on the health of the business and publish what we are building in our weekly changelog.

About the Role

This is an opportunity to join Rootly as an early SRE leader and shape our technical foundation. You will experience the balance of being scrappy and operating at scale. What you’ll be doing one day could look very different the next. You will be empowered to identify opportunities that will help us grow and own it. In short, this role is designed for individuals that crave ownership, stimulating technical challenges, love shipping fast, and are mission driven. We won’t sugarcoat it, the work will be challenging, but it will also be one of the most rewarding learning experiences of your career.

Responsibilities

  • Embed with product teams to enhance observability, reliability, and performance of their services.
  • Own our CI/CD pipelines, observability tooling, monitoring systems, and incident response processes.
  • Build tools and automation to eliminate manual toil, improve engineering velocity and developer experience, and improve system reliability.
  • Collaborate deeply across engineering to understand systems at the code level and surface cross‑cutting reliability, performance, and scaling concerns.
  • Architect and scale our infrastructure, ensuring best‑in‑class performance, availability, and operational excellence.
  • Drive capacity planning efforts to ensure our infrastructure is resilient and scalable as we grow.
  • Define and manage SLOs and error budgets in partnership with Engineering teams who own production services.
  • Be vocal – act as a strong voice and force of reliability, quality, performance, and scalability.

About You'll Need

  • 5+ years of experience in an SRE, Platform, or Infrastructure Engineering role.
  • 5+ years of experience writing software in a production environment.
  • Strong technical knowledge of cloud infrastructure, distributed systems, and reliability practices.
  • Strong understanding of observability, performance tuning, and scaling strategies.
  • Deep familiarity with incident response, monitoring, and CI/CD systems.
  • Hands‑on experience supporting web or RPC services at meaningful scale.
  • You write code to solve infrastructure problems; not shell scripts alone, but production‑grade software.

Preferred Qualifications

  • You have a big‑picture systems mindset and a proactive approach to reliability.
  • You’ve embedded with product teams and influenced design and architecture decisions.
  • You’re comfortable taking ownership of complex problems—and seeing them through.
  • Experience with Ruby and Go is a plus.

Why Rootly?

We’re not just another startup. We’re building something category‑defining and want teammates who crave ownership, love solving hard problems, and thrive in a high‑bar, high‑impact environment.

Benefits and Compensation

  • Competitive compensation and early equity in a fast‑growing, venture‑backed company.
  • Comprehensive medical, dental, and vision coverage.
  • 3 weeks of vacation, plus unlimited sick and mental health days, and a company‑wide end‑of‑year shutdown to recharge.
  • $500 stipend for home office setup.
  • Unlimited token usage and access to AI tools.
  • A fast‑moving, high‑impact environment where your leadership and ideas directly shape the future of the company.

If this sounds like the kind of challenge and opportunity you’re looking for, apply now and let’s build something great together.

Rootly is an equal opportunity employer. We aim to create an environment where every team member at Rootly feels like they belong so they can have a greater impact on our business and customers. We do not discriminate on the basis of race, religion, colour, national origin, gender, sexual orientation, age, marital status, veteran status, or disability status.

#J-18808-Ljbffr
Create a job alert for this search

Senior Site Reliability Engineer • Toronto, ON, Canada

Similar jobs

Senior Site Reliability Engineer

Guidewire Softwaretoronto, on, Canada
Full-time

At Guidewire, we make software that offers Property and Casualty (P&C) Insurance companies the tools to take care of their customers when they need it the most, whether that’s a time of crisis, a n... Show more

 • Promoted

Senior Staff Site Reliability Engineer

CerebrasToronto, ON, CA
Full-time

Become a Staff Site Reliability Engineer at Cerebras Systems, revolutionizing AI inference service reliability.Design innovative solutions for operational challenges.This position is crucial for en... Show more

 • Promoted

Senior Site Reliability Engineer - Hybrid & Leadership - C$120,000 - C$140,000 A Year

AgriTechToronto County, Canada
Full-time

Senior Site Reliability Engineer needed to lead infrastructure, improve product resilience, and mentor a team in Vancouver.Requires 5+ years of cloud experience (AWS/GCP). Show more

 • Promoted

Senior I Site Reliability Engineer (Canada) - C$148,988 - C$200,644 A Year

Bit CompleteEast York, Canada
Full-time

Senior Site Reliability Engineer to design, build, and scale cloud-native systems, focusing on Kubernetes, infrastructure-as-code, databases, and observability, while mentoring peers. Show more

 • Promoted

Senior Site Reliability Engineer

ThinkificToronto, ON, CA
Full-time

Senior Site Reliability Engineer.Senior Site Reliability Engineer.Are you an experienced Site Reliability Engineer looking for a new challenge?.Senior Site Reliability Engineer.Senior Site Reliabil... Show more

 • Promoted

Remote Senior Site Reliability Engineer Role

ViafouraToronto, ON, CA
Remote
Full-time

Advance your career as a Senior Site Reliability Engineer at Viafoura, specializing in Kubernetes and AWS infrastructure.This remote role positions you to improve our platform's performance and sca... Show more

 • Promoted

Senior Site Reliability Engineer, Kong Konnect

Kong Inc.Toronto, ON, CA
Full-time

Senior Site Reliability Engineer, Kong Konnect.Join to apply for the Senior Site Reliability Engineer, Kong Konnect role at Kong Inc.Are you ready to power the World's connections?.If you don’t thi... Show more

 • Promoted

Site Reliability Engineer

TELUS DigitalToronto, ON, CA
Full-time

Welcome to TELUS Digital — where innovation drives impact at a global scale.As an award-winning digital product consultancy and the digital division of TELUS, one of Canada’s largest telecommunicat... Show more

 • Promoted

IBM Site Reliability Engineering Expert

LeadingtalentMarkham
Full-time

Step into a career as a Site Reliability Engineer at IBM, focused on enhancing system reliability and performance.Engage directly with production systems and optimize customer experience.In this ro... Show more

 • Promoted

Senior Site Reliability Engineer

Sage Recruiting Inc.Toronto, Ontario, Canada
Full-time

This range is provided by Sage Recruiting Inc.Your actual pay will be based on your skills and experience — talk with your recruiter to learn more.Senior Site Reliability Engineer (Founding Role).A... Show more

 • Promoted

Senior Site Reliability Engineer

RootlyToronto, ON, CA
Full-time

At Rootly, we are on a mission to be the go‑to way companies respond when things go wrong, helping every organization be more reliable.We do this by building an industry‑leading incident management... Show more

 • Promoted

Site Reliability Engineer (SRE)

Tangerine BankToronto
Full-time +1

Press Tab to Move to Skip to Content Link.Select how often (in days) to receive an alert:.Tangerine is Canada’s leading direct bank.We offer flexible and accessible banking options, innovative prod... Show more

 • Promoted

Senior Site Reliability Engineer

Morningstar Credit Ratings, LLCToronto, Ontario, Canada
Full-time

Investment Services is Morningstar’s internal product group focused on building and maintaining the platforms that power our global data operations.We enable the Managed Investment Data (MID), Refe... Show more

 • Promoted

Senior Site Reliability Engineer

iManageToronto, Ontario, Canada
Full-time

SRE is part of a global organization that leverages the latest technology to communicate with our colleagues across the globe.We organize ourselves into distributed teams – SRE teams are anchored t... Show more

 • Promoted

Senior Site Reliability Engineer - $130,000 - $150,000 A Year

FableToronto, Canada
Full-time

About FableGlobal enterprises work with Fable to make products more accessible for over one billion people who live with disabilities.Our customers include global leaders like Walmart, Slack, and S... Show more

 • Promoted

Senior Site Reliability Engineer - C$180,000 - C$200,000 A Year

Sage Recruiting Inc.Toronto County, Canada
Full-time

Seeking a Senior Site Reliability Engineer to build and own the SRE program from scratch for a new fintech platform.Responsibilities include incident response, system design, and automation. Show more

 • Promoted

Senior Site Reliability Engineer

MorningstarToronto
Full-time

Investment Services is Morningstar’s internal product group focused on building and maintaining the platforms that power our global data operations.We enable the Managed Investment Data (MID), Refe... Show more

 • Promoted

Senior Site Reliability Engineer- Remote

ClickHouseToronto, ON, CA
Remote
Full-time

Senior Site Reliability Engineer- Remote.Recognized on the 2025 Forbes Cloud 100 list, ClickHouse is one of the most innovative and fast-growing private cloud companies.With more than 3,000 custome... Show more

 • Promoted

Site Reliability Engineer

Momentum Financial Services GroupToronto
Full-time

At Momentum Financial Services Group, we help people move forward by reimagining how money works for those who need it most.With more than 40 years of experience, we’re the team behind Money Mart—C... Show more

 • Promoted

Senior Site Reliability Engineer - C$130,000 - C$180,000 A Year

Acquird.ioToronto, Canada
Full-time

Senior SRE needed to maintain cloud infrastructure, CI/CD pipelines, and ensure system uptime and reliability.Requires experience with Azure and cloud engineering. Show more