Talent.com
MongoDB
Staff Site Reliability Engineer, FabricMongoDB • Vancouver, British Columbia, CA
Staff Site Reliability Engineer, Fabric

Staff Site Reliability Engineer, Fabric

MongoDB • Vancouver, British Columbia, CA
30+ days ago
Job type
  • Full-time
Job description

The Team

Platform Engineering is the department within SRE that is responsible for a range of critical infrastructure and operational functions that support the broader engineering organization. Among these are our multi-cloud-provider Kubernetes infrastructure, deployment machinery, and observability and alerting systems.

The Fabric team manages the infrastructure that enables secure communication between systems and from the public internet. Their responsibilities encompass network architecture, service mesh, and edge load balancing, ensuring customer data remains safe in transit. The team plays a crucial role in developing and maintaining the reliable and globally connected multi-cloud network that supports MongoDB products.

This role can sit in our Toronto or Vancouver offices, or fully remote from anywhere in North America. When based in an office, we provide hybrid work accommodation.

Role Overview

We are seeking a talented Site Reliability Engineer (SRE) with a strong networking background to join the Fabric team. This role is pivotal in building and maintaining the robust infrastructure necessary for secure and efficient communication between our services. As an SRE on the Fabric team, you will leverage your expertise in networking, distributed systems, and automation to ensure our systems are resilient, scalable, and reliable.

The ideal candidate should

  • Have 10+ years of experience working on software and operating distributed systems, with deep expertise in networking fundamentals and a good understanding of how the internet works, e.g. TCP/IP (including IPv6), DNS, TLS/mTLS, BGP, tunnels, overlays, and SDN principles.
  • Possess a customer-focused mindset, driving improvements that benefit end‑users.
  • Value efficiency in processes and operations, and display a strong preference for automation over manual processes (“allergic to ops work”).
  • Be intimately familiar with modern cloud-based infrastructure and the network design primitives of at least one of AWS, Azure, or GCP, e.g. VPCs, subnetting, routing, VPNs, peering, private link / private service connect, and CDNs.
  • Have a strong knowledge of service mesh and load‑balancing concepts, and be eager to implement these in a multi‑cloud environment.

Expectations

  • Participate in the development of a reliable and resilient multi‑cloud globally‑connected network that is crucial for MongoDB’s services.
  • Collaborate with service‑owning teams to provide internal support, addressing technical issues and offering guidance on best practices for service‑to‑service connectivity.
  • Participate in a 24/7 on‑call rotation to swiftly resolve issues related to network architecture and service‑to‑service connectivity, ensuring minimal disruption and high availability.

MongoDB is an equal opportunities employer. MongoDB is committed to providing any necessary accommodations for individuals with disabilities within the application and interview process. To request an accommodation due to a disability, please inform your recruiter.

MongoDB’s base salary range for this role in Canada is: $144,000—$200,000 CAD. Other benefits for eligible employees may include equity, participation in the employee stock purchase program, flexible paid time off, 20 weeks fully‑paid gender‑neutral parental leave, fertility and adoption assistance, Registered Retirement Savings Plan (RRSP) with employer match, mental health counseling, backup child and elder care, and health, dental, and vision benefits.

Req ID: 426185

#J-18808-Ljbffr
Create a job alert for this search

Staff Site Reliability Engineer, Fabric • Vancouver, British Columbia, CA

Similar jobs

Staff Site Reliability Engineer, Fabric

MongoDBvancouver, metro vancouver regional district, Canada
Full-time

Platform Engineering is the department within SRE that is responsible for a range of critical infrastructure and operational functions that support the broader engineering organization.Among these ... Show more

 • Promoted

Senior Site Reliability Engineer

ThinkificVancouver, Metro Vancouver Regional District, CA
Full-time

Senior Site Reliability Engineer.Senior Site Reliability Engineer.Are you an experienced Site Reliability Engineer looking for a new challenge?.Senior Site Reliability Engineer.Senior Site Reliabil... Show more

 • Promoted

Staff Site Reliability Engineer - C$124,200 - C$166,700 A Year

Walt Disney Animation StudiosVancouver, Canada
Full-time

Seeking a Staff Site Reliability Engineer with expertise in Linux, software development, CI tools, Git, cloud hosting, and container computing to optimize service deployments and improve system ava... Show more

 • Promoted

Site Reliability Engineer Ii - C$100,000 - C$139,500 A Year

Electronic ArtsWest End, Canada
Full-time

The Site Reliability Engineer will design, build, and optimize systems.They will develop features and improve the platform for hosting EA's games and create automation tools. Show more

 • Promoted

Site Reliability Engineer

TELUS DigitalVancouver, Metro Vancouver Regional District, CA
Full-time

Welcome to TELUS Digital — where innovation drives impact at a global scale.As an award-winning digital product consultancy and the digital division of TELUS, one of Canada’s largest telecommunicat... Show more

 • Promoted

Senior Site Reliability Engineer

Arista NetworksVancouver, Canada
Full-time

Company DescriptionArista Networks is an industry leader in data‐driven, client‐to‐cloud networking for large data center, campus and routing environments.What sets us apart is our relentless pursu... Show more

 • Promoted

Senior Site Reliability Engineer- Remote

ClickHouseVancouver, Metro Vancouver Regional District, Canada
Remote
Full-time

Senior Site Reliability Engineer- Remote.Recognized on the 2025 Forbes Cloud 100 list, ClickHouse is one of the most innovative and fast-growing private cloud companies.With more than 3,000 custome... Show more

 • Promoted

Senior Site Reliability Engineer - $120,000 - $140,000 A Year

Semios GroupVancouver, Canada
Full-time

Who we are:Founded in 2010, Semios Group is a leading agricultural technology company helping growers, agronomists, and ag retailers manage over 200 million acres across five countries.Semios pione... Show more

 • Promoted

Site Reliability Engineer

Applevancouver, metro vancouver regional district, Canada
Full-time

The Apple Service Engineering - SRE team is looking for Site Reliability Engineers with experience in developing processes, tools, and automation for managing distributed systems in production envi... Show more

 • Promoted

Principal Site Reliability Engineer

SaviyntVancouver
Full-time

Work on a mission‑critical SaaS platform used by global enterprises.Solve complex reliability challenges at scale.Influence architecture and engineering culture at a company level.Competitive compe... Show more

 • Promoted

Senior Site Reliability Engineer - Hybrid & Leadership - $120,000 - $140,000 A Year

A Leading Agricultural Technology CompanyVancouver, Canada
Full-time

Seeking a Senior Site Reliability Engineer to lead infrastructure, improve product resilience, and mentor a team in Vancouver.Requires 5+ years of experience in cloud environments. Show more

 • Promoted

Staff Platform Engineer

RivianVancouver, Metro Vancouver Regional District, CA
Full-time

Rivian is on a mission to keep the world adventurous forever.This goes for the emissions‑free Electric Adventure Vehicles we build, and the curious, courageous souls we seek to attract.As a company... Show more

 • Promoted

Site Reliability Consultant

PythianVancouver, Metro Vancouver Regional District, CA
Full-time

Operate and optimize Kubernetes clusters, Istio service mesh, and Linux-based systems.Automate workflows using Go, Python, and Shell scripting.Build monitoring and observability solutions with Prom... Show more

 • Promoted

Site Reliability Engineer Vancouver, BC

LayerZerovancouver, metro vancouver regional district, Canada
Full-time

Founded in 2021, LayerZero’s vision is to create a community of cross-chain developers, building dApps that are no longer constrained by individual blockchain capabilities.With LayerZero's simple, ... Show more

 • Promoted

Site Reliability Engineer - $86,105 - $128,150 A Year

SmartThingsNorth Vancouver, Canada
Full-time

Seeking an SRE/DevOps Engineer to maintain cloud systems, automate infrastructure, optimize workflows, and ensure operational excellence.Collaborate with dev teams and mentor junior engineers. Show more

 • Promoted

Senior Site Reliability Engineer Focused on Kubernetes Infrastructure

Chainlink LabsVancouver, Metro Vancouver Regional District, Canada
Full-time

Elevate decentralized architecture as a Senior Site Reliability Engineer.Spearhead Kubernetes-based infrastructure for decentralized applications, driving scalability, security, and operational eff... Show more

 • Promoted

CFD Staff Engineer: Remote Client Solutions & Training

SimuTech GroupVancouver, Metro Vancouver Regional District, CA
Remote
Full-time

A leading engineering simulation company is seeking a professional to provide technical support and training in CFD tools like ANSYS Fluent.This role requires a Master's degree and 3-6 years of eng... Show more

 • Promoted

Member of Technical Staff, Model Efficiency

CohereVancouver, Metro Vancouver Regional District, CA
Full-time

Member of Technical Staff, Model Efficiency.Be among the first 25 applicants.Get AI-powered advice on this job and more exclusive features.Our mission is to scale intelligence to serve humanity.We’... Show more

 • Promoted

Senior Site Reliability Engineer - $113,900 - $147,400 A Year

SamsungNorth Vancouver, Canada
Full-time

Lead a Cloud Engineering team to ensure the reliability, scalability, and performance of critical systems.Architect cloud-native solutions, minimize downtime, and maximize operational efficiency. Show more

 • Promoted

Site Reliability Engineer

ArbitrumVancouver, British Columbia, Canada
Full-time

LayerZero The Future is Omnichain.Founded in 2021, LayerZero’s vision is to create a community of cross-chain developers, building dApps that are no longer constrained by individual blockchain capa... Show more