Talent.com
Chelsea Avondale
Remote Reliability EngineerChelsea Avondale • Saint-Laurent, Quebec
Remote Reliability Engineer

Remote Reliability Engineer

Chelsea Avondale • Saint-Laurent, Quebec
30+ days ago
Job type
  • Full-time
  • Remote
Job description

Chelsea Avondale is the world’s most cutting-edge home insurance group. We have developed sophisticated risk modeling and insurance pricing technologies for home insurance and deploy that technology through our own insurance company.


Our team consists of some of the brightest minds in insurance, software development, finance, and operations. Our group includes our scientific research & engineering division (Skynet Software) and Canadian property & casualty insurance company (Max Insurance).


Together, our group is transforming the Canadian and global insurance landscape.


JOB DESCRIPTION:


Chelsea Avondale is looking for a Reliability Engineer with a background in infrastructure system engineering to support the growth of a secure, dynamic, and scalable IT environment across the group. Our business is going through rapid growth, and it is essential that our systems infrastructure keeps pace.


The Reliability Engineer will play a crucial role in ensuring the reliability, scalability, and performance of our systems, enabling the continuous delivery of our products and services. They will be accountable for ensuring overall availability, as well as enhancing Engineering teams’ capability to design, build and operate robust systems at scale.


This position is ideal for candidates who have an extraordinary sense of responsibility and are not afraid to roll up their sleeves. Our IT environment is not toolkit rich. What we are NOT looking for is someone who wants to take months installing a large number of tools from their preferred toolkit. We take pride in maintaining a fundamental stack of technologies, much of it in Python, and we are looking for someone who shares this mentality.


If you are someone who thrives in a high-performance culture and is eager for work that is both challenging and constantly evolving, this role is perfect for you. We strongly encourage and help our team members to improve and enhance their personal skill sets within our organization. On your journey with us, you will have the ability to learn and grow rapidly, taking on more responsibilities.


RESPONSIBILITIES:



  • Play an integral role in the design, implementation & maintenance of AWS cloud server environments.

  • Design, implement, and maintain robust monitoring and alerting systems in Python to detect and respond to incidents in a timely manner.

  • Collaborate with cross-functional teams to enhance reliability of our systems and services.

  • Design, configure, deploy, and maintain infrastructure on AWS using best practices and industry standards.

  • Conduct post-incident analysis to identify root causes, implement corrective actions, and prevent similar issues in the future.

  • Assist in capacity planning & optimize services to provide scalable, stable, & secure systems.

  • Implement high availability and disaster recovery solutions to provide data redundancy, resilience, and data loss prevention.

  • Assist with the implementation of select network engineering solutions including firewalls, load balancing, VPNs & LANs, where necessary.


PREFERRED EXPERIENCE & SKILLS:



  • Bachelor’s degree in Computer Science, Computer Engineering, Electrical Engineering, or related field.

  • 1+ years of experience as a Reliability Engineer or similar role, with a focus on maintaining high-performance, scalable, and reliable web systems.

  • We also encourage highly motivated new grads to apply.

  • Hands-on experience with AWS cloud environments – instances, CloudWatch, EFS, etc.

  • Proficiency at Python is a must.

  • Experience using NGINX for reverse proxy, load balancing, and caching.

  • Experience with Unix / Windows server configuration, administration, performance tuning and troubleshooting.

  • Working knowledge of web technologies (web servers, DNS, SSL, Browsers).

  • Working knowledge of web development processes (source control, deployment, etc.).

  • Experience load testing, pen testing, and providing security for cloud resources is beneficial.


Skynet Software welcomes and encourages applications from people with disabilities. Accommodations are available on request for candidates taking part in all aspects of the selection process.

Create a job alert for this search

Remote Reliability Engineer • Saint-Laurent, Quebec

Similar jobs

Senior Site Reliability Engineer — Kubernetes, AWS & Observability

ThinkificMontreal (administrative region), QC, CA
Full-time

A leading e-learning provider in Canada is seeking a Senior Site Reliability Engineer to enhance and secure their infrastructure supporting online course creators.This role involves improving perfo... Show more

 • Promoted

Senior Full-Stack Engineer - Accessibility & Inclusive Tech (Remote)

Accessibility Partners CanadaMontreal (administrative region), QC, CA
Remote
Full-time

A leader in accessible technology is seeking a Senior Full-Stack Developer to create equitable and accessible digital systems.This role involves developing both front-end and back-end systems, focu... Show more

 • Promoted

Site Reliability Engineer

Hunter BondMontréal, Canada
Full-time

Role: DevOps EngineerClient: Most Elite Tech Firm in CanadaCompensation: Up to $200k CAD + Bonus + PackageLocation: MontrealOverviewAn Elite FinTech Firm is looking for a highly talented DevOps Eng... Show more

 • Promoted

Senior Platform Engineer - Remote, Scale & Reliability

Lillio (formerly HiMama)Montreal (administrative region), QC, CA
Remote
Full-time

A leading EdTech company in Canada is seeking a Senior Platform Engineer to enhance system reliability and performance while contributing to impactful software solutions.The role involves making te... Show more

 • Promoted

Director of Engineering — Platform & Reliability (Remote)

CliniaMontreal (administrative region), QC, CA
Remote
Full-time

A tech-driven health company in Canada is seeking a Director of Engineering to lead an engineering team of 25.You will manage delivery, ensure platform reliability, and set engineering standards wh... Show more

 • Promoted

Senior Site Reliability Engineer (Remote-First)

VySystemsMontreal (administrative region), QC, CA
Remote
Full-time

A leading technology company is seeking a Senior Site Reliability Engineer with robust Kubernetes knowledge to work remotely.Ideal candidates have over 6 years of experience in IT disciplines, prof... Show more

 • Promoted

Senior Site Reliability Engineer Focused on Kubernetes Infrastructure

Chainlink LabsMontreal (administrative region), QC, CA
Full-time

Elevate decentralized architecture as a Senior Site Reliability Engineer.Spearhead Kubernetes-based infrastructure for decentralized applications, driving scalability, security, and operational eff... Show more

 • Promoted

Platform Engineer (Remote Initially)

My RecruiterMontreal (administrative region), QC, CA
Remote
Full-time

About the job Platform Engineer (Remote Initially).My recruiter Inc caters to grow the business to next level by providing professional services to empower the organizational values with right suit... Show more

 • Promoted

Senior AI Engineer - Backend-Driven Agentic (Remote)

BayOne SolutionsMontreal (administrative region), QC, CA
Remote
Full-time

A technology firm is hiring a Senior Full Stack Engineer (Backend-Leaning) with a strong focus on backend services and AI applications.The ideal candidate will have experience in building agent-bas... Show more

 • Promoted

Senior Site Reliability Engineer- Remote

ClickHouseMontreal (administrative region), QC, CA
Remote
Full-time

Senior Site Reliability Engineer- Remote.Recognized on the 2025 Forbes Cloud 100 list, ClickHouse is one of the most innovative and fast-growing private cloud companies.With more than 3,000 custome... Show more

 • Promoted

Site Reliability Engineer

TELUS DigitalMontreal (administrative region), QC, CA
Full-time

Welcome to TELUS Digital — where innovation drives impact at a global scale.As an award-winning digital product consultancy and the digital division of TELUS, one of Canada’s largest telecommunicat... Show more

 • Promoted

Senior Distributed Diffusion ML Engineer — Remote

Bagel LabsMontreal (administrative region), QC, CA
Remote
Full-time

A leading machine learning research lab is looking for an expert in distributed systems to design scalable infrastructure for diffusion models.The ideal candidate will have substantial experience i... Show more

 • Promoted

Site Reliability Engineer (Linux / Cloud Infrastructure)

Atlantis IT GroupMontréal, Montreal (administrative region), Canada
Full-time

Site Reliability Engineer (Linux / Cloud Infrastructure) role with hands-on experience across Linux, distributed systems, scripting, databases, monitoring, containers, cloud SaaS integrations, mess... Show more

 • Promoted

Site Reliability Engineer

MaintainXmontreal (administrative region), qc, Canada
Full-time

MaintainX is the world's leading AI-powered maintenance and asset management platform, serving 13,000+ customers including Duracell, Shell, Cintas, and Brenntag.We raised $150M in Series D funding ... Show more

 • Promoted

Senior Site Reliability Engineer

SecurityScorecardmontreal (administrative region), qc, Canada
Full-time

SecurityScorecard is the global leader in cybersecurity ratings, with over 12 million companies continuously rated, operating in 64 countries.Founded in 2013 by security and risk experts Dr.Alex Ya... Show more

 • Promoted

Intact Hybrid Site Reliability Engineer

IntactMontreal (administrative region), QC, CA
Full-time

Join Intact as a Site Reliability Engineer and elevate operational reliability across cloud platforms.This hands-on role employs Azure, AWS, and GCP expertise to manage incidents effectively.The SR... Show more

 • Promoted

Remote Ai Agent Engineer - Llm & Workflow Orchestration - $100,000 - $120,000 A Year - Remote

PR technology companyLe Plateau, Canada
Remote
Full-time

Design and implement AI applications and agentic workflows using AI frameworks and prompt engineering. Show more

 • Promoted

Reliability Engineer - C$100,000 - C$150,000 A Year

Epiroc CanadaLe Plateau, Canada
Full-time

Reliability Engineer responsible for developing and implementing processes for product serviceability, reliability, and maintainability, performing data analytics, and communicating with stakeholders. Show more

 • Promoted

Site Reliability Engineer - Tech Talent International

Tech Talent InternationalMontreal
Full-time

Join Tech Talent International as a Senior Site Reliability Engineer, specializing in Automation & Observability, located in Montreal.This hybrid role focuses on enhancing production efficiency and... Show more

 • Promoted

Enterprise Resiliency Engineer

Intact Financial CorporationMontreal
Full-time

Lead resiliency initiatives as an Enterprise Resiliency Engineer.Drive architecture and production reliability practices across multi-cloud platforms in a hybrid environment.You will define the sta... Show more