Talent.com
Software Mind
[8SN] Site Reliability Engineer (SRE) – UIUXSoftware Mind • Montreal, Quebec, Canada
[8SN] Site Reliability Engineer (SRE) – UIUX

[8SN] Site Reliability Engineer (SRE) – UIUX

Software Mind • Montreal, Quebec, Canada
1 day ago
Job type
  • Full-time
Job description

About the Role

We are looking for a Site Reliability Engineer (SRE) UI/UX to support the deployment operations and ongoing maintenance of a production UI service running on Kubernetes.

This role focuses on monitoring service health troubleshooting production issues investigating incidents and ensuring reliable service delivery. You will work closely with engineering and client teams to support production operations and complete work based on a client-directed backlog.

While this role supports a UI-based service it is not a frontend development position. Working knowledge of Web Components is required to perform first-level debugging of UI-related issues but deep frontend development expertise is not expected.

What Youll Do

  • Support the deployment operations and ongoing maintenance of production services running on Kubernetes.
  • Monitor service health availability and performance.
  • Investigate and troubleshoot production incidents using logs monitoring and debugging tools.
  • Perform log analysis and incident debugging using Splunk.
  • Identify service issues and collaborate with engineering teams to support timely resolution.
  • Participate in incident response and production support activities.
  • Perform first-level debugging of UI-related issues involving Web Components.
  • Support service reliability and continuous improvement initiatives.
  • Assist with CI/CD pipelines and cloud-native application operations when needed.
  • Work effectively within a client-directed backlog and established priorities.

Qualifications :

Required Qualifications

  • 4 years of experience in Site Reliability Engineering DevOps Platform Engineering Production Support or a related role.
  • Hands-on experience supporting the deployment operations and ongoing maintenance of production services running on Kubernetes.
  • Experience monitoring service health troubleshooting production issues and supporting service reliability.
  • Proficiency with Splunk for log analysis and incident debugging.
  • Experience participating in production incident response and root-cause analysis.
  • Working knowledge of Web Components and the ability to perform first-level debugging of UI-related issues.
  • Strong troubleshooting analytical and problem-solving skills.
  • Experience collaborating with software engineering and cross-functional teams.
  • Ability to work independently and effectively within a client-directed backlog.
  • Excellent written and spoken English at least B2 level.


Additional Information :

Preferred Qualifications

  • Experience supporting CI/CD pipelines.
  • Familiarity with multi-tenant services.
  • Experience with cloud-native application operations.
  • Experience supporting high-availability enterprise or SaaS platforms.
  • Familiarity with additional monitoring and observability tools.
  • Experience with cloud platforms such as AWS Azure or GCP.
  • Familiarity with container and deployment technologies such as Docker and Helm.

What We Offer

  • Competitive salary and laptop
  • Professional development and training opportunities
  • Work with cutting-edge cloud and container technologies
  • Flexible work arrangements and collaborative team environment
  • Impact on organization-wide digital transformation initiatives


Remote Work :

Yes


Employment Type :

Full-time


Department / Functional Area: Software Development
Experience: years
Vacancy: 1
Create a job alert for this search

[8SN] Site Reliability Engineer (SRE) – UIUX • Montreal, Quebec, Canada

Similar jobs

Site Reliability Engineer

Vertex Elite LLCRivière-Des-Prairies-Pointe-Aux-Trembles, Canada
Full-time

Duration: ContractKey Skills:Monitoring / Observability tools - Dynatrace, ELK etc.Platform/ cloud Observability - OpenShift, Prometheus / Azure Cloud etc.Key Responsibilities:Collaborate with vari... Show more

 • Promoted • New!

Sr. Engineer

TechDoQuestmontreal (administrative region), qc, Canada
Full-time

Perform icing numerical simulations on complex aerodynamic configurations.Prepare, execute, and analyze high‑lift and icing wind tunnel test campaigns, including CFD and certification.Architect and... Show more

 • Promoted

VIA Rail Senior Reliability Engineer Role

VIA Rail Canada Inc.Montreal (administrative region), QC, CA
Full-time

Elevate your career with VIA Rail as a Senior Reliability Engineer specializing in mechanical systems and design.Contribute to the future of rail travel in Canada.In this key position, you will ens... Show more

 • Promoted

Senior Full-Stack Engineer - Accessibility & Inclusive Tech (Remote)

Accessibility Partners CanadaMontreal (administrative region), QC, CA
Remote
Full-time

A leader in accessible technology is seeking a Senior Full-Stack Developer to create equitable and accessible digital systems.This role involves developing both front-end and back-end systems, focu... Show more

 • Promoted

Site Reliability Engineer

Hunter BondMontréal, Canada
Full-time

Role: DevOps EngineerClient: Most Elite Tech Firm in CanadaCompensation: Up to $200k CAD + Bonus + PackageLocation: MontrealOverviewAn Elite FinTech Firm is looking for a highly talented DevOps Eng... Show more

 • Promoted

Intact Hybrid Site Reliability Engineer

IntactMontreal
Full-time

Join Intact as a Site Reliability Engineer and elevate operational reliability across cloud platforms.This hands-on role employs Azure, AWS, and GCP expertise to manage incidents effectively.The SR... Show more

 • Promoted

Senior Site Reliability Engineer Focused on Kubernetes Infrastructure

Chainlink LabsMontreal (administrative region), QC, CA
Full-time

Elevate decentralized architecture as a Senior Site Reliability Engineer.Spearhead Kubernetes-based infrastructure for decentralized applications, driving scalability, security, and operational eff... Show more

 • Promoted

Senior Site Reliability Engineer (Remote-First)

VySystemsMontreal (administrative region), QC, CA
Remote
Full-time

A leading technology company is seeking a Senior Site Reliability Engineer with robust Kubernetes knowledge to work remotely.Ideal candidates have over 6 years of experience in IT disciplines, prof... Show more

 • Promoted

Site Reliability Engineer

ApTaskMontréal, Canada
Full-time

Direct message the job poster from ApTask Looking for an intermediate between 2 to 5 years' experience.The Application Infrastructure (Al) department is seeking a Site Reliability Engineer (SRE... Show more

 • Promoted

Senior Ii Site Reliability Engineer

Akamai TechnologiesRivière-Des-Prairies-Pointe-Aux-Trembles, Canada
Full-time

Job DescriptionJoin our SRE team! Our team uses large datasets to analyze and measure the performance and reliability of our platform.We are networking data scientists: we combine our knowledge of ... Show more

 • Promoted • New!

Senior Site Reliability Engineer

ThinkificMontreal (administrative region), QC, CA
Full-time

Senior Site Reliability Engineer.Senior Site Reliability Engineer.Are you an experienced Site Reliability Engineer looking for a new challenge?.Senior Site Reliability Engineer.Senior Site Reliabil... Show more

 • Promoted

Site Reliability Engineer

BasetenMontréal, Canada
Full-time

About BasetenBaseten powers mission‐critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer.By uniting applied AI resear... Show more

 • Promoted

SAR Payload Systems Engineering Advisor

Sky Systems, Inc. (SkySys)montreal (administrative region), qc, Canada
Full-time

SAR Payload Systems Engineering Advisor.Duration: 12 months – 40 hours per week.Location: Montreal – West Island, 4 days per week in office.Salary: CAD $125K - $140K Annually with Standard Benefits... Show more

 • Promoted

Site Reliability Engineer

TELUS DigitalMontreal (administrative region), QC, CA
Full-time

Welcome to TELUS Digital — where innovation drives impact at a global scale.As an award-winning digital product consultancy and the digital division of TELUS, one of Canada’s largest telecommunicat... Show more

 • Promoted

Senior Site Reliability Engineer- Remote

ClickHouseMontreal (administrative region), QC, CA
Remote
Full-time

Senior Site Reliability Engineer- Remote.Recognized on the 2025 Forbes Cloud 100 list, ClickHouse is one of the most innovative and fast-growing private cloud companies.With more than 3,000 custome... Show more

 • Promoted

Engineering Sr. Advisor_1ENGEK

TRACTEBELmontreal (administrative region), qc, Canada
Full-time

TRACTEBEL, ranked 1st in Hydropower design & 4th in Nuclear design by Engineering News-Record’s (ENR) 2025 annual ranking, is a global engineering and consulting company dedicated to engineering a ... Show more

 • Promoted

Senior Infrastructure Reliability Engineer

ShippoMontreal (administrative region), QC, CA
Full-time

Enhance shipping solutions as a Senior Site Reliability Engineer in a remote setting.Focus on infrastructure integrity, scalability, and performance in a collaborative environment.This position inv... Show more

 • Promoted

Senior Site Reliability Engineer

SecurityScorecardmontreal (administrative region), qc, Canada
Full-time

SecurityScorecard is the global leader in cybersecurity ratings, with over 12 million companies continuously rated, operating in 64 countries.Founded in 2013 by security and risk experts Dr.Alex Ya... Show more

 • Promoted

Site Reliability Engineer

MaintainXMontreal
Full-time

MaintainX is the world's leading AI-powered maintenance and asset management platform, serving 13,000+ customers including Duracell, Shell, Cintas, and Brenntag.We raised $150M in Series D funding ... Show more

 • Promoted

Site Reliability Engineer (Linux / Cloud Infrastructure)

Atlantis IT GroupMontréal, Quebec, Canada
Full-time

Site Reliability Engineer (Linux / Cloud Infrastructure) role with hands-on experience across Linux, distributed systems, scripting, databases, monitoring, containers, cloud SaaS integrations, mess... Show more