Talent.com
SAP
DevOps Engineer - BTP Site Reliability Engineering teamSAP • Montreal, Quebec, Canada
DevOps Engineer - BTP Site Reliability Engineering team

DevOps Engineer - BTP Site Reliability Engineering team

SAP • Montreal, Quebec, Canada
21 days ago
Job type
  • Full-time
  • Permanent
Job description

We help the world run better
At SAP, we keep it simple: you bring your best to us, and we'll bring out the best in you. We're builders touching over 20 industries and 80% of global commerce, and we need your unique talents to help shape what's next. The work is challenging - but it matters. You'll find a place where you can be yourself, prioritize your wellbeing, and truly belong. What's in it for you? Constant learning, skill growth, great benefits, and a team that wants you to grow and succeed.

Important information:

  • This is a hybrid role based out of SAP Montreal office, working in-office with the team 3 days per week.
  • Candidates must be legally entitled to work in Canada at the time of application. This position is not eligible for employer-sponsored work authorization (e.g., LMIA or other immigration support).

What you'll do

We are looking for an engineer to join an already established SRE team for the SAP Business AI Platform.

As a Site Reliability Engineer, you will have the opportunity to operate and support business critical Cloud services. As part of your daily job, you will proactively monitor the service behavior and identify areas for improvement. You will participate in the development of tools for monitoring and troubleshooting cloud services built on latest open source and SAP technologies, following SRE principles.

Responsibilities:

  • Act as technical expert during Live site incidents (downtimes of supported services in scope), investigate and solve incidents on a deep technical level.
  • Drive root cause analysis and follow-up improvements to prevent issues from reoccurring.
  • Perform in-depth troubleshooting and log analysis to identify and solve complex issues in accordance with internal and external SLAs.
  • Build software-based solutions to address improvements in service reliability and stability.
  • Enhance infrastructure and platform monitoring by gathering system metrics (4 Golden Signals) and implementing tools for recovery.
  • Integrate and collaborate closely with development teams and work with them on outputs from Postmortems and product improvements.
  • Learn new technologies and keep up to date with latest development increments.
  • Create and maintain technical documentation.
  • Define, advocate, apply SRE best practices.
  • Participate in the on-call rotation (follow the sun approach) to react to major incidents. On-call has a special compensation package.

What you bring

  • Experience with Kubernetes and good understanding of container technologies.
  • Understanding of modern cloud architectures (experience with Cloud Platforms such as AWS, Azure, GCP are a plus).
  • Experience with Unix/Linux operating system
  • Scripting skills, CI/CD (ArgoCD, Concourse, Github Actions and are a plus) - enthusiasm for automation - make the computers do the work for you.
  • Experience using AI-assisted engineering tools (e.g., Claude Code CLI, GitHub Copilot, or similar) to improve troubleshooting, automation, documentation, root cause analysis, and operational efficiency.
  • 2+ years experience in SRE.
  • Working efficiently in emergency situations. Affinity to quickly analyze and solve problems in a global team setup.
  • Excellent team player, passionate about his/her work, self-motivated and driven.
  • Excellent communication skills - precise, based on facts.
  • Fluency in English.
  • Preferred Additional Skills and Competencies:
    • Coding experience with Python, GO, Bash
    • CKA/CKAD/CKS certifications
    • Experience with modern monitoring, logging, and alerting tools (Grafana, Prometheus, Kibana, Loki, Splunk On-Call, Dynatrace)
    • Security best practices for application development and operations in a public Cloud Environment
    • Contribution to open-source projects

Meet the team

The Reliability Engineering organization provides multitude of products and services related to operations and continuity of business delivery.

The Site Reliability Engineering teams make the SAP Business AI Platform run better by providing 24x7 deep technical coverage for Incident Management (Outages and other incidents with major customer impact) applying SRE principles. We share a "Live-Site" First culture and care for the business continuity of our customers running mission critical applications in the Cloud.

#LI-GL1

Bring out your best
SAP innovations help more than four hundred thousand customers worldwide work together more efficiently and use business insight more effectively. Originally known for leadership in enterprise resource planning (ERP) software, SAP has evolved to become a market leader in end-to-end business application software and related services for database, analytics, intelligent technologies, and experience management. As a cloud company with two hundred million users and more than one hundred thousand employees worldwide, we are purpose-driven and future-focused, with a highly collaborative team ethic and commitment to personal development. Whether connecting global industries, people, or platforms, we help ensure every challenge gets the solution it deserves. At SAP, you can bring out your best.

We win with inclusion
SAP's culture of inclusion, focus on health and well-being, and flexible working models help ensure that everyone - regardless of background - feels included and can run at their best. At SAP, we believe we are made stronger by the unique capabilities and qualities that each person brings to our company, and we invest in our employees to inspire confidence and help everyone realize their full potential. We ultimately believe in unleashing all talent and creating a better world.

SAP is committed to the values of Equal Employment Opportunity and provides accessibility accommodations to applicants with physical and/or mental disabilities. If you are interested in applying for employment with SAP and are in need of accommodation or special assistance to navigate our website or to complete your application, please send an e-mail with your request to Recruiting Operations Team: Careers@sap.com.

For SAP employees: Only permanent roles are eligible for the SAP Employee Referral Program , according to the eligibility rules set in the SAP Referral Policy. Specific conditions may apply for roles in Vocational Training.

Qualified applicants will receive consideration for employment without regard to their age, race, religion, national origin, ethnicity, gender (including pregnancy, childbirth, et al), sexual orientation, gender identity or expression, protected veteran status, or disability, in compliance with applicable federal, state, and local legal requirements.

SAP believes the value of pay transparency contributes towards an honest and supportive culture and is a significant step toward demonstrating SAP's commitment to pay equity. SAP provides the annualized compensation range inclusive of base salary and variable incentive target for the career level applicable to the posted role. The targeted combined range for this position is 97,800 - 166,200 . The actual amount to be offered to the successful candidate will be within that range, dependent upon the key aspects of each case which may include education, skills, experience, scope of the role, location, etc. as determined through the selection process. Any SAP variable incentive includes a targeted dollar amount, and any actual payout amount is dependent on company and personal performance. A summary of benefits and eligibility requirements can be found by clicking this link: www.SAPNorthAmericaBenefits.com.

Due to the nature of the role, which involves global interactions with SAP entities, as well as with employees and stakeholders in Canada, functional proficiency in English is required for positions based in the Quebec.

AI Usage in the Recruitment Process

For information on the responsible use of AI in our recruitment process, please refer to our Guidelines for Ethical Usage of AI in the Recruiting Process .

Please note that any violation of these guidelines may result in disqualification from the hiring process.

Requisition ID: 458700 | Work Area: Software-Development Operations | Expected Travel: 0 - 10% | Career Status: Professional | Employment Type: Regular Full Time | Additional Locations: #LI-Hybrid

Requisition ID: 458700

Posted Date: Aug 14, 2026

Work Area: Software-Development Operations

Career Status: Professional

Employment Type: Regular Full Time

Expected Travel: 0 - 10%

Location:
Montreal, Quebec, CA, H3B 0B3

Create a job alert for this search

DevOps Engineer - BTP Site Reliability Engineering team • Montreal, Quebec, Canada

Similar jobs

Senior Site Reliability Engineer — Kubernetes, AWS & Observability

ThinkificMontreal (administrative region), QC, CA
Full-time

A leading e-learning provider in Canada is seeking a Senior Site Reliability Engineer to enhance and secure their infrastructure supporting online course creators.This role involves improving perfo... Show more

 • Promoted

Site Reliability Engineer

ApTaskMontréal, Quebec, Canada
Full-time

Direct message the job poster from ApTask Looking for an intermediate between 2 to 5 years' experience.The Application Infrastructure (Al) department is seeking a Site Reliability Engineer (SRE) to... Show more

 • Promoted

SRE - Reliable Infrastructure (Hybrid/Remote, Montreal)

Hunter Bondmontreal (administrative region), qc, Canada
Remote
Full-time

A leading global quant fund in Montreal is seeking a skilled DevOps Engineer to enhance the reliability of critical services.You will be part of a dynamic team responsible for maintaining essential... Show more

 • Promoted

Experienced Site Reliability Engineer Remote

Tecsys Inc.Montreal (administrative region), QC, CA
Remote
Full-time

Join Tecsys as an Experienced Site Reliability Engineer and elevate our cloud infrastructure reliability.Work remotely and focus on automation and system health.At Tecsys, we are searching for a Si... Show more

 • Promoted

Site Reliability Engineer

Tecsys Inc.Montreal (administrative region), QC, CA
Permanent

Having recognized the advantages of remote work, including employee morale, productivity, reduced commuting on employee wellbeing and the environment, we are proud to be a digital-first company.The... Show more

 • Promoted

Senior Site Reliability Engineer (Remote-First)

VySystemsMontreal (administrative region), QC, CA
Remote
Full-time

A leading technology company is seeking a Senior Site Reliability Engineer with robust Kubernetes knowledge to work remotely.Ideal candidates have over 6 years of experience in IT disciplines, prof... Show more

 • Promoted

DevOps Engineer - AWS, CI/CD & Reliability Focus

NewtonMontreal (administrative region), QC, CA
Full-time

A leading cryptocurrency firm in Canada is seeking a DevOps Engineer to improve CI/CD workflows and manage infrastructure.The ideal candidate will have experience with AWS, automation, and operatio... Show more

 • Promoted

Elite DevOps & SRE Engineer (Hybrid) — Montreal

Hunter BondMontreal (administrative region), QC, CA
Full-time

A leading technology company in Montreal is seeking a DevOps/Site Reliability Engineer to join an elite infrastructure team.This role offers the opportunity for hands-on engineering, focusing on la... Show more

 • Promoted

Low-Latency Platform Engineer | Trading DevOps (Hybrid)

Hunter Bondmontreal (administrative region), qc, Canada
Full-time

A leading recruitment agency is looking for a Team Leader – Infrastructure in Montreal for a hybrid role.The ideal candidate will be experienced in low latency Linux environments and possess skills... Show more

 • Promoted

DevOps Engineer

Insight Globalmontreal, montreal (administrative region), Canada
Full-time

Get AI-powered advice on this job and more exclusive features.This range is provided by Insight Global.Your actual pay will be based on your skills and experience — talk with your recruiter to lear... Show more

 • Promoted

Innovative Platform Engineer Combines DevOps and Linux Expertise

Hunter Bondmontreal (administrative region), qc, Canada
Full-time

Unlock your potential as a Platform Engineer in a hybrid workspace! Design and implement cutting-edge automated solutions for a low latency environment using Docker, Ansible, and CI/CD.This positio... Show more

 • Promoted

Site Reliability Engineer (Linux / Cloud Infrastructure)

Atlantis IT GroupMontreal, Montreal (administrative region), CA
Full-time

Site Reliability Engineer (Linux / Cloud Infrastructure) role with hands-on experience across Linux, distributed systems, scripting, databases, monitoring, containers, cloud SaaS integrations, mess... Show more

 • Promoted

Site Reliability Engineer

BasetenMontréal, Canada
Full-time

About Baseten Baseten powers mission‑critical inference for the world’s most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer.By uniting applied AI research,... Show more

 • Promoted

DevOps Engineer

VySystemsMontreal (administrative region), QC, CA
Full-time

Senior Site Reliability Engineer (Remote First).Direct message the job poster from VySystems.IT disciplines, including technical architecture, network management, application development, middlewar... Show more

 • Promoted

DevOps/Site Reliability Engineer - Up to $200k CAD + Bonus - Elite Tech Firm

Hunter BondMontreal (administrative region), QC, CA
Full-time

DevOps/Site Reliability Engineer.Most Elite Tech Firm in Canada.Up to $200k CAD + Bonus + Full Package.One of Canada’s most elite tech firms is hiring a Site Reliability Engineer to join a seriousl... Show more

 • Promoted

Lead Platform Engineer Enhancing DevOps and System Reliability

Lillio (formerly HiMama)Montreal (administrative region), QC, CA
Full-time

Transform early childhood education as a Senior Platform Engineer focused on system performance and collaborative tooling.Drive key initiatives for scalable, reliable digital platforms.In this pivo... Show more

 • Promoted

Site Reliability Engineer

TELUS DigitalMontreal (administrative region), QC, CA
Full-time

Welcome to TELUS Digital — where innovation drives impact at a global scale.As an award-winning digital product consultancy and the digital division of TELUS, one of Canada’s largest telecommunicat... Show more

 • Promoted

Senior Site Reliability Engineer- Remote

ClickHouseMontreal (administrative region), QC, CA
Remote
Full-time

Senior Site Reliability Engineer- Remote.Recognized on the 2025 Forbes Cloud 100 list, ClickHouse is one of the most innovative and fast-growing private cloud companies.With more than 3,000 custome... Show more

 • Promoted

Platform Engineer / DevOps Engineer - Trading - $130,000-$250,000 CAD + Bonus

Hunter Bondmontreal (administrative region), qc, Canada
Full-time

My client are seeking a knowledgeable Platform Engineer to work on their low latency Linux estate.The role sits between Platform Engineering, DevOps, Linux Systems Administration and SRE, and incor... Show more

 • Promoted

Remote Site Reliability Engineer - Scale Crypto Systems

NewtonMontreal (administrative region), QC, CA
Remote
Full-time

A leading innovative tech company in Toronto is looking for a Site Reliability Engineer.In this pivotal role, you will enhance the reliability and resilience of critical services, manage incidents,... Show more