Talent.com
Atlantis IT Group
Site Reliability Engineer (Linux / Cloud Infrastructure)Atlantis IT Group • Montreal, Montreal (administrative region), CA
Site Reliability Engineer (Linux / Cloud Infrastructure)

Site Reliability Engineer (Linux / Cloud Infrastructure)

Atlantis IT Group • Montreal, Montreal (administrative region), CA
30+ days ago
Job type
  • Full-time
Job description

Overview


Site Reliability Engineer (Linux / Cloud Infrastructure) role with hands-on experience across Linux, distributed systems, scripting, databases, monitoring, containers, cloud SaaS integrations, messaging, load balancers, security, and incident management.


Responsibilities



  • Provide hands-on administration of Linux 7.x and related infrastructure.

  • Work with Service Oriented Architecture, distributed systems, and scripting (Python, shell).

  • Manage relational databases (e.g., Sybase, DB2, SQL, Postgres) and application integration, configuration, and troubleshooting.

  • Operate observability and monitoring tools: Open Telemetry, Prometheus, Grafana, Splunk, Ansible.

  • Manage web servers (Apache, Nginx) and application servers (Tomcat, JBoss) for integration and troubleshooting.

  • Work with Docker containers, Kubernetes, and SaaS platform integration.

  • Understand messaging systems (e.g., Kafka) and their role in the architecture.

  • Design and implement load balancing, web proxies, and storage platforms (NAS/SAN) from an implementation perspective.

  • Apply basic security policies for secure hosting solutions, including Kerberos and encryption methods (SSL/TLS).

  • Experience in managing large web-based, multi-tier (n-tier) applications in secure cloud environments.

  • Apply SRE principles with appropriate tooling approach; strong Linux/Unix admin, storage, networking, and web technologies knowledge.

  • Troubleshoot application issues and manage incidents effectively.

  • Exhibit excellent verbal and written communication skills.


Qualifications



  • Hands-on experience with Linux 7.x operating system (5+ years) at an advanced level.

  • Hands-on experience with SOA, distributed systems, and scripting (Python, shell).

  • Experience with relational databases (Sybase, DB2, SQL, Postgres).

  • Exposure to tools: Open Telemetry, Prometheus, Grafana, Splunk, Ansible.

  • Hands-on experience with web servers (Apache, Nginx) and application servers (Tomcat, JBoss).

  • Experience with Docker, Kubernetes, and SaaS platform integration.

  • Experience with Kafka and messaging technologies.

  • Understanding of load balancers, web proxies, and NAS/SAN storage from an implementation perspective.

  • Familiar with security policies for secure hosting, Kerberos, SSL/TLS.

  • Experience managing large web-based n-tier applications in secure cloud environments.

  • Strong knowledge of SRE principles and tooling.

  • Strong infrastructure knowledge in Linux/Unix administration, storage, networking, and web technologies.

  • Excellent troubleshooting and incident management capabilities.


Senioriry level


Mid-Senior level


Employment type


Contract


Job function


Information Technology


Industries


IT Services and IT Consulting

#J-18808-Ljbffr
Create a job alert for this search

Site Reliability Engineer (Linux / Cloud Infrastructure) • Montreal, Montreal (administrative region), CA

Similar jobs

Senior Site Reliability Engineer — Kubernetes, AWS & Observability

ThinkificMontreal (administrative region), QC, CA
Full-time

A leading e-learning provider in Canada is seeking a Senior Site Reliability Engineer to enhance and secure their infrastructure supporting online course creators.This role involves improving perfo... Show more

 • Promoted

Site Reliability Engineer

ApTaskMontréal, Quebec, Canada
Full-time

Direct message the job poster from ApTask Looking for an intermediate between 2 to 5 years' experience.The Application Infrastructure (Al) department is seeking a Site Reliability Engineer (SRE) to... Show more

 • Promoted

Site Reliability Engineer

Vertex Elite LLCRivière-Des-Prairies-Pointe-Aux-Trembles, Canada
Full-time

Duration: ContractKey Skills:Monitoring / Observability tools - Dynatrace, ELK etc.Platform/ cloud Observability - OpenShift, Prometheus / Azure Cloud etc.Key Responsibilities:Collaborate with vari... Show more

 • Promoted

Open-Source Container & Virtualization Engineer (Remote)

CanonicalMontreal (administrative region), QC, CA
Remote
Full-time

A leading open source software firm in Canada is seeking exceptional software engineers specializing in virtualisation and container technology.The role emphasizes working in Go, Rust, or C/C++, wi... Show more

 • Promoted

Senior Site Reliability Engineer (Remote-First)

VySystemsMontreal (administrative region), QC, CA
Remote
Full-time

A leading technology company is seeking a Senior Site Reliability Engineer with robust Kubernetes knowledge to work remotely.Ideal candidates have over 6 years of experience in IT disciplines, prof... Show more

 • Promoted

Senior Site Reliability Engineer Focused on Kubernetes Infrastructure

Chainlink LabsMontreal (administrative region), QC, CA
Full-time

Elevate decentralized architecture as a Senior Site Reliability Engineer.Spearhead Kubernetes-based infrastructure for decentralized applications, driving scalability, security, and operational eff... Show more

 • Promoted

Lead Software Engineer in Cloud Solutions

Autodeskmontreal (administrative region), qc, Canada
Full-time

Join Autodesk as a Lead Software Engineer focusing on Cloud Solutions.Innovate and drive product strategies while enhancing your leadership skills in a collaborative setting.In this role, you will ... Show more

 • Promoted

Remote Platform Engineer — Cloud & Kubernetes Ops

PlanetMontreal (administrative region), QC, CA
Remote
Full-time

A leading global space and data company is seeking a Software Engineer in Platform Operations.This full-time remote role prioritizes building and operating cloud infrastructure supporting engineeri... Show more

 • Promoted

Site Reliability Engineer

Hunter Bondmontreal (administrative region), qc, Canada
Full-time

Most Elite Tech Firm in Canada.Up to $200k CAD + Bonus + Package.An Elite FinTech Firm is looking for a highly talented DevOps Engineer/Systems SRE to join a talented flat-structured team within a ... Show more

 • Promoted

Specialist Site Reliability Engineer

Global Talent Alliance, CanadaMontreal (administrative region), QC, CA
Full-time

About the job Specialist Site Reliability Engineer.The role of the Specialist Site Reliability Engineer (SRE) is to execute RAM analysis and engineering in support of the I&T solutions.The overall ... Show more

 • Promoted

Site Reliability Engineer

TELUS DigitalMontreal (administrative region), QC, CA
Full-time

Welcome to TELUS Digital — where innovation drives impact at a global scale.As an award-winning digital product consultancy and the digital division of TELUS, one of Canada’s largest telecommunicat... Show more

 • Promoted

Senior Site Reliability Engineer- Remote

ClickHouseMontreal (administrative region), QC, CA
Remote
Full-time

Senior Site Reliability Engineer- Remote.Recognized on the 2025 Forbes Cloud 100 list, ClickHouse is one of the most innovative and fast-growing private cloud companies.With more than 3,000 custome... Show more

 • Promoted

Lead Platform Engineer Enhancing DevOps and System Reliability

Lillio (formerly HiMama)Montreal (administrative region), QC, CA
Full-time

Transform early childhood education as a Senior Platform Engineer focused on system performance and collaborative tooling.Drive key initiatives for scalable, reliable digital platforms.In this pivo... Show more

 • Promoted

Senior Site Reliability Engineer

SecurityScorecardmontreal (administrative region), qc, Canada
Full-time

SecurityScorecard is the global leader in cybersecurity ratings, with over 12 million companies continuously rated, operating in 64 countries.Founded in 2013 by security and risk experts Dr.Alex Ya... Show more

 • Promoted

Intact Hybrid Site Reliability Engineer

IntactMontreal (administrative region), QC, CA
Full-time

Join Intact as a Site Reliability Engineer and elevate operational reliability across cloud platforms.This hands-on role employs Azure, AWS, and GCP expertise to manage incidents effectively.The SR... Show more

 • Promoted

Lasso Informatics Infrastructure Engineer

Lasso InformaticsMontreal (administrative region), QC, CA
Full-time

Join Lasso Informatics as a Senior Cloud Infrastructure Engineer to innovate data management solutions in cloud environments.Oversee AWS and GCP infrastructures with a focus on security and scalabi... Show more

 • Promoted

Senior DevOps Engineer - AWS (Remote)

LumenaltaMontreal (administrative region), QC, CA
Remote
Full-time

At Lumenalta, we create impactful software solutions that drive innovation and transform businesses.Since 2000, we’ve partnered with visionary leaders to build cutting‑edge tech, solve complex chal... Show more

 • Promoted

Site Reliability Engineer

MaintainXMontreal
Full-time

MaintainX is the world's leading AI-powered maintenance and asset management platform, serving 13,000+ customers including Duracell, Shell, Cintas, and Brenntag.We raised $150M in Series D funding ... Show more

 • Promoted

Lead AWS Cloud DevOps Engineer (Remote - Namer)

JobgetherMontreal (administrative region), QC, CA
Remote
Full-time

Lead AWS Cloud DevOps Engineer.Location: North America (Remote).As a Lead AWS Cloud DevOps Engineer, you will oversee and build cloud infrastructure that powers mission‑critical platforms, ensuring... Show more

 • Promoted

Senior Cloud & DevOps Engineer - Remote | Unlimited PTO

Lazer TechnologiesMontreal (administrative region), QC, CA
Remote
Full-time

A world-class digital product studio is seeking a Senior Infrastructure/DevOps Engineer to support a remote-first team.The ideal candidate will have over 5 years of experience, mastery in Docker an... Show more