Talent.com
Tata Consultancy Services
SRE Observability EngineerTata Consultancy Services • Toronto, ON, Canada
SRE Observability Engineer

SRE Observability Engineer

Tata Consultancy Services • Toronto, ON, Canada
30+ days ago
Salary
CA$90,000.00 yearly
Job type
  • Full-time
Job description

Tata Consultancy Services (TCS) is an equal opportunity employer, and embraces diversity in race, nationality, ethnicity, gender, age, physical ability, neurodiversity, and sexual orientation, to create a workforce that reflects the societies we operate in. Our continued commitment to Culture and Diversity is reflected in our people stories across our workforce and implemented through equitable workplace policies and processes.

About TCS

TCS is an IT services, consulting, and business solutions organization that has been partnering with many of the world’s largest businesses in their transformation journeys for over 55 years. Its consulting‑led, cognitive‑powered portfolio of business, technology, and engineering services and solutions is delivered through its unique Location Independent Agile delivery model, recognized as a benchmark of excellence in software development. A part of the Tata group, India's largest multinational business group, TCS operates in 55 countries and employs over 607,000 highly skilled individuals, including more than 10,000 in Canada. The company generated consolidated revenues of US $30 billion in the fiscal year ended March 31, 2025, and is listed on the BSE and the NSE in India. TCS’ proactive stance on climate change and award‑winning work with communities across the world have earned it a place in leading sustainability indices such as the MSCI Global Sustainability Index and the FTSE4Good Emerging Index.

Required Skill Set

  • We are looking for a Mid‑Level Observability Engineer to help implement, operate, and improve observability capabilities across our applications and platforms.
  • This role focuses on hands‑on onboarding, instrumentation, dashboarding, and alerting, working under established standards and guidance from senior engineers.
  • You will collaborate with application, SRE, and operations teams to ensure systems are observable, supportable, and production ready.
  • Observability Implementation: Implement and maintain metrics, logs, and traces for applications and infrastructure.
  • Assist with onboarding applications into observability platforms (e.g., Dynatrace, ELK, Datadog).
  • Configure dashboards, alerts, and basic anomaly detection.
  • Work with development teams to enable structured logging, basic distributed tracing, and core metrics.
  • Validate observability requirements during Production Readiness Reviews (PRR).
  • Troubleshoot missing or low‑quality telemetry.
  • Configure alerts based on golden signals (latency, errors, traffic, saturation).
  • Help reduce alert noise by tuning thresholds and alert logic.
  • Support incident response by gathering logs, metrics, and traces; perform root‑cause analysis using observability tools.
  • Maintain dashboards and documentation used by on‑call and support teams.
  • Participate in on‑call rotations (as applicable).
  • Automation / Continuous Improvement: Assist in automating observability onboarding and validation tasks.
  • Create and maintain reusable dashboards and alert templates.
  • Follow established observability standards and best practices.

Required Qualifications

  • Good years of experience in Observability or SRE.
  • Working knowledge of metrics, logs, and basic tracing concepts.
  • Hands‑on experience with at least one observability platform (Dynatrace, Elastic ELK, Datadog, New Relic, etc.).
  • Basic understanding of SLIs, SLOs, and service health indicators.
  • Experience with cloud platforms or hybrid environments.
  • Ability to write scripts (Python, Bash, PowerShell) for automation and troubleshooting.

Preferred Qualifications

  • Experience with OpenTelemetry or APM agents.
  • Familiarity with Kubernetes or containerized workloads.
  • Experience working with incident management tools (PagerDuty, ServiceNow).
  • Exposure to Dynatrace, Kibana, ELK, or similar cloud‑native monitoring.
  • Experience in regulated or enterprise environments.

Salary Range - CA$ 90,000 - CA$ 120,000 Per Year

Accessibility Statement

Tata Consultancy Services Canada Inc. is committed to meeting the accessibility needs of all individuals in accordance with the Accessibility for Ontarians with Disabilities Act (AODA) and the Ontario Human Rights Code (OHRC). Should you require accommodation during the recruitment and selection process, please inform Human Resources.

#J-18808-Ljbffr

Create a job alert for this search

SRE Observability Engineer • Toronto, ON, Canada

Similar jobs

Observability Engineer (Sre) – Opentelemetry Platform - C$110,000 - C$130,000 A Year

Major Online Travel PlatformToronto County, Canada
Full-time

Develops and enhances observability solutions for a major online travel platform, focusing on system reliability and monitoring standards.Requires SRE/DevOps experience and APM tool skills. Show more

 • Promoted

Senior SRE Leader: Scale Reliability & Observability

RootlyToronto, ON, CA
Full-time

A fast-growing tech startup in Toronto is seeking an experienced Site Reliability Engineer.The role involves enhancing service performance, owning CI/CD pipelines, and building automation tools.Ide... Show more

 • Promoted

Observability Lead Software Engineer Role

WaabiToronto, ON, CA
Full-time

Take the lead in observability engineering at Waabi as a Software Engineer focused on SRE practices.Design and optimize systems that ensure the health of autonomous technology solutions.Waabi is at... Show more

 • Promoted

Sr. Solutions Engineer

Menlo VenturesToronto, ON, CA
Full-time

We are looking for pre-sales professionals who have a successful track record helping large enterprises become more data-driven.Working with the Enterprise Account Executive (AE), the Sr.Solutions ... Show more

 • Promoted

S0i3/PMM Emulation Engineer

TekWissen ®Markham, York Region, CA
Full-time

Be among the first 25 applicants.This range is provided by TekWissen ®.Your actual pay will be based on your skills and experience — talk with your recruiter to learn more.TekWissen is a global wor... Show more

 • Promoted

Dynatrace SRE — End-to-End Observability (Hybrid)

DexianToronto, ON, CA
Full-time

A leading staffing and IT solutions provider in Toronto is seeking a Site Reliability Engineer with strong Dynatrace expertise.This role focuses on ensuring the reliability, performance, and observ... Show more

 • Promoted

Senior Satellite Systems Designer - Payload Lead

Kepler Communications Inc.Toronto, ON, CA
Full-time

At Kepler Communications, we're not just imagining the future of on-demand space connectivity - we're leading it!.Our mission is to provide real-time Internet access for space-based assets, enablin... Show more

 • Promoted

Sr Engineer, End User Services

MonerisToronto, ON, CA
Full-time

As a Senior Engineer, End User Services, you will lead the strategy and execution of Moneris’ modern endpoint management and end‑user enablement technologies.You’ll own device standards, automation... Show more

 • Promoted

Senior SRE/DevOps Engineer - Kubernetes & Observability

Infotek Consulting Inc.Toronto, ON, CA
Full-time

A consulting firm in Canada is seeking a skilled Site Reliability / DevOps Engineer for a contract role in Toronto.The ideal candidate will have over 10 years of experience in SRE/DevOps, strong ex... Show more

 • Promoted

Remote Aerospace AI Engineer — Train & Improve Models

DataAnnotationMarkham, York Region, CA
Remote
Full-time

An innovative company is seeking a skilled Aerospace Engineer to enhance AI models through the application of physics.This role allows you to work remotely, choose your projects, and set your own s... Show more

 • Promoted

Senior SRE

ViafouraToronto, ON, CA
Full-time

Senior Site Reliability Engineer.Viafoura is a leading audience engagement platform that powers real-time conversations and community experiences for digital publishers and brands worldwide.We're s... Show more

 • Promoted

SRE Developer I — Flexible, Remote/Hybrid, Observability Focus

Vena SolutionsToronto, ON, CA
Remote
Full-time

A tech-driven solutions provider is seeking a Site Reliability Developer in Toronto.This flexible position allows for in-office, hybrid, or remote working.Responsibilities include supporting IT pro... Show more

 • Promoted

Observability Engineer — SRE for Metrics, Logs & Traces

Tata Consultancy ServicesToronto, ON, CA
Full-time

A leading IT services firm is seeking a Mid-Level Observability Engineer in Toronto, Ontario, to help improve observability capabilities across applications and platforms.The role involves working ... Show more

 • Promoted

Senior SRE Focused on Automation and Cloud

Morningstar Credit Ratings, LLCToronto, ON, CA
Full-time

Explore a Senior Site Reliability Engineer role with Morningstar in Toronto, ON, emphasizing AWS and CI/CD automation.This hybrid position is key to ensuring robust investment data operations.You w... Show more

 • Promoted

SRE & Production Reliability Engineer — Hybrid

Tangerine BankToronto, ON, CA
Full-time

A leading digital bank in Toronto is seeking a qualified SRE & Production Support professional to enhance their technology solutions.You will manage team workflows, ensure timely resolution of prod... Show more

 • Promoted

RQ07954 - Sr. Solutions Designer

Rubicon PathToronto, ON, CA
Full-time

Undertakes the design of hosting technology solutions based on the clients service specifications, standards, policies, best practices and cost models, in order to meet client application business ... Show more

 • Promoted

Research Engineer - Agentic Software Systems Engineering

Huawei Technologies Canada Co., Ltd.Markham, ON, CA
Permanent

Huawei Canada has an immediate permanent opening for a Research Engineer.The Intelligent Complex Systems Team, currently a part of the Waterloo Research Centre, examines recent advancements in arti... Show more

 • Promoted

RQ10771 - Sr. Solutions Designer

Source CodeToronto, ON, CA
Full-time

Design, develop and enhance large scale software systems using RESTful & Micro Services based architecture and design.Design containerized based solutions/architecture.Prepare and continuously enha... Show more

 • Promoted

IBM Site Reliability Engineering Expert

LeadingtalentMarkham, ON, CA
Full-time

Step into a career as a Site Reliability Engineer at IBM, focused on enhancing system reliability and performance.Engage directly with production systems and optimize customer experience.In this ro... Show more

 • Promoted

Senior Site Reliability Engineer — Kubernetes, AWS & Observability

ThinkificToronto, ON, CA
Full-time

A leading e-learning provider in Canada is seeking a Senior Site Reliability Engineer to enhance and secure their infrastructure supporting online course creators.This role involves improving perfo... Show more