Talent.com
Amazon.com.ca, ULC
Senior Performance Engineer, Efficiency Red TeamAmazon.com.ca, ULC • Toronto, Ontario, CAN
Senior Performance Engineer, Efficiency Red Team

Senior Performance Engineer, Efficiency Red Team

Amazon.com.ca, ULC • Toronto, Ontario, CAN
11 hours ago
Job type
  • Full-time
Job description

Are you the kind of engineer who obsesses over performance? Does the idea of hunting for hidden waste across the world's largest infrastructure and turning those findings into hundreds of millions of dollars in freed capacity sound like the most exciting job you can imagine? If so, keep reading.

This role offers something rare. Deep technical research freedom, immediate access to the largest compute infrastructure in the world, and direct line of sight to impact that reshapes how Amazon builds and operates its infrastructure. The Efficiency Red Team has already driven billions in cumulative capacity savings, reclaimed petabytes of DRAM through automated profiling, and built GPU efficiency frameworks adopted across Amazon. And we are just getting started. The surface area of opportunity grows with every new workload, every new chip generation, and every new AI model Amazon deploys. You would be joining a team with a proven foundation and an expanding frontier.

Compute demand is rising with no sign of slowing down. The explosive growth of Generative AI has created unprecedented demand for GPUs, and the shock-waves are now constraining DRAM availability for traditional compute across the entire industry. Silicon, power, and physical space are not infinite. When capacity is constrained, efficiency becomes the new capacity. At Amazon's scale, a single efficiency pattern discovered and applied can unlock resources equivalent to building entirely new data centers. A modest reduction in memory footprint across a widely deployed library translates into millions of dollars in freed capacity. An optimization that improves response latency by milliseconds across billions of requests directly improves the experience of hundreds of millions of customers and drives business growth.

We are looking for experienced performance engineers to join a focused group whose mission is discovering and eliminating waste across every layer of the compute stack. You will work at the intersection of hardware and software, from GPU memory management in generative AI inference pipelines to DRAM allocation patterns deep inside language runtime, from storage subsystem inefficiencies to CPU scheduling behaviors at hyper-scale.

Key job responsibilities
- Apply first-principles analysis across CPU, DRAM, GPU, storage, I/O, and networking layers to uncover optimization opportunities that others overlook
- Design and lead rigorous investigations that establish root causes and produce actionable findings with broad impact
- Build proof-of-concept implementations for novel efficiency approaches such as memory compression, smart over-subscription, and language-level rewrites
- Develop repeatable patterns and practices that translate individual findings into fleet-wide optimization playbooks
- Quantify business impact at Amazon scale, connecting every technical insight to capacity freed, dollars saved, and customer experience improved
- Collaborate across organizational boundaries where the largest opportunities often live at the intersection of teams, systems, and technology stacks
- Contribute findings to internal profiling tools, automated optimization agents, and engineering guidance that become force multipliers across Amazon
- Directly influence how Amazon reasons about performance engineering and capacity efficiency at scale as a founding member of the Efficiency Red Team


A day in the life
Your day could include profiling a widely deployed library to quantify its memory allocation overhead, then building a prototype that demonstrates a significant reduction in DRAM footprint and calculating the impact in freed servers and dollars. You might be deep in GPU utilization data, investigating why a generative AI training cluster shows substantial idle time during peak hours and designing an over-subscription model that safely reclaims that capacity. You could be translating complex performance findings into a clear narrative for senior leadership, showing them that a single optimization pattern applied across Amazon's fleet frees capacity worth more than most companies spend on infrastructure in a year.

Some weeks you will be reading CPU performance counters and cache miss rates. Other weeks you will be analyzing token throughput economics in large language model serving infrastructure. The variety is deliberate because waste hides everywhere, and finding it requires curiosity that refuses to stay in a single lane.

About the team
The Efficiency Red Team is a small group of performance engineers who move fluidly across the entire compute stack. The same engineer might profile CPU cache behavior in a traditional workload one week and analyze token throughput in a GPU inference cluster the next. You will have the freedom to pursue deep technical research, choose your own investigations, and follow the data wherever it leads. This is a founding opportunity to help shape the mission, methods, and culture of a team designed to be Amazon's center of excellence in performance engineering. Every pattern this team discovers feeds directly into detection and remediation tools that operate continuously across the fleet, turning individual insights into lasting, compounding impact.

BASIC QUALIFICATIONS

- 5+ years of leading design or architecture (design patterns, reliability and scaling) of new and existing systems experience
- Experience as a mentor, tech lead or leading an engineering team
- 7+ years of professional experience in systems programming, performance engineering, or infrastructure optimization
- Deep proficiency in at least one systems language (C, C++, Rust, or Java)
- Hands-on experience with performance profiling and analysis tools
- Demonstrated ability to conduct root cause analysis across multiple layers of the compute stack (CPU, memory, storage, networking)
- Track record of delivering measurable performance improvements in production systems at scale

PREFERRED QUALIFICATIONS

- Bachelor's degree in computer science or equivalent
- Experience with GPU profiling and optimization (CUDA, GPU memory management, inference pipeline tuning)
- Familiarity with generative AI infrastructure including model serving, training pipelines, and token economics
- Experience with language runtime internals (JVM garbage collection tuning, memory management design, or equivalent)
- Background in capacity planning, fleet management, or infrastructure economics at hyper-scale
- Contributions to open source performance tools or published research in systems performance

Amazon is an equal opportunity employer and does not discriminate on the basis of protected veteran status, disability, or other legally protected status.

Our inclusive culture empowers Amazonians to deliver the best results for our customers.

Create a job alert for this search

Senior Performance Engineer, Efficiency Red Team • Toronto, Ontario, CAN

Similar jobs

Senior Performance Engineer - Scale & Tuning - C$118,000 - C$161,700 A Year

Cloud Identity ProviderToronto, Canada
Full-time

Senior Performance Engineer to join a cloud identity provider in Toronto, focusing on Node.Js, Golang, and distributed systems to identify and resolve performance bottlenecks. Show more

 • Promoted

Senior Full-Stack Engineer

Chexy careerToronto, ON, CA
Full-time

At Chexy, we’re helping tenants take back control of their biggest monthly expense - rent.Chexy is the first-ever Canadian rewards program that allows renters to pay rent digitally, earn rewards on... Show more

 • Promoted

Experienced Solutions Engineer Role

CodexToronto, ON, CA
Full-time

Shape enterprise sales outcomes as a Solutions Engineer in the Greater Toronto Area.This full-time position combines technical expertise with consultative sales processes in a hybrid model.Our clie... Show more

 • Promoted

S0i3/PMM Emulation Engineer

TekWissen ®Markham, York Region, CA
Full-time

Be among the first 25 applicants.This range is provided by TekWissen ®.Your actual pay will be based on your skills and experience — talk with your recruiter to learn more.TekWissen is a global wor... Show more

 • Promoted

Performance Engineer

Tata Consultancy ServicesToronto, ON, CA
Full-time

Tata Consultancy Services (TCS) is an equal opportunity employer, and embraces diversity in race, nationality, ethnicity, gender, age, physical ability, neurodiversity, and sexual orientation, to c... Show more

 • Promoted

Senior Full-Stack Engineer: Microservices, Apis & Ci/Cd - C$94,300 - C$141,500 A Year

Leading Financial InstitutionToronto, Canada
Full-time

A leading financial institution seeks an Applications Development Senior Programmer Analyst in Mississauga, Canada.The candidate will lead software development phases from design to support, focusi... Show more

 • Promoted

Empowered Senior Full-Stack Engineer Driving Healthcare Innovation

NimbleRxToronto, Ontario, Canada
Full-time

Lead the development of impactful technologies as a Senior Full-Stack Engineer in a mission-driven healthtech environment.Collaborate on building scalable products that enhance patient and pharmaci... Show more

 • Promoted

Senior Software Engineer in Performance Engineering

SentryToronto, Ontario, Canada
Full-time

Drive innovation in software performance at Sentry as a Senior Software Engineer.Tackle complex issues in a hybrid work environment designed for collaboration.Join Sentry’s Issue Workflow team, whe... Show more

 • Promoted

Performance Engineer - Resiliency & Dr Expert - C$90,000 - C$120,000 A Year

IT services companyNorth York, Canada
Full-time

Performance Engineer specializing in resiliency and disaster recovery testing, experienced with LoadRunner, JMeter, Dynatrace, and Splunk, ideally with banking domain knowledge. Show more

 • Promoted

Senior Solutions Engineer | REMOTE (ONTARIO)

GatekeeperToronto, ON, CA
Remote
Full-time

Senior Solutions Engineer | REMOTE (ONTARIO).Gatekeeper is the ONLY unified contract & third-party risk management platform that protects compliance‑centric organisations by delivering company‑wide... Show more

 • Promoted

Senior Application Engineer Hybrid Role

Honeywell TechnologiesMarkham, ON, CA
Full-time

Join Honeywell as a Senior Application Engineer focusing on Advanced Applications, particularly in Blending.This hybrid role offers you the chance to work in Markham or Calgary while enhancing cust... Show more

 • Promoted

Hopper Senior Engineer Payments Team

HopperToronto, Ontario, Canada
Full-time

Become a key player in Hopper's Payments Team as a Senior Engineer! Leverage your backend engineering skills to enhance payment experiences for millions of users.As part of Hopper’s mission to opti... Show more

 • Promoted

Senior Rust Engineer for High-Performance Solana Contracts

Apeing DEXToronto, ON, CA
Full-time

A cutting-edge tech company in Toronto is seeking a top-tier Rust engineer who has already shipped production smart contracts.Your role will include designing, building, and deploying high-performa... Show more

 • Promoted

Innovative Principal Engineer Role Enercare

Enercare Inc.Markham, Ontario, Canada
Full-time

Become a Principal Engineer with Enercare Inc.Lead the charge in developing applications that enhance the customer experience through technology.This role is a full-time opportunity reporting to th... Show more

 • Promoted

Senior Quality Engineer, Infrastructure - C$104,000 - C$123,500 A Year - Remote

StackAdaptNorth York, Canada
Remote
Full-time

Leads design and implementation of testing infrastructure to boost engineering velocity and product reliability.Collaborates with teams to build robust testing systems and frameworks for confident ... Show more

 • Promoted

Senior ML Kernel Performance Engineer

Amazon Web Services (AWS)Toronto, ON, CA
Full-time

Annapurna Labs team at Amazon builds Neuron, the software development kit used to accelerate deep learning and GenAI workloads on Amazon’s custom machine learning accelerators, Inferentia and Train... Show more

 • Promoted

Docebo Senior Full-Stack Engineer

DoceboToronto, ON, CA
Full-time

Join Docebo as a Senior Full-Stack Engineer, owning feature lifecycles from ideation to deployment.Drive impact with AI-assisted development tools and close customer engagement.As a Senior Full-Sta... Show more

 • Promoted

Senior Digital/FPGA Engineer

Per Vices CorporationToronto, Ontario, Canada
Full-time

We are looking for an engineer to help us build Software Defined Radios (SDRs) for mission critical infrastructure.Our products use FPGAs to provide a high performance interface between data provid... Show more

 • Promoted

Senior Full-Stack Engineer, Trust & Safety – Remote - $150,000 - $200,000 A Year - Remote

AffirmNorth York, Canada
Remote
Full-time

Seeking a Senior Full-Stack Engineer for a fintech company to improve customer authentication and verification processes through high-availability systems.Role requires 4+ years of software enginee... Show more

 • Promoted

Senior Quality Engineer

ScotiabankToronto, Ontario, Canada
Full-time

Select how often (in days) to receive an alert:.Join a purpose driven winning team, committed to results, in an inclusive and high-performing culture.Contributes to the overall success of Client Ma... Show more