Talent.com
Huawei Technologies Canada Co., Ltd.
Agentic RL Researcher – Distributed ComputingHuawei Technologies Canada Co., Ltd. • Markham, Ontario, CA
Agentic RL Researcher – Distributed Computing

Agentic RL Researcher – Distributed Computing

Huawei Technologies Canada Co., Ltd. • Markham, Ontario, CA
30+ days ago
Job type
  • Full-time
  • Permanent
Job description

Job description

Huawei Canada has an immediate permanent opening for a Researcher.

About the team:

The Distributed Data Storage and Management Lab leads research in distributed data systems, aiming to develop next-generation cloud serverless products that encompass core infrastructure and databases. This lab addresses various data challenges, including cloud-native disaggregated databases, pay-by-query user models, and optimizing low-level data transfers via RDMA. Teams within this lab create advanced cloud serverless data infrastructure and implement cutting-edge networking technologies for Huawei's global AI infrastructure.


About the job:

  • Design and develop advanced Agentic Reinforcement Learning (RL) and Multi-Agent Reinforcement Learning (MARL) algorithms for cooperative, competitive, and mixed-agent environments, including CTDE, decentralized learning, and hierarchical agent systems.

  • Build scalable simulation and training platforms for large-scale agent systems, supporting self-play, population-based training, curriculum learning, and emergent behavior analysis.

  • Optimize multi-agent learning performance on distributed compute clusters, improving sample efficiency, credit assignment, agent coordination, communication learning, and training stability.

  • Research and prototype new approaches for multi-agent intelligence, including communication protocols, credit assignment, game-theoretic learning dynamics, meta-learning, and adaptive agent populations.

  • Translate cutting-edge research in agentic AI and MARL into production-ready systems for real-world or high-fidelity simulated environments.

  • Develop benchmarking frameworks and evaluation metrics for agent coordination, robustness, scalability, and safety.

  • Collaborate with research, infrastructure, and product teams to deploy scalable agentic learning systems in real-world applications.

  • Contribute to technical leadership and innovation through publications, patents, open-source contributions, and conference presentations.

The total target annual compensation for this position ranges from $106,000 to $156,000 depending on education, experience, and demonstrated expertise.


Job requirements

About the ideal candidate:

  • MS or PhD in Computer Science, Electrical Engineering, or a related field, with a focus on Reinforcement Learning, Multi-Agent Systems, Agentic AI, or Distributed AI.

  • Strong expertise in reinforcement learning algorithms, particularly in multi-agent settings (e.g., policy gradients, value-based methods, CTDE, credit assignment, and coordination in non-stationary environments).

  • Solid foundations in optimization, probability, and game theory, with the ability to design and analyze complex learning systems.

  • Experience building scalable RL training infrastructure, including distributed rollouts, large-scale simulation, and experiment pipelines.

  • Strong programming skills in Python and/or C++, with experience developing high-performance or distributed ML systems.

  • Demonstrated impact through research publications, open-source contributions, patents, or production ML systems in reinforcement learning, multi-agent learning, or large-scale AI systems.

Additional Information:

Huawei Canada is committed to a fair, inclusive, and accessible recruitment process. If you require accommodation during any stage of the hiring process, please let us know and we will work with you to meet your needs.

All applications for this position are reviewed directly by our hiring team, we do not use artificial intelligence tools to screen or select candidates.

Create a job alert for this search

Agentic RL Researcher – Distributed Computing • Markham, Ontario, CA

Similar jobs

Remote AI Developer Co-op — Nuclear Tech

Nuclear Promise XToronto, ON, CA
Remote
Full-time

A leading nuclear technology firm is seeking an AIDeveloper co-op to design and develop AI agents that enhance nuclear operations.This remote role requires an individual eager to take ownership of ... Show more

 • Promoted

CoStar Group Field Researcher Role

CoStar Group, Inc.Toronto, ON, CA
Full-time

Join CoStar Group as a Field Researcher in Toronto, Canada, and play a pivotal role in enhancing real estate analytics.Your expert data collection will contribute to our valued client insights.In t... Show more

 • Promoted

Contract Artificial Intelligence Engineer Onsite

Pacer GroupToronto
Full-time

Join Pacer Group as a Contract Artificial Intelligence Engineer in the BFSI domain, engaging onsite 2-3 days a week.This role is ideal for highly experienced individuals looking to innovate.We are ... Show more

 • Promoted

AI Researcher in Emerging Risks Team

MSCI Inc.Toronto
Full-time

Become part of MSCI’s Emerging Risks R&D team as an AI Researcher and innovate how investors navigate risks.Utilize advanced AI technologies to quantify evolving market challenges.This role focuses... Show more

 • Promoted

Remote AI Engineer - LLM Specialist

PulsoraToronto, ON, CA
Remote
Full-time

Discover an exciting career as a Remote AI Engineer specializing in Large Language Models.Play a key role in developing AI-driven software solutions in a fully remote setting.Your expertise in AI/M... Show more

 • Promoted

Principal Machine Learning Infrastructure Researcher

LightmatterToronto, ON, CA
Full-time

Lightmatter is leading the revolution in AI data center infrastructure, enabling the next giant leaps in human progress.The company invented the world’s first 3D-stacked photonics engine, Passage™,... Show more

 • Promoted

Agentic RL Researcher – Distributed Computing

Huawei CanadaMarkham, York region, Canada
Permanent

Huawei Canada has an immediate permanent opening for a Researcher.The Distributed Data Storage and Management Lab leads research in distributed data systems, aiming to develop next-generation cloud... Show more

 • Promoted

Bioinformatician: AMR Platform & Tools (Remote)

BugSeqToronto, ON, CA
Remote
Full-time

A bioinformatics technology company is seeking a talented bioinformatician to design and evaluate tools and pipelines focused on antimicrobial resistance.This role involves developing core features... Show more

 • Promoted

Lead Researcher: Energy-Efficient AI Systems

FujitsuToronto
Full-time

Take on the Senior Researcher role at the University of Toronto, exploring energy-efficient AI platforms.Collaborate globally to advance AI accelerator research and innovation.This position situate... Show more

 • Promoted

Artificial Intelligence Engineer

Tata Consultancy ServicesToronto, ON, CA
Full-time

Tata Consultancy Services (TCS) is an equal opportunity employer, and embraces diversity in race, nationality, ethnicity, gender, age, physical ability, neurodiversity, and sexual orientation to cr... Show more

 • Promoted

Huawei Co-op Researcher in Distributed Systems

Huawei Technologies Canada Co., Ltd.Markham
Full-time

Join Huawei Canada as a Co-op Researcher specializing in distributed data systems and AI infrastructure.Collaborate on innovative cloud technologies that support smarter networks.Huawei Canada's la... Show more

 • Promoted

IT Sourcer / Recruitment Researcher (Prospecting Only)

Jobs for HumanityToronto, ON, CA
Full-time

IT Sourcer / Recruitment Researcher (Prospecting Only).Canadian IT consulting and professional services firm specializing in the placement of highly qualified technology consultants for public‑sect... Show more

 • Promoted

Computational Design Engineer: EDA & Scientific Computing

Axiomatic-AIToronto, ON, CA
Full-time

Axiomatic_AI is launching with the aim to accelerate R&D by "Automated Interpretable Reasoning" (AIR) – a verifiably truthful AI model built for reasoning in science and engineering.Axiomatic_AI is... Show more

 • Promoted

Agentic AI Developer

Kumaran SystemsToronto, ON, CA
Full-time

AI agents that automate business processes, enhance customer experiences, and drive operational efficiency.The ideal candidate will have strong expertise in Agentic AI systems, AI agent development... Show more

 • Promoted

Remote Clinical Researcher for AI-Driven Health Research

HelixRecruitToronto, ON, CA
Remote
Part-time

A leading recruiting firm is seeking a Clinical Researcher with a PhD for a part-time remote role.You will contribute expertise to AI projects, including designing legal workflows and evaluating AI... Show more

 • Promoted

Research and Development Engineer

Computer Talk Technology Inc.Markham, ON, CA
Full-time

Research and Development Engineer.Working as a key member of the Product Engineering Team, the R & D Engineer will develop and commercialize new products and technologies related to our unique comm... Show more

 • Promoted

LLM Engineer

MindlanceToronto
Full-time

Direct message the job poster from Mindlance.Hiring: LLM Engineers in Toronto! Be part of a fast-scaling team building Smarter, Next-Generation AI Agents alongside World-Leading AI Labs.Location: T... Show more

 • Promoted

Agentic AI Systems Developer - Remote

NTT DATA, Inc.Toronto, ON, CA
Remote
Full-time

We are currently seeking a Agentic AI Systems Developer - Remote to join our team in Toronto, Ontario (CA-ON), Canada (CA).You will design and build agentic AI systems for healthcare using the Neur... Show more

 • Promoted

AI and NLP Researcher - Emerging Risks at MSCI

MSCI IncToronto, ON, CA
Full-time

Enhance MSCI's risk evaluation capabilities as an AI and NLP Researcher focused on Emerging Risks.Be at the forefront of analyzing trends such as climate change and supply chain disruption.As part ... Show more

 • Promoted

Remote Mathematics Researcher (PhD) - 34877

TuringToronto, ON, CA
Remote
Full-time

Remote contract for PhDs in Mathematics, Statistics, or related fields.Work on cutting-edge projects with top AI labs while earning upto $150/hr, fully remote, with flexible weekly hours.We seek ma... Show more