Co-op Researcher – LLM Post-Training & Agentic Coding

Posted 16 Days Ago
Be an Early Applicant
Edmonton, AB, CAN
In-Office
Entry level
Information Technology • Other
The Role
Conduct research on LLM post-training, reinforcement learning, reasoning, agentic coding, and self-improving agents. Develop training methods, synthetic data, environments, evaluation systems, and scalable infrastructure for intelligent coding agents. Design experiments, implement and validate research ideas, optimize GPU workflows, analyze model behavior, and communicate findings while collaborating with research and engineering teams.
Summary Generated by Built In

Huawei Canada has an immediate Co-op opening for a Researcher.

About the team:

Founded in 2012, the Noah’s Ark lab has evolved into a prominent research organization with notable achievements in academia and industry. The lab’s mission focuses on advancing artificial intelligence and related fields to benefit the company and society. Driven by impactful, long-term projects, the aim is to enhance state-of-the-art research while integrating innovations into the company's products and services, including Next Generation Foundation Model, Agentic Models, Physical AI and Reinforcement Learning.

About the job:

  • Conduct cutting-edge research on LLM post-training, reinforcement learning, reasoning, agentic coding, and self-improving agents, with the opportunity to take technically challenging ideas from initial hypothesis to rigorous experimental validation.

  • Develop new approaches for agent training, task and data synthesis, environment generation, scalable learning from interaction, and agent evaluation, with particular interest in terminal, repository-level, and software-engineering environments.

  • Build and evaluate intelligent agents capable of reasoning, tool use, code generation and modification, and solving complex real-world tasks, while studying their capabilities, failure modes, and generalization behavior.

  • Own research problems end-to-end, including literature review, problem formulation, experiment design, implementation, large-scale evaluation, analysis, and communicating high-quality research findings.

  • Build reliable and scalable infrastructure for LLM training, inference, agent rollout, evaluation, and synthetic data generation, supporting rapid experimentation across large models and datasets.

  • Improve the efficiency and reliability of GPU-based post-training and inference workflows, including distributed execution, resource utilization, debugging, automation, and day-to-day research compute operations.

  • Work closely with researchers and engineering teams across Huawei to rapidly prototype new ideas, reproduce and extend state-of-the-art methods, and translate promising research advances into practical LLM systems and downstream projects.

  • Contribute in a fast-moving, high-ownership research environment where a focused Co-op term can lead to meaningful research, engineering, open-source, or publication-quality contributions.

About the ideal candidate:

  • Currently enrolled in a senior Bachelor’s, Master’s, or PhD program in Computer Science, Machine Learning, Artificial Intelligence, or a related technical field, with strong fundamentals in machine learning and modern deep learning.

  • Strong understanding of large language models and modern post-training techniques, including Transformer-based models, supervised fine-tuning, reinforcement learning, preference optimization, distillation, reasoning, or synthetic-data-driven training.

  • Strong programming and software-engineering skills, particularly in Python and PyTorch, with the ability to independently implement, debug, profile, and iterate on complex ML or agent systems.

  • Hands-on experience in one or more relevant areas, such as LLM post-training, coding or tool-using agents, reinforcement learning, large-scale model training/inference, distributed systems, or ML infrastructure.

  • Demonstrated ability to take an open-ended and ambiguous technical problem, identify the important questions, design effective experiments or systems, and independently drive the work toward measurable results.

  • Strong analytical and experimental judgment, including the ability to design controlled experiments, evaluate models rigorously, diagnose unexpected behaviors, and distinguish genuine improvements from experimental noise.

  • Highly self-motivated, intellectually curious, and execution-oriented, with strong communication and collaboration skills and the ability to operate effectively in a rapid research iteration cycle with substantial ownership.

  • Strong assets include top-tier AI/ML publications or substantial research experience, impactful open-source or agent/LLM projects, and hands-on experience with frameworks and systems such as vLLM, SGLang, FSDP, DeepSpeed, Hugging Face, Docker, Linux, Git, or multi-GPU compute environments; we welcome candidates who are especially strong in either research or systems engineering, and particularly value those who can bridge both.

Additional Information:

Huawei Canada is committed to a fair, inclusive, and accessible recruitment process. If you require accommodation during any stage of the hiring process, please let us know and we will work with you to meet your needs.

All applications for this position are reviewed directly by our hiring team, we do not use artificial intelligence tools to screen or select candidates.

Skills Required

  • Currently enrolled in a senior Bachelor's, Master's, or PhD program in Computer Science, Machine Learning, Artificial Intelligence, or a related technical field
  • Strong fundamentals in machine learning and modern deep learning
  • Strong understanding of large language models and post-training techniques, including supervised fine-tuning, reinforcement learning, preference optimization, distillation, reasoning, or synthetic-data-driven training
  • Strong programming and software-engineering skills, particularly in Python and PyTorch
  • Ability to independently implement, debug, profile, and iterate on complex machine learning or agent systems
  • Hands-on experience in LLM post-training, coding or tool-using agents, reinforcement learning, large-scale model training or inference, distributed systems, or ML infrastructure
  • Ability to solve open-ended technical problems, design experiments or systems, and independently drive work toward measurable results
  • Strong analytical and experimental judgment, including rigorous evaluation and diagnosing unexpected model behavior
  • Strong communication and collaboration skills, with self-motivation, intellectual curiosity, and execution focus
  • Top-tier AI or ML publications or substantial research experience
  • Impactful open-source or agent/LLM projects
  • Hands-on experience with vLLM, SGLang, FSDP, DeepSpeed, Hugging Face, Docker, Linux, Git, or multi-GPU computing environments
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Shenzhen
1,770 Employees
Year Founded: 1987

What We Do

Founded in 1987, Huawei is a leading global provider of information and communications technology (ICT) infrastructure and smart devices. We are committed to bringing digital to every person, home and organization for a fully connected, intelligent world. We have approximately 197,000 employees and we operate in over 170 countries and regions, serving more than three billion people around the world. In Canada, Huawei conducts innovative and leading edge research in 5G technologies, along with advanced development of emerging cloud, device and network technologies & services. While our renowned Canada Research Centre in the thriving technology landscape of Ottawa, Ontario continues to grow rapidly in size and strategic product initiatives, additional presence has also been established across Canada with R&D facilities in Vancouver, Edmonton, Waterloo, Markham, Montreal, and a R&D office in Quebec City.

Similar Jobs

UL Solutions Logo UL Solutions

Electrical Engineer

Automotive • Professional Services • Software • Consulting • Energy • Chemical • Renewable Energy
Remote or Hybrid
Canada
15000 Employees
104K-130K Annually

Block Logo Block

Software Engineer

Blockchain • eCommerce • Fintech • Payments • Software • Financial Services • Cryptocurrency
In-Office or Remote
8 Locations
12000 Employees
153K-270K Annually

Rokt Logo Rokt

Customer Success Manager

Artificial Intelligence • Digital Media • eCommerce • Marketing Tech • Software • Automation
In-Office
Calgary, AB, CAN
800 Employees
73K-112K Annually

Pfizer Logo Pfizer

Director R&D EHS Program Lead

Artificial Intelligence • Healthtech • Machine Learning • Natural Language Processing • Biotech • Pharmaceutical
In-Office or Remote
36 Locations
121990 Employees
177K-294K Annually

Similar Companies Hiring

Rosendin Thumbnail
Other • Manufacturing
San Jose, CA
6219 Employees
OmniCable Thumbnail
Other
Houston, Texas
815 Employees
Golden Pet Brands Thumbnail
Digital Media • eCommerce • Information Technology • Marketing Tech • Pet • Retail • Social Media
El Segundo, California
178 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account