Intern - ML Inference Performance Engineer

Posted 5 Days Ago
Be an Early Applicant
Eindhoven, NLD
Hybrid
Internship
Artificial Intelligence • Hardware • Software
The Role
Evaluate and optimize machine learning inference performance across hardware and software platforms. Develop benchmarking tools, reproducible evaluation procedures, dashboards, and standardized reporting for throughput, latency, power, and accuracy. Analyze end-to-end computer vision pipelines, host-device overhead, compiler toolchains, SDKs, and runtime execution. Set up evaluation hardware and lab infrastructure, assess accelerator capabilities, and communicate findings that inform engineering and product roadmap decisions.
Summary Generated by Built In

About Us

Axelera AI is not your regular deep-tech company. We are creating the next-generation AI platform to support anyone who wants to help advancing humanity and improve the world around us.

In just five years, we have raised a total of $370 million and have built a world-class team of 250+ employees (including 60+ PhDs with more than 40,000 citations), both remotely from 20 different countries and with offices in Belgium, France, Switzerland, Italy, the UK, headquartered at the High Tech Campus in Eindhoven, Netherlands.

We have also launched our Metis™ AI Platform, which achieves a 3-5x increase in efficiency and performance, and have visibility into a strong business pipeline exceeding $100 million.

Our unwavering commitment to innovation has firmly established us as a global industry pioneer.

Are you up for the challenge?

Position Overview

We're looking for a curious, rigorous engineer to join our team and dig into the performance of ML inference systems. You'll work across the full inference stack — from model export and compiler toolchains to runtime execution on silicon — to build a clear, evidence-based picture of how different platforms perform and why.

Your work will go beyond running benchmarks: you'll develop a repeatable evaluation methodology, investigate performance bottlenecks at the hardware and software level, and build the tooling that transforms raw measurements into actionable insight. The findings you produce can directly shape our product decisions.

Key responsibilities:

  • Benchmarking & Tooling: Develop a thorough understanding of internal benchmarking tools covering throughput, latency, power, and accuracy across device-level, host-transaction, and end-to-end pipeline scenarios. Improve existing tooling, define reproducible procedures, and establish a standardised results format for rigorous cross-platform comparisons. Maintain a dedicated dashboard for performance visualisations.

  • Platform Evaluation: Research and evaluate AI accelerator products from various vendors, gaining hands-on experience with their SDKs, toolchains, flexibility, and limitations through a structured evaluation process. Track model support across platforms to identify strengths, gaps, and areas for improvement.

  • Pipeline Analysis: Characterise full inference pipelines, capturing host-device transaction overhead and end-to-end performance metrics. Ensure equivalent pipeline configurations across platforms using frameworks such as GStreamer to maintain methodological consistency.

  • Lab & Infrastructure: Set up and maintain lab hosts across multiple hardware platforms and support the onboarding of new evaluation hardware.

  • Reporting: Synthesise findings into clear, structured reports that directly inform engineering and roadmap decisions.

Requirements:

  • Currently enrolled in the final years of a Bachelor's programme or in a Master's programme in Computer Engineering, Electrical Engineering, Computer Science, or a related field. This position may also be carried out as a Master's thesis project.

  • Python development experience

  • C/C++ knowledge

  • Experience with end-to-end computer vision pipelines

  • Familiarity with benchmarking concepts (performance, latency, etc.)

  • Experience with inference tools, APIs, or SDKs (e.g., TensorRT)

  • Familiarity with deep learning model concepts (quantization, ONNX, PyTorch, etc.)

  • Development experience using agentic AI

  • Knowledge of version control (Git)

  • Familiarity with LLM benchmarking concepts

  • Proficiency with Linux, Bash scripting, and Docker

  • Hands-on experience with embedded hosts

  • Proficient written and verbal communication skills in English, with the ability to document findings clearly and precisely.

  • Good organizational skills

Nice to have:

  • GStreamer knowledge

  • Basic GUI design experience

Location

Work from our Axelera AI office in Eindhoven (Netherlands).

What we offer 

This is your chance to shape and be part of a dynamic, fast-growing, international organization. We offer an attractive compensation package, including a pension plan, extensive employee insurances and the option to get company shares.   

An open culture that supports creativity and continual innovation is awaiting you. Collaborative ownership and freedom with responsibility is characteristic for the way we act and work as a team. 

At Axelera AI, we wholeheartedly embrace equal opportunity and hold diversity in the highest regard. Our steadfast commitment is to cultivate a warm and inclusive environment that empowers and celebrates every member of our team. We welcome applicants from all backgrounds to join us in shaping the future of AI.

Skills Required

  • Currently enrolled in the final years of a Bachelor's program or in a Master's program in Computer Engineering, Electrical Engineering, Computer Science, or a related field
  • Python development experience
  • C/C++ knowledge
  • Experience with end-to-end computer vision pipelines
  • Familiarity with benchmarking concepts, including performance and latency
  • Experience with inference tools, APIs, or SDKs such as TensorRT
  • Familiarity with deep learning model concepts, including quantization, ONNX, and PyTorch
  • Development experience using agentic AI
  • Knowledge of version control using Git
  • Familiarity with LLM benchmarking concepts
  • Proficiency with Linux, Bash scripting, and Docker
  • Hands-on experience with embedded hosts
  • Proficient written and verbal communication skills in English
  • Ability to document findings clearly and precisely
  • Good organizational skills
  • GStreamer knowledge
  • Basic GUI design experience
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Eindhoven
147 Employees
Year Founded: 2021

What We Do

Axelera AI delivers an AI-native game-changing hardware and software platform to accelerate artificial intelligence. Headquartered in the AI Innovation Center of the High Tech Campus in Eindhoven, Axelera AI has R&D offices in Belgium, Switzerland, UK and Italy and operations in 15 European countries . Its team of experts in AI software and hardware hail from top AI firms and Fortune 500 companies.

Similar Jobs

Pfizer Logo Pfizer

Senior Manager, HTA, Value and Evidence (HV&E), Genitourinary Cancer

Artificial Intelligence • Healthtech • Machine Learning • Natural Language Processing • Biotech • Pharmaceutical
In-Office or Remote
30 Locations
121990 Employees
139K-232K Annually

Pfizer Logo Pfizer

Staff Software Engineer

Artificial Intelligence • Healthtech • Machine Learning • Natural Language Processing • Biotech • Pharmaceutical
In-Office or Remote
36 Locations
121990 Employees

HiBob Logo HiBob

Associate Account Executive

HR Tech • Information Technology • Professional Services • Sales • Software
Remote or Hybrid
Netherlands
1350 Employees
50K-50K Annually

Pfizer Logo Pfizer

Sustainability Senior Manager

Artificial Intelligence • Healthtech • Machine Learning • Natural Language Processing • Biotech • Pharmaceutical
Remote or Hybrid
30 Locations
121990 Employees
112K-207K Annually

Similar Companies Hiring

Revel Thumbnail
Aerospace • Hardware • Robotics • Software
Marina Del Rey, California
60 Employees
Blee Thumbnail
Artificial Intelligence • Marketing Tech • Software
New York, New York
30 Employees
Vega Thumbnail
Artificial Intelligence • Automotive • Insurance • Transportation
US
43 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account