Performance Architect, AI HW

Reposted 27 Days Ago
Be an Early Applicant
Santa Clara, CA, USA
In-Office
100K-500K Annually
Mid level
Hardware • Manufacturing
The Role
As an AI Performance Architect, you will analyze and optimize AI workloads on Tensix architecture, linking architecture with software and hardware to enhance efficiency and scalability.
Summary Generated by Built In

Tenstorrent is leading the industry on cutting-edge AI technology, revolutionizing performance expectations, ease of use, and cost efficiency. With AI redefining the computing paradigm, solutions must evolve to unify innovations in software models, compilers, platforms, networking, and semiconductors. Our diverse team of technologists have developed a high performance RISC-V CPU from scratch, and share a passion for AI and a deep desire to build the best AI platform possible. We value collaboration, curiosity, and a commitment to solving hard problems. We are growing our team and looking for contributors of all seniorities.

The Tensix team is building the high-performance compute fabric that powers Tenstorrent’s AI and ML workloads. As an AI Performance Architect, you will model, analyze, and optimize how real AI workloads run on the Tensix architecture, shaping future hardware features and ensuring every design decision delivers measurable performance gains. This role connects architecture, software, and RTL to push the limits of efficiency and scalability across next-generation AI systems.

This role is hybrid, based out of Toronto, ON; Austin, TX; Santa Clara, CA or remote.

We welcome candidates at various experience levels for this role. During the interview process, candidates will be assessed for the appropriate level, and offers will align with that level, which may differ from the one in this posting.


Who You Are

  • Deeply analytical engineer with strong intuition for AI workload behavior and system-level performance bottlenecks.
  • Experienced in C++ and Python for simulation, modeling, and performance analysis across heterogeneous compute systems.
  • Adept at bridging software and hardware teams—translating deep learning workloads into architectural insight and measurable design tradeoffs.
  • Curious, data-driven, and comfortable pushing the limits of efficiency, scalability, and accuracy in high-performance AI systems.

What We Need

  • Benchmark and analyze complex AI workloads across single and multi-node hardware configurations to guide next-gen architecture.
  • Develop and maintain performance models, simulators, and micro-benchmark suites to drive feature evaluation and design optimization.
  • Conduct detailed PPA (Performance, Power, Area) studies to assess design tradeoffs and inform hardware-software co-design decisions.
  • Collaborate closely with RTL, Compiler, and Runtime teams to instrument and correlate performance models with silicon results.

What You’ll Learn

  • Advanced modeling techniques for large-scale AI systems, including multi-chip and distributed performance analysis.
  • How architectural choices propagate through the software stack—from compiler and runtime layers down to custom AI accelerators.
  • Emerging deep learning trends and their impact on compute architecture design and performance tuning.
  • How to define and validate performance features that directly translate to measurable gains across real-world AI workloads.

Compensation for all engineers at Tenstorrent ranges from $100k - $500k including base and variable compensation targets. Experience, skills, education, background and location all impact the actual offer made.

Tenstorrent offers a highly competitive compensation package and benefits, and we are an equal opportunity employer.

This offer of employment is contingent upon the applicant being eligible to access U.S. export-controlled technology.  Due to U.S. export laws, including those codified in the U.S. Export Administration Regulations (EAR), the Company is required to ensure compliance with these laws when transferring technology to nationals of certain countries (such as EAR Country Groups D:1, E1, and E2).   These requirements apply to persons located in the U.S. and all countries outside the U.S.  As the position offered will have direct and/or indirect access to information, systems, or technologies subject to these laws, the offer may be contingent upon your citizenship/permanent residency status or ability to obtain prior license approval from the U.S. Commerce Department or applicable federal agency.  If employment is not possible due to U.S. export laws, any offer of employment will be rescinded.

Skills Required

  • Experience in C++
  • Experience in Python
  • Understanding of AI workload behavior and system performance
  • Ability to develop performance models
  • Experience with benchmarking AI workloads
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Toronto, ON
389 Employees
Year Founded: 2016

What We Do

Tenstorrent is a next-generation computing company that builds computers for AI. Headquartered in Toronto, Canada, with U.S. offices in Austin, Texas, and Silicon Valley, and global offices in Belgrade and Bangalore, Tenstorrent brings together experts in the field of computer architecture, ASIC design, advanced systems, and neural network compilers. Join us: www.tenstorrent.com/careers

Similar Jobs

PwC Logo PwC

Consultant

Artificial Intelligence • Professional Services • Business Intelligence • Consulting • Cybersecurity • Generative AI
Remote or Hybrid
64 Locations
370000 Employees
99K-232K Annually

PwC Logo PwC

San Francisco - Tax - Associate - Summer/Fall 2027

Artificial Intelligence • Professional Services • Business Intelligence • Consulting • Cybersecurity • Generative AI
Hybrid
San Francisco, CA, USA
370000 Employees
54K-142K Annually

PwC Logo PwC

CTIO -Activation & Customer Success - Experienced Associate

Artificial Intelligence • Professional Services • Business Intelligence • Consulting • Cybersecurity • Generative AI
Remote or Hybrid
62 Locations
370000 Employees
51K-140K Annually

PwC Logo PwC

San Diego - Tax - Intern - Summer 2028 - Destination CPA

Artificial Intelligence • Professional Services • Business Intelligence • Consulting • Cybersecurity • Generative AI
Hybrid
San Diego, CA, USA
370000 Employees
29K-48K Hourly

Similar Companies Hiring

Rosendin Thumbnail
Other • Manufacturing
San Jose, CA
6219 Employees
Amalgamated Sugar Thumbnail
Food • Greentech • Agriculture • Industrial • Manufacturing
Boise, Idaho
768 Employees
Revel.io Thumbnail
Aerospace • Hardware • Robotics • Software
US
50 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account