Forward Deployment Engineer

Posted 4 Days Ago
Be an Early Applicant
Tokyo, JPN
In-Office
105K-130K Annually
Senior level
Artificial Intelligence • Hardware • Machine Learning • Natural Language Processing • Software • Generative AI
SambaNova is the #1 platform for business AI.
The Role
Designs, builds, and deploys production generative AI applications for strategic enterprise customers using SambaNova hardware and software. Responsibilities include architecting RAG and multi-agent workflows, optimizing inference performance, troubleshooting across model, software, and hardware layers, translating customer needs into product requirements, supporting technical sales engagements, creating reusable deployment accelerators, and presenting technical findings to customers and internal teams.
Summary Generated by Built In

SambaNova is a leader in next-generation AI infrastructure, delivering a full-stack inference platform for customers worldwide. At the core of SambaNova's technology is the RDU (Reconfigurable Dataflow Unit) — a chip built on a dataflow architecture rather than the traditional GPU model. Its decode performance is especially strong for agentic workloads like multi-turn agents, code generation, and long-running applications. RDUs are packaged into SambaRack, rack-scale hardware that lets customers deploy state-of-the-art models with better performance, greater energy efficiency, and faster time to value.

Join the company that's building the future of AI computing. SambaNova is disrupting the AI and high-performance computing space with an integrated hardware and software platform.

Our SambaStack inference serving platform is pushing the boundaries of inference serving for generative AI and large language models. We are a team of passionate innovators tackling some of the world's most challenging computational problems. We help enterprises and service providers host their own AI inference platforms, powered by our state-of-the-art RDU (Reconfigurable Dataflow Unit) hardware architecture. Our cloud-agnostic, enterprise-grade inference serving platform enables seamless deployment, management, and scaling of foundation model workloads at production scale.


SambaNova is hiring a Forward Deployed Engineer (FDE)  for our SambaStack based product portfolio in our Customer Success Organization. 

Responsibilities: 

  • Embed directly with strategic enterprise customers to design, build, and deploy production GenAI applications on SambaNova's SN40L platform and SambaStack based product portfolio
  • Architect and implement LLM-powered workflows — including RAG pipelines, multi-agent systems, fine-tuning workflows, and coding solutions — tailored to each customer's data, infrastructure, and business goals.
  • Optimize AI inference performance on SambaNova hardware; benchmark model throughput, latency, and accuracy against customer requirements and competitor baselines.
  • Troubleshoot and resolve production issues end-to-end across model, software, and hardware layers — acting as the first and last line of technical escalation in the field.
  • Translate customer needs into clear product requirements and engineering feedback; serve as the primary voice of field reality to SambaNova's Product and Engineering teams.
  • Partner with Account Executives and Solutions Engineers to shape technical sales strategy, scope engagements, and demonstrate platform differentiation during evaluations and proof-of-concepts.
  • Develop reusable accelerators, reference architectures, and internal playbooks that scale learnings from one deployment to many.
  • Present technical findings, architecture decisions, and roadmap input at customer executive briefings and internal forums; represent SambaNova at industry conferences and events.

Qualifications: 

  • 5+ years of hands-on engineering experience, with a strong record of shipping production AI/ML systems.
  • Deep expertise in GenAI application development: LLM orchestration, RAG, agentic frameworks (LangChain, LlamaIndex, DSPy), prompt engineering, and evaluation pipelines.
  • Strong foundations in ML fundamentals — model training, fine-tuning, inference optimization, quantization, and performance benchmarking.
  • Proficiency in Python (required); working knowledge of C++ or CUDA a strong plus for hardware-layer debugging.
  • Experience deploying AI workloads on cloud infrastructure (AWS, Azure, GCP) and familiarity with containerization, orchestration (Kubernetes, Docker), and MLOps tooling.
  • Comfortable engaging directly with customers: able to run technical discovery, set expectations, push back constructively, and present to executive and practitioner audiences alike.
  • Bachelor's or graduate degree in Computer Science, Electrical Engineering, Mathematics, Physics, or equivalent practical experience.
  • Willingness to travel up to 50% to customer sites — flexible based on engagement needs.

Bonus qualifications: 

  • Experience with AI accelerators or custom silicon (TPUs  etc.)
  • CUDA / low-level GPU programming
  • Familiarity with VLLM / SGLang
  • Enterprise AI deployments in regulated industries

Base Salary Range:

Base Pay Range
¥10,500,000¥13,000,000 JPY

Submission Guidelines
Please note that in order to be considered an applicant for any position at SambaNova Systems, you must submit an application form for each position for which you believe you are qualified. 

EEO Policy
SambaNova Systems is an Equal Opportunity/Affirmative Action Employer. All qualified applicants will receive consideration for employment without regard basis of age (40 and over), color, disability, gender identity, genetic information, marital status, military or veteran status, national origin/ancestry, race, religion, creed, sex (including pregnancy, childbirth, breastfeeding), sexual orientation, and any other applicable status protected by federal, state, or local laws.

Benefits Summary for US-Based, Full-Time Employment Positions
SambaNova offers a competitive total rewards package, including the base salary, plus equity and benefits. We cover 95% premium coverage for employee medical insurance, and 77% premium coverage for dependents and offer a Health Savings Account (HSA) with employer contribution. We also offer Dental, Vision, Short/Long term Disability, Basic Life, Voluntary Life, and AD&D insurance plans in addition to Flexible Spending Account (FSA) options like Health Care, Limited Purpose, and Dependent Care. Our library of well-being benefits available to you and your dependents includes a full subscription to Headspace, Gympass+ membership with access to physical gyms, One Medical membership, counseling services with an Employee Assistance Program, and much more.

Skills Required

  • 5+ years of hands-on engineering experience with a record of shipping production AI/ML systems
  • Expertise in generative AI application development, including LLM orchestration, RAG, agentic frameworks, prompt engineering, and evaluation pipelines
  • Knowledge of model training, fine-tuning, inference optimization, quantization, and performance benchmarking
  • Proficiency in Python
  • Experience deploying AI workloads on cloud infrastructure such as AWS, Azure, or GCP
  • Familiarity with Kubernetes, Docker, and MLOps tooling
  • Ability to engage directly with customers, conduct technical discovery, manage expectations, and present to executive and technical audiences
  • Bachelor's or graduate degree in Computer Science, Electrical Engineering, Mathematics, Physics, or equivalent practical experience
  • Willingness to travel up to 50% to customer sites
  • Working knowledge of C++ or CUDA for hardware-layer debugging
  • Experience with AI accelerators or custom silicon, such as TPUs
  • CUDA or low-level GPU programming experience
  • Familiarity with vLLM or SGLang
  • Experience with enterprise AI deployments in regulated industries
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Palo Alto, CA
500 Employees
Year Founded: 2017

What We Do

AI is changing the world and at SambaNova, we believe that you don’t need unlimited resources to take advantage of the most advanced, valuable AI capabilities - capabilities that are helping organizations explore the universe, find cures for cancer, and giving companies access to insights that provide a competitive edge. We deliver the world’s fastest and only complete AI solution for enterprises and governments with world-record inference performance and accuracy. Powered by the SambaNova SN40L Reconfigurable Dataflow Unit (RDU), organizations can build a technology backbone for the next decade of AI innovation with SambaNova Suite. Our fully integrated hardware-software system, DataScale®, enables organizations to train, fine-tune, and deploy the most demanding AI workloads using the largest and most challenging models. Most recently, with the launch of our newest offering, SambaNova Cloud, developers can supercharge AI-powered applications on Llama 3.2 models. SambaNova was founded in 2017 in Palo Alto, California, by a group of industry luminaries, business leaders, and world-class innovators who understand AI. Today, we’ve built an incredibly smart and motivated team dedicated to making a lasting impact on the industry and equipping our customers to thrive in the new era of AI.

Why Work With Us

As a talent first company, we aim to hire the greatest and most innovative minds in the industry- driving the next generation of AI computing where no barrier is too high and the possibilities are truly limitless. We encourage our peers to take risks and take the initiative to make a lasting impact on the AI and ML industries.

Gallery

Gallery

Similar Jobs

ai& Logo ai&

Compute - Forward Deployment Engineer

Artificial Intelligence • Cloud • Software • Infrastructure as a Service (IaaS)
Remote or Hybrid
2 Locations

ai& Logo ai&

AI - Forward Deployment Engineer

Artificial Intelligence • Cloud • Software • Infrastructure as a Service (IaaS)
Remote or Hybrid
2 Locations

ServiceNow Logo ServiceNow

Mid-market Account Executive

Artificial Intelligence • Cloud • HR Tech • Information Technology • Productivity • Software • Automation
Hybrid
Tokyo, JPN
29000 Employees

Mastercard Logo Mastercard

Manager, Communications

Blockchain • Fintech • Payments • Consulting • Cryptocurrency • Cybersecurity • Quantum Computing
Hybrid
Tokyo, JPN
38800 Employees

Similar Companies Hiring

Revel.io Thumbnail
Aerospace • Hardware • Robotics • Software
US
50 Employees
Blee Thumbnail
Artificial Intelligence • Marketing Tech • Software
New York, New York
30 Employees
Vega Thumbnail
Artificial Intelligence • Automotive • Insurance • Transportation
US
43 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account