Sr Staff Engineer

Posted 8 Days Ago
Be an Early Applicant
San Jose, CA, USA
In-Office
221K-221K Annually
Senior level
Semiconductor • Manufacturing
The Role
Develop, fine-tune, evaluate, optimize, and deploy large language and multimodal models in local, on-premises, edge, and restricted environments. Build scalable training and inference pipelines, integrate models into production APIs, monitor performance and resource usage, and ensure security and compliance. Collaborate with engineering and business teams, define KPIs, document workflows, evaluate emerging AI techniques, and mentor other engineers.
Summary Generated by Built In
Lattice Overview

There is energy here…energy you can feel crackling at any of our international locations. It’s an energy generated by enthusiasm for our work, for our teams, for our results, and for our customers. Lattice is a worldwide community of engineers, designers, and manufacturing operations specialists in partnership with world-class sales, marketing, and support teams, who are developing programmable logic solutions that are changing the industry. Our focus is on R&D, product innovation, and customer service, and to that focus, we bring total commitment and a keenly sharp competitive personality.

Energy feeds on energy. If you flourish in a fast paced, results-oriented environment, if you want to achieve individual success within a “team first” organization, and if you believe you can contribute and succeed in a demanding yet collegial atmosphere, then Lattice may well be just what you’re looking for.

Job Description:

We are looking for a Sr. Staff Engineer to join our team.


Key Responsibilities


  • Design and architect enterprise-wide Generative AI solutions, including reference architectures, integration patterns, and technical standards.
  • Design, build, and maintain an enterprise AI gateway to centralize access, governance, and monitoring of AI model consumption across the organization.
  • Develop and implement intelligent routing techniques to direct requests across multiple large language models (LLMs) and AI providers based on cost, latency, accuracy, and availability requirements.
  • Evaluate and integrate local/on-premises models as alternatives to third-party hosted models, with a focus on reducing operational costs.
  • Establish frameworks for measuring and demonstrating cost savings achieved through model selection, routing optimization, and infrastructure decisions.
  • Collaborate with engineering, security, and compliance stakeholders to ensure AI architecture adheres to organizational governance and regulatory requirements.
  • Define best practices for prompt orchestration, caching strategies, and fallback mechanisms within the AI gateway.
  • Provide technical leadership and mentorship to engineering teams adopting Generative AI capabilities.
  • Provide mentorship and lead a team of AI/ML engineers.

Required Qualifications


  • Demonstrated experience architecting Generative AI solutions at an enterprise scale.
  • Research and integrate cutting-edge LLMs and autonomous AI agent architecture into development processes
  • Develop RAG pipelines that enhance AI‘s ability to retrieve relevant knowledge and generate context-aware responses.
  • Build and optimize agentic AI systems that can interact with APIs, databases, and development environments (such as LangChain, OpenAI APIs, etc.)
  • Fine-tune LLMs (GPT, Llama, Mistral, Claude, Gemini etc.) for domain-specific applications.
  • Optimize models for local inference through quantization, pruning, and distillation.
  • Deploy models on-prem or at the edge using frameworks such as PyTorch, TensorRT, ONNX, vLLM, or llama.cpp.
  • Build and maintain training and inference pipelines for reproducibility and scalability.
  • Integrate locally deployed models into production systems via APIs and internal services.
  • Monitor model performance, drift, latency, and resource utilization in production
  • Optimize retrieval mechanisms to enhance response accuracy, grounding AI outputs in real-world data
  • Hands-on experience designing and implementing AI gateway solutions and model routing techniques.
  • Proven track record of achieving measurable cost savings through the use of local/open-source models or alternative optimization techniques.
  • Strong understanding of LLM provider ecosystems, API integration patterns, and multi-model orchestration.
  • Experience with cloud infrastructure and enterprise architecture frameworks.
  • Solid grasp of AI governance, security, and compliance considerations in enterprise environments.
  • Excellent communication skills, with the ability to present technical concepts to both technical and non-technical stakeholders.
  • Architect and deploy scalable AI models and retrieval pipelines using cloud-based MLOps pipelines (AWS/GCP/Azure, Docker, Kubernetes)
  • Optimize LLMs for real-time AI inferencing, ensuring low latency and high-performance AI solutions
  • Education: Master's or Ph.D. in Computer Science, AI, Machine Learning, or a related field.
  • Experience: 5+ years of experience in AI and machine learning, with at least 2 years of experience working on LLMs, code generation, RAG, or AI-powered automation

Pay & Benefits

Consistent with Lattice Semiconductor values and applicable law, we provide the following information to promote pay transparency and equity. We have a market-based pay structure which varies by location.  Please note that the base pay range is a guideline, and our compensation range reflects the cost of labor in the U.S. geographic market based on the location of the role. Pay within these ranges varies and depends on job-related knowledge, skills, and relevant work experience. 

For candidates who receive and offer, the starting salary will vary based on various factors including, but not limited to, such qualifications as, skill level, competencies, and work location.  The range provided may represent a candidate range and may not reflect the full range for an individual tenured employee.

Base Pay Range

220800

In addition to base pay, this role may be eligible for variable/ incentive compensation and/ or equity.  In addition, this role is eligible for a comprehensive, competitive benefits package which may include healthcare and retirement plans, paid time off, and more! 

Additional Information:

This position requires a successful background and reference checks and satisfactory proof of your right to work in the United States.

Lattice recognizes that employees are its greatest asset and the driving force behind success in a highly competitive, global industry.  Lattice continually strives to provide a comprehensive compensation and benefits program to attract, retain, motivate, reward and celebrate the highest caliber employees in the industry. 


Lattice is an international, service-driven developer of innovative low cost, low power programmable design solutions.  Our global workforce, some 1,000 strong, shares a total commitment to customer success and an unbending will to win.  For more information about how our FPGA, CPLD and programmable power management  devices help our customers unlock their innovation, visit 
www.latticesemi.com.  You can also follow us via Twitter, Facebook, or RSS. At Lattice, we value the diversity of individuals, ideas, perspectives, insights and values, and what they bring to the workplace.  Applications are welcome from all qualified candidates.

 

As an E-Verify employer, we use this system to confirm the employment eligibility of all new hires in accordance with federal law. All applicants will be required to complete a Form I-9, Employment Eligibility Verification, upon hire. We do not use E-Verify to pre-screen job candidates and will comply with all E-Verify regulations.

Skills Required

  • Master's or Ph.D. in computer science, engineering, or a related field, or equivalent practical experience
  • 8+ years of experience in artificial intelligence and machine learning
  • At least 3 years of experience with LLMs, code generation, large-scale neural networks, RAG, or AI-powered automation
  • Hands-on experience with LLMs and open-weight models
  • Proficiency in Python and machine learning frameworks such as PyTorch or TensorFlow
  • Expertise with vector databases and retrieval models
  • Experience deploying models in local, on-premises, or resource-constrained environments
  • Experience with multi-agent AI systems for autonomous coding tasks
  • Experience developing multimodal and ensemble models with business stakeholders
  • Understanding of model optimization techniques including quantization, batching, and memory optimization
  • Familiarity with Linux, Docker, and basic cloud or on-premises infrastructure
  • Commitment to continuous learning in AI and machine learning
  • Ability to propose innovative solutions to complex problems
  • Mentoring experience or ability to mentor engineers
  • Excellent verbal and written communication skills
  • Strong interpersonal and teamwork skills
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Hillsboro, OR
975 Employees
Year Founded: 1983

What We Do

Lattice Semiconductor (NASDAQ: LSCC) is the low power programmable leader. We solve customer problems across the network, from the Edge to the Cloud, in the growing communications, computing, industrial, automotive and consumer markets. Our technology, long-standing relationships, and commitment to world-class support lets our customers quickly and easily unleash their innovation to create a smart, secure and connected world. Lattice was founded in 1983 and is headquartered in Hillsboro, Oregon with major operations in San Jose, California, Shanghai, China, and Manila, Philippines.

Similar Jobs

Capital One Logo Capital One

Artificial Intelligence Engineer

Fintech • Machine Learning • Payments • Software • Financial Services
Remote or Hybrid
5 Locations
55000 Employees
286K-392K Annually

Capital One Logo Capital One

Staff Software Engineer

Fintech • Machine Learning • Payments • Software • Financial Services
Hybrid
5 Locations
55000 Employees
286K-392K Annually

Clear Street Logo Clear Street

Senior / Staff AI Model Engineer

Fintech • Software • Financial Services
Easy Apply
Remote or Hybrid
USA
581 Employees
200K-350K Annually

Shield AI Logo Shield AI

Design Engineer

Aerospace • Artificial Intelligence • Machine Learning • Robotics • Software
In-Office or Remote
2 Locations
170K-260K Annually

Similar Companies Hiring

Fortune Brands Innovations Thumbnail
Manufacturing
Deerfield, IL
10000 Employees
Rosendin Thumbnail
Other • Manufacturing
San Jose, CA
6219 Employees
Amalgamated Sugar Thumbnail
Food • Greentech • Agriculture • Industrial • Manufacturing
Boise, Idaho
768 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account