Embedded AI Engineer

Posted Yesterday
Be an Early Applicant
Tempe, AZ, USA
In-Office
100K-150K Annually
Senior level
Artificial Intelligence • Information Technology • Software • Consulting
The Role
Design, optimize, and deploy ML models for resource-constrained edge devices. Apply compression, quantization, and hardware-aware tuning; build cross-platform inference runtimes; ensure secure, privacy-preserving on-device workflows; collaborate with hardware/firmware teams; develop benchmarking, telemetry, staged rollout, and documentation for production edge AI.
Summary Generated by Built In
Embedded AI Engineer - Remote 
Bright Vision Technologies is a technology consulting and software development company delivering cloud, AI, data, and enterprise solutions across the United States. 
This is a fantastic opportunity to join an established and well-respected organization offering tremendous career growth potential. 
 
Job Title: Embedded AI Engineer
Location: 100% Remote (U.S.) 
Position Type: Full-time, Direct W2 
Salary Range: $100,000–$150,000 Annually 
Experience Required: 6+ years 
Sponsorship: U.S. Citizens, Green Card Holders, EAD Holders, and H-1B transfer candidates are encouraged to apply. We are unable to sponsor new H-1B visa petitions for this position. 
Job Summary 
We are looking for an Embedded AI Engineer to design, optimize, and deploy machine learning models that run efficiently on resource-constrained edge devices, including mobile platforms, embedded systems, and specialized accelerators. The role requires deep expertise in model compression, quantization, and hardware-aware optimization, along with strong systems engineering skills to ship reliable AI capabilities outside the data center. The ideal candidate has shipped edge AI in production environments where compute, memory, energy, and connectivity constraints fundamentally shape the engineering trade-offs.
Key Responsibilities
  • Design and implement edge AI solutions optimized for diverse hardware including mobile SoCs, NPUs, and embedded accelerators.
  • Apply quantization, pruning, distillation, and architectural optimization to fit models within edge constraints.
  • Tune model performance for latency, energy efficiency, and memory footprint on target hardware.
  • Build cross-platform inference runtimes leveraging frameworks such as TensorFlow Lite, ONNX Runtime, and Core ML.
  • Optimize models for specific accelerator backends including DSPs, NPUs, and mobile GPUs.
  • Implement on-device model update, versioning, and rollback workflows that allow safe staged rollouts to large device populations and rapid recovery if a model release behaves unexpectedly in the field.
  • Design hybrid edge-cloud architectures that gracefully degrade based on connectivity and device capability.
  • Build telemetry pipelines that respect privacy while enabling continuous improvement.
  • Collaborate with hardware, firmware, and product teams to align AI capabilities with device constraints.
  • Implement secure execution paths, model protection, and integrity verification on edge devices.
  • Develop benchmarking suites that characterize accuracy, latency, and energy trade-offs across devices.
  • Drive responsible AI considerations including on-device privacy and bias evaluation.
  • Maintain comprehensive, current technical documentation — including architecture diagrams, design decisions, configuration references, runbooks, and operational procedures — so that the system remains supportable, auditable, and easy to onboard new engineers onto over time.
  • Stay current with edge AI hardware and software developments, regularly review release notes and community discussions, and translate noteworthy advances into concrete recommendations and adoption proposals for the team.
Required Qualifications
  • Bachelor’s or Master’s degree in Computer Science, Computer Engineering, or a related field.
  • Six or more years of experience in ML engineering, with significant work on edge or mobile AI.
  • Strong proficiency in Python and C++.
  • Hands-on experience with model compression, quantization, and pruning techniques.
  • Experience with at least one major edge inference framework.
  • Solid understanding of mobile and embedded hardware architectures.
  • Experience deploying ML models to production on mobile or embedded platforms.
  • Strong performance engineering and profiling skills.
  • Familiarity with on-device privacy and security considerations.
  • Strong communication and cross-functional collaboration skills.
Preferred Qualifications
  • Experience with custom NPU or DSP toolchains.
  • Familiarity with federated learning or on-device personalization.
  • Exposure to safety-critical or industrial edge deployments.
  • Open-source contributions to edge AI frameworks.
  • Experience optimizing LLMs for on-device inference.
How to Apply 
Would you like to know more about this opportunity? For immediate consideration, please send your resume to [email protected] or contact us at (908) 505-3544. Learn more about Bright Vision Technologies at www.bvteck.com
Bright Vision Technologies is an Equal Opportunity Employer. 
 

Skills Required

  • Bachelor's or Master's degree in Computer Science, Computer Engineering, or related field
  • Six or more years of experience in ML engineering with significant edge or mobile AI work
  • Strong proficiency in Python and C++
  • Hands-on experience with model compression, quantization, and pruning techniques
  • Experience with at least one major edge inference framework (e.g., TensorFlow Lite, ONNX Runtime, Core ML)
  • Solid understanding of mobile and embedded hardware architectures
  • Experience deploying ML models to production on mobile or embedded platforms
  • Strong performance engineering and profiling skills
  • Familiarity with on-device privacy and security considerations
  • Strong communication and cross-functional collaboration skills
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
53 Employees
Year Founded: 2020

What We Do

Bright Vision Technologies is a minority-owned organization founded in July 2020 and based in New Jersey, USA. The company specializes in delivering top-tier staffing and IT consulting services, including custom computer programming and systems design. Additionally, they are a product engineering firm with a flagship AI-powered talent intelligence and enterprise automation platform called Lumina, which helps transform IT into a strategic asset for their valued partners.

Similar Jobs

Lessen LLC Logo Lessen LLC

District Lead

Cloud • Real Estate • Software • PropTech
Remote or Hybrid
Scottsdale, Arizona, USA
713 Employees

monday.com Logo monday.com

Enterprise Account Executive

Artificial Intelligence • Productivity • Sales • Software
Remote or Hybrid
US
3048 Employees
130K-180K Annually

monday.com Logo monday.com

Enterprise Account Manager

Artificial Intelligence • Productivity • Sales • Software
Remote or Hybrid
US
3048 Employees
155K-180K Annually

The Aerospace Corporation Logo The Aerospace Corporation

Operations Specialist

Aerospace • Artificial Intelligence • Cloud • Machine Learning • Software • Cybersecurity • Defense
Hybrid
Phoenix, AZ, USA
4600 Employees

Similar Companies Hiring

Golden Pet Brands Thumbnail
Digital Media • eCommerce • Information Technology • Marketing Tech • Pet • Retail • Social Media
El Segundo, California
178 Employees
Kepler  Thumbnail
Fintech • Software
New York, New York
6 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account