AI Platform Engineer

Posted 3 Days Ago
Be an Early Applicant
Raleigh, NC, USA
In-Office
100K-180K Annually
Expert/Leader
Artificial Intelligence • Information Technology • Software • Consulting
The Role
Design, build, and operate scalable, cloud-native AI inference platforms for production ML workloads. Optimize GPU utilization and inference performance, implement autoscaling, monitoring, security, and collaborate with ML teams to deploy and support enterprise AI models.
Summary Generated by Built In
Bright Vision Technologies is a technology consulting and software development company delivering cloud, AI, data, and enterprise solutions across the United States. This is a fantastic opportunity to join an established and well-respected organization offering tremendous career growth potential.
Job Title: AI Platform Engineer
Location: 100% Remote (Continental United States)
Position Type: Full-time, Direct W2
Salary Range: $100,000 – $150,000 per annum
Experience: 6+ years
Sponsorship: U.S. Citizens, Green Card Holders, EAD Holders, and H-1B transfer candidates are encouraged to apply. We are unable to sponsor new H-1B visa petitions for this position.
 

Job Summary:
We are seeking an AI Platform Engineer to design, build, and operate scalable AI inference platforms for production ML workloads. The ideal candidate will have expertise in distributed systems, LLM serving, GPU optimization, autoscaling, and cloud-native infrastructure, with a strong focus on performance, reliability, and observability.

Key Responsibilities:

  • Design and maintain scalable AI model serving platforms.
  • Optimize inference performance, GPU utilization, and request routing.
  • Build autoscaling, deployment, and monitoring solutions.
  • Implement caching, security, and high-availability strategies.
  • Collaborate with ML teams to deploy and support production AI models.

Required Qualifications:

  • 6+ years of experience in distributed systems, infrastructure, or ML platform engineering.
  • Strong proficiency in Python and Go, Rust, or C++.
  • Experience with LLM inference frameworks (vLLM, TensorRT-LLM), Kubernetes, cloud platforms, and GPU optimization.

Preferred Qualifications:

Bright Vision Technologies is a technology consulting and software development company delivering cloud, AI, data, and enterprise solutions across the United States. This is a fantastic opportunity to join an established and well-respected organization offering tremendous career growth potential.

Job Title

AI Platform Engineer

Location: 100% Remote (Continental United States)
Position Type: Full-time, Direct W2
Salary Range: $130,000–$180,000 Annually (based on experience)
Experience Required: 10+ Years

Sponsorship: U.S. Citizens, Green Card Holders, EAD Holders, and H-1B transfer candidates are encouraged to apply. We are unable to sponsor new H-1B visa petitions for this position.

Job Summary

Bright Vision Technologies is seeking a highly experienced AI Platform Engineer with 10+ years of experience in distributed systems, cloud-native infrastructure, and AI platform engineering to design, build, and operate enterprise-scale AI inference and machine learning platforms. The ideal candidate will possess deep expertise in LLM serving, GPU optimization, Kubernetes, cloud infrastructure, distributed systems, and MLOps, with a proven ability to deliver highly scalable, reliable, secure, and cost-efficient AI platforms supporting production machine learning workloads.

Key Responsibilities
  • Design, build, and maintain scalable AI inference and model-serving platforms for enterprise production environments.
  • Architect highly available, cloud-native infrastructure supporting Large Language Models (LLMs), foundation models, and machine learning services.
  • Optimize inference latency, throughput, GPU utilization, memory management, and request scheduling across distributed AI workloads.
  • Design autoscaling, workload orchestration, traffic management, and intelligent request routing strategies for AI services.
  • Implement model deployment, versioning, rollback, and lifecycle management using modern MLOps practices.
  • Develop monitoring, observability, logging, distributed tracing, and alerting solutions to ensure platform reliability and performance.
  • Implement caching strategies, API gateways, security controls, authentication, authorization, and high-availability architectures.
  • Collaborate with AI researchers, ML engineers, DevOps teams, and software engineers to deploy and support production AI models.
  • Drive cloud infrastructure optimization, resource utilization, FinOps initiatives, and operational excellence.
  • Mentor engineering teams, conduct architecture reviews, and establish best practices for AI platform engineering and cloud-native development.
  • Evaluate emerging AI infrastructure technologies, model-serving frameworks, and GPU acceleration techniques to drive continuous innovation.
Required Qualifications
  • Bachelor's or Master's degree in Computer Science, Computer Engineering, Artificial Intelligence, or a related technical discipline.
  • 10+ years of professional experience in distributed systems, infrastructure engineering, cloud platforms, or machine learning platform engineering.
  • Strong programming skills in Python and at least one systems programming language such as Go, Rust, or C++.
  • Extensive experience with Large Language Model (LLM) serving, model inference optimization, and production AI infrastructure.
  • Hands-on experience with vLLM, TensorRT-LLM, Triton Inference Server, Ray Serve, or similar AI serving frameworks.
  • Strong expertise in Kubernetes, container orchestration, Docker, and cloud-native application architectures.
  • Experience optimizing GPU workloads using CUDA, NVIDIA GPU technologies, distributed inference, and high-performance AI infrastructure.
  • Experience with cloud platforms including AWS, Microsoft Azure, or Google Cloud Platform (GCP).
  • Strong understanding of distributed systems, networking, scalability, observability, and security best practices.
  • Excellent analytical, communication, collaboration, and technical leadership skills.
Preferred Qualifications
  • Experience designing and operating multi-region AI platforms and globally distributed inference services.
  • Knowledge of model optimization techniques such as quantization, pruning, compression, speculative decoding, KV cache optimization, and mixed-precision inference.
  • Experience with MLOps, GitOps, Infrastructure as Code (Terraform, Bicep, CloudFormation), and CI/CD automation.
  • Familiarity with service mesh technologies such as Istio or Linkerd, API gateways, and event-driven architectures.
  • Contributions to open-source AI infrastructure projects, technical publications, patents, or conference presentations.
  • Experience implementing FinOps strategies, cloud cost optimization, and enterprise AI governance.
  • Experience with multi-region AI deployments and AI infrastructure.
  • Familiarity with model optimization techniques such as quantization or compression.
  • Open-source contributions or experience supporting large-scale AI APIs.

Interested in this opportunity? Apply today for immediate consideration!

Email your updated resume to [email protected]
Call or Text (908) 505-3545 if you have any questions.
Learn more about Bright Vision Technologies at www.bvteck.com.

We look forward to connecting with talented professionals and helping you take the next step in your career.

Bright Vision Technologies is an Equal Opportunity Employer.

Skills Required

  • 10+ years professional experience in distributed systems, infrastructure, or ML platform engineering
  • Bachelor's or Master's degree in Computer Science, Computer Engineering, AI, or related field
  • Strong programming skills in Python and at least one systems language (Go, Rust, or C++)
  • Experience with LLM serving and model inference optimization
  • Hands-on experience with vLLM, TensorRT-LLM, Triton Inference Server, Ray Serve or similar frameworks
  • Experience with Kubernetes, Docker, and cloud-native architectures
  • Experience optimizing GPU workloads using CUDA and NVIDIA GPU technologies
  • Experience with cloud platforms (AWS, Microsoft Azure, or Google Cloud Platform)
  • Strong understanding of distributed systems, networking, scalability, observability, and security best practices
  • Experience designing and operating multi-region AI platforms, model optimization (quantization, pruning), MLOps, GitOps, and IaC (Terraform, Bicep, CloudFormation)
  • Familiarity with service mesh technologies (Istio, Linkerd), API gateways, FinOps, and enterprise AI governance
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
53 Employees
Year Founded: 2020

What We Do

Bright Vision Technologies is a minority-owned organization founded in July 2020 and based in New Jersey, USA. The company specializes in delivering top-tier staffing and IT consulting services, including custom computer programming and systems design. Additionally, they are a product engineering firm with a flagship AI-powered talent intelligence and enterprise automation platform called Lumina, which helps transform IT into a strategic asset for their valued partners.

Similar Jobs

Remote or Hybrid
United States
300 Employees

CrowdStrike Logo CrowdStrike

Staff Engineer

Cloud • Computer Vision • Information Technology • Sales • Security • Cybersecurity
Remote or Hybrid
USA
11000 Employees
195K-290K Annually
Remote or Hybrid
United States
300 Employees

Bright Vision Technologies Logo Bright Vision Technologies

Platform Engineer

Artificial Intelligence • Information Technology • Software • Consulting
In-Office
2 Locations
53 Employees
100K-180K Annually

Similar Companies Hiring

Golden Pet Brands Thumbnail
Digital Media • eCommerce • Information Technology • Marketing Tech • Pet • Retail • Social Media
El Segundo, California
178 Employees
Kepler  Thumbnail
Fintech • Software
New York, New York
6 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account