Agentic AI Engineer (LLM)

Posted Yesterday
Be an Early Applicant
Hà Nội, VNM
In-Office
Mid level
Automotive • Greentech • Transportation • Manufacturing
The Role
Designs and deploys production-grade LLM agents, multi-agent systems, RAG pipelines, model fine-tuning workflows, function-calling systems, and AI evaluation frameworks. Optimizes inference through routing, caching, batching, quantization, and serving improvements. Builds APIs, observability, and reusable AI services while collaborating with product, engineering, and infrastructure teams. Researches emerging AI technologies, documents architectures and best practices, reviews technical work, and mentors junior engineers.
Summary Generated by Built In

VINFAST is a pioneering electric vehicle (EV) company committed to revolutionizing the automotive  industry with sustainable and innovative mobility solutions. As a leading player in the EV market,  VinFast is dedicated to delivering high-quality, cutting-edge electric vehicles that redefine the driving  experience. Our team consists of passionate professionals driven by a shared vision of creating a greener and  more sustainable future through innovation, technology, and excellence. 

In this role, you will be instrumental in SLP Center, using your skills to process and analyze  large volumes of data. You will collaborate with diverse teams, including Engineering, Quality Control,  VINFAST's suppliers, to create cutting-edge solutions that will drive the future of transportation. 

  • Design, develop, and optimize LLM-based AI Agents for real-world applications with a focus on accuracy, latency, scalability, and reliability.  
  • Research, evaluate, and integrate state-of-the-art foundation models (OpenAI, Gemini, Claude, open-source LLMs, etc.) for different AI tasks.  
  • Fine-tune, align, and optimize large language models using techniques such as SFT, LoRA/QLoRA, DPO, preference optimization, and model distillation.  
  • Design multi-agent systems, agent orchestration, planning, memory, tool calling, and workflow execution architectures.  
  • Develop Retrieval-Augmented Generation (RAG) pipelines, including retrieval, reranking, indexing, embedding optimization, and knowledge integration.  
  • Design and optimize function calling pipelines for structured task execution and external tool integration.  
  • Build evaluation frameworks for AI systems, including automated benchmarking, offline evaluation, online A/B testing, and hallucination detection.  
  • Optimize model inference performance through prompt engineering, routing, caching, batching, quantization, speculative decoding, and serving optimization.  
  • Develop and maintain APIs and reusable AI services for rapid integration into products.  
  • Build observability capabilities including logging, tracing, monitoring, and performance analytics for AI agents in production.  
  • Continuously research emerging AI technologies, models, and agent frameworks, rapidly validating and applying promising approaches to production systems.  
  • Collaborate with product managers, backend engineers, frontend engineers, and infrastructure teams to deliver production-ready AI features.  
  • Produce technical documentation, architecture designs, best practices, and reusable components for the AI platform.  
  • Mentor junior engineers, conduct technical reviews, and contribute to engineering standards and AI best practices


Requirements
  • Bachelor's or Master's degree in Computer Science, Artificial Intelligence, Machine Learning, Data Science, Software Engineering, or a related field.  
  • 3+ years of experience in AI/ML engineering, NLP, LLM applications, or a related software engineering role. (Adjust based on seniority.)  
  • Strong understanding of Large Language Models (LLMs), Transformer architectures, and modern generative AI techniques.  
  • Hands-on experience building AI agents using frameworks such as LangGraph, LangChain, LlamaIndex, Semantic Kernel, AutoGen, CrewAI, or equivalent.  
  • Experience designing and implementing Retrieval-Augmented Generation (RAG) systems, including embedding models, vector databases, retrieval, reranking, and knowledge indexing.  
  • Experience fine-tuning LLMs using techniques such as LoRA, QLoRA, Supervised Fine-Tuning (SFT), Direct Preference Optimization (DPO), or other parameter-efficient training methods.  
  • Strong programming skills in Python and familiarity with AI/ML libraries such as PyTorch, Hugging Face Transformers, PEFT, vLLM, or equivalent.  
  • Experience integrating foundation models through APIs (OpenAI, Gemini, Anthropic, or open-source LLMs) and evaluating model performance.  
  • Solid understanding of prompt engineering, function calling, structured output generation, tool integration, and agent orchestration.  
  • Experience deploying AI services in production environments, including model serving, inference optimization, containerization (Docker), and cloud platforms.  
  • Knowledge of observability for AI systems, including logging, tracing, monitoring, evaluation pipelines, and performance analysis.  
  • Familiarity with software engineering best practices, including Git, CI/CD, API development, testing, and code review.  
  • Strong analytical and problem-solving skills with the ability to rapidly evaluate and adopt emerging AI technologies.  
  • Excellent communication skills and ability to collaborate effectively in cross-functional teams.  
    Preferred Qualifications 
  • Experience with distributed model training and inference optimization (DeepSpeed, FSDP, TensorRT-LLM, vLLM, SGLang, etc.).  
  • Experience with multi-agent systems, workflow orchestration, planning, and memory architectures.  
  • Knowledge of model evaluation methodologies, benchmarking, hallucination detection, and AI safety.  
  • Experience building voice agents, conversational AI, or virtual assistants.  
  • Contributions to open-source AI projects, research publications, Kaggle competitions, or personal AI projects are a plus.  
  • Familiarity with Kubernetes, GPU infrastructure, distributed systems, and MLOps platforms is preferred. 


Benefits
  • Competitive salary
  • Premium healthcare package, including PVI insurance & annual health check-ups
  • 13th-month salary & performance bonuses to reward your contributions
  • Enjoy preferential pricing for services within the Vingroup ecosystem including Vinmec, Vinpearl, and Vinschool...
  • Opportunity to collaborate with and learn from industry-leading professionals in the automotive domain
Work Location: Technopark Tower, Vinhomes Ocean Park, Gia Lam, Hanoi, Vietnam

With respect to all your personal data shared to VinFast in the application and the entire recruitment process of VinFast, by clicking “Apply”, submitting your resumé/CV and/or participating in VinFast's recruitment process, you agree that you have read VinFast's Personal Data Protection Policy ("Policy") posted at https://vinfastauto.com/vn_vi/dieu-khoan-phap-ly or https://vinfast.vn/privacy-policy/, you agree to the Policy and consent for VinFast to process your personal data in accordance with the Policy and the applicable regulations on personal data protection.
To all recruitment agencies: VinFast does not accept agency resumes. Please do not forward resumes to our careers alias or other VinFast employees. VinFast is not responsible for any fees related to unsolicited resumes.

Skills Required

  • Bachelor's or Master's degree in Computer Science, Artificial Intelligence, Machine Learning, Data Science, Software Engineering, or a related field
  • 3+ years of experience in AI/ML engineering, NLP, LLM applications, or related software engineering
  • Strong understanding of large language models, Transformer architectures, and modern generative AI techniques
  • Hands-on experience building AI agents with LangGraph, LangChain, LlamaIndex, Semantic Kernel, AutoGen, CrewAI, or equivalent
  • Experience designing and implementing Retrieval-Augmented Generation systems, including embeddings, vector databases, retrieval, reranking, and knowledge indexing
  • Experience fine-tuning LLMs using LoRA, QLoRA, SFT, DPO, or other parameter-efficient training methods
  • Strong Python programming skills and familiarity with PyTorch, Hugging Face Transformers, PEFT, vLLM, or equivalent
  • Experience integrating foundation models through OpenAI, Gemini, Anthropic, or open-source model APIs and evaluating performance
  • Understanding of prompt engineering, function calling, structured output generation, tool integration, and agent orchestration
  • Experience deploying AI services in production, including model serving, inference optimization, Docker, and cloud platforms
  • Knowledge of AI system observability, including logging, tracing, monitoring, evaluation pipelines, and performance analysis
  • Familiarity with Git, CI/CD, API development, testing, and code review
  • Strong analytical and problem-solving skills
  • Excellent communication and cross-functional collaboration skills
  • Experience with distributed model training and inference optimization using DeepSpeed, FSDP, TensorRT-LLM, vLLM, SGLang, or similar
  • Experience with multi-agent systems, workflow orchestration, planning, and memory architectures
  • Knowledge of model evaluation, benchmarking, hallucination detection, and AI safety
  • Experience building voice agents, conversational AI, or virtual assistants
  • Contributions to open-source AI projects, research publications, Kaggle competitions, or personal AI projects
  • Familiarity with Kubernetes, GPU infrastructure, distributed systems, and MLOps platforms
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
29,900 Employees
Year Founded: 2017

What We Do

VinFast is a Vietnamese automotive manufacturer focused on electric mobility. The company designs and produces smart electric vehicles, including cars, SUVs, e-buses and e-scooters, combining advanced technology with highly automated manufacturing. Its mission is to make sustainable transportation accessible worldwide and accelerate the transition to an all-electric future through innovative, environmentally friendly products, charging solutions, warranties, and customer services.

Similar Jobs

Mastercard Logo Mastercard

Analyst, Business Development

Blockchain • Fintech • Payments • Consulting • Cryptocurrency • Cybersecurity • Quantum Computing
Hybrid
Hà Nội, VNM
38800 Employees

UL Solutions Logo UL Solutions

Intern, Chemical Safety Testing

Automotive • Professional Services • Software • Consulting • Energy • Chemical • Renewable Energy
Remote or Hybrid
Việt Nam
15000 Employees

UL Solutions Logo UL Solutions

Intern, Toy/Lab Tools Testing

Automotive • Professional Services • Software • Consulting • Energy • Chemical • Renewable Energy
Remote or Hybrid
Việt Nam
15000 Employees

UL Solutions Logo UL Solutions

Intern, Textile Testing (Softline)

Automotive • Professional Services • Software • Consulting • Energy • Chemical • Renewable Energy
Remote or Hybrid
Việt Nam
15000 Employees

Similar Companies Hiring

Fortune Brands Innovations Thumbnail
Manufacturing
Deerfield, IL
10000 Employees
Rosendin Thumbnail
Other • Manufacturing
San Jose, CA
6219 Employees
Amalgamated Sugar Thumbnail
Food • Greentech • Agriculture • Industrial • Manufacturing
Boise, Idaho
768 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account