MLOps Engineer Internship

Sorry, this job was removed at 06:21 p.m. (UTC) on Monday, Aug 10, 2026
Be an Early Applicant
Hiring Remotely in United States
Remote
Internship
Artificial Intelligence • Cloud • Information Technology • Consulting
The Role
Build and maintain MLOps infrastructure: CI/CD for models, monitoring, deployment into a NestJS monolith, feature store/versioning, retraining workflows, A/B experimentation, and LLM operationalization.
Summary Generated by Built In
Overview: Operationalizing AI and Personalization

The MLOps Engineer is crucial for bridging the gap between data science and production, responsible for the reliable, scalable, and secure deployment of machine learning models. You will operationalize the models powering our AI agents and the e-commerce personalization systems, ensuring continuous integration, delivery, and monitoring of our predictive analytics and recommendation engines.

Internship Details

Duration: 3 months
Start Date: Immediate
Location: Remote
Stipend: None initially. Based on your first-quarter performance, you may be offered a paid full-time opportunity, or even be absorbed directly by the client as an FTE.

Key Responsibilities & Core Projects

You will build the automation infrastructure that turns static models into continuously improving production systems.

  • Model CI/CD Pipelines: Design and build robust CI/CD pipelines dedicated to the machine learning lifecycle: automated model training, validation, and deployment using tools integrated with our main Makefile CI/CD setup.

  • Model Monitoring & Tracking: Implement comprehensive monitoring and alerting for model performance (e.g., drift detection, prediction accuracy, latency) and track experiments and model artifacts using version control tools.

  • Production Deployment: Operationalize the deployment of ML models powering AI agents and e-commerce services, ensuring they integrate seamlessly into the NestJS modular monolith architecture.

  • Versioning & Feature Stores: Manage model versioning and lineage. Collaborate on the design and maintenance of a centralized Feature Store to ensure consistent data for training and serving.

  • Experimentation Infrastructure: Implement and manage the infrastructure necessary for A/B testing different model versions or personalization strategies in a production environment (e-commerce storefront).

  • Retraining Workflows: Define and automate the model retraining workflows based on data drift or performance degradation triggers, ensuring models remain relevant to the dynamic supply chain and customer behavior.

Required Technologies & Tools

Candidates must possess hands-on expertise in the tools and methodologies used for production ML and MLOps:

  • MLOps Tools: Experience with model registries, experiment tracking, and serving platforms (e.g., MLflow, Kubeflow, Sagemaker).

  • CI/CD & Automation: Proficiency in building pipelines (using Python/Bash scripting) and experience with Docker and Terraform.

  • Data & Compute: Experience managing data pipelines for ML (ETL/ELT) and optimizing compute resources for training and inference.

  • Programming: Strong proficiency in Python and familiarity with TypeScript/Node.js for deployment integration.

  • Methodology: Deep understanding of MLOps best practices, responsible AI principles, and monitoring concepts.

AI Agent Focus

You will ensure the reliability and continuous improvement of the core AI layer.

  • LLM Operationalization: Implement specific pipelines for the fine-tuning, validation, and deployment of Large Language Models (LLMs) used in our AI agents.

  • Agent Performance Tracking: Develop metrics and tracking systems to measure the business impact and operational efficiency of multi-agent systems and recommendation engines.

  • Framework Integration: Operationalize models built using frameworks like LangChain or LlamaIndex, ensuring they are secure, versioned, and scalable in a production environment.

Success Metrics & Career Path

Performance will be measured by:

  • Deployment Velocity: Speed and reliability of deploying new or retrained models to production.

  • Model Performance: Maintaining model accuracy and minimizing performance drift in production.

  • Pipeline Automation: Percentage of the ML lifecycle (training, validation, deployment) that is fully automated.

Mentorship Structure: Reports to the Solution Architect or Head of Technology, collaborating closely with Data Architects, Data Scientists, and SREs to maintain a reliable AI ecosystem.

Skills Required

  • Experience with model registries, experiment tracking, and serving platforms (e.g., MLflow, Kubeflow, SageMaker).
  • Proficiency in building CI/CD pipelines using Python and Bash scripting.
  • Experience with Docker and Terraform.
  • Experience managing data pipelines for ML (ETL/ELT) and optimizing compute resources for training and inference.
  • Strong proficiency in Python.
  • Familiarity with TypeScript/Node.js for deployment integration.
  • Deep understanding of MLOps best practices, responsible AI principles, and model monitoring concepts (drift detection, metrics, alerting).
  • Experience operationalizing LLMs, including fine-tuning, validation, and deployment.
  • Experience integrating frameworks like LangChain or LlamaIndex.
  • Experience implementing experimentation infrastructure and A/B testing in production.
  • Experience with model versioning, lineage, and feature store design/maintenance.

Similar Jobs

Kustomer Logo Kustomer

Head of Revenue Operations

Artificial Intelligence • Enterprise Web • Machine Learning • Natural Language Processing • Software • Conversational AI • Automation
Remote or Hybrid
New York City, NY, USA
200 Employees
200K-240K Annually

Headway Logo Headway

Staff Software Engineer

Consumer Web • Healthtech • Professional Services • Social Impact • Software
In-Office or Remote
3 Locations
819 Employees
265K-331K Annually

Headway Logo Headway

Business Operations Manager

Consumer Web • Healthtech • Professional Services • Social Impact • Software
Remote
USA
819 Employees
122K-190K Annually

Headway Logo Headway

Product Manager

Consumer Web • Healthtech • Professional Services • Social Impact • Software
Remote
USA
819 Employees
265K-332K Annually
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company

What We Do

Metasys is a global management and technology consulting firm that helps organizations improve performance through digital transformation. It develops strategies and technology solutions spanning AI, cloud, emerging technologies, digital engineering, intelligent manufacturing, supply chain, and managed services. The company serves industries including aerospace and defense, automotive, financial services, healthcare, software, retail, and utilities, combining technology, data, and industry expertise to deliver measurable impact.

Similar Companies Hiring

Hanover Park Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
42 Employees
Golden Pet Brands Thumbnail
Digital Media • eCommerce • Information Technology • Marketing Tech • Pet • Retail • Social Media
El Segundo, California
178 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account