Senior MLOps Engineer

Posted Yesterday
Be an Early Applicant
Hiring Remotely in Ukraine
Remote
Senior level
Information Technology • Consulting
The Role
Build and operate enterprise MLOps and LLMOps platforms on AWS using SageMaker, Bedrock, MLflow, GPU infrastructure, CI/CD, and infrastructure as code. The role covers model deployment, agent orchestration, RAG, voicebot STT/TTS pipelines, monitoring, drift detection, cost controls, data tokenization, privacy, and governance. Collaboration includes data engineering, AI architecture, security, and cloud teams supporting hybrid telecom AI initiatives.
Summary Generated by Built In

Client Overview:
Our client is an Azerbaijani telecommunications company, the largest mobile network operator in Azerbaijan. The main products are: Fixed telephony, Mobile telephony, Internet services, Wireless broadband, and Value-added services.

Project Objectives:
The primary goal is to accelerate the client’s Data & AI initiatives via a secure, hybrid cloud foundation on AWS while systematically modernizing the IT estate as part of the cloud migration.

Key Project Objectives include:

  • Cloud Foundation & Landing Zone: Deploy target hybrid network architectures, establishing a secure Landing Zone and hybrid Data/AI platforms on AWS.
  • Security, Compliance & Governance: Operationalize on-prem tokenization (achieving zero raw PII in the cloud), resolve policy blockers to include AWS in the ISMS, and establish a Cloud Center of Excellence (CCoE) to govern Cloud adoption.
  • AI Chatbot & Voicebot Design & Implementation: Develop and operationalize a flagship Customer Care Chatbot and Voicebot as the first hybrid-setup consumer.
Responsibilities:
  • Build, operationalize, and automate end-to-end MLOps pipelines using Amazon SageMaker Pipelines and MLflow for experiment tracking, model versioning, and registry lifecycle management.
  • Design, deploy, and manage production SageMaker inference endpoints (real-time, serverless, and batch) and Amazon Bedrock API integrations for LLM/SLM deployment with cost controls and latency optimization (Bedrock API Gatekeeper).
  • Implement AgentOps / LLMOps frameworks (AgentCore, Bedrock Guardrails, Promptfoo) to manage multi-agent orchestration, prompt evaluation, safety guardrails, and RAG retrieval pipelines.
  • Operationalize real-time STT / TTS (Speech-to-Text / Text-to-Speech) voicebot pipelines and low-latency speech inference on hybrid/cloud GPU node pools for the flagship Customer Care Voicebot.
  • Optimize specialized GPU node pools (NVIDIA A100/L40S / EC2 GPU instance types) for Azerbaijani SLM/LLM model training, fine-tuning, and scalable inference workloads.
  • Establish automated CI/CD for Machine Learning using GitLab CI/CD pipelines and Infrastructure-as-Code (Terraform or AWS CDK) to enforce security-gated MLOps promotion workflows (from SageMaker Canvas/Sandbox to production).
  • Integrate data de-identification, Format Preserving Encryption (FPE), and tokenization wrappers into ML data pipelines to ensure zero raw PII enters AWS cloud environments during model training and inference.
  • Set up telemetry, performance monitoring, model drift detection, and cost anomaly alerting for AI/ML workloads using Amazon CloudWatch, Splunk, and FinOps spend control frameworks.
  • Collaborate with Data Engineering, AI Architects, and Cloud Teams to integrate vector storage/retrieval (RAG), Apache Spark/EMR-on-EKS runtimes, and local tokenization databases.
  • Author technical MLOps runbooks, model deployment procedures, governance documentation, and disaster recovery playbooks.
Requirements:
  • 4+ years of hands-on experience in MLOps, DataOps, or Platform Engineering with a primary focus on enterprise Amazon SageMaker (Pipelines, Feature Store, Model Registry, Endpoints).
  • Proven experience deploying and operating Generative AI, LLM/SLM models, and Amazon Bedrock services alongside agentic frameworks and RAG pipelines.
  • Hands-on expertise with MLflow for experiment tracking, model registry, and lifecycle management.
  • Solid experience in GPU optimization and orchestration (NVIDIA A100/L40S, AWS EC2 GPU instances) for model training, fine-tuning, and low-latency real-time inference (STT/TTS voice pipelines).
  • Proficient in building CI/CD for Machine Learning (GitLab CI/CD, GitHub Actions) and Infrastructure-as-Code (Terraform or AWS CDK).
  • Practical knowledge of LLMOps / AgentOps tools and methodologies (AgentCore, prompt evaluations, Bedrock Guardrails, vector databases for RAG).
  • Strong understanding of data security, privacy, and tokenization (FPE, handling sensitive/PII data within ML pipelines).
  • Proficient in Python, PySpark, Docker, and Kubernetes/EKS fundamentals for containerized ML workloads.
Nice-to-Have Skills:
  • AWS Certified Machine Learning – Specialty certification.
  • AWS Certified Solutions Architect – Associate/Professional or AWS Certified DevOps Engineer – Professional.
  • Experience in telecom domain AI/ML applications, low-latency real-time voice/chat processing (ASR/TTS), or hybrid cloud data sovereignty architectures.
  • Experience with EMR-on-EKS, Starburst/Athena, or Apache Iceberg data lake integrations.
Soft Skills & Team Fit:
  • Strong critical thinking, problem-solving, and analytical skills.
  • Excellent communication and collaboration skills to work closely with cross-functional teams (Data Engineering, AI/GenAI Engineers, Security, Cloud/Infrastructure).
  • Results-oriented, proactive mindset with strong ownership of deliverables within an Agile / Scrum framework.
  • Upper-Intermediate+ English level (written and spoken).
What we propose:
  • Opportunity to lead critical, high-impact Data & AI platform delivery for a major telecommunications operator.
  • Hands-on work with modern MLOps and GenAI stack (Amazon SageMaker, Amazon Bedrock, MLflow, AgentCore, STT/TTS voicebot pipelines).
  • Flexible remote work options with structured, predictable collaboration within a well-balanced team.

We offer*:

  • Flexible working format - remote, office-based or flexible
  • A competitive salary and good compensation package
  • Personalized career growth
  • Professional development tools (mentorship program, tech talks and trainings, centers of excellence, and more)
  • Active tech communities with regular knowledge sharing
  • Education reimbursement
  • Memorable anniversary presents
  • Corporate events and team buildings
  • Other location-specific benefits

*not applicable for freelancers

Skills Required

  • 4+ years of hands-on experience in MLOps, DataOps, or Platform Engineering, primarily with Amazon SageMaker.
  • Experience with SageMaker Pipelines, Feature Store, Model Registry, and Endpoints.
  • Experience deploying and operating Generative AI, LLM/SLM models, Amazon Bedrock, agentic frameworks, and RAG pipelines.
  • Hands-on expertise with MLflow for experiment tracking, model registry, and lifecycle management.
  • Experience optimizing and orchestrating NVIDIA A100/L40S and AWS EC2 GPU instances for training, fine-tuning, and inference.
  • Experience building machine learning CI/CD pipelines using GitLab CI/CD or GitHub Actions.
  • Experience with Infrastructure as Code using Terraform or AWS CDK.
  • Knowledge of LLMOps and AgentOps tools and methodologies, including AgentCore, prompt evaluation, Bedrock Guardrails, and vector databases.
  • Strong understanding of data security, privacy, tokenization, and Format Preserving Encryption in ML pipelines.
  • Proficiency in Python, PySpark, Docker, and Kubernetes/EKS fundamentals.
  • AWS Certified Machine Learning Specialty certification.
  • AWS Certified Solutions Architect Associate/Professional or AWS Certified DevOps Engineer Professional certification.
  • Experience with telecom AI/ML applications, low-latency voice processing, or hybrid cloud data sovereignty architectures.
  • Experience with EMR-on-EKS, Starburst/Athena, or Apache Iceberg.
  • Upper-Intermediate or higher English proficiency.
  • Strong critical thinking, problem-solving, analytical, communication, collaboration, ownership, and Agile/Scrum skills.
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Valletta
2,135 Employees
Year Founded: 2002

What We Do

N-iX is a global software solutions and engineering services company that helps world’s leading organizations turn challenges into lasting business value, operational efficiency, and revenue growth using advanced technology. Whether you need to build a custom solution, modernize your digital product or acquire extra tech expertise - we have the experience and capabilities to ensure your success. With over 2,000 professionals in 25 countries across Europe and the Americas, N-iX offers expert solutions in cloud, data analytics, embedded software, IoT, AI, machine learning, and other tech domains. Being in business for over two decades, we have worked with dozens of industry-leading enterprises and Fortune 500 companies creating value across a wide variety of sectors, including finance, manufacturing, supply chain, retail, e-commerce, healthcare, and more. Our unique combination of business domain expertise and technical know-how enables us to effectively collaborate with ISVs, tech companies, and enterprises of all sizes. Thanks to the strong tech ecosystem and partnerships with AWS, GCP, Microsoft, SAP, OpenText, Snowflake, and others, we bring extra speed, scale and efficiency to more than 160 organizations across the globe. N-iX is recognized by numerous industry awards, such as CRN Solution Provider 500, Global Outsourcing 100 by IAOP, ISG Provider Lens™, Modern Application Development services providers by Forrester, etc

Similar Jobs

Point Wild Logo Point Wild

Senior MLOps Engineer

Software • Cybersecurity
Remote
Ukraine
105 Employees

Boeing Logo Boeing

Engineering Manager

Aerospace • Information Technology • Software • Cybersecurity • Design • Defense • Manufacturing
Remote or Hybrid
Kyiv City, UKR
170000 Employees

Superhuman Logo Superhuman

Software Engineer

Artificial Intelligence • Information Technology • Machine Learning • Natural Language Processing • Productivity • Software • Generative AI
Remote or Hybrid
Ukraine
1500 Employees

Pfizer Logo Pfizer

Sustainability Senior Manager

Artificial Intelligence • Healthtech • Machine Learning • Natural Language Processing • Biotech • Pharmaceutical
Remote or Hybrid
30 Locations
121990 Employees
112K-207K Annually

Similar Companies Hiring

Axle Health Thumbnail
Artificial Intelligence • Healthtech • Information Technology • Logistics
Santa Monica, CA
25 Employees
NODA AI Thumbnail
Artificial Intelligence • Information Technology • Software • Cybersecurity
Sydney, AU
54 Employees
Golden Pet Brands Thumbnail
Digital Media • eCommerce • Information Technology • Marketing Tech • Pet • Retail • Social Media
El Segundo, California
178 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account