AI Ops Engineer

Posted 2 Days Ago
Be an Early Applicant
Chennai, Tamil Nadu, IND
In-Office
Mid level
Information Technology • Marketing Tech • Analytics
The Role
Build and operate secure, scalable cloud infrastructure for AI and generative AI applications. Responsibilities include MLOps and LLMOps pipelines, CI/CD automation, infrastructure as code, container orchestration, observability, reliability engineering, security governance, and developer enablement. The role supports AI engineers, software engineers, and analytical modelers while managing production deployments, model versioning, evaluation, monitoring, drift detection, and hallucination mitigation.
Summary Generated by Built In

OneMagnify is an AI native, platform-enabled B2B digital agency operating at the intersection of data, technology, and creativity. We help complex organizations drive measurable business outcomes by building smarter customer experiences and delivering highly integrated solutions across digital, media, and technology. By combining deep industry expertise with advanced analytics and artificial intelligence, we enable our clients to make better decisions, move faster, and compete more effectively in dynamic markets.

We are seeking a motivated and talented AI Operations Engineer with strong cloud infrastructure and AI operations expertise to build, automate, and secure the platform that powers advanced artificial intelligence solutions.

The Impact You’ll Have:

In this role, you will support quantitative analytics and AI engineering teams by designing, automating, and operating end-to-end cloud and AI infrastructure. Responsibilities include managing CI/CD pipelines, containerized microservices, observability platforms, and governance controls to ensure AI applications and models run safely, reliably, and at scale in production environments.

What you’ll do:

  • Cloud Platform Engineering: Architect and operate highly available, multi-service AI infrastructure on the cloud, managing the full lifecycle of compute, storage, networking, and security for AI workloads.
  • AI-Ops / LLMOps Pipelines: Establish robust MLOps and LLMOps pipelines covering the end-to-end lifecycle of Generative AI tools — including model deployment, version control, prompt and artifact tracking, automated evaluation, and continuous monitoring for performance, drift, and hallucination mitigation.
  • CI/CD Automation: Design and maintain automated build, test, and deployment pipelines for full-stack AI applications, ensuring seamless and secure continuous integration and delivery across front-end, back-end, and AI components.
  • Infrastructure as Code: Manage cloud infrastructure using Infrastructure as Code (e.g., Terraform, Cloud Build, Kubernetes manifests) to deliver reproducible, auditable, and scalable environments.
  • Containerization & Orchestration: Build and operate containerized microservices (Docker/Kubernetes), managing scaling, rolling deployments, resource optimization, and service resilience for AI workloads.
  • Observability & Reliability: Implement comprehensive monitoring, logging, tracing, alerting, and SRE practices to ensure platform reliability, availability, and performance of AI applications in production.
  • Security & Governance: Embed security across the platform — managing user identities and controlling access rights, safeguarding sensitive credentials, network policies, and data protection — ensuring all AI workloads meet Ford's strict data privacy, security, and compliance standards.
  • Developer Enablement: Work closely with AI engineers, software engineers, and analytical modelers to provide self-service tooling, environments, and automated workflows that remove friction from development to production.

What you’ll need:

  • Education: Master's or Bachelor's degree in Computer Science, Software Engineering, Cloud Computing, Data Engineering, or a related technical discipline.
  • DevOps/Cloud Engineering Experience: 3–5 years of overall experience in cloud engineering, DevOps, or SRE, with at least 1–2 years of dedicated, hands-on experience deploying and operating AI, ML, and Generative AI applications in production.
  • AI/MLOps Engineering: Proven track record of implementing CI/CD for AI workloads, containerization (Docker/Kubernetes), and cloud infrastructure management (Terraform, Cloud Build, GKE).
  • Automation & Reliability: Demonstrated experience with infrastructure automation, incident response, and building observable, self-healing production systems.
  • Cloud Platform: Extensive hands-on experience with a leading cloud provider (e.g., Google Cloud Platform), including Cloud Run, Cloud Build, GKE, GCS, BigQuery, IAM, VPC networking, and Secret Manager.
  • DevOps / SRE Practices: Strong proficiency in CI/CD tooling, GitOps, containerization (Docker), orchestration (Kubernetes), Infrastructure as Code (Terraform), and cloud-native monitoring and logging.
  • AI-Ops / LLMOps: Proficiency in tools and platforms for model deployment, prompt/model versioning, evaluation, tracing, and monitoring of LLM and GenAI outputs.
  • Scripting & Automation: Hands-on experience with a programming/scripting language such as Python, or Bash for automating infrastructure and operational tasks, and for building tooling that serves collaborators.
  • Software Engineering Fundamentals: Solid understanding of full-stack application architecture and modern deployment patterns for AI tools, with the ability to integrate front-end, back-end, and AI services reliably.
  • Security & Compliance: Familiarity with cloud security best practices, identity and access management, secrets management, and compliance standards in regulated environments.
  • Analytics Workflow Understanding: Awareness of the typical workflows of data scientists and modelers (data wrangling, feature engineering, model validation) so you can build reliable platforms and pipelines that serve them.

Future-Ready Skills (Nice to Have):

  • Experience in integrated marketing, digital agency, marketing services, or consulting environments preferred.
  • Previous exposure to the Banking, Financial Services, or Credit Analytics industries. Experience with Machine Learning engineering and model serving frameworks, or relevant cloud certifications (e.g., Google Cloud Professional DevOps Engineer / Cloud Architect).

Benefits

We offer a comprehensive benefits package including Medical Insurance, PF, Gratuity, paid holidays, and more.

We are an equal opportunity employer

We believe that Innovative ideas and solutions start with unique perspectives. That’s why we’re committed to providing every employee a workplace that’s free of discrimination and intolerance. We’re proud to be an equal opportunity employer and actively search for like-minded people to join our team.

Skills Required

  • Bachelor’s or master’s degree in Computer Science, Software Engineering, Cloud Computing, Data Engineering, or a related technical field.
  • 3–5 years of experience in cloud engineering, DevOps, or SRE.
  • 1–2 years of hands-on experience deploying and operating AI, ML, or generative AI applications in production.
  • Experience implementing CI/CD for AI workloads, containerization with Docker and Kubernetes, and cloud infrastructure management using Terraform, Cloud Build, and GKE.
  • Experience with infrastructure automation, incident response, and observable, self-healing production systems.
  • Hands-on experience with Google Cloud Platform, including Cloud Run, Cloud Build, GKE, GCS, BigQuery, IAM, VPC networking, and Secret Manager.
  • Proficiency with CI/CD tooling, GitOps, Docker, Kubernetes, Terraform, and cloud-native monitoring and logging.
  • Proficiency with AI-Ops or LLMOps tools for model deployment, prompt and model versioning, evaluation, tracing, and monitoring of LLM and generative AI outputs.
  • Hands-on experience with Python or Bash for infrastructure automation and operational tooling.
  • Understanding of full-stack application architecture and modern deployment patterns for AI tools.
  • Familiarity with cloud security, identity and access management, secrets management, and compliance standards in regulated environments.
  • Awareness of data science and modeling workflows, including data wrangling, feature engineering, and model validation.
  • Experience in integrated marketing, digital agency, marketing services, or consulting environments.
  • Exposure to banking, financial services, or credit analytics industries.
  • Experience with machine learning engineering, model serving frameworks, or relevant Google Cloud certifications.

OneMagnify Compensation & Benefits Highlights

The following summarizes recurring compensation and benefits themes identified from responses generated by popular LLMs to common candidate questions about OneMagnify and has not been reviewed or approved by OneMagnify.

  • Parental & Family Support — Parental support is positioned as a defined benefit with six weeks of 100% paid parental leave plus employer-paid short-term disability. Parental leave sentiment is characterized as middling, but the policy itself is clearly specified as part of the package.
  • Wellbeing & Lifestyle Benefits — Work flexibility is framed as a meaningful part of the total rewards experience through hybrid/remote options and flexible time off with paid sick leave. Added lifestyle-oriented perks such as volunteer PTO, pet insurance, legal assistance, ERGs, recognition programs, and partner discounts broaden the offering beyond core insurance.
  • Flexible Benefits — A broad slate of optional add-ons and programs is described, including tuition reimbursement up to $5,250/year and other elective-style perks. This variety can help different employee needs be met even when base pay competitiveness is seen as only average.

OneMagnify Insights

Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Detroit, MI
610 Employees
Year Founded: 1967

What We Do

We are technologists, strategists, and creative minds dedicated to moving brands forward with work that delivers results. Our messaging, communications and design are driven by human insights and fresh ideas because starting a conversation is one thing, but building a lasting relationship is another.

Similar Jobs

In-Office
Chennai, Tamil Nadu, IND
17787 Employees
In-Office
Chennai, Tamil Nadu, IND
35 Employees

OneMagnify Logo OneMagnify

Artificial Intelligence Engineer

Information Technology • Marketing Tech • Analytics
In-Office
Chennai, Tamil Nadu, IND
610 Employees

Similar Companies Hiring

NODA AI Thumbnail
Artificial Intelligence • Information Technology • Software • Cybersecurity
Sydney, AU
54 Employees
Golden Pet Brands Thumbnail
Digital Media • eCommerce • Information Technology • Marketing Tech • Pet • Retail • Social Media
El Segundo, California
178 Employees
Blee Thumbnail
Artificial Intelligence • Marketing Tech • Software
New York, New York
30 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account