Senior Azure ML Infrastructure Engineer

Posted 4 Days Ago
Be an Early Applicant
New York, NY, USA
Hybrid
160K-180K Annually
Senior level
Insurance • Professional Services • Real Estate • Financial Services
The Role
Leads the architecture, development, and optimization of production-grade machine learning infrastructure on Azure. Designs scalable training and inference environments, implements MLOps practices, automates ML workflows, and establishes reusable deployment frameworks. Partners with data science, DevOps, and platform engineering teams to productionize models, strengthen security and observability, and ensure reliable, compliant systems. Mentors junior engineers and drives ML infrastructure strategy across multiple use cases.
Summary Generated by Built In

Simpson Thacher & Bartlett LLP is looking for a Senior Azure ML Infrastructure Engineer to lead the design, development, and optimization of scalable ML infrastructure on Microsoft Azure. In this role, you will be the technical lead for deploying and maintaining robust (Machine Learning) ML Ops frameworks, ensuring efficient collaboration between data science, engineering, and DevOps teams. You’ll be instrumental in scaling our machine learning capabilities from experimentation to production across multiple use cases.

ESSENTIAL JOB DUTIES & RESPONSIBILITIES

Infrastructure Architecture & Engineering

  • Lead the architecture and implementation of production-grade ML infrastructure using Azure Machine Learning, AKS, Azure Data Lake, Azure Databricks, and related services.

  • Design scalable training and inference environments for deep learning and traditional ML workloads, optimizing performance and cost.

MLOps Strategy & Execution

  • Define and implement MLOps best practices: versioning, CI/CD for ML pipelines, monitoring, and model governance.

  • Automate end-to-end ML workflows using tools such as MLFlow, Azure ML Pipelines, or Kubeflow.

  • Build reusable templates and frameworks to standardize ML deployment across teams.

Cross-Functional Leadership

  • Collaborate with data scientists to productionize models, offering guidance on infrastructure, deployment strategies, and performance optimization.

  • Partner with DevOps and platform engineering teams to align infrastructure with broader cloud strategies and compliance standards.

  • Mentor junior ML and platform engineers, sharing best practices and driving engineering excellence.

Security, Reliability, and Observability

  • Implement enterprise-grade security and compliance controls using Azure Active Directory, RBAC, and data encryption strategies.

  • Integrate observability tooling (e.g., Azure Monitor, Prometheus, Grafana) for end-to-end monitoring of ML systems.

  • Ensure systems are highly available, reliable, and scalable to meet the demands of production ML workloads.

EDUCATION

  • Bachelor’s degree in Computer Science, Information Systems, or a related field, or equivalent practical experience in lieu of formal education.

  • Legal IT experience a plus but not required

SKILLS AND EXPERIENCE

Required

  • 5+ years of experience in ML infrastructure, cloud engineering, or MLOps

  • 2+ years of experience working in Azure environments.

  • Deep hands-on experience with Azure cloud services relevant to ML, including Azure Machine Learning, AKS, Blob Storage, Databricks, Azure Data Factory, and Synapse.

  • Strong expertise in containerization (Docker) and orchestration (Kubernetes, preferably AKS).

  • Proficient in Python and scripting languages (e.g., Bash, PowerShell).

  • Advanced knowledge of CI/CD tools such as Azure DevOps, GitHub Actions, or Jenkins for ML workloads.

  • Solid understanding of IaC tools: Terraform, Bicep, or ARM templates.

Preferred

  • Microsoft Azure certifications (e.g., Azure AI Engineer Associate, Azure Solutions Architect Expert, or DevOps Engineer Expert).

  • Experience designing ML infrastructure in regulated industries (finance, healthcare, etc.).

  • Familiarity with feature stores, distributed training, and model monitoring frameworks.

  • Leadership experience in building infrastructure for ML at scale.

WHY YOU WILL LOVE THIS ROLE

  • Lead the development of high-impact ML platforms that support real-world AI applications.

  • Influence the direction of our ML and cloud infrastructure strategy.

  • Work in a forward-thinking, collaborative team that values experimentation, clean architecture, and automation.

  • Competitive salary, equity opportunities, and comprehensive benefits.

  • Continuous learning budget and Azure certification support.

Salary Information

NY Only: The estimated base salary range for this position is $160,000 to $180,000 at the time of posting.

The actual salary offered will depend on a variety of factors, including without limitation, the qualifications of the individual applicant for the position, years of relevant experience, level of education attained, certifications or other professional licenses held, and if applicable, the location in which the applicant lives and/or from which they will be performing the job. This role is exempt meaning it is not overtime pay eligible.


Simpson Thacher will not sponsor applicants for work visas for this position.

Privacy Notice

For information about how Simpson Thacher & Bartlett LLP collects and processes your personal information, please refer to our Privacy Notice available at https://www.stblaw.com/other/privacy-notice.

Simpson Thacher & Bartlett is committed to a collegial work environment in which all individuals are treated with respect and dignity. The Firm prohibits discrimination or harassment based upon race, color, religion, gender, gender identity or expression, age, national origin, citizenship status, disability, marital or partnership status, sexual orientation, veteran’s status or any other legally protected status. This Policy pertains to every aspect of an individual’s relationship with the Firm, including but not limited to recruitment, hiring, compensation, benefits, training and development, promotion, transfer, discipline, termination, and all other privileges, terms and conditions of employment.

#LI-Hybrid

Skills Required

  • Bachelor's degree in Computer Science, Information Systems, or a related field, or equivalent practical experience
  • 5+ years of experience in ML infrastructure, cloud engineering, or MLOps
  • 2+ years of experience working in Azure environments
  • Hands-on experience with Azure Machine Learning, AKS, Blob Storage, Databricks, Azure Data Factory, and Synapse
  • Strong expertise in Docker and Kubernetes, preferably AKS
  • Proficiency in Python and scripting languages such as Bash or PowerShell
  • Advanced knowledge of Azure DevOps, GitHub Actions, or Jenkins for ML workloads
  • Solid understanding of Terraform, Bicep, or ARM templates
  • Microsoft Azure certification, such as Azure AI Engineer Associate, Azure Solutions Architect Expert, or DevOps Engineer Expert
  • Experience designing ML infrastructure in regulated industries such as finance or healthcare
  • Familiarity with feature stores, distributed training, and model monitoring frameworks
  • Leadership experience building infrastructure for ML at scale
  • Legal IT experience
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
3,888 Employees
Year Founded: 1884

What We Do

Simpson Thacher & Bartlett LLP is one of the world's leading international law firms. Established in 1884, the firm employs approximately 2,000 lawyers across 14 global offices. Headquartered in New York, it provides coordinated legal advice and transactional capability to clients worldwide, focusing on the success of its clients through smart solutions to critical commercial challenges.

Similar Jobs

Hybrid
New York, NY, USA
289097 Employees

Enverus Logo Enverus

Consulting Geologist, Anadarko Basin - Contract - 26330

Big Data • Information Technology • Software • Analytics • Energy
In-Office or Remote
2 Locations
1800 Employees
60-75 Hourly

MetLife Logo MetLife

AVP, HR Technology

Fintech • Information Technology • Insurance • Financial Services • Big Data Analytics
Hybrid
New York, NY, USA
43000 Employees
164K-219K Annually
Hybrid
New York, NY, USA
205000 Employees
120K-196K Annually

Similar Companies Hiring

Hanover Park Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
42 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees
Vega Thumbnail
Artificial Intelligence • Automotive • Insurance • Transportation
US
43 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account