AI DevOps Engineer

Posted 20 Days Ago
Be an Early Applicant
London, Greater London, England, GBR
In-Office
Senior level
Financial Services
The Role
Build, deliver, and operate AI-focused DevOps environments across hybrid infrastructure, Azure, and AWS. Develop RAG and data pipelines, MLOps workflows, AI agents, GPU workloads, infrastructure automation, and containerized platforms. Support live, business-critical systems while improving scalability, reliability, observability, and standardization. Collaborate across teams, modernize infrastructure, and implement emerging AI and cloud technologies.
Summary Generated by Built In

Department Overview:

Our Technology Infrastructure team operates globally and is responsible for every aspect of the firm's hybrid platform. This ranges from our EUC/Office environments to Trading and Core service Co-Location Data Centres, and extends to Public Cloud, delivering top-tier technology services to a dynamic and demanding Trading organisation.
In addition to meeting the round-the-clock operational demands of the platforms, we continuously evolve and transform our platforms to maintain a competitive edge that our business requires. We innovate to provide valuable solutions and leverage our skilled Technology teams to deliver against rapidly changing business requirements.

Role Overview: 

The AI DevOps Engineer is a very technical, hands-on role, responsible for end to end delivery and operations of our DevOps environments. They will be joining the team as we lead an ambitious technology modernisation journey, moving to greater adoption of AI, microservices and public cloud. We’re continuing to evolve our global infrastructure - spanning multiple global data centres - into a modern hybrid architecture that leverages both existing on-prem and multi-cloud environments and capabilities across Azure and AWS.

Over the past four years, we've modernised our core trading systems and introduced a range of key technologies including AI, Kubernetes, Kafka, ELK, Prometheus, and more. As we expand and enhance the platforms, our focus is on automating and standardising to ensure consistency, scalability and reliability across all environments.

Experiences and Skills Required: 

Experience, skills required:

Must have:

  • Very good understanding of AIOps and DevOps:
    • AgentOps/DataOps frameworks (e.g. LangGraph, Dagster) and OpenTelemetry
    • Predictive operations tools (e.g. Prophet, Dynatrace)
    • AI tool integration, including embedding LLM APIs and agents into pipelines
  • Strong expertise across Azure and AWS, with a solid understanding of AI-native services:
    • Experience with Azure AI Foundry and Amazon Bedrock
    • Cloud automation experience
  • AI Automation:
    • Experience building and deploying AI agents
  • Infrastructure as Code (IaC):
    • Experience with Terragrunt, Terraform, and Ansible
  • Strong Python skills
  • Excellent understanding of containerisation technologies, ideally Kubernetes
  • Experience with GitHub Actions
  • Delivery-focused mindset, with priorities aligned to business needs.

Nice to have, but willing to learn otherwise:

  • Experience integrating DevOps infrastructure with MCPs
  • Experience with infrastructure testing frameworks (Terratest, Terraform Test, Molecule, JUnit, pytest)
  • Kubernetes engineering and support across on-prem, VMware, AKS and EKS environments
  • Linux and Windows administration and network automation
  • Java and/or C# troubleshooting capability
  • CI/CD experience with GitHub Actions and/or TeamCity
  • Artifact repository management (Nexus, Artifactory)
  • Observability, monitoring and tracing (OpenTelemetry, Prometheus, ELK, Tempo, Jaeger)
  • Platform engineering on Azure and/or AWS
  • Data platform experience (Kafka, Redis)
  • Chaos engineering and resilience testing (Gremlin, LitmusChaos)

About You:

We’re looking for a very strong individual contributor who is passionate about emerging technologies. You’ll need to be comfortable working in a very dynamic, fast-paced environment. Enabling the business is our number one priority – this means working on any part of the delivery stack, from routine o/s IaC, to documentation, instrumentation all the way to automating work through Python microservices or LangChain based agents.

You should be passionate about AI and other emerging technology trends, innovations and directions, and eager to suggest and implement new solutions to enhance our technology performance. You should also understand the need for rapid delivery while appreciating business risk and constraints to ensure adherence to service levels.
The candidate should have extensive experience working both independently and as part of a diverse team, meeting both broad and specific project/BAU objectives. This role requires excellent organisational skills, open communication, and a collaborative approach.
Your commitment to continuous improvement should be evident, both in learning from others and sharing knowledge to enhance our team's capability and function.



BlueCrest is committed to providing an inclusive environment for its workforce. As an employer, we provide equal opportunities to all people regardless of their gender, marital or civil partnership status, race, religion or ethnicity, disability, age, sexual orientation or nationality.

Skills Required

  • Strong understanding of AI infrastructure and DevOps
  • Experience with RAG pipelines, MLOps frameworks, data pipelines, and versioning
  • Experience with Kubeflow, MLflow, Argo, or Airflow
  • Experience with predictive operations tools such as Prophet or Dynatrace
  • Experience integrating AI tools, LLM APIs, or agents into pipelines
  • Strong AI-focused Azure and AWS expertise
  • Experience orchestrating GPU workloads
  • Experience with Azure AI Foundry and Amazon Bedrock
  • Experience building or using AI agents
  • Experience integrating DevOps infrastructure with MCPs
  • Experience with Terragrunt and Terraform
  • Python experience
  • Strong understanding of containerization technologies
  • Experience with infrastructure testing frameworks such as Terratest, Terraform Test, Molecule, JUnit, or pytest
  • Kubernetes engineering and support across on-premises, VMware, AKS, and EKS environments
  • Linux and Windows administration and network automation experience
  • Java or C# troubleshooting capability
  • CI/CD experience with GitHub Actions or TeamCity
  • Artifact repository management with Nexus or Artifactory
  • Observability, monitoring, and tracing experience with OpenTelemetry, Prometheus, ELK, Tempo, or Jaeger
  • Azure or AWS platform engineering experience
  • Data platform experience with Kafka or Redis
  • Chaos engineering and resilience testing experience with Gremlin or LitmusChaos
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: London
491 Employees

What We Do

BlueCrest Capital Management was founded in 2000, focused on fixed income macro trading. The firm has now developed into one of the largest global alternative asset managers, with offices in London, Geneva, Jersey, New York, Miami and Singapore.

Similar Jobs

Hybrid
West Haddon, Northamptonshire, England, GBR
500 Employees
35K-45K Annually
In-Office
London, Greater London, England, GBR
4632 Employees
In-Office
London, Greater London, England, GBR
In-Office
London, Greater London, England, GBR
2600 Employees

Similar Companies Hiring

Granted Thumbnail
Artificial Intelligence • Healthtech • Insurance • Mobile • Financial Services
New York, New York
23 Employees
Hanover Park Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
42 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account