Distinguished Data Engineer

Posted Yesterday
Be an Early Applicant
Hiring Remotely in Office, Machaze, Manica, MOZ
Remote
Senior level
Software • Financial Services
The Role
Build and operate production-grade data pipelines and AI-ready data foundations for machine learning, LLMs, RAG, semantic search, and agentic workflows. Develop datasets, embeddings, vector indexes, retrieval layers, APIs, metadata structures, and governed data products using Snowflake, AWS, Kafka, Python, Spark, and SQL. Ensure data quality, lineage, security, provenance, observability, and production reliability while collaborating with AI/ML engineers, research analysts, sustainability specialists, and investment stakeholders.
Summary Generated by Built In
About the OpportunityJob Type: Permanent

Application Deadline: 24 September 2026

Job Description

Title                  Distinguished Data Engineer

Department       AMP

Location           Bengaluru

Reports To        Research & Sustainable Investing Data Engineering Lead

Level                 5

We share a commitment to making things better for clients and each other. We continually explore new technology and different ways of working to put our clients first. Bring your boldest ideas to our Research & Sustainable Investing Technology team and feel like you are making progress.

About your team

AMP Delivery is responsible for the design and delivery of all changes in business process and/or technology solutions that support the growth for Fidelity’s Global Investment Solutions & Services business. We partner with Investment Management, Asset Management Operations and Distribution teams across London, Hong Kong, Tokyo, Toronto, Australia, Singapore, and China.  

The Research & Sustainability team delivers strategic initiatives that enhance investment decision-making through modern research workflows, investment data platforms, sustainability capabilities, advanced analytics, and AI-powered solutions. We are increasingly leveraging Large Language Models (LLMs), frontier AI models, and agentic workflows to transform how investment professionals discover insights, conduct research, and make investment decisions across Equities, Fixed Income, and Multi-Asset. 

About your role

This is a specialist, hands-on AI data engineering role within the AMP - Research & Sustainable Investing team. Reporting to the Data Engineering Lead, you will build the trusted data foundations that power machine learning, Large Language Model (LLM), retrieval-augmented generation (RAG), search and agentic AI workflows for investment research and sustainability use cases.

You will engineer structured, semi-structured and unstructured data through the full AI data lifecycle - ingestion, transformation, enrichment, feature and evaluation dataset creation, embedding generation, vector indexing, retrieval and governed delivery. You will work primarily across Snowflake, AWS and Kafka while integrating with enterprise sources including Oracle and Microsoft SQL Server.

You will work closely with AI/ML engineers, data engineers, research analysts, sustainability specialists and investment teams to make AI data accurate, discoverable, traceable, secure and production-ready. The role directly contributes to Accelerated Innovation, Cost Optimisation, Risk Mitigation and Business Enablement.

Key Responsibilities

  • Design, build, test, deploy and operate production-grade data pipelines for structured, semi-structured and unstructured research, sustainability and investment data.
  • Build AI-ready datasets for machine learning and generative AI, including feature, training and evaluation datasets, embeddings, vector indexes and retrieval-augmented generation workflows.               
  • Engineer retrieval data flows that curate, enrich, index and serve trusted content for semantic search, vector search and LLM-based applications, with appropriate metadata and provenance.
  • Help to develop metadata and knowledge structures, including taxonomies, ontologies, entity resolution and knowledge graphs where appropriate, to improve data discovery, retrieval and contextual understanding.
  • Build reusable data services and APIs that expose governed investment data to AI/ML models, applications, analytics and agentic workflows.
  • Apply data quality, lineage, data contracts, security, privacy, entitlements, auditability and provenance controls across data pipelines and AI-ready data products.
  • Use Snowflake, AWS, Kafka and enterprise data platforms to ingest and prepare data, integrating safely with Oracle and Microsoft SQL Server where source or legacy data is required.
  • Apply strong software and data engineering practices to data pipelines, including Python/SQL development, Git, CI/CD, automated testing, observability, monitoring, performance tuning and secure production support.
  • Partner with AI/ML engineers and investment stakeholders to translate AI use cases into reusable data products, trusted retrieval layers and measurable production outcomes.
  • Investigate complex data, embedding, indexing, retrieval and pipeline issues, perform root-cause analysis and implement sustainable fixes with the Data Engineering Lead and wider team.

About You

Core Technical Skills

  • Data Engineering: Strong hands-on experience preparing structured, semi-structured and unstructured data for machine learning and generative AI, including feature, training and evaluation datasets.
  • Embeddings, Vector Search and RAG: Practical experience building embedding pipelines, vector indexes, retrieval workflows and RAG data foundations for applications.
  • Evaluation Data: Experience creating and curating evaluation datasets and applying data-quality, relevance, traceability and provenance checks to support reliable retrieval and model evaluation.
  • Data Pipelines: Strong experience designing production-grade ingestion, transformation, enrichment and delivery pipelines using batch, near-real-time, streaming, API and event-based patterns.
  • Programming: Strong proficiency in Python, Spark and SQL, with the ability to build maintainable, tested production code for data preparation, embeddings, indexing, retrieval and data services.
  • Data Platforms and Integration: Experience with Snowflake, AWS and Kafka, plus enterprise relational sources such as Oracle or Microsoft SQL Server; able to integrate modern data and retrieval workloads with existing platforms.
  • Data Services and Orchestration: Experience building reusable APIs and data services and using orchestration tools such as Airflow or Control-M to operate dependable production data workflows.
  • Engineering Practices: Experience with Git, CI/CD, automated testing, infrastructure as code, data observability, monitoring, performance tuning and secure production engineering.
  • Governance and Responsible AI: Good understanding of data quality, lineage, data contracts, access control, privacy, entitlements, auditability, provenance and responsible AI controls across governed data products and AI-ready datasets.
  • Data Modelling and Semantic Layers: Good understanding of canonical and dimensional modelling, semantic layers and governed business concepts that can be consumed consistently by analytics and AI-enabled solutions.

Professional Experience

  • Demonstrated hands-on experience delivering production data engineering solutions for AI/ML use cases, particularly data pipelines and data foundations supporting machine learning, LLMs, search, retrieval and generative AI.
  • Experience working closely with AI/ML engineers, data engineers, architects, analysts and business stakeholders to move data and retrieval workflows from experimentation into secure, reliable and supportable production services.
  • Ability to own assigned data engineering components from design through build, testing, deployment, monitoring and production support.                       
  • Strong communication skills, with the ability to explain data engineering, retrieval, metadata, data quality and governance topics clearly and translate investment and AI use cases into practical engineering solutions.

Key Soft Skills

  • Technical Ownership: Takes accountability for the quality, security, resilience, provenance and maintainability of data products and AI-ready data components from build through production support.
  • Problem-Solving: Applies strong analytical judgement to ambiguous data, retrieval and integration problems, investigates root causes and drives issues to sustainable resolution.
  • Collaboration: Works effectively with AI/ML engineers, data engineers and investment stakeholders, contributes constructively to reviews and shares knowledge across the team.
  • Communication: Explains complex data engineering, retrieval and governance topics clearly to technical and non-technical audiences and documents decisions and solutions effectively.
  • Learning and Adaptability: Keeps current with evolving data engineering and AI-enablement approaches, experiments responsibly and applies new techniques where they create measurable value.

Feel rewarded

For starters, we'll offer you a comprehensive benefits package. We'll value your wellbeing and support your development. And we'll be as flexible as we can about where and when you work - finding a balance that works for all of us. It's all part of our commitment to making you feel motivated by the work you do and happy to be part of our team. For more about our work, our approach to dynamic working and how you could build your future here, visit careers.fidelityinternational.com.

Skills Required

  • Strong hands-on experience preparing structured, semi-structured, and unstructured data for machine learning and generative AI
  • Practical experience building embedding pipelines, vector indexes, retrieval workflows, and RAG data foundations
  • Experience creating and curating evaluation datasets with data-quality, relevance, traceability, and provenance checks
  • Strong experience designing production-grade ingestion, transformation, enrichment, and delivery pipelines using batch, near-real-time, streaming, API, and event-based patterns
  • Strong proficiency in Python, Spark, and SQL
  • Experience with Snowflake, AWS, Kafka, Oracle, or Microsoft SQL Server
  • Experience building reusable APIs and data services and using orchestration tools such as Airflow or Control-M
  • Experience with Git, CI/CD, automated testing, infrastructure as code, data observability, monitoring, performance tuning, and secure production engineering
  • Understanding of data quality, lineage, data contracts, access control, privacy, entitlements, auditability, provenance, and responsible AI controls
  • Understanding of canonical and dimensional modeling, semantic layers, and governed business concepts
  • Hands-on experience delivering production data engineering solutions for AI/ML use cases
  • Experience moving data and retrieval workflows from experimentation into secure, reliable, supportable production services
  • Ability to own data engineering components from design through build, testing, deployment, monitoring, and production support
  • Strong communication skills for explaining data engineering, retrieval, metadata, data quality, and governance topics

Fidelity International Compensation & Benefits Highlights

The following summarizes recurring compensation and benefits themes identified from responses generated by popular LLMs to common candidate questions about Fidelity International and has not been reviewed or approved by Fidelity International.

  • Parental & Family Support Family leave and carers’ support are emphasized via equalized paid parental leave globally and enhanced maternity/adoption policies. Inclusive provisions also cover carers’ leave and compassionate leave to support diverse family needs.
  • Healthcare Strength Healthcare benefits and private medical insurance are consistently highlighted as part of the core package across locations. Wellbeing resources, including an Employee Assistance Programme and menopause support, reinforce the depth of health coverage.
  • Retirement Support Pension and retirement savings are positioned as a strong element of the total package in multiple markets. Retirement design is frequently cited alongside paid time off and flexibility as part of a solid overall offer.

Fidelity International Insights

Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: London
9,919 Employees
Year Founded: 1969

What We Do

Fidelity International offers investment solutions and services and retirement expertise to more than 2.5 million customers globally. As a privately held, purpose-driven company with a 50-year heritage, we think generationally and invest for the long term. Operating in more than 25 countries and with $739.9 billion* in total assets, our clients range from central banks, sovereign wealth funds, large corporates, financial institutions, insurers and wealth managers, to private individuals. Our Workplace & Personal Financial Health business provides individuals, advisers and employers with access to world-class investment choices, third-party solutions, administration services and pension guidance. Together with our Investment Solutions & Services business, we invest $567 billion on behalf of our clients. By combining our asset management expertise with our solutions for workplace and personal investing, we work together to build better financial futures. *Data as of 31 March 2021

Similar Jobs

Mondelēz International Logo Mondelēz International

R&D Manager - Cote d'Or Chocolate

Big Data • Food • Hardware • Machine Learning • Retail • Automation • Manufacturing
Remote or Hybrid
4 Locations
90000 Employees

Invenergy Logo Invenergy

Senior PI Administrator

Greentech • Real Estate • Social Impact • Energy • Industrial • Solar • Renewable Energy
Remote or Hybrid
18 Locations
2500 Employees
125K-155K Annually

Auror Logo Auror

Senior Engineering Lead (Risk Detection)

Artificial Intelligence • Big Data • Retail • Security • Social Impact • Software • Business Intelligence
Remote or Hybrid
Office, Machaze, Manica, MOZ
212 Employees
143K-190K Annually

Compa Logo Compa

Software Engineer

Artificial Intelligence • HR Tech • Software • Business Intelligence
Remote or Hybrid
4 Locations
75 Employees
125K-180K Annually

Similar Companies Hiring

Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees
Revel.io Thumbnail
Aerospace • Hardware • Robotics • Software
US
50 Employees
Blee Thumbnail
Artificial Intelligence • Marketing Tech • Software • Productivity
US
15 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account