Data Engineer

Posted 17 Days Ago
Be an Early Applicant
3 Locations
In-Office
Senior level
Cloud • Information Technology • Software • Consulting
The Role
Design and build production data pipelines, integrate operational, contractual and financial clinical-trial data, create governed versioned data products and a source-of-truth using ontologies/knowledge graphs, and deliver trusted datasets for forecasting and downstream analytics.
Summary Generated by Built In

About the role

We are seeking an experienced Data Engineerto build the foundational data infrastructure supporting a large-scale clinical trial operational and financial forecasting programme. You will design and deliver reliable, scalable data products that enable forecasting across the full trial lifecycle, including study, site, subject, visit and procedure-level activity. Your work will underpin downstream analysis of investigator fees, laboratory kits, resource demand and revenue recognition.

This is a hands-on engineering role operating across complex clinical trial data domains. You will work with contract, budget, operational and actuals data from multiple enterprise source systems, ensuring that data is integrated, reconciled and available at the right level of granularity. You will build solutions that support multi-currency, multi-site and multi-country studies, while maintaining the governance, traceability and version control required for trusted forecasting outputs.

Key responsibilities

  • Develop automated ingestion pipelines for contract and budget data, together with structured consumption layers that make trusted data available to forecasting products and business users;

  • Integrate site-level rate and budget data, build feedback pipelines that return operational actuals at forecast grain, and publish versioned datasets for use by several forecasting initiatives;

  • Define and implement a common data ontology and knowledge graph, establishing consistent identifiers and relationships across upstream and downstream systems;

  • Create and maintain a governed source-of-truth register with clear versioning and amendment tracking.

  • Partner with stakeholders to ensure the subject-level visit projection engine—the primary downstream consumer—receives accurate, timely and well-documented data;

MUST-HAVE SKILLS

Data Pipeline Development and Data Product Delivery:

  • Strong, hands-on experience designing, building and maintaining robust production data pipelines. You can automate ingestion, transformation and publication of data from multiple source systems, and you understand how to produce reusable, versioned datasets that meet the needs of several downstream products;

  • Experience working with data at different operational grains and delivering reliable data products in a governed enterprise environment is essential;

Enterprise Data Integration:

  • Integrate complex datasets from operational, contractual and financial systems, resolving differences in data models, identifiers and business definitions. You are comfortable building a shared integration layer that supports data across countries, sites and currencies, while preserving the detail required at study, site, subject, visit and procedure level;

Data Ontology, Knowledge Graphs and Master Data:

  • Practical experience in designing data ontologies, canonical models or knowledge graphs that create consistent entities, identifiers and relationships across disparate systems. You understand how this capability supports data reconciliation, lineage and trusted cross-domain analytics. You can apply these principles to establish a governed source of truth with clear version control and amendment history;

Clinical Trial, Life Sciences or Complex Operational Data

  • Experience working with complex operational data, ideally within clinical trials, life sciences or a similarly regulated environment. You understand the importance of data quality, auditability and traceability where data informs operational and financial decision-making;

  • Exposure to clinical trial contracts, budgets, site data, subject activity or visit-level data will be particularly valuable;

NICE TO HAVE

  • Experience supporting forecasting, planning or revenue-recognition data products — useful for understanding the needs of the programme’s downstream consumers.

  • Experience with actuals-versus-forecast feedback loops — valuable for improving forecast accuracy and operational learning over time.

  • Familiarity with subject-level visit projection or patient journey models — beneficial for supporting the programme’s principal forecasting engine.

  • Experience operating in multi-country, multi-currency environments — useful when designing scalable and consistent data solutions across global trials.

Skills Required

  • Design, build and maintain robust production data pipelines for automated ingestion, transformation and publication
  • Deliver reliable, versioned data products across different operational grains in a governed enterprise environment
  • Integrate complex datasets from operational, contractual and financial systems; resolve data model, identifier and business definition differences
  • Practical experience designing data ontologies, canonical models or knowledge graphs to establish consistent entities and relationships
  • Establish and maintain a governed source-of-truth register with version control and amendment tracking
  • Experience with complex operational data in clinical trials, life sciences, or similarly regulated environments, including data quality, auditability and traceability
  • Exposure to clinical trial contracts, budgets, site data, subject activity or visit-level data
  • Experience supporting forecasting, planning or revenue-recognition data products
  • Experience with actuals-versus-forecast feedback loops
  • Familiarity with subject-level visit projection or patient journey models
  • Experience operating in multi-country, multi-currency environments
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
80 Employees
Year Founded: 2019

What We Do

BoatyardX is a Dublin-headquartered custom software development and technology business that partners with organizations to design, build, and support bespoke cloud products. Its services span solution discovery, design, development, and longer-term support. The company serves large enterprises and larger startups, using distributed teams in Ireland, Romania, Colombia, and the United States to deliver secure, cloud-portable digital solutions for global clients.

Similar Jobs

AllCloud Logo AllCloud

Data Engineer

Information Technology • Professional Services
Remote or Hybrid
România
474 Employees

Boatyardx Logo Boatyardx

Data Engineer

Cloud • Information Technology • Software • Consulting
In-Office
3 Locations
80 Employees

Ruby Labs Logo Ruby Labs

Data Engineer

Information Technology • Software
In-Office or Remote
25 Locations
28 Employees

Growe Talents Logo Growe Talents

Data Engineer

Agency • Gaming • HR Tech • Professional Services
In-Office or Remote
28 Locations
19 Employees

Similar Companies Hiring

Kepler  Thumbnail
Artificial Intelligence • Fintech • Software
New York, New York
9 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees
Revel.io Thumbnail
Aerospace • Hardware • Robotics • Software
US
50 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account