Data Engineer (Clinical/Laboratory/Scientific data + AWS)

Posted 3 Days Ago
Be an Early Applicant
Hiring Remotely in CZE
Remote
Senior level
Information Technology • Analytics • Consulting • Cybersecurity
The Role
Design, build, and maintain scalable AWS data pipelines and platforms using Python, SQL, Spark, PySpark, Databricks, and Dataiku. Develop ETL/ELT workflows, optimize Redshift and S3 solutions, automate infrastructure with Terraform and CloudFormation, and maintain CI/CD pipelines. Monitor and troubleshoot data workloads, document solutions, and collaborate with global stakeholders. The role supports regulated clinical, laboratory, scientific, and pharmaceutical data environments.
Summary Generated by Built In

This is a remote position.

We are looking for an experienced Senior Data Engineer to join a data-focused project for a global pharmaceutical company. You will design, build, and maintain scalable cloud data pipelines and platforms using AWS, Python, SQL, Spark, and Databricks.

Key Responsibilities
  • Design, develop, and maintain scalable data pipelines and data-processing solutions.
  • Build cloud-based data solutions using AWS services.
  • Develop ETL/ELT workflows using Python, SQL, Spark, and PySpark.
  • Work with Databricks and Dataiku for data engineering, transformation, and analytics use cases.
  • Develop and optimize solutions involving Amazon Redshift and Amazon S3.
  • Automate infrastructure deployments using Terraform and CloudFormation.
  • Build and maintain CI/CD pipelines using Jenkins and Git.
  • Containerize applications and data workloads using Docker.
  • Monitor data workflows and troubleshoot performance or reliability issues.
  • Produce clear technical documentation in accordance with SDLC standards.
  • Collaborate with data engineers, architects, analysts, and global business stakeholders.


Requirements
  • 5–8 years of relevant data engineering experience.
  • Strong hands-on experience with AWS, particularly:
    • Amazon S3
    • IAM
    • Amazon Redshift
    • AWS Glue
    • AWS Lambda
    • AWS Step Functions
    • Amazon CloudWatch
    • Amazon SageMaker
  • Strong programming skills in Python or Java.
  • Advanced SQL knowledge, ideally including Amazon Redshift.
  • 2–3 years of hands-on experience with Apache Spark or PySpark.
  • Experience with Databricks and/or Dataiku.
  • Experience with Terraform and AWS CloudFormation.
  • Practical knowledge of Jenkins, Git, and Docker.
  • Understanding of software development lifecycle processes and technical documentation.
  • Strong communication skills and the ability to work within an international team.
  • Previous experience in the pharmaceutical, healthcare, laboratory, or life-sciences sector.
  • Knowledge of DSCS and DPTM platforms or related data-management environments.
  • Experience working with regulated or validated data solutions.
  • Familiarity with data originating from laboratory or scientific systems.


Benefits
  • Location: European Union
  • Contract Type: Freelance / Contract
  • Start date: September, 2026
  • Time Allocation: 40 hours/week


Skills Required

  • 5-8 years of relevant data engineering experience
  • Hands-on AWS experience, including Amazon S3, IAM, Amazon Redshift, AWS Glue, AWS Lambda, AWS Step Functions, Amazon CloudWatch, and Amazon SageMaker
  • Strong programming skills in Python or Java
  • Advanced SQL knowledge
  • Amazon Redshift SQL experience
  • 2-3 years of hands-on Apache Spark or PySpark experience
  • Experience with Databricks and/or Dataiku
  • Experience with Terraform and AWS CloudFormation
  • Practical knowledge of Jenkins, Git, and Docker
  • Understanding of software development lifecycle processes and technical documentation
  • Strong communication skills and ability to work within an international team
  • Previous experience in pharmaceutical, healthcare, laboratory, or life-sciences sectors
  • Knowledge of DSCS and DPTM platforms or related data-management environments
  • Experience working with regulated or validated data solutions
  • Familiarity with data originating from laboratory or scientific systems
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
Year Founded: 2015

What We Do

futureproof s.r.o. is a Czech consulting and staffing firm focused on data, analytics, cybersecurity, and IT infrastructure. It provides contract and permanent staffing, team augmentation, time-and-material resources, and specialist or lead placements, while also offering expert consulting through a network of architects and project leaders. The company emphasizes niche technical expertise, trusted relationships, continuous learning, and long-term value for clients.

Similar Jobs

Remote
Czech Republic
575 Employees
948K-1M Annually

Pfizer Logo Pfizer

Director R&D EHS Program Lead

Artificial Intelligence • Healthtech • Machine Learning • Natural Language Processing • Biotech • Pharmaceutical
In-Office or Remote
36 Locations
121990 Employees
177K-294K Annually

SailPoint Logo SailPoint

Consultant

Artificial Intelligence • Cloud • Sales • Security • Software • Cybersecurity • Data Privacy
Remote or Hybrid
2 Locations
2461 Employees

Mondelēz International Logo Mondelēz International

Program Manager

Big Data • Food • Hardware • Machine Learning • Retail • Automation • Manufacturing
Remote or Hybrid
9 Locations
90000 Employees
4K-4K Annually

Similar Companies Hiring

Milestone Systems Thumbnail
Artificial Intelligence • Security • Software • Analytics • Big Data Analytics
Lake Oswego, OR
1500 Employees
NODA AI Thumbnail
Artificial Intelligence • Information Technology • Software • Cybersecurity
Sydney, AU
54 Employees
Golden Pet Brands Thumbnail
Digital Media • eCommerce • Information Technology • Marketing Tech • Pet • Retail • Social Media
El Segundo, California
178 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account