Middle Data Engineer

Posted Yesterday
Be an Early Applicant
Hiring Remotely in Ukraine
Remote
Mid level
Information Technology • Consulting
The Role
Build and operate scalable AWS data lakehouse pipelines using Python, PySpark, Airflow, S3, Iceberg, Glue, and EMR. Integrate diverse data sources, implement transformations, dimensional models, data quality, testing, governance, metadata, lineage, and documentation. Prepare curated data in Snowflake for analytics and reporting. Collaborate with engineering, architecture, DevOps, product, and BI teams on CI/CD and Terraform infrastructure. Monitor production workflows, troubleshoot issues, and improve reliability and performance.
Summary Generated by Built In

N-iX is looking for a Middle Data Engineer to help build and operate a new AWS-based data lakehouse for our customer in the e-commerce domain. The ideal candidate has hands-on experience with AWS data services, modern data pipelines, and lakehouse architectures.

Our Client is a global full-service e-commerce and subscription billing platform on a mission to simplify software sales everywhere. For nearly two decades, we’ve helped SaaS, digital goods, and subscription-based businesses grow by managing payments, global tax compliance, fraud prevention, and recurring revenue at scale. Our flexible, cloud-based platform combined with consultative services helps clients accelerate growth, reach new markets, and build long-term customer relationships.

Data is at the heart of everything the client does—powering insights, driving innovation, and shaping business decisions. The client is building a next-generation data platform and is looking for a Middle Data Engineer to contribute to its development.

Role Overview:

As a Middle Data Engineer, you will contribute to building and operating a scalable and cost-efficient AWS data lakehouse. You will work closely with data engineers, DevOps engineers, architects, and product leads to implement reliable data pipelines, data models, and data products that support business decision-making.

Your Responsibilities:

  • Develop and maintain scalable data pipelines using S3, Apache Iceberg, Glue Data Catalog, EMR, Airflow (MWAA), Python, and PySpark.
  • Work with data across medallion layers, from source ingestion to reusable business entities, dimensional models, and analytics-ready datasets.
  • Integrate data from various source systems, including databases, APIs, business applications, and file-based sources.
  • Implement data transformations and contribute to building end-to-end data products, including scheduled data deliveries and serverless workloads using services such as AWS Lambda and S3.
  • Apply automated testing and data quality practices across data pipelines and contribute to metadata management, cataloguing, lineage, and documentation using DataHub and AWS services.
  • Prepare and make curated data available in Snowflake for analytics, reporting, and SQL-based use cases.
  • Collaborate with architects, DevOps engineers, product owners, engineering managers, and Power BI engineers on solution implementation, CI/CD, and Terraform-based infrastructure.
  • Monitor data pipelines and production workflows, investigate issues, improve performance and reliability, and maintain technical documentation and runbooks.

What We’re Looking For:

  • 3+ years of experience in data engineering, ideally with AWS, lakehouse, or hybrid data platforms.
  • Strong Python, PySpark, and SQL skills, with experience developing production data pipelines and data transformations.
  • Practical experience with Airflow and AWS data services, including S3, Apache Iceberg, Glue Data Catalog, Glue Data Quality, Athena, and EMR.
  • Good understanding of data modelling and harmonisation, including dimensional modelling, SCD Type 1 and 2, schema evolution, and working with data from multiple source systems.
  • Experience integrating data from databases, APIs, business applications, and file-based sources.
  • Understanding of data governance concepts, including metadata management, cataloguing, lineage, and documentation. Experience with DataHub is a plus.
  • Familiarity with Snowflake, BI platforms, semantic layers, and SQL- or dbt-based transformation approaches.
  • Experience with Git, GitLab CI/CD, automated testing, and Infrastructure as Code. Terraform experience is a plus.
  • Good problem-solving and troubleshooting skills, with a proactive approach to learning and taking ownership of assigned tasks.
  • Good communication and documentation skills.
  • Upper-Intermediate English or higher to collaborate effectively with English-speaking stakeholders on the client side.

We offer*:

  • Flexible working format - remote, office-based or flexible
  • A competitive salary and good compensation package
  • Personalized career growth
  • Professional development tools (mentorship program, tech talks and trainings, centers of excellence, and more)
  • Active tech communities with regular knowledge sharing
  • Education reimbursement
  • Memorable anniversary presents
  • Corporate events and team buildings
  • Other location-specific benefits

*not applicable for freelancers

Skills Required

  • 3+ years of experience in data engineering, ideally with AWS, lakehouse, or hybrid data platforms
  • Strong Python, PySpark, and SQL skills
  • Experience developing production data pipelines and data transformations
  • Practical experience with Airflow and AWS data services, including S3, Apache Iceberg, Glue Data Catalog, Glue Data Quality, Athena, and EMR
  • Understanding of dimensional modeling, SCD Type 1 and 2, schema evolution, and data harmonization
  • Experience integrating data from databases, APIs, business applications, and file-based sources
  • Understanding of data governance, metadata management, cataloguing, lineage, and documentation
  • Familiarity with Snowflake, BI platforms, semantic layers, and SQL- or dbt-based transformation approaches
  • Experience with Git, GitLab CI/CD, automated testing, and Infrastructure as Code
  • Good problem-solving and troubleshooting skills, with a proactive approach to learning and ownership
  • Good communication and documentation skills
  • Upper-Intermediate English or higher
  • Experience with DataHub
  • Terraform experience
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Valletta
2,135 Employees
Year Founded: 2002

What We Do

N-iX is a global software solutions and engineering services company that helps world’s leading organizations turn challenges into lasting business value, operational efficiency, and revenue growth using advanced technology. Whether you need to build a custom solution, modernize your digital product or acquire extra tech expertise - we have the experience and capabilities to ensure your success. With over 2,000 professionals in 25 countries across Europe and the Americas, N-iX offers expert solutions in cloud, data analytics, embedded software, IoT, AI, machine learning, and other tech domains. Being in business for over two decades, we have worked with dozens of industry-leading enterprises and Fortune 500 companies creating value across a wide variety of sectors, including finance, manufacturing, supply chain, retail, e-commerce, healthcare, and more. Our unique combination of business domain expertise and technical know-how enables us to effectively collaborate with ISVs, tech companies, and enterprises of all sizes. Thanks to the strong tech ecosystem and partnerships with AWS, GCP, Microsoft, SAP, OpenText, Snowflake, and others, we bring extra speed, scale and efficiency to more than 160 organizations across the globe. N-iX is recognized by numerous industry awards, such as CRN Solution Provider 500, Global Outsourcing 100 by IAOP, ISG Provider Lens™, Modern Application Development services providers by Forrester, etc

Similar Jobs

Mondelēz International Logo Mondelēz International

o9 Change Readiness Lead

Big Data • Food • Hardware • Machine Learning • Retail • Automation • Manufacturing
Remote or Hybrid
11 Locations
90000 Employees

Pfizer Logo Pfizer

Senior Director, Applied AI - US Commercial

Artificial Intelligence • Healthtech • Machine Learning • Natural Language Processing • Biotech • Pharmaceutical
In-Office or Remote
28 Locations
121990 Employees
215K-358K Annually

DraftKings Logo DraftKings

Lead NetOps Engineer

Digital Media • Gaming • Information Technology • Software • Sports • Esports • Big Data Analytics
Remote or Hybrid
Ukraine
6400 Employees

Superhuman Logo Superhuman

Software Engineer

Artificial Intelligence • Information Technology • Machine Learning • Natural Language Processing • Productivity • Software • Generative AI
Remote or Hybrid
Ukraine
1500 Employees

Similar Companies Hiring

Standard Template Labs Thumbnail
Artificial Intelligence • Information Technology • Software
New York, NY
25 Employees
NODA AI Thumbnail
Artificial Intelligence • Information Technology • Software • Cybersecurity
Sydney, AU
54 Employees
Golden Pet Brands Thumbnail
Digital Media • eCommerce • Information Technology • Marketing Tech • Pet • Retail • Social Media
El Segundo, California
178 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account