L1 Data Engineer - Remote

Posted Yesterday
6 Locations
Remote
Mid level
Artificial Intelligence • Computer Vision • Software
The Role
Design, build, and maintain scalable data pipelines and ETL/ELT workflows using Azure Data Factory, Databricks, PySpark, Python, and SQL. Transform large datasets, optimize data ingestion and queries, monitor pipeline performance, resolve data quality issues, support data modeling, document engineering standards, and contribute to cloud data infrastructure improvements. The role also involves collaboration with architects and senior engineers and requires familiarity with Azure Synapse, Delta Lake, governance, and security.
Summary Generated by Built In

We are looking for a motivated and technically solid L1 Data Engineer to join our growing Data & Analytics team. In this role, you will be responsible for designing, building, and maintaining the data architecture and infrastructure that supports our organization's data strategy. You will work hands-on to develop, test, and deploy reliable data solutions — ensuring pipelines are scalable, efficient, and aligned with business requirements.

This is an ideal opportunity for a data professional who is eager to deepen their expertise in cloud-native data platforms, particularly within the Microsoft Azure and Databricks ecosystem, and who thrives in a collaborative, fast-paced environment.

KEY RESPONSIBILITIES

• Design, develop, and maintain scalable data pipelines and ETL/ELT workflows to support business intelligence and analytics use cases.

• Build and optimize data ingestion processes using Azure Data Factory and Databricks, ensuring data quality and consistency across all layers of the data platform.

• Transform and process large datasets using PySpark and Python, applying best practices for performance and maintainability.

• Write and optimize complex SQL queries to support analytical reporting and data validation requirements.

• Collaborate with data architects and senior engineers to implement and maintain data models aligned with organizational standards.

• Monitor, troubleshoot, and resolve pipeline failures and data quality issues, applying root-cause analysis to prevent recurrence.

• Contribute to documentation of data pipelines, data dictionaries, and engineering standards.

• Support the team in exploring and evaluating new tools and approaches to continuously improve the data infrastructure.


Requirements
  • 3+ years of professional experience in a Data Engineering or closely related role.
  • Strong proficiency in Python for data processing, transformation, and automation tasks.
  • Hands-on experience with Pandas for data manipulation and PySpark for distributed data processing.
  • Practical experience with Databricks, including notebook development, clusters, and job orchestration.
  • Experience building and managing data pipelines with Azure Data Factory.
  • Working knowledge of Azure Synapse Analytics, particularly Spark pool integration.
  • Solid SQL skills, including query writing, optimization, and performance tuning.
  • Familiarity with data engineering principles including incremental loading, data lake architecture, and Delta Lake.
  • Understanding of data governance and security concepts within a cloud data platform.

NICE TO HAVE

  • Experience with SQL Server migration projects, including schema conversion and data movement.
  • Exposure to Terraform for Azure infrastructure provisioning and management.
  • Familiarity with CI/CD practices applied to data engineering workflows.
  • Experience with Delta Sharing or Lakehouse Federation concepts.

CERTIFICATION REQUIREMENT

  • Candidates are expected to hold or be actively working toward the Databricks Certified Data Engineer Associate certification. This certification validates foundational knowledge across the following domains:
  • Databricks Lakehouse Platform architecture and capabilities
  • ETL and ELT workflows using Spark SQL and PySpark
  • Incremental data processing and structured streaming
  • Production pipeline development and orchestration
  • Data governance and security within the Databricks environment

Skills Required

  • 3+ years of professional experience in data engineering or a closely related role
  • Strong proficiency in Python for data processing, transformation, and automation
  • Hands-on experience with Pandas
  • Hands-on experience with PySpark
  • Practical experience with Databricks, including notebooks, clusters, and job orchestration
  • Experience building and managing pipelines with Azure Data Factory
  • Working knowledge of Azure Synapse Analytics and Spark pool integration
  • Solid SQL skills, including query writing, optimization, and performance tuning
  • Understanding of incremental loading, data lake architecture, and Delta Lake
  • Understanding of data governance and security concepts within cloud data platforms
  • Databricks Certified Data Engineer Associate certification, or active pursuit of the certification
  • Experience with SQL Server migration projects
  • Exposure to Terraform for Azure infrastructure provisioning and management
  • Familiarity with CI/CD practices for data engineering workflows
  • Experience with Delta Sharing or Lakehouse Federation
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: San Francisco, CA
10 Employees
Year Founded: 2020

What We Do

DeepSource stands as a trusted partner for businesses seeking cutting-edge AI services in computer vision, natural language processing, and predictive analytics. With a particular focus on Arabic NLP and ChatGPT bot development, DeepSource is dedicated to empowering companies with groundbreaking solutions that streamline operations, optimize workflows, and enhance user experiences. Our commitment to excellence is evident in our approach to addressing a wide range of AI needs, from hiring top talent and managing end-to-end AI projects to providing tailored consulting and comprehensive training programs. DeepSource's team of experts is equipped with extensive knowledge and experience in various AI technologies, which enables them to develop and deploy advanced solutions across multiple industries. Our adaptive strategies and innovative methodologies allow businesses to stay competitive in today's rapidly evolving digital landscape

Similar Jobs

Mondelēz International Logo Mondelēz International

Sourcing Manager - Marketing Services

Big Data • Food • Hardware • Machine Learning • Retail • Automation • Manufacturing
Remote or Hybrid
Cairo, EGY
90000 Employees

Mondelēz International Logo Mondelēz International

FP&A Senior Analyst

Big Data • Food • Hardware • Machine Learning • Retail • Automation • Manufacturing
Remote or Hybrid
Cairo, EGY
90000 Employees

Mondelēz International Logo Mondelēz International

Procurement Manager

Big Data • Food • Hardware • Machine Learning • Retail • Automation • Manufacturing
Remote or Hybrid
Cairo, EGY
90000 Employees

Ericsson Logo Ericsson

Integration Engineer

Cloud • Information Technology • Internet of Things • Machine Learning • Software • Cybersecurity • Infrastructure as a Service (IaaS)
In-Office or Remote
3 Locations
88000 Employees

Similar Companies Hiring

Kepler  Thumbnail
Artificial Intelligence • Fintech • Software
New York, New York
9 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees
Revel.io Thumbnail
Aerospace • Hardware • Robotics • Software
US
50 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account