Azure Data Lead

Posted 2 Hours Ago
Be an Early Applicant
Jersey City, NJ, USA
In-Office
Expert/Leader
Information Technology • Consulting
The Role
Designs, develops, and supports scalable Azure data engineering solutions, including batch and near-real-time ingestion, PySpark and Python transformations, data lake and warehouse integration, and Cosmos DB workloads. Responsibilities include building reliable pipelines, implementing data quality and monitoring controls, optimizing performance and costs, supporting CI/CD and production operations, and collaborating with technical and business stakeholders.
Summary Generated by Built In
Company Description

Derex Technologies Inc specializes in providing IT consulting, staffing solutions and software services. Globally headquartered in Harrison New Jersey since 1996 Derex delivers the highest quality technology professionals and an array of customized IT talent solutions designed to improve productivity and drive results to global clients throughout North America.

With over two decades of unparalleled experience, Derex provides supports to its clientele, across such industries as Systems Integration, Banking and Finance, Telecommunications, Pharmaceutical and Life Sciences, Energy, Healthcare, Technology, Transportation, and local and federal Government agencies.

Job Description

Title: Azure Data Lead

Location: NJ (Day 1 Onsite)

 

Overview

This role is for an experienced Azure Data Lead who can design, build, and support scalable data engineering solutions on Microsoft Azure. The individual will work on modern data platforms involving batch and near-real-time ingestion, data transformation, data lake and warehouse integration, and operational data workloads using Azure-native services.

The role requires strong hands-on engineering capability in PySpark, Python, SQL, Azure Data Factory, Azure Databricks, Azure Data Lake Storage, Azure Synapse Analytics, and Azure Cosmos DB. The candidate should be able to convert business and data requirements into reliable, secure, performant, and production-ready data pipelines.

 

Required skills:

  • 10+ years of experience in data engineering, cloud data platforms, ETL/ELT development, or large-scale data processing
  • Strong hands-on experience in designing, developing, testing, and maintaining Azure-based data pipelines and data processing solutions
  • Must have strong hands-on experience with PySpark and Python for large-scale data transformation, automation, data quality checks, and reusable data engineering frameworks
  • Azure Data Factory for data ingestion, orchestration, parameterized pipelines, triggers, and monitoring
  • Azure Databricks and Apache Spark for scalable data processing using PySpark notebooks, jobs, workflows, and optimized Spark transformations
  • Azure Data Lake Storage Gen2 for lakehouse-style storage, folder structures, file formats, access control, and lifecycle management
  • Azure Synapse Analytics or Azure SQL for analytical workloads, SQL development, data modeling, performance tuning, and reporting integration
  • Azure Cosmos DB for NoSQL data modeling, partition key design, indexing strategy, throughput optimization, change feed processing, and integration with analytics pipelines

Required technical skills:

  • Strong Python programming skills, including data structures, functions, exception handling, logging, reusable modules, API integration, and automation scripts
  • Strong PySpark development experience using DataFrame APIs, joins, aggregations, window functions, UDFs, partitioning, caching, broadcast joins, and performance optimization
  • Good SQL skills for querying, transformation, data validation, stored procedures, performance tuning, and troubleshooting data issues
  • Experience with Git, Azure DevOps, CI/CD practices, unit testing, deployment pipelines, monitoring, and production support for data engineering workloads

 

 

Responsibilities:

 

Data Engineering Design & Development

Design, develop, and maintain scalable Azure data engineering solutions across:

  • Batch, incremental, and near-real-time data ingestion from databases, APIs, files, applications, and streaming sources
  • Azure Data Factory, Azure Databricks, ADLS Gen2, Azure Synapse Analytics, Azure SQL, and Azure Cosmos DB

Build and optimize data pipelines for:

  • Data extraction, cleansing, transformation, enrichment, validation, and loading into curated data layers
  • Reusable PySpark frameworks, parameterized notebooks, modular Python components, and metadata-driven processing patterns
  • Data quality controls, exception handling, audit logging, reconciliation, restartability, and operational monitoring

Develop Cosmos DB-based data solutions by:

  • Designing containers, partition keys, indexing policies, consistency levels, TTL, and throughput configuration based on access patterns
  • Implementing ingestion and integration patterns between Cosmos DB, Azure Data Factory, Databricks, ADLS, and analytical stores
  • Using Cosmos DB change feed, bulk operations, query tuning, partition-aware design, and cost optimization practices

 

Pipeline Delivery, Optimization & Support

Own hands-on delivery across:

  • PySpark-based ETL/ELT jobs for large-scale structured, semi-structured, and unstructured data processing
  • Python-based automation, data validation utilities, reusable transformation logic, and integration scripts
  • Azure Data Factory pipelines, Databricks jobs, Synapse SQL workloads, Cosmos DB integrations, and downstream analytics data products

Drive engineering discipline through:

  • Code reviews, unit testing, version control, CI/CD, deployment automation, and environment configuration management
  • Pipeline monitoring, failure handling, performance tuning, cost optimization, and production incident resolution

Preferred Qualifications:

  • Microsoft Azure Data Engineer certification or equivalent hands-on Azure project experience
  • Experience with Delta Lake, lakehouse patterns, medallion architecture, and data warehouse modeling
  • Exposure to event-driven or streaming patterns using Event Hubs, Kafka, Stream Analytics, or Databricks Structured Streaming
  • Understanding of data security, RBAC, managed identities, private endpoints, encryption, and compliance-driven data handling

Key Attributes:

  • Strong analytical and problem-solving skills with the ability to troubleshoot complex data and pipeline issues
  • Ability to work with business analysts, architects, QA teams, and client stakeholders to clarify requirements and deliver reliable data solutions
  • Good communication skills with the ability to explain technical designs, pipeline behavior, and production issues clearly
  • Ownership mindset with focus on quality, maintainability, performance, security, and operational stability

 

Must Have skills:

  • Cosmos DB
  • Azure Data Lake Storage
  • Azure Synapse Analytics
  • Azure Data Factory
  • Postgre SQL

 

 

 

 

Regards,

 

Manoj Goud

Derex Technologies INC

Contact : 973-834-5005 Ext 206

Additional Information

All your information will be kept confidential according to EEO guidelines.

Skills Required

  • 10+ years of experience in data engineering, cloud data platforms, ETL/ELT development, or large-scale data processing
  • Hands-on experience designing, developing, testing, and maintaining Azure-based data pipelines and processing solutions
  • Strong PySpark and Python development experience for large-scale data transformation, automation, data quality checks, and reusable frameworks
  • Experience with Azure Data Factory, including ingestion, orchestration, parameterized pipelines, triggers, and monitoring
  • Experience with Azure Databricks and Apache Spark for scalable data processing
  • Experience with Azure Data Lake Storage Gen2, including storage structures, file formats, access control, and lifecycle management
  • Experience with Azure Synapse Analytics or Azure SQL for analytical workloads, SQL development, modeling, tuning, and reporting integration
  • Experience with Azure Cosmos DB data modeling, partition keys, indexing, throughput optimization, change feed processing, and analytics integration
  • Strong Python programming skills, including data structures, exception handling, logging, reusable modules, API integration, and automation
  • Strong PySpark DataFrame API experience, including joins, aggregations, window functions, UDFs, partitioning, caching, broadcast joins, and optimization
  • Good SQL skills for querying, transformation, validation, stored procedures, performance tuning, and troubleshooting
  • Experience with Git, Azure DevOps, CI/CD, unit testing, deployment pipelines, monitoring, and production support
  • Experience with PostgreSQL
  • Microsoft Azure Data Engineer certification or equivalent hands-on Azure project experience
  • Experience with Delta Lake, lakehouse patterns, medallion architecture, and data warehouse modeling
  • Exposure to event-driven or streaming technologies such as Event Hubs, Kafka, Stream Analytics, or Databricks Structured Streaming
  • Understanding of data security, RBAC, managed identities, private endpoints, encryption, and compliance-driven data handling
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Harrison, NJ
24 Employees
Year Founded: 1996

What We Do

DEREX Technologies, Inc. established in 1996 is engaged nationally in providing professional computer services, including management consulting firm. We provide expert services nationwide to Fortune 500 companies and other private and public organizations in the United States. Derex provides a multi-faceted portfolio of products and services to its clients, including complete IT solutions.

Similar Jobs

BuildOps Logo BuildOps

Account Executive

Cloud • Mobile • Software
Easy Apply
Remote or Hybrid
United States
500 Employees
140K-160K Annually

Sprout Social Logo Sprout Social

Product Marketing Strategist

Marketing Tech • Social Media • Software • Analytics • Business Intelligence
Easy Apply
Remote or Hybrid
US
1400 Employees
94K-156K Annually
Hybrid
Jersey City, NJ, USA
289097 Employees

Similar Companies Hiring

Standard Template Labs Thumbnail
Artificial Intelligence • Information Technology • Software
New York, NY
25 Employees
NODA AI Thumbnail
Artificial Intelligence • Information Technology • Software • Cybersecurity
Sydney, AU
54 Employees
Golden Pet Brands Thumbnail
Digital Media • eCommerce • Information Technology • Marketing Tech • Pet • Retail • Social Media
El Segundo, California
178 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account