Lead Data Engineer - Databricks

Posted 14 Days Ago
Be an Early Applicant
Coimbatore, Tamil Nadu, IND
Hybrid
Expert/Leader
Blockchain • Database • Analytics
The Role
Lead the design, development, and optimization of scalable Databricks ETL/ELT pipelines. Integrate data from databases, S3, files, and REST APIs; implement Delta Lake medallion architecture, Unity Catalog, dimensional models, data quality controls, and Spark performance tuning. Schedule and monitor Databricks workflows, collaborate cross-functionally, and guide engineers while driving technical decisions.
Summary Generated by Built In
Position Overview

We are looking for a Data Engineer with hands-on experience in building scalable data pipelines and data engineering solutions on the Databricks Lakehouse Platform. The ideal candidate should have strong expertise in Python, PySpark, SQL, Databricks, AWS, and REST API integrations for data ingestion, managing large volumes of data, and data export
ShyftLabs is a growing data product company that was founded in early 2020 and works primarily with Fortune 500 companies. We deliver digital solutions built to help accelerate the growth of businesses in various industries, by focusing on creating value through innovation.
 

Job Responsibilities:

    Design, develop, and maintain scalable ETL/ELT pipelines using Databricks,
    PySpark, and SQL.
    ● Integrate data from multiple sources, including databases, Amazon S3, files, and REST APIs.
    ● Build data pipelines with Databricks Unity Catalog.
    ● Implement business logic, data transformations, and dimensional data models.
    ● Create, schedule, monitor, and optimize Databricks Jobs and Workflows.
    ● Design and manage Delta Lake tables using Medallion Architecture (Bronze, Silver,Gold).
    ● Ensure data quality through validations, error handling, logging, and monitoring.
    ● Optimize Spark workloads for performance, scalability, and reliability.
    ● Collaborate with cross-functional teams to deliver production-ready data solutions.

Basic Qualification:

    Strong expertise in Python, PySpark, and Advanced SQL.
    ● Hands-on experience with the Databricks Lakehouse Platform.
    ● Good understanding of Unity Catalog, Delta Lake, Databricks Workflows/Jobs,
    Clusters, Notebooks, Repos, and Medallion Architecture.
    ● Experience integrating with REST APIs for data ingestion and data export.
    ● Strong knowledge of ETL/ELT development, batch processing, incremental loading,
    and data transformation.
    ● Experience with data modeling (Star Schema, Snowflake Schema, Fact & Dimension
    tables, SCD concepts).
    ● Understanding of data warehousing concepts and best practices.
    ● Experience working with structured and semi-structured data (CSV, JSON, Parquet,
    Delta).
    ● Knowledge of partitioning, file optimization, Spark performance tuning, and query
    optimization.
    ● Experience with Git and CI/CD best practices

Preferred Qualifications:

  • 9+ years of experience in Data Engineering, including 3+ years of hands-on experience with Databricks.
  • Prior experience in a Lead Data Engineer / Technical Lead role, with experience guiding engineers and driving technical decisions.
  • Strong hands-on experience with Databricks, Apache Spark, and SQL.
  • Experience designing, developing, and optimizing ETL/ELT data pipelines.
  •  Experience with Auto Loader, Spark Declarative pipelines, Kafka, Airflow, or dbt is a plus. 
  • Databricks certification is an added advantage. Give me Jd for lead role 

We are proud to offer a competitive salary alongside a strong insurance package. We pride ourselves on the growth of our employees, offering extensive learning and development resources.

Skills Required

  • Strong expertise in Python, PySpark, and advanced SQL
  • Hands-on experience with the Databricks Lakehouse Platform
  • Experience with Unity Catalog, Delta Lake, Databricks Workflows and Jobs, clusters, notebooks, repos, and Medallion Architecture
  • Experience integrating REST APIs for data ingestion and data export
  • Strong knowledge of ETL/ELT development, batch processing, incremental loading, and data transformation
  • Experience with data modeling, including star schema, snowflake schema, fact and dimension tables, and SCD concepts
  • Understanding of data warehousing concepts and best practices
  • Experience working with structured and semi-structured data, including CSV, JSON, Parquet, and Delta
  • Knowledge of partitioning, file optimization, Spark performance tuning, and query optimization
  • Experience with Git and CI/CD best practices
  • 9+ years of experience in data engineering
  • 3+ years of hands-on Databricks experience
  • Prior Lead Data Engineer or Technical Lead experience
  • Experience guiding engineers and driving technical decisions
  • Strong hands-on experience with Databricks, Apache Spark, and SQL
  • Experience designing, developing, and optimizing ETL/ELT data pipelines
  • Experience with Auto Loader, Spark Declarative Pipelines, Kafka, Airflow, or dbt
  • Databricks certification
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Toronto, Ontario
110 Employees
Year Founded: 2018

What We Do

We provide customized data and analytics consulting services, including automation and software development for a sustainable and intuitive digital transformation.

Similar Jobs

Toast Logo Toast

Senior Software Engineer – (Java Full Stack)

Cloud • Fintech • Food • Information Technology • Software • Hospitality
Hybrid
Chennai, Tamil Nadu, IND
5000 Employees

Pfizer Logo Pfizer

Sr. Associate, Basic Results Author

Artificial Intelligence • Healthtech • Machine Learning • Natural Language Processing • Biotech • Pharmaceutical
Hybrid
Chennai, Tamil Nadu, IND
121990 Employees

Pfizer Logo Pfizer

Manager, PKI Engineer

Artificial Intelligence • Healthtech • Machine Learning • Natural Language Processing • Biotech • Pharmaceutical
Hybrid
Chennai, Tamil Nadu, IND
121990 Employees

Pfizer Logo Pfizer

Manager, Data Sharing & Disclosure

Artificial Intelligence • Healthtech • Machine Learning • Natural Language Processing • Biotech • Pharmaceutical
In-Office
Chennai, Tamil Nadu, IND
121990 Employees

Similar Companies Hiring

Northslope Thumbnail
Artificial Intelligence • Information Technology • Software • Analytics • Consulting • Generative AI
London, GB
100 Employees
Scotch Thumbnail
Artificial Intelligence • eCommerce • Fintech • Payments • Retail • Software • Analytics
US
35 Employees
Milestone Systems Thumbnail
Artificial Intelligence • Security • Software • Analytics • Big Data Analytics
Lake Oswego, OR
1500 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account