Databricks Data Engineer - India

Posted Yesterday
Hiring Remotely in IN
Remote
Senior level
Artificial Intelligence • Information Technology • Machine Learning • Consulting
The Role
Support, maintain, troubleshoot, and optimize production Databricks applications, Spark jobs, SQL queries, Delta Lake pipelines, clusters, workflows, and workspace configurations. Implement data quality checks, logging, alerting, governance, security, and cost-management practices. Collaborate on enhancements and bug fixes, reduce technical debt, document operational processes, and participate in EST-aligned support rotations.
Summary Generated by Built In

Location: Remote

Work Hours: EST (Eastern Standard Time) aligned

Experience: 6–9 years

Role Overview

We are looking for an experienced Databricks Data Engineer to support, maintain, and enhance existing Databricks-based data applications and pipelines. The role focuses on ensuring reliability, performance, and scalability of production Databricks workloads rather than building net-new platforms from scratch. You will work closely with data, analytics, and engineering teams to keep critical data applications stable, optimized, and aligned with business needs.

Key Responsibilities

  • Support and maintain existing Databricks applications, notebooks, jobs, and Delta Lake pipelines in production.

  • Monitor, troubleshoot, and resolve issues related to job failures, performance degradation, data quality, and cluster utilization.

  • Optimize existing Spark jobs, SQL queries, and Delta tables for cost, performance, and reliability.

  • Manage and improve Databricks workspace configurations, including clusters, job scheduling, access controls, and Unity Catalog (where applicable).

  • Implement and maintain data quality checks, logging, alerting, and basic observability for Databricks workloads.

  • Collaborate with stakeholders to understand requirements for enhancements or bug fixes on existing applications.

  • Perform incremental improvements, refactoring, and technical debt reduction on current Databricks solutions.

  • Ensure adherence to best practices around security, governance, and cost management within the Databricks environment.

  • Document existing pipelines, dependencies, and operational runbooks.

  • Participate in on-call or support rotations as needed to maintain production stability (within EST working hours).

Required Qualifications

  • 6–9 years of overall experience in data engineering, with strong hands-on experience in Databricks.

  • Solid proficiency in Apache Spark (PySpark and/or Scala) and SQL.

  • Proven experience supporting and optimizing production Databricks workloads (jobs, notebooks, Delta Lake, workflows).

  • Strong understanding of Delta Lake concepts (ACID transactions, time travel, optimization techniques such as Z-ordering, vacuum, optimize).

  • Experience with Databricks Job clusters, Interactive clusters, and performance tuning (partitioning, caching, shuffle optimization, autoscaling).

  • Familiarity with data modeling, ETL/ELT patterns, and production data pipeline support.

  • Experience working with cloud platforms (preferably Azure, AWS, or GCP) in the context of Databricks.

  • Ability to troubleshoot complex Spark and Databricks issues independently.

  • Strong communication skills and ability to work effectively in a remote, EST-aligned team.

Preferred Qualifications

  • Experience with Unity Catalog, Databricks SQL, or Lakehouse architecture.

  • Knowledge of CI/CD practices for Databricks (e.g., Databricks Asset Bundles, Git integration, Terraform/ARM templates).

  • Familiarity with orchestration tools (Airflow, Azure Data Factory, or Databricks Workflows).

  • Exposure to data quality frameworks, monitoring tools, or cost optimization initiatives on Databricks.

  • Experience supporting analytics or BI teams consuming Databricks data products.

Work Arrangement

  • Fully remote

  • Must be available and productive during EST business hours

  • Collaborative remote environment with regular syncs and support responsibilities

Why Join UsWhy Join Us?

  • Join a team of industry veterans from Google, Meta, and top-tier tech companies.

  • Work on impactful, high-scale projects with leading global clients.

  • Enjoy a flexible, remote-first culture focused on innovation and excellence.

  • Competitive salary, equity options, and continuous learning opportunities.

  • Shape the future of cloud and AI infrastructure at a rapidly growing company.

Perks And Benefits Of Working With Us

  • Internet allowance

  • Laptop

  • PF

  • Paid PTO

  • Annual Bonus

  • Graduity

  • Yearly Team building experiences

  • Mentorship and sponsorship opportunities

  • Manager resources and support

  • Life & accidental insurance for additional protection.

We are an equal opportunity employer. We celebrate diversity and are committed to creating an inclusive environment for all employees. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability, or any other protected characteristic.

Skills Required

  • 6-9 years of overall data engineering experience
  • Strong hands-on experience with Databricks
  • Proficiency in Apache Spark, including PySpark and/or Scala
  • Proficiency in SQL
  • Experience supporting and optimizing production Databricks workloads, including jobs, notebooks, Delta Lake, and workflows
  • Strong understanding of Delta Lake concepts, including ACID transactions, time travel, Z-ordering, VACUUM, and OPTIMIZE
  • Experience with Databricks job and interactive clusters and performance tuning
  • Familiarity with data modeling, ETL/ELT patterns, and production data pipeline support
  • Experience with cloud platforms, preferably Azure, AWS, or GCP, in the context of Databricks
  • Ability to troubleshoot complex Spark and Databricks issues independently
  • Strong communication skills and ability to work effectively in a remote, EST-aligned team
  • Experience with Unity Catalog, Databricks SQL, or Lakehouse architecture
  • Knowledge of CI/CD practices for Databricks, including Databricks Asset Bundles, Git integration, Terraform, or ARM templates
  • Familiarity with orchestration tools such as Airflow, Azure Data Factory, or Databricks Workflows
  • Exposure to data quality frameworks, monitoring tools, or Databricks cost optimization initiatives
  • Experience supporting analytics or BI teams consuming Databricks data products
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
Year Founded: 2025

What We Do

Cogniify is a Bay Area-based AI execution firm that designs, builds, and deploys custom AI systems for Fortune 500 and Global 2000 companies. The company helps enterprises move from AI pilots to industrialized impact and enterprise-scale production, utilizing deep expertise in AI, advanced analytics, data engineering, and domain consulting.

Similar Jobs

Quillbot Logo Quillbot

Manager, Paid Media

Artificial Intelligence • Edtech • Mobile • Natural Language Processing • Productivity • Software
Easy Apply
Remote
India
232 Employees

JPMorganChase Logo JPMorganChase

Data Scientist

Financial Services
Remote or Hybrid
2 Locations
289097 Employees

Atlassian Logo Atlassian

Principal Engineer

Cloud • Information Technology • Productivity • Security • Software • App development • Automation
In-Office or Remote
Bengaluru, Bengaluru Urban, Karnataka, IND
11000 Employees

Atlassian Logo Atlassian

Software Engineer

Cloud • Information Technology • Productivity • Security • Software • App development • Automation
In-Office or Remote
Bengaluru, Bengaluru Urban, Karnataka, IND
11000 Employees

Similar Companies Hiring

Golden Pet Brands Thumbnail
Digital Media • eCommerce • Information Technology • Marketing Tech • Pet • Retail • Social Media
El Segundo, California
178 Employees
Kepler  Thumbnail
Artificial Intelligence • Fintech • Software
New York, New York
9 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account