Senior Software Engineer - Ingestion

Reposted 16 Days Ago
Be an Early Applicant
Bengaluru, Bengaluru Urban, Karnataka, IND
In-Office
Senior level
Big Data • Machine Learning • Software • Analytics • Big Data Analytics
The Role
Design and build highly scalable, fault-tolerant data ingestion connectors and engines for the Lakehouse. Implement incremental data capture and log parsing, optimize performance on large clusters, debug low-level systems, own architecture and roadmap, and lead complex cross-team technical projects from design through operations.
Summary Generated by Built In

P-1403

At Databricks, we are passionate about enabling data teams to solve the world's toughest problems - from making the next mode of transportation a reality to accelerating the development of medical breakthroughs. We do this by building and running the world's best data and AI infrastructure platform so our customers can use deep data insights to improve their business.

Ingesting data into the Lakehouse is a strategic area of investment for Databricks and a key enabler for Data and AI workflows. Lakeflow Connect is looking to solve this problem by providing ready-to-use, point-and-click connectors for a wide variety of sources, including enterprise applications (like Salesforce, Workday, ServiceNow, SharePoint), databases (e.g., SQL Server), cloud storage, message queues, and local files.

In addition to being an important part of Lakeflow and Data Engineering, Connect is also a key platform capability. Every surface in Databricks (Dashboards, Notebooks, SQL, AI) requires ingestion capabilities and the lead for this role will need to work closely with other products to embed Connect into these surfaces.

We are looking for engineers with experience in core Database internals to join our Lakeflow Connect team. A key part of Connect is to extract data from OLTP systems while imposing minimal load on production systems. To do this efficiently we are building systems that use techniques such as incremental data capture, log parsing, etc. We are looking for engineers who continue to be hands on and are looking to make a large impact on an important problem for the company.

The Impact you will have:

  • Solve real business needs at large scale by applying your software engineering.
  • Deliver a highly scalable, available, and fault-tolerant engine processing hundreds of TB of data daily across thousands of customers
  • Low level systems debugging, performance measurement & optimization on large production clusters.
  • Build architecture design, influence product roadmap, and take ownership and responsibility over new projects
  • Use your deep experience to help prevent and investigate production issues.
  • Plan and lead complicated technical projects that work with several teams within the company.
  • Break down complex problems quickly into potential solutions, knowns, and unknowns, and de-risk (through prototyping/validation).

What we look for:

  • BS (or higher) in Computer Science, or a related field
  • 5+ years of production level experience in one of: Python, Java, Scala, C++, or similar language.
  • Experience developing large-scale distributed systems from scratch
  • Experience in areas like Database replication, backup, transaction recovery at one of the major database vendors (Microsoft SQL Server , Oracle, IBM etc) is a plus to have.
  • Hands-on experience in developing and operating backend systems.
  • Ability to contribute effectively throughout all project phases, from initial design and development to implementation and ongoing operations, with guidance from senior team members.

About Databricks

Databricks is the Data and AI company. More than 20,000 organizations worldwide — including adidas, AT&T, Bayer, Block, Mastercard, Rivian, Unilever, and 70% of the Fortune 500 — rely on the Databricks Data + AI Platform to build and scale data and AI apps, analytics and agents. Headquartered in San Francisco with 30+ offices around the globe, Databricks offers a unified platform that includes Genie, Lakebase, Agent Bricks, Lakeflow, Lakehouse, and Unity Catalog. To learn more, follow Databricks on LinkedIn, X, YouTube, and Instagram.
Benefits
At Databricks, we strive to provide comprehensive benefits and perks that meet the needs of all of our employees. For specific details on the benefits offered in your region click here.

Our Commitment to Diversity and Inclusion

At Databricks, we are committed to fostering a diverse and inclusive culture where everyone can excel. We take great care to ensure that our hiring practices are inclusive and meet equal employment opportunity standards. Individuals looking for employment at Databricks are considered without regard to age, color, disability, ethnicity, family or marital status, gender identity or expression, language, national origin, physical and mental ability, political affiliation, race, religion, sexual orientation, socio-economic status, veteran status, and other protected characteristics.

Compliance

If access to export-controlled technology or source code is required for performance of job duties, it is within Employer's discretion whether to apply for a U.S. government license for such positions, and Employer may decline to proceed with an applicant on this basis alone.

Skills Required

  • BS or higher in Computer Science or related field
  • 5+ years production experience in Python, Java, Scala, C++ or similar language
  • Experience developing large-scale distributed systems
  • Hands-on experience developing and operating backend systems
  • Experience with core database internals, incremental data capture, log parsing, replication, backup, or transaction recovery
  • Experience specifically with database vendors (Microsoft SQL Server, Oracle, IBM) for replication/transaction systems
  • Ability to lead technical projects across teams and contribute through design, implementation, and operations

Databricks Compensation & Benefits Highlights

The following summarizes recurring compensation and benefits themes identified from responses generated by popular LLMs to common candidate questions about Databricks and has not been reviewed or approved by Databricks.

  • Healthcare Strength Company materials highlight comprehensive medical, dental, and vision coverage alongside mental-health resources, wellness reimbursements, and business travel insurance. Offerings are described as broad and modern, with core health coverage consistently emphasized.
  • Parental & Family Support Paid parental leave is explicitly called out, with details such as up to 20 weeks for birthing parents and up to 12 weeks for non-birthing parents in the U.S. Public materials also reference family-forming support, reinforcing the focus on families.
  • Wellbeing & Lifestyle Benefits Wellness programs and perks include gym reimbursement, periodic wellness events (e.g., yoga, massages), and in-office meals and snacks in many locations. Personal development funds and discounts further enhance lifestyle and growth support.

Databricks Insights

Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: San Francisco, CA
2,200 Employees
Year Founded: 2013

What We Do

As the leader in Unified Data Analytics, Databricks helps organizations make all their data ready for analytics, empower data science and data-driven decisions across the organization, and rapidly adopt machine learning to outpace the competition. By providing data teams with the ability to process massive amounts of data in the Cloud and power AI with that data, Databricks helps organizations innovate faster and tackle challenges like treating chronic disease through faster drug discovery, improving energy efficiency, and protecting financial markets.

Similar Jobs

Cargill Logo Cargill

Software Engineer

Food • Greentech • Logistics • Sharing Economy • Transportation • Agriculture • Industrial
In-Office
Bengaluru, Bengaluru Urban, Karnataka, IND
155000 Employees

Cargill Logo Cargill

Consultant

Food • Greentech • Logistics • Sharing Economy • Transportation • Agriculture • Industrial
In-Office
Bengaluru, Bengaluru Urban, Karnataka, IND
155000 Employees

Optum Logo Optum

Senior Data Analyst

Artificial Intelligence • Big Data • Healthtech • Information Technology • Machine Learning • Software • Analytics
In-Office
Bengaluru, Bengaluru Urban, Karnataka, IND
160000 Employees

Optum Logo Optum

Machine Learning Engineer

Artificial Intelligence • Big Data • Healthtech • Information Technology • Machine Learning • Software • Analytics
In-Office
Bengaluru, Bengaluru Urban, Karnataka, IND
160000 Employees

Similar Companies Hiring

Hanover Park Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
42 Employees
Kepler  Thumbnail
Artificial Intelligence • Fintech • Software
New York, New York
9 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account