Data Engineer

Posted 5 Hours Ago
Be an Early Applicant
Bengaluru, Bengaluru Urban, Karnataka, IND
Hybrid
Senior level
Healthtech • Software • Analytics • Biotech • Pharmaceutical • Manufacturing
Takeda exists to create better health for people, brighter future for the world.
The Role
Build and maintain scalable batch and streaming data pipelines using Databricks, PySpark, and SQL. Develop reliable datasets for analytics, BI, and downstream systems; troubleshoot and optimize production pipelines; apply data quality, testing, documentation, and engineering standards. Collaborate with analytics, product, platform, architecture, security, and DevOps teams to deliver cloud-based data solutions within established enterprise frameworks.
Summary Generated by Built In

By clicking the “Apply” button, I understand that my employment application process with Takeda will commence and that the information I provide in my application will be processed in line with Takeda’s Privacy Notice and Terms of Use.  I further attest that all information I submit in my employment application is true to the best of my knowledge.

Job Description

PRIMARY OBJECTIVES: 


  • Build and maintain scalable data pipelines and datasets that support analytics, reporting, and downstream business systems.
  • Develop data solutions on Databricks using established engineering patterns, reusable frameworks, and enterprise standards.
  • Ensure reliable, high-quality, and performant data delivery across batch and, where relevant, streaming use cases.
  • Support Takeda’s data transformation journey through strong engineering practices, collaboration, and scalable platform-aligned development.

                                   

RESPONSIBILITIES: 

  • Design, develop, test, and maintain scalable data pipelines and integrations using Databricks, PySpark, and SQL.
  • Build datasets optimized for analytics, BI, and downstream consumption while ensuring data quality, reconciliation, and production reliability.
  • Work within established data frameworks, design patterns, and reusable components created by other engineering teams.
  • Read, understand, troubleshoot, and extend existing codebases and pipeline logic in line with engineering standards.
  • Collaborate with analytics, product, and business teams to support data models and data products for enterprise use cases.
  • Contribute to unit, integration, and performance testing, documentation, and engineering best practices.
  • Partner with platform, architecture, security, and DevOps teams to deploy and support pipeline solutions in cloud environments.
  • Troubleshoot data and pipeline issues and drive continuous improvement in performance, scalability, and maintainability.

SCOPE OF SUPERVISION:

NUMBER SUPERVISED WORKERS

Direct

Indirect

Employees

0-3

0-3

Non-Employees

0-3

0-3


 

 

EDUCATION AND EXPERIENCE:  

  • Bachelor’s or Master’s degree in Computer Science, Engineering, Information Systems, or related field.
  • 5+ years of experience in data engineering, data warehousing, or large-scale data platform development.
  • Strong hands-on experience with Databricks and distributed data processing.
  • Strong hands-on experience with PySpark for pipeline development and transformation of large datasets.
  • Strong hands-on experience with SQL, including joins, aggregations, optimization, and analytical data processing.
  • Experience building and maintaining data pipelines for batch processing; exposure to streaming is a plus.
  • Experience working with existing enterprise frameworks, shared libraries, and engineering standards.
  • Experience reading, understanding, debugging, and enhancing existing code developed by other teams.
  • Experience with cloud data platforms such as AWS or Azure.
  • Experience working in agile, cross-functional engineering environments.

KEY SKILLS AND COMPETENCIES:  

  • Strong proficiency in PySpark and SQL; Python alone is not sufficient for this role.
  • Strong understanding of distributed data processing, performance optimization, and scalable pipeline design.
  • Ability to work effectively within predefined patterns, frameworks, and architectural guardrails.
  • Strong code reading and code comprehension skills across shared enterprise codebases.
  • Good understanding of data modeling, schema design, and data quality controls.
  • Strong engineering discipline in testing, version control, documentation, and maintainable development.
  • Strong problem-solving skills and ability to troubleshoot production data issues.
  • Effective communication and collaboration with technical and non-technical stakeholders.

NICE TO HAVE:

·         Experience with streaming technologies such as Spark Structured Streaming or Kafka.

·         Experience with orchestration and workflow tools in enterprise data environments.

·         Experience with Infrastructure as Code, preferably Terraform.

·         Experience designing and developing API-based integrations.



LICENSES/CERTIFICATIONS:

  • Preferred - Databricks Certified Data Engineer Associate / Professional
  • Preferred - AWS or Azure Data Engineering certification

 

PHYSICAL DEMANDS: 

·         N/A


TRAVEL REQUIREMENTS:

·         Access to transportation to attend meetings.

·         Ability to fly to meetings regionally and globally.


LocationsIND - Bengaluru

Worker TypeEmployee

Worker Sub-TypeRegular

Time TypeFull time

Skills Required

  • Bachelor's or Master's degree in Computer Science, Engineering, Information Systems, or a related field
  • 5+ years of experience in data engineering, data warehousing, or large-scale data platform development
  • Strong hands-on experience with Databricks and distributed data processing
  • Strong hands-on experience with PySpark for pipeline development and large-dataset transformation
  • Strong hands-on experience with SQL, including joins, aggregations, optimization, and analytical data processing
  • Experience building and maintaining batch-processing data pipelines
  • Experience with enterprise frameworks, shared libraries, and engineering standards
  • Experience reading, debugging, and enhancing existing code developed by other teams
  • Experience with cloud data platforms such as AWS or Azure
  • Experience working in agile, cross-functional engineering environments
  • Experience with streaming technologies such as Spark Structured Streaming or Kafka
  • Experience with orchestration and workflow tools in enterprise data environments
  • Experience with Infrastructure as Code, preferably Terraform
  • Experience designing and developing API-based integrations
  • Databricks Certified Data Engineer Associate or Professional certification
  • AWS or Azure Data Engineering certification

Takeda Compensation & Benefits Highlights

The following summarizes recurring compensation and benefits themes identified from responses generated by popular LLMs to common candidate questions about Takeda and has not been reviewed or approved by Takeda.

  • Retirement Support — Employer-funded retirement is described as notably strong, combining a dollar-for-dollar 401(k) match with an additional company contribution that scales with age and service. Access to an employee stock purchase plan further supports long-term wealth building.
  • Parental & Family Support — Paid bonding leave for all parents, substantial adoption/surrogacy reimbursement, and robust caregiver resources (backup care and Maven family-forming support) are emphasized as core strengths. These offerings create a comprehensive safety net for a range of family situations.
  • Healthcare Strength — Multiple medical plan options (nationwide PPO/HSA and regional HMOs), employer HSA funding, and integrated mental-health and well-being programs signal depth in coverage. Preventive care is covered in-network, and plan choices by state expand access.

Takeda Insights

Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Cambridge, MA
50,000 Employees
Year Founded: 1781

What We Do

For over 240 years, Takeda’s propensity to evolve has driven the next generation of innovation, and as a future-focused organization, we’re continuing to drive forward with endurance in our steadfast pursuit to achieve the best outcomes for our patients in a rapidly changing world.  We have been preparing for this period of value creation by investing in data, digital and technology, and we’re proud of our employees and their commitment to turning groundbreaking ideas into life-changing impacts.   Since our founding in Japan, integrity and putting patients first have been at the heart of our identity, and we will emerge ready for our future as one of the most trusted and science-driven digital biopharmaceutical companies. Join a team where your innovation impacts lives.   Together, we’ll realize improved outcomes by improving data quality, enhancing launch execution and improving the patient journey. You’ll play a critical role in accelerating data collection and increasing accuracy across all parts of the business. Patients across the globe will benefit from access to treatments afforded by greater opportunities and efficiency in our research and development.  

Why Work With Us

We connect to our history and Japanese heritage through everything we do to bring our purpose, values, vision, and imperatives to life. We are committed to bringing better health and a brighter future to patients. Being a part of Takeda means having the opportunity to be a part of something bigger than yourself.

Gallery

Gallery

Similar Jobs

Takeda Logo Takeda

Data Engineer

Healthtech • Software • Analytics • Biotech • Pharmaceutical • Manufacturing
Hybrid
Bengaluru, Bengaluru Urban, Karnataka, IND
50000 Employees

Optum Logo Optum

Data Engineer

Artificial Intelligence • Big Data • Healthtech • Information Technology • Machine Learning • Software • Analytics
In-Office
Bengaluru, Bengaluru Urban, Karnataka, IND
160000 Employees

Hewlett Packard Enterprise Logo Hewlett Packard Enterprise

Data Scientist

Artificial Intelligence • Cloud • Information Technology • Consulting
In-Office
Bengaluru, Bengaluru Urban, Karnataka, IND
85422 Employees
Hybrid
Bengaluru, Bengaluru Urban, Karnataka, IND
289097 Employees

Similar Companies Hiring

Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees
Revel Thumbnail
Aerospace • Hardware • Robotics • Software
Marina Del Rey, California
60 Employees
Blee Thumbnail
Artificial Intelligence • Marketing Tech • Software
New York, New York
30 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account