JD – Snr Software Engineer (ETL)
Experience & Expectations :
- Leverage extensive experience (4 to 8 years overall ETL experience, to assist in solution design and delivery along with build of new ETLs).
- We are seeking an experienced ETL Developer with strong expertise in Big Data (Spark, Cloudera).
- Experience orchestrating workflows using AWS Step Functions (state machines) for reliable and scalable data pipelines.
- Ability to implement end-to-end serverless data architectures integrating Glue, Lambda, S3, and Redshift
Core Responsibilities :
- Build and maintain high volume ETL/ELT pipelines across Hadoop (HDFS, Hive, Spark, Kafka) and AWS (Glue, EMR, Lambda, Step Functions, Redshift).
- Develop distributed data processing solutions using PySpark, Spark SQL, and scalable cloud serverless patterns.
- Implement reusable data ingestion frameworks for batch, ability to design & implement Orchestration process and Leverage AI
- Optimize data workflows using partitioning, bucketing, compression, file formats (Parquet/ORC).
- Understanding hybrid data lake architectures using S3 + HDFS, ensuring data governance and best practices are adheres
- Experience to deliver complex projects in an Agile environment
- Assist in Design and build the robust, scalable and secure software solutions across the having no/least adoption
- Define clear technical specifications and make architecture decisions that align with business goals and long-term scalability.
- Implement best practices (including secure code guidelines) through the implementation of unit tests, automation, leverage and code reviews. Drive continuous improvement in code quality and maintainability.
- Troubleshooting issues and proactively solving problems as they arise, ensuring the smooth operation of full stack applications
- Ability to understand the data flow diagram, data modelling and Lineages
- Job orchestration using Airflow, Control M, Step Functions, or event-driven triggers.
- Ensure data is protected and compliant with regulatory standards.
- Work closely with business stakeholders to enable high quality datasets.
- Work on best practice adoption and provide guidance to peers/juniors in team.
- Ability to respond on incidents, and troubleshooting Spark performance issues, job failures, and cluster bottlenecks.
- Collaborate closely with team members, QA and cross product teams to streamline release processes.
- Collaborate with business stakeholders to gather, analyse, and translate data into technical solutions
Technical Skills :
- Strong experience with the AWS data stack (S3, Glue, EMR, Lambda, Kinesis, Redshift, Step Functions etc.,).
- Strong hands-on expertise in Scala, PySpark, Spark optimization techniques, HiveQL, and distributed computing.
- Good understanding of Hadoop ecosystem (HDFS, Hive, Spark, YARN, Kafka).
- Good work experience in SQL in hive and impala
- Proficiency in at least one scripting/programming language: Python, Shell scripting.
- Strong experience with CI/CD, GitHub, Git commands.
- Expertise in ETL and Data Warehousing and cloud concepts.
- Good understanding of data modelling (star/snowflake), partitioning strategies, and schema evolution.
- Expertise in data profiling and decision making.
- Able to understand, design and create data flow diagrams.
- Able to understand the architecture and design end-to-end data flow.
- Hands-on experience with Airflow, or Control‑M, or other orchestrators.
- To monitor and support BAU and year end activities, if needed.
- Exposure to security and compliance aspects in Cloud.
- Familiarity with serverless patterns and containerization (Docker, ECS/EKS).
Other Requirements
- Strong logical and analytical, problem-solving, and communication skills.
- Communicate effectively and concisely with multiple stakeholders and coordinate and collaborate with cross functional teams.
- AWS certifications (Data Engineer, or Developer) are a plus.
Detail-Oriented and proactive in problem-solving and issue resolution
We offer you a competitive total rewards package, continuing education & training, and tremendous potential with a growing worldwide organization.
DISCLAIMER:
Nothing in this job description restricts management's right to assign or reassign duties and responsibilities of this job to other entities; including but not limited to subsidiaries, partners, or purchasers of Alight business units.
Skills Required
- 4 to 8 years ETL/Big Data experience
- Experience with Spark and PySpark (development and optimization)
- Hands-on expertise in Scala
- Experience with AWS data stack: S3, Glue, EMR, Lambda, Redshift, Kinesis
- Experience orchestrating workflows using AWS Step Functions (state machines)
- Build and maintain ETL/ELT pipelines across Hadoop (HDFS, Hive, YARN) and AWS
- Strong SQL skills including HiveQL and experience with Impala
- Knowledge of Hadoop ecosystem components (HDFS, Hive, YARN, Kafka)
- Proficiency in Python and Shell scripting
- Experience with data modeling (star/snowflake), partitioning strategies, schema evolution
- Experience with ETL, data warehousing and cloud data concepts
- Experience with orchestration tools such as Airflow or Control-M
- Experience troubleshooting Spark performance, job failures, and cluster bottlenecks
- Familiarity with CI/CD practices, GitHub and Git commands
- Understanding of data governance, security and compliance in cloud environments
- Familiarity with serverless patterns and containerization (Docker, ECS/EKS)
- Experience with Parquet/ORC and file-format optimization (partitioning, bucketing, compression)
- Experience delivering complex projects in an Agile environment
- AWS certifications (Data Engineer or Developer)
Alight Solutions Compensation & Benefits Highlights
The following summarizes recurring compensation and benefits themes identified from responses generated by popular LLMs to common candidate questions about Alight Solutions and has not been reviewed or approved by Alight Solutions.
-
Leave & Time Off Breadth — Leave offerings are described as generous, including multiple vacation weeks alongside wellness days, floating holidays, and paid holidays. Time-off flexibility is frequently positioned as a meaningful part of the overall rewards package.
-
Retirement Support — Retirement benefits are framed as a notable strength, anchored by a 401(k) match structure and an additional retirement account contribution once eligible. Day-one participation and the employer contribution design are presented as differentiators versus many entry-level packages.
-
Wellbeing & Lifestyle Benefits — Wellbeing perks are positioned as a real addition to total rewards, including dedicated wellness days and mental health support such as premium access to Calm. Remote-work enablement is also reinforced through company-provided equipment, which reduces out-of-pocket setup costs.
Alight Solutions Insights
What We Do
Alight is a leading cloud-based human capital technology and services provider that powers confident health, wealth and wellbeing decisions for 36 million people and dependents. Our Alight Worklife® platform combines data and analytics with a simple, seamless user experience. Supported by our global delivery capabilities, Alight Worklife is transforming the employee experience for people around the world. With personalized, data-driven health, wealth, pay and wellbeing insights, Alight brings people the security of better outcomes and peace of mind throughout life’s big moments and most important decisions. Learn how Alight unlocks growth for organizations of all sizes at alight.com.








