Junior Data Engineer

Posted 16 Days Ago
Be an Early Applicant
Amman, JOR
Hybrid
Junior
Software
The Role
Supports the development, maintenance, monitoring, and troubleshooting of ETL/ELT pipelines and Big Data workloads. Responsibilities include SQL development, data validation, quality checks, workflow automation, production monitoring, anomaly analysis, and collaboration with engineering, DevOps, QA, database, and analytics teams. The role uses Hadoop, Spark, Airflow, Kafka, Kubernetes, Python, Shell scripting, and various data formats while contributing to code reviews, documentation, and continuous improvement.
Summary Generated by Built In

Junior Data Engineer - Jordan Office

Job Overview

As a Junior Data Engineer at Ligadata, you will support the development, maintenance, and monitoring of data pipelines and Big Data solutions. You will work with senior engineers and cross-functional teams to ensure reliable data processing, data quality, and timely delivery.
The role requires a good foundation in SQL, Linux, Shell scripting, data analysis, and Big Data technologies, with a strong willingness to learn and troubleshoot within a production data environment.
Responsibilities

  • Develop and maintain ETL/ELT data pipelines.
  • Write and optimize SQL queries for data processing, validation, and analysis.
  • Support data workflows using Apache Airflow.
  • Work with Big Data technologies such as Hadoop, HDFS, Hive, Spark, Presto/Trino, Kafka, and HBase.
  • Perform data validation, reconciliation, and data-quality checks.
  • Develop scripts and automation using Shell/Bash and Python.
  • Monitor data pipelines and assist in troubleshooting job failures and production issues.
  • Work with structured and semi-structured data formats such as Parquet, JSON, CSV, and Avro.
  • Support applications and data workloads running on Kubernetes (K8s).
  • Analyze data to identify inconsistencies, anomalies, and operational issues.
  • Participate in code reviews, documentation, and continuous improvement activities.
  • Collaborate with senior engineers, DevOps, QA, database, and analytics teams.
  • Use AI-assisted engineering tools to support development, troubleshooting, documentation, and data analysis while validating generated results before use.

Qualifications

  • Bachelor’s degree in Computer Science, Engineering, Information Technology, or a related field.
  • 1–3 years of experience in Data Engineering, Big Data, Software Engineering, or a related role.
  • Good knowledge of SQL and relational database concepts.
  • Good understanding of Linux and Shell/Bash scripting.
  • Basic knowledge of the Hadoop ecosystem, including HDFS and Hive.
  • Familiarity with Spark, Presto/Trino, and Apache Airflow.
  • Basic understanding of Kafka and distributed data-processing concepts.
  • Familiarity with Kubernetes and containerized environments.
  • Knowledge of at least one programming language, preferably Python, Scala, or Java.
  • Basic understanding of ETL/ELT, data warehousing, data quality, and data analysis.
  • Familiarity with Git and software-development practices.
  • Knowledge and practical experience using AI tools such as ChatGPT, GitHub Copilot, or similar tools for engineering tasks.
  • Good analytical, troubleshooting, and problem-solving skills.
  • Strong willingness to learn and develop technical skills.
  • Good communication and teamwork skills.
  • Self-motivated with a strong sense of ownership.

Skills Required

  • Bachelor’s degree in Computer Science, Engineering, Information Technology, or a related field
  • 1–3 years of experience in Data Engineering, Big Data, Software Engineering, or a related role
  • Knowledge of SQL and relational database concepts
  • Understanding of Linux and Shell/Bash scripting
  • Basic knowledge of Hadoop, HDFS, and Hive
  • Familiarity with Apache Spark, Presto/Trino, and Apache Airflow
  • Basic understanding of Kafka and distributed data-processing concepts
  • Familiarity with Kubernetes and containerized environments
  • Knowledge of at least one programming language
  • Python, Scala, or Java programming experience
  • Understanding of ETL/ELT, data warehousing, data quality, and data analysis
  • Familiarity with Git and software-development practices
  • Practical experience using AI tools such as ChatGPT, GitHub Copilot, or similar tools for engineering tasks
  • Analytical, troubleshooting, and problem-solving skills
  • Strong willingness to learn and develop technical skills
  • Good communication and teamwork skills
  • Self-motivation and a strong sense of ownership
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Palo Alto, CA
153 Employees
Year Founded: 2014

What We Do

LigaData is a big-data product startup whose founding members were instrumental in building powerful data technologies including Hadoop, at Yahoo. LigaData champions an open source project named Kamanja, which enables continuous decisioning, the notion of deploying data mining and real time decisioning models on an unlimited volume of data sets. Kamanja is in use at scale by top financial and healthcare institutions for use cases ranging from threat detection to optimizing customer contact. Our dynamic work environment encourages and rewards innovators who bring outside-the-box thinking and leadership skills. Do you have the entrepreneurial vision and ambition to be a part of our journey?

Similar Jobs

Hilton Logo Hilton

Director - Food & Beverage

Software • Hospitality
In-Office
Amman, JOR
121228 Employees

Hilton Logo Hilton

Reservations Agent

Software • Hospitality
In-Office
Amman, JOR
121228 Employees

Ericsson Logo Ericsson

Architect

Cloud • Information Technology • Internet of Things • Machine Learning • Software • Cybersecurity • Infrastructure as a Service (IaaS)
In-Office or Remote
6 Locations
88000 Employees

Capco Logo Capco

Business Consulting Opportunities - Middle East

Fintech • Professional Services • Consulting • Energy • Financial Services • Cybersecurity • Generative AI
Remote or Hybrid
10 Locations
6000 Employees

Similar Companies Hiring

Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees
Revel Thumbnail
Aerospace • Hardware • Robotics • Software
Marina Del Rey, California
60 Employees
Blee Thumbnail
Artificial Intelligence • Marketing Tech • Software
New York, New York
30 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account