Senior DataOps Engineer

Posted Yesterday
Be an Early Applicant
27 Locations
Remote
Senior level
Information Technology • Consulting
The Role
Build and operate secure AWS-based data platforms and DataOps pipelines using S3, Iceberg, Glue, EMR, Spark, Flink, Kafka, and MSK. Implement medallion lakehouse architecture, batch and streaming processing, PII tokenization, data migration, governance, schema-drift detection, CI/CD automation, observability, FinOps monitoring, and disaster recovery procedures. Collaborate with data, cloud, security, and governance teams to support hybrid cloud modernization and AI initiatives.
Summary Generated by Built In

N-iX is looking for Senior DataOps Engineer to join the team

Client Overview:
Our client is an Azerbaijani telecommunications company, the largest mobile network operator in Azerbaijan. The main products are: Fixed telephony, Mobile telephony, Internet services, Wireless broadband, and Value-added services.

Project Objectives:
The primary goal is to accelerate the client’s Data & AI initiatives via a secure, hybrid cloud foundation on AWS while systematically modernizing the IT estate as part of the cloud migration.


Key Project Objectives include:

  • Cloud Foundation & Landing Zone: Deploy target hybrid network architectures, establishing a secure Landing Zone and hybrid Data/AI platforms on AWS.
  • Security, Compliance & Governance: Operationalize on-prem tokenization (achieving zero raw PII in the cloud), resolve policy blockers to include AWS in the ISMS, and establish a Cloud Center of Excellence (CCoE) to govern Cloud adoption.
  • AI Chatbot & Voicebot Design & Implementation: Develop and operationalize a flagship Customer Care Chatbot and Voicebot as the first hybrid-setup consumer.

Key Responsibilities:

  • Design, implement, and maintain the AWS Data Platform foundation, including Amazon S3 lake layout, Apache Iceberg table format standardization, and AWS Glue Data Catalog integration.
  • Build and optimize scalable batch and streaming data pipelines using Amazon EMR (Serverless and EMR-on-EKS), Apache Spark/PySpark, Apache Flink, and Amazon Athena with workgroup cost caps.
  • Implement data tokenization and de-identification pipelines on the data side (PA-T) using Protegrity/Spark UDFs for batch, Spark Streaming for micro-batch, and Kafka Connect SMT for streaming PII masking prior to cloud transit.
  • Build and manage real-time streaming architectures with Amazon MSK (Managed Streaming for Apache Kafka), MSK Replicator/MirrorMaker2, and Schema Registry integration.
  • Implement Medallion Architecture (Bronze, Silver, Gold layers) for lakehouse data modeling, automating Iceberg table registration and backfill frameworks.
  • Execute data migration waves across non-PII and PII datasets using AWS DataSync, automated register steps, and streaming migration configs.
  • Configure data access control and governance models using AWS Lake Formation, cross-engine authorization, AWS Macie for PII detection, and Informatica Data Catalog (Axon, EDC, BDQ) integration.
  • Build automated schema-drift identification components, GitLab CI/CD pipeline integration for dataset synchronization, and automated data contracts/circuit breakers.
  • Implement DataOps observability, centralizing logging via Amazon CloudWatch, configuring FinOps cost & anomaly monitoring, and setting up automated alerts.
  • Author technical documentation, operational runbooks, disaster recovery (DR) procedures, and cutover/rollback playbooks for data platform hardening.

Requirements:

Mandatory Technical Skills:

  • 4+ years of hands-on experience as a Data Engineer or DataOps Engineer building enterprise-grade data platforms and pipelines.
  • Strong expertise with AWS Data Analytics stack: Amazon S3, AWS Glue Data Catalog, AWS Lake Formation, Amazon EMR (EMR-on-EKS / Serverless), Amazon Athena, AWS DataSync, and Amazon MSK.
  • Deep experience with Apache Iceberg table format, cataloging, compaction, and schema evolution.
  • Proficient in Apache Spark / PySpark and Spark Streaming for batch, micro-batch, and real-time data processing.
  • Strong experience in Medallion Lakehouse Architecture design and implementation (Bronze, Silver, Gold layers).
  • Practical experience in implementing data tokenization and encryption at scale (e.g., Protegrity, Thales, FPE, or Spark UDF-based de-identification pipelines).
  • Solid knowledge of event-driven architectures & streaming: Apache Kafka / Amazon MSK, Kafka Connect (SMT), Schema Registry, and MirrorMaker2 / MSK Replicator.
  • Expertise with Relational Databases (Amazon RDS, PostgreSQL, Oracle) and data synchronization techniques.
  • Hands-on experience with DataOps CI/CD & Automation: GitLab CI/CD, Infrastructure-as-Code (Terraform / AWS CDK), schema-drift detection, and data contract validation.
  • Familiarity with data governance tools and enterprise data catalogs (e.g., Informatica Axon/EDC, AWS Lake Formation).

Strong Plus (Nice-to-Have Skills):

  • AWS Certified Data Analytics – Specialty or AWS Certified Data Engineer – Associate.
  • Experience in telecom domain data models, CDR processing, and high-throughput real-time telemetry.
  • Experience with cloud-side tokenization/detokenization via Athena UDFs / AWS Lambda.
  • Familiarity with containerization (Docker, EKS, Kubernetes) for big data runtimes.
  • Experience with AWS Macie and FinOps cost-allocation/anomaly-detection frameworks.

Soft Skills & Team Fit:

  • Strong critical thinking, problem-solving, and analytical skills.
  • Excellent communication and collaboration skills to work closely with cross-functional teams (Data Science, Cloud/Platform, Security, Governance).
  • Results-oriented, proactive mindset with strong ownership of deliverables within an Agile / Scrum framework.
  • Upper-Intermediate+ English level (written and spoken).

We offer*:

  • Flexible working format - remote, office-based or flexible
  • A competitive salary and good compensation package
  • Personalized career growth
  • Professional development tools (mentorship program, tech talks and trainings, centers of excellence, and more)
  • Active tech communities with regular knowledge sharing
  • Education reimbursement
  • Memorable anniversary presents
  • Corporate events and team buildings
  • Other location-specific benefits

*not applicable for freelancers

Skills Required

  • 4+ years of hands-on experience as a Data Engineer or DataOps Engineer building enterprise-grade data platforms and pipelines
  • Strong expertise with AWS S3, Glue Data Catalog, Lake Formation, EMR, Athena, DataSync, and MSK
  • Deep experience with Apache Iceberg, including cataloging, compaction, and schema evolution
  • Proficiency with Apache Spark, PySpark, and Spark Streaming
  • Strong experience designing and implementing Medallion Lakehouse Architecture
  • Practical experience implementing data tokenization and encryption at scale
  • Knowledge of Apache Kafka, Amazon MSK, Kafka Connect SMT, Schema Registry, and MirrorMaker2 or MSK Replicator
  • Expertise with relational databases including Amazon RDS, PostgreSQL, or Oracle, and data synchronization techniques
  • Hands-on experience with GitLab CI/CD, Terraform or AWS CDK, schema-drift detection, and data contract validation
  • Familiarity with Informatica Axon, Informatica EDC, Informatica BDQ, or AWS Lake Formation
  • AWS Certified Data Analytics Specialty or AWS Certified Data Engineer Associate
  • Experience with telecom data models, CDR processing, and high-throughput real-time telemetry
  • Experience with cloud-side tokenization or detokenization using Athena UDFs or AWS Lambda
  • Familiarity with Docker, EKS, or Kubernetes for big data runtimes
  • Experience with AWS Macie and FinOps cost-allocation or anomaly-detection frameworks
  • Strong critical thinking, problem-solving, analytical, communication, and collaboration skills
  • Results-oriented, proactive ownership within an Agile or Scrum framework
  • Upper-Intermediate or higher English proficiency
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Valletta
2,135 Employees
Year Founded: 2002

What We Do

N-iX is a global software solutions and engineering services company that helps world’s leading organizations turn challenges into lasting business value, operational efficiency, and revenue growth using advanced technology. Whether you need to build a custom solution, modernize your digital product or acquire extra tech expertise - we have the experience and capabilities to ensure your success. With over 2,000 professionals in 25 countries across Europe and the Americas, N-iX offers expert solutions in cloud, data analytics, embedded software, IoT, AI, machine learning, and other tech domains. Being in business for over two decades, we have worked with dozens of industry-leading enterprises and Fortune 500 companies creating value across a wide variety of sectors, including finance, manufacturing, supply chain, retail, e-commerce, healthcare, and more. Our unique combination of business domain expertise and technical know-how enables us to effectively collaborate with ISVs, tech companies, and enterprises of all sizes. Thanks to the strong tech ecosystem and partnerships with AWS, GCP, Microsoft, SAP, OpenText, Snowflake, and others, we bring extra speed, scale and efficiency to more than 160 organizations across the globe. N-iX is recognized by numerous industry awards, such as CRN Solution Provider 500, Global Outsourcing 100 by IAOP, ISG Provider Lens™, Modern Application Development services providers by Forrester, etc

Similar Jobs

Mastercard Logo Mastercard

Consultant

Blockchain • Fintech • Payments • Consulting • Cryptocurrency • Cybersecurity • Quantum Computing
Remote or Hybrid
Athens, GRC
38800 Employees

Pfizer Logo Pfizer

Staff Software Engineer

Artificial Intelligence • Healthtech • Machine Learning • Natural Language Processing • Biotech • Pharmaceutical
In-Office or Remote
36 Locations
121990 Employees

Mondelēz International Logo Mondelēz International

Brand Manager

Big Data • Food • Hardware • Machine Learning • Retail • Automation • Manufacturing
Remote or Hybrid
Athens, GRC
90000 Employees

Pfizer Logo Pfizer

Sustainability Senior Manager

Artificial Intelligence • Healthtech • Machine Learning • Natural Language Processing • Biotech • Pharmaceutical
Remote or Hybrid
30 Locations
121990 Employees
112K-207K Annually

Similar Companies Hiring

Axle Health Thumbnail
Artificial Intelligence • Healthtech • Information Technology • Logistics
Santa Monica, CA
25 Employees
NODA AI Thumbnail
Artificial Intelligence • Information Technology • Software • Cybersecurity
Sydney, AU
54 Employees
Golden Pet Brands Thumbnail
Digital Media • eCommerce • Information Technology • Marketing Tech • Pet • Retail • Social Media
El Segundo, California
178 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account