Job Title: Data Platform engineer
Experience: 6 to 7 years
Location: Hybrid - Scottsdale, Arizona
No. of Positions: 1
Job Summary
We are seeking a data platform engineer who will be responsible for responsible for designing, deploying, automating, and supporting enterprise data platform infrastructure built on Cloudera Data Platform (CDP). The role emphasizes platform reliability, automation, and operational excellence by leveraging Ansible to provision, configure, patch, and manage large-scale Hadoop and cloud-based data environments. Working closely with infrastructure, security, DevOps, and data engineering teams, the engineer ensures the platform remains secure, highly available, scalable, and optimized to support critical financial data processing and analytics workloads.
Key Responsibilities
- Deploy, administer, and maintain Cloudera Data Platform (CDP) clusters across development, testing, and production environments.
- Automate infrastructure provisioning, software installation, configuration management, and patching using Ansible playbooks and roles.
- Monitor platform health, troubleshoot performance issues, and optimize Hadoop ecosystem services.
- Perform cluster upgrades, capacity planning, disaster recovery testing, and lifecycle management.
- Implement security controls including Kerberos, LDAP/Active Directory integration, TLS, and role-based access control.
- Collaborate with data engineering and application teams to support ingestion, processing, and analytics workloads.
- Develop operational automation and CI/CD workflows to improve platform efficiency and reduce manual effort.
- Maintain system documentation, standard operating procedures, and automation artifacts.
- Ensure platform compliance with enterprise security, governance, and regulatory standards.
- Participate in 24 hrs on-call support in rotation (secondary and primary Pager duty).
Required Skills & Qualifications
- Strong experience in SQL and data modeling
- Experience with workflow orchestration tools like Airflow
- Good understanding of ETL/ELT processes and data pipeline architecture
- Familiarity with cloud platforms (AWS/Azure/GCP) is a plus
- Strong problem-solving and analytical skills
- Excellent Communication skills
- Ability to effectively collaborate with global stakeholders and teams
- Cloudera Data Platform (CDP) and Cloudera Manager
- Ansible automation, playbook development, and configuration management
- Linux system administration (RHEL/CentOS)
- Hadoop ecosystem components (HDFS, YARN, Hive, Impala, Spark, Solr, Kafka, Oozie)
- Shell scripting (Bash) and Python
- Git, Jenkins, and CI/CD pipelines
- Monitoring tools such as Prometheus, Grafana, or Splunk
- Networking, DNS, load balancing, and storage concepts
- Cassandra (docker) knowledge is a plus.
- AI knowledge will be welcomed
- Experience in fintech and banking would also count
Responsibilities
Key Responsibilities
- Deploy, administer, and maintain Cloudera Data Platform (CDP) clusters across development, testing, and production environments.
- Automate infrastructure provisioning, software installation, configuration management, and patching using Ansible playbooks and roles.
- Monitor platform health, troubleshoot performance issues, and optimize Hadoop ecosystem services.
- Perform cluster upgrades, capacity planning, disaster recovery testing, and lifecycle management.
- Implement security controls including Kerberos, LDAP/Active Directory integration, TLS, and role-based access control.
- Collaborate with data engineering and application teams to support ingestion, processing, and analytics workloads.
- Develop operational automation and CI/CD workflows to improve platform efficiency and reduce manual effort.
- Maintain system documentation, standard operating procedures, and automation artifacts.
- Ensure platform compliance with enterprise security, governance, and regulatory standards.
- Participate in 24 hrs on-call support in rotation (secondary and primary Pager duty).
Graduate in Computer Science, or related field. 6+ years of experience in data engineering or a related field.
Base Compensation Range: $110,000 - $130,000
The posted range is the hiring range for this role — a subset of the broader range available to employees over time — and reflects base salary across our national hiring scale. Final offers are based on several factors, including the candidate's skills and experience, internal pay equity, work location, market conditions for the role, and the specific scope and responsibilities of the position. The top of the range is reserved for candidates who notably exceed the requirements; the lower end applies to those with less experience or fewer preferred qualifications. For positions based in higher-cost zones (e.g., California, New York, New Jersey), actual compensation may exceed the posted range; your recruiter will share specifics during the process.
About UsSkills Required
- 6+ years of experience in data engineering or related field
- Graduate in Computer Science or related field
- Cloudera Data Platform (CDP) and Cloudera Manager experience
- Ansible automation, playbook development, and configuration management
- Linux system administration (RHEL/CentOS)
- Hadoop ecosystem components (HDFS, YARN, Hive, Impala, Spark, Solr, Kafka, Oozie)
- Strong SQL skills and data modeling
- Experience with workflow orchestration tools (Airflow)
- Understanding of ETL/ELT processes and data pipeline architecture
- Shell scripting (Bash) and Python
- Git, Jenkins, and CI/CD pipelines
- Monitoring tools such as Prometheus, Grafana, or Splunk
- Implement security controls including Kerberos, LDAP/Active Directory integration, TLS, and role-based access control
- Networking, DNS, load balancing, and storage concepts
- Participate in 24-hour on-call support rotation (PagerDuty)
- Strong problem-solving, analytical, and communication skills; ability to collaborate with global stakeholders
- Familiarity with cloud platforms (AWS/Azure/GCP)
- Cassandra (docker) knowledge
- AI knowledge
- Experience in fintech and banking
What We Do
Choosing a digital partner is about more than capabilities — it’s about collaboration and character. Unrealistic overhauls and off-the-shelf products ignore what matters most — your unique needs, culture, goals, and your legacy data and technology environments. At EXL, our collaboration is built on ongoing listening and learning to adapt our methodologies. We’re your business evolution partner—tailoring solutions that make the most of data to make better business decisions and drive more intelligence into your increasingly digital operations. Whether your goals are scaling the use of AI and digital, redesign operating models, or driving better and faster decisions, we’re here to partner with you to help you gain—and maintain—competitive advantage with efficient, sustainable models at scale. Our expertise in transformation, data science, and change management helps make your business more efficient and effective, improve customer relationships and enhance revenue growth. Instead of focusing on multi-year, resource- and time-intensive platform designs or migrations, we look deeper at your entire value chain to integrate strategies with impact. We use our specialization in analytics, digital interventions, and operations management—alongside deep industry expertise — to deliver solutions that help you outperform the competition. At EXL, it’s all about outcomes—your outcomes—and delivering success on your terms. Share your goals with us and together, we’ll optimize how you leverage data to drive your business forward. For more information, visit www.exlservice.com.


.png)





