Problem solvers who don’t just move data but build scalable, reliable data platforms that drive intelligent business decisions.
As a Lead Data Engineer, you will design, build, and scale data platforms that power AI-driven products and analytics solutions. You will work closely with data scientists, product managers, and engineering teams to enable seamless data flow, storage, and processing for enterprise use cases.
Data Platform Development – Design and build scalable data pipelines and data lake architectures for high-performance analytics.
End-to-End Engineering – Develop robust data ingestion, transformation, and processing systems from raw data to production-ready datasets.
Cloud-Native Development – Architect and deploy data solutions on AWS, Azure, or GCP with strong DevOps and automation practices.
Big Data Expertise – Work with distributed data processing frameworks like Spark, Hadoop, and Hive to handle large-scale data workloads.
Collaboration – Partner with cross-functional teams to deliver reliable and impactful data solutions.
Best Practices – Ensure data quality, governance, security, and performance through standardization and continuous improvement.
Strong experience in data engineering with proficiency in Python, Scala, or Java.
Hands-on expertise in building scalable data pipelines and distributed systems.
Good understanding of data lake architecture, data warehousing, and relational databases (Oracle, SQL, DB2, Teradata).
Experience with Big Data technologies such as Hadoop, Spark, Hive, and HBase.
Familiarity with workflow orchestration tools like Airflow.
Experience with cloud platforms (AWS, Azure, or GCP).
Strong problem-solving mindset with the ability to take ownership and lead initiatives.
Passionate about data, innovation, and solving real-world enterprise challenges.
Chennai-based with flexibility to collaborate with global teams and stakeholders.
Lead and deliver data engineering and analytics projects that create business impact.
Build and optimize data pipelines and infrastructure for scalability and efficiency.
Collaborate with teams to understand requirements and translate them into technical solutions.
Communicate complex data concepts to both technical and business stakeholders.
Drive data strategy and support analytics-driven decision-making.
Design and develop data-driven features and components for product platforms.
Travel to client locations when required.
Skills Required
- 5+ years in Data Engineering or Software Engineering
- Proficiency in Python, Scala, or Java
- Hands-on experience building scalable data pipelines and distributed systems
- Understanding of data lake architecture, data warehousing, and relational databases (Oracle, SQL, DB2, Teradata)
- Experience with Big Data technologies (Hadoop, Spark, Hive, HBase)
- Familiarity with workflow orchestration tools like Airflow
- Experience with cloud platforms (AWS, Azure, or GCP) and cloud-native deployment
- Strong problem-solving mindset, ownership, and leadership of initiatives
- Experience with DevOps and automation practices for data platforms
- Chennai-based and able to collaborate with global teams
What We Do
Crayon Data is a leading provider of AI-led revenue acceleration solutions, headquartered in Singapore with a local presence in India and the UAE. The company was founded in 2012 with the vision of simplifying the world’s choices. Our flagship platform, maya.ai, helps enterprises across the Banking, Fintech, and Travel industries create and capture sustainable revenue streams by unlocking the value of data. maya.ai's capability is driven by four “as a Service” components - Data, Recommendation, Customer Experience, and Marketplace - that work individually and together to create tangible results. Crayon Data recently won the E50 awards organized by KPMG and the Business Times in Singapore. Crayon was featured in HFS Hot Vendors Compendium in 2021. They were also among the top 15 finalists at Emerging Enterprise Awards 2019, Singapore.
.jpeg)








