The Role
Build and maintain scalable ETL pipelines, data warehouses, data lakes, and database systems. Ensure data quality, reliability, security, and governance through validation, monitoring, access controls, and privacy policies. Collaborate with analysts, data scientists, and product teams to deliver analysis-ready datasets. Optimize big data queries, storage, workflows, latency, and infrastructure costs using cloud, orchestration, and distributed processing technologies.
Summary Generated by Built In
Role Summary
We are looking for a skilled Data Engineer to design, build, and maintain our core data infrastructure. In this role, you will make data clean, reliable, and accessible. You will work closely with data scientists, analysts, and product teams to power business intelligence and machine learning initiatives.
Key Responsibilities
- Pipeline Development: Build, test, and maintain scalable ETL (Extract, Transform, Load) pipelines to move data from various sources into centralized storage.
- Data Warehousing: Design and optimize data warehouses, data lakes, and database systems for high performance.
- Data Quality: Implement validation checks, error-handling routines, and monitoring tools to ensure high data integrity and reliability.
- Collaboration: Partner with data analysts and data scientists to understand data needs and deliver clean, analysis-ready datasets.
- Optimization: Fine-tune big data queries, storage structures, and processing workflows to reduce latency and infrastructure costs.
- Security & Governance: Apply security protocols, access controls, and data privacy policies to keep enterprise data safe.
Required Skills & QualificationsEducation:
- Bachelor’s degree or Master's degree in CS, IT, EC, Data Engineering, or a related technical field (or equivalent work experience).
- Programming: Strong proficiency in Python, plus expert-level SQL skills.
- Database Knowledge: Experience with relational databases (PostgreSQL, MySQL, SQL Server, etc)
- Big Data & Cloud Tools: Must have AWS knowledge. Familiarity with other cloud platforms (GCP, & Azure) and big data frameworks (Apache Spark).
- ETL: Must have Databricks knowledge. Familiarity with AWS Glue.
- Workflow Automation: Experience with orchestration tools like AWS Lambda, Step Function, Airflow, etc.
- Problem-Solving: Strong analytical mindset with an ability to troubleshoot complex data bottlenecks.
- Minimum 5 years of Relevent Experience.
- Experience of having worked on education project(s) or eGovernance projects at State / Central govt. programs, preferred
Skills Required
- Bachelor's or Master's degree in Computer Science, Information Technology, Electronics and Communication, Data Engineering, or a related technical field, or equivalent work experience
- Strong proficiency in Python
- Expert-level SQL skills
- Experience with relational databases such as PostgreSQL, MySQL, or SQL Server
- AWS knowledge
- Databricks knowledge
- At least 5 years of relevant experience
- Familiarity with GCP and Azure
- Familiarity with Apache Spark
- Familiarity with AWS Glue
- Experience with workflow orchestration tools such as AWS Lambda, AWS Step Functions, or Airflow
- Strong analytical and problem-solving skills for troubleshooting complex data bottlenecks
- Experience working on education or eGovernance projects for state or central government programs
Am I A Good Fit?
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.
Success! Refresh the page to see how your skills align with this role.
The Company
What We Do
ConveGenius is an education-technology company focused on making learning accessible and engaging for every child. It develops technology-enabled education solutions, including SwiftChat, SwiftPAL, Swift Insights, and Vidya Samiksha Kendra initiatives, and collaborates with governments and partners to support educational transformation. Its work combines digital learning, AI-powered tools, and large-scale outreach, with deployments spanning Indian states and a user base exceeding 150 million.








