Top Data Engineer Jobs in Mexico City

YesterdaySaved
Remote
México
Mid level
Mid level
Information Technology • Software • Business Intelligence
Designs, builds, deploys, and optimizes cloud-based data pipelines on GCP. Integrates data from databases, APIs, and streaming platforms using ETL and ELT processes; manages BigQuery and NoSQL databases; performs data modeling and query optimization; and uses Airflow and modern transformation tools. Collaborates with data scientists, analysts, clients, and business stakeholders in distributed multinational teams.
Top Skills: AlteryxApache AirflowSparkBigQueryConfluenceDatabricksDataformDbtEltETLGitGoogle Cloud Platform (Gcp)JIRALookerAzureNoSQLPower BIPythonSQLTableauTalend
YesterdaySaved
Remote
MEX
Entry level
Entry level
Information Technology • Logistics • Professional Services • Consulting
Designs and maintains scalable cloud-native ETL/ELT pipelines, data models, Lakehouse architectures, APIs, and real-time streaming solutions. Uses Azure, Databricks, Spark, Python, SQL, and distributed data technologies to support analytics and business intelligence. Ensures data quality, governance, security, performance, and reliability while collaborating with analysts, data scientists, application teams, and stakeholders. Supports DataOps, CI/CD, monitoring, troubleshooting, and automated deployment of enterprise data platforms.
Top Skills: Amazon KinesisApache AirflowApache KafkaSparkAuto LoaderAzureAzure Api ManagementAzure Data FactoryAzure Data Lake StorageAzure DatabricksAzure DevopsAzure Event HubsAzure FunctionsAzure Synapse AnalyticsChange Data CaptureCi/CdDatabricks LakehouseDataopsDelta LakeDelta Live TablesDelta TablesDockerHadoopInfrastructure As CodeJavaKubernetesMicroservicesMySQLOpenshiftOraclePl/SqlPysparkPythonRestful ApisScalaSQLSQL Server
YesterdaySaved
Remote
MEX
Expert/Leader
Expert/Leader
Information Technology • Logistics • Professional Services • Consulting
Leads the design and modernization of enterprise data platforms, pipelines, and lakehouse architectures supporting credit risk, decisioning, analytics, and AI initiatives. Architects scalable ETL/ELT and streaming solutions, establishes data quality, governance, lineage, security, and observability standards, and optimizes distributed workloads. Partners with cross-functional stakeholders, drives technical roadmaps and architectural decisions, mentors engineers, supports compliance and audits, and leads complex initiatives across multiple teams.
Top Skills: AirflowSparkAvroAWSAzureCi/CdDatabricksDelta LakeDevOpsGCPGitHadoopInfrastructure As CodeJavaKafkaOrcParquetPysparkPythonSQL
YesterdaySaved
In-Office or Remote
2 Locations
Mid level
Mid level
Insurance
Build and maintain scalable batch and streaming ETL/ELT pipelines using Python, SQL, PySpark, Databricks, and Delta Lake. Design medallion data models, ingest diverse sources, optimize Spark and SQL performance, implement data quality monitoring, orchestrate workflows, and apply testing, version control, and CI/CD practices. Partner with analysts, data scientists, and stakeholders to create reliable datasets, while documenting lineage and design decisions.
Top Skills: Apache AirflowAuto LoaderAWSAzureCi/CdDatabricksDatabricks SqlDatabricks WorkflowsDbtDelta LakeDelta Live TablesEvent HubsGCPGitKafkaLakeflow Declarative PipelinesPysparkPythonSQLStructured StreamingTerraformUnity Catalog
YesterdaySaved
Remote
4 Locations
Entry level
Entry level
Information Technology • Software • Consulting
Migrate insurance policy data from legacy administration systems to a modern platform. Analyze, trace, reconcile, map, transform, cleanse, and validate data across databases and files. Build SQL queries and ETL processes, resolve data quality issues, and support production data loads. Collaborate with technical and non-technical stakeholders throughout the migration lifecycle while using AI coding tools, Agile methods, and secure data-handling practices.
Top Skills: Amazon S3Aws RdsAws Secrets ManagerClaudeETLGitGithub CopilotJIRAKanbanExcelScrumSQLSQL Server
Reposted 7 Days AgoSaved
Remote
MX
Mid level
Mid level
Cloud • Information Technology • Software • Consulting
Support migration of archived Concur data to AWS S3 using AWS Glue Catalog and Amazon Athena. Upload and organize data in S3, create Glue Catalog tables from DDL, configure Athena, develop and run SQL queries for data retrieval and validation, interpret relational schemas, join and validate datasets, and perform basic file handling, extraction, and structured data analysis.
Top Skills: Amazon AthenaAmazon S3Aws Glue CatalogSQL
7 Days AgoSaved
Remote
México
Senior level
Senior level
Artificial Intelligence • Cloud • Information Technology • Machine Learning • Natural Language Processing • Consulting
Lead the development, optimization, and scaling of production-grade data ingestion and transformation pipelines. Design bulk onboarding, backfills, migrations, and event-driven data processes with monitoring, idempotency, rollback plans, data quality controls, reconciliation, anomaly detection, and auditability. Collaborate with Platform and SRE teams to improve production safety and observability, while producing runbooks, cutover checklists, validation reports, and post-mortem follow-ups.
Top Skills: AWSIamKafkaPostgresSnowflakeSQL
8 Days AgoSaved
Remote
2 Locations
Mid level
Mid level
Information Technology
Designs and develops scalable data pipelines, automated ingestion processes, and data warehousing solutions. Builds dimensional and physical data models, transforms raw data using SQL and Python, and implements data quality validation and monitoring. Collaborates with analysts, data scientists, and business stakeholders to translate requirements into technical specifications. Applies data governance, access controls, security measures, and regulatory compliance practices.
Top Skills: BigQueryDatabricksDbtPythonRedshiftSnowflakeSQL
9 Days AgoSaved
Remote
México
Senior level
Senior level
Professional Services • Software
Designs and operates production data platforms supporting warehouses, streaming and batch pipelines, semantic layers, RAG systems, embeddings, vector stores, and AI feature pipelines. Responsibilities include dimensional modeling, data contracts, lineage, observability, testing, governance, cost optimization, CI/CD, cloud infrastructure, and API delivery. Partners with data scientists and stakeholders, mentors engineers, and ensures reliable, secure, and maintainable data systems.
Top Skills: AirflowAksApache IcebergAWSAzureChromaCi/CdDagsterDbtDelta LakeEksFastapiFeature StoresFlinkGCPGkeGoGreat ExpectationsGrpcHudiJavaKafkaKinesisKubernetesMilvusMlflowOcrPgvectorPineconePrefectPythonRustScalaSodaSpark Structured StreamingSQLTypescriptWeaviateWeights & Biases
Reposted 22 Days AgoSaved
In-Office
Mexico City, Cuauhtémoc, Mexico City, MEX
Senior level
Senior level
News + Entertainment
Own and build batch and real-time data pipelines to support marketing and fandom initiatives. Collaborate with stakeholders to model data, deliver high-quality datasets, enable analytics and ML, source data from APIs, and partner with data scientists and analytics engineers to drive business insights.
Top Skills: BigQueryFlinkGreenplumHiveIcebergKafkaPythonRedshiftSnowflakeSparkSQLTeradataVertica
14 Days AgoSaved
Remote
MEX
Entry level
Entry level
Information Technology • Logistics • Professional Services • Consulting
Leads the design and development of scalable on-premises and cloud-native data platforms. Builds production-grade ETL/ELT pipelines using big data technologies, supports governance and data quality, and contributes to architecture and planning. Mentors engineers, reviews code, solves complex data issues, and promotes maintainable, testable solutions. The role requires Python, Spark, Hadoop, SQL, Databricks, orchestration tools, and DevOps practices, with Java experience preferred.
Top Skills: Apache AirflowApache NifiSparkCi/CdDatabricksGitHadoopHiveImpalaJavaPythonSnowflakeSQLTalend
14 Days AgoSaved
Remote
México
Entry level
Entry level
Information Technology • Consulting
Lead the design and implementation of scalable on-premises and cloud data platforms. Build and maintain ETL/ELT pipelines, support data governance and quality, participate in architecture and planning, and mentor engineers through code reviews and technical guidance. The role requires strong Python, Spark, Hadoop, SQL, Databricks, orchestration, DevOps, and CI/CD expertise, along with effective communication in English and Spanish.
Top Skills: Apache AirflowApache NifiSparkCi/CdDatabricksEltETLGitHadoopHiveImpalaPythonSnowflakeSQLTalend
New

Cut your apply time in half.

Use ourAI Assistantto automatically fill your job applications.

Use For Free
Application Tracker Preview
14 Days AgoSaved
Remote
México
Senior level
Senior level
Information Technology • Consulting
Build and maintain large-scale ETL/ELT pipelines using Databricks, Spark, Python, Hadoop, and cloud data platforms. Develop and optimize PySpark data processing jobs, improve pipeline performance and cost, and support CI/CD, testing, deployment automation, and observability. Collaborate with Product Managers to define user-focused data modules and features. Debug complex data issues and communicate effectively with technical and non-technical stakeholders.
Top Skills: Apache AirflowApache NifiSparkCi/CdClouderaDatabricksGitHadoopHiveImpalaJavaPysparkPythonSnowflakeSQLTalend
14 Days AgoSaved
Remote
2 Locations
Senior level
Senior level
Information Technology • Software
Builds the data and knowledge infrastructure powering AI agents, including production RAG pipelines, embeddings, vector search, reranking, ingestion, ETL/ELT, document processing, and knowledge architecture. Integrates Salesforce, ServiceNow, enterprise repositories, PostgreSQL, vector stores, and Snowflake while ensuring data quality, security, lineage, scalability, and retrieval performance. Develops multi-tenant indexing and feedback systems and collaborates with GenAI engineers in a remote Brazil-based role requiring fluent English and occasional travel.
Top Skills: AirflowBgeCi/CdCohereConfluenceDagsterDbtE5ElasticsearchFull-Text SearchGitGraphQLHaystackJsonbLangchainLlamaindexNeo4JOcrOpenai EmbeddingsOpensearchPandasPdfplumberPgvectorPineconePostgresPrefectPymupdfPythonQdrantRestSalesforceServicenowSharepointSnowflakeSnowflake Cortex AiSQLWeaviate
Reposted 25 Days AgoSaved
In-Office
Mexico City, Cuauhtémoc, Mexico City, MEX
Senior level
Senior level
Cloud • Software
Design, architect, and operate an agentic data engineering ecosystem: build and govern autonomous agents for ETL, synthetic data, QA, and modeling; integrate secure context via MCP servers to Snowflake, Salesforce, and AWS; validate agent outputs and translate ambiguous business requests into production data products.
Top Skills: AgentforceAirflowSparkAWSClaude CodeClaude EnterpriseCodexCursorDbtDockerGitJIRAKubernetesLanggraphModel Context Protocol (Mcp) ServersPythonServerlessSlackSnowflakeSQLTableau
26 Days AgoSaved
In-Office
Mexico City, Cuauhtémoc, Mexico City, MEX
Mid level
Mid level
Cloud • Software
Build and maintain ETL pipelines and recurring reports for content analytics, learner feedback, and business-impact cohorts. Maintain dashboards/headless analytics, run data quality checks, operate AI query tools, and support stakeholders with data validation and ad hoc requests, following established data architecture and standards.
Top Skills: AirflowClaude CodeCursorDbtSQLTableau
Reposted 17 Days AgoSaved
Remote
3 Locations
66K-82K Annually
Mid level
66K-82K Annually
Mid level
Artificial Intelligence • Digital Media • Information Technology • Internet of Things • Software
Build and maintain scalable batch and streaming data pipelines, develop and optimize data models in a cloud warehouse, implement ELT/ETL and orchestration (Airflow/Dagster), ensure data quality and monitoring, handle sensitive data with HIPAA-aligned practices, and evolve the platform's data architecture while partnering with analytics, product, and engineering teams.
Top Skills: AirflowAWSAzureBigQueryDagsterDbtGCPPythonRedshiftSnowflakeSQL
17 Days AgoSaved
Remote
México
Mid level
Mid level
Agency • Marketing Tech
Build and maintain end-to-end data pipelines, models, data marts, and semantic layers supporting agency, client, BI, product, and AI consumers. Integrate fragmented marketing and customer data across advertising platforms, e-commerce, analytics, and CRM systems. Manage dbt projects, multi-tenant Snowflake infrastructure, data quality, performance, and client-specific modeling. Use AI coding agents, cloud infrastructure, CI/CD, and automated testing while collaborating across product, engineering, analytics, and client teams.
Top Skills: Amazon Ads ApiCi/CdClaude CodeCursorDbtGa4GCPGitGithub CopilotGoogle Ads ApiKlaviyoLinkedin Ads ApiMeta Ads ApiMicrosoft Ads ApiPythonShopifySnowflakeSQLTiktok Ads Api
Reposted 19 Days AgoSaved
Remote
México
Senior level
Senior level
Artificial Intelligence • HR Tech • Professional Services • Consulting
Design, build, and scale data pipelines and analytics solutions across the full data lifecycle. Implement ETL, data modeling, and high-volume data systems using Spark/Scala, Python, Databricks, Airflow, and AWS. Collaborate with stakeholders, contribute to architecture and CI/CD practices, and deliver projects on schedule.
Top Skills: AirflowApache ParquetApache Spark (Scala)AWSDatabricksGithub ActionsMySQLPython
Reposted 19 Days AgoSaved
Remote
México
Junior
Junior
Artificial Intelligence • HR Tech • Professional Services • Consulting
Build, monitor, and troubleshoot data pipelines and transformations for healthcare data. Extract, cleanse, and load data; implement data quality checks; document datasets; support API and EFT integrations; triage incidents; and participate in Agile ceremonies to estimate and deliver data solutions.
Top Skills: APIsEdi X12EftEncryptionExcelKafkaMicrostrategyPower BISQLTableau
Reposted 19 Days AgoSaved
Remote
México
Mid level
Mid level
Information Technology • Software • Analytics
Design, build, and operationalize scalable data solutions and pipelines for customers. Analyze and transform data, ensure data quality/governance/security, test data movements, troubleshoot incidents, and collaborate with teams to support AI/ML and BI use cases.
Top Skills: AWSAzureDatabricksDssGCPHadoopHpoKafkaOlapOltpPandasScikit-LearnSparkSQL
Reposted 8 Days AgoSaved
In-Office or Remote
Mexico City, Cuauhtémoc, Mexico City, MEX
Senior level
Senior level
Software
This role involves finding audio data sources, managing cloud infrastructure for data ingestion, and collaborating on dataset roadmaps for AI models.
Top Skills: BashDockerGCPPythonTerraform
YesterdaySaved
Remote
México
110K-110K Annually
Expert/Leader
110K-110K Annually
Expert/Leader
Information Technology • Professional Services • Software • Cybersecurity
Lead scalable data engineering initiatives across on-premises and cloud platforms. Build production-grade ETL/ELT pipelines and distributed data solutions using Python, Spark, SQL, Databricks, and Hadoop technologies. Provide technical leadership through architecture decisions, code reviews, mentoring, testing, and engineering standards. Support data quality, governance, lineage, cataloging, access management, troubleshooting, optimization, and cross-functional planning.
Top Skills: Apache AirflowApache NifiSparkCi/CdDatabricksEltETLGitHadoopHiveImpalaJavaPysparkPythonSnowflakeSQLTalend
YesterdaySaved
Remote
México
90K-90K Annually
Senior level
90K-90K Annually
Senior level
Information Technology • Professional Services • Software • Cybersecurity
Designs, develops, maintains, and troubleshoots scalable Talend ETL/ELT pipelines and production data integration solutions. Responsibilities include data transformation, SQL development, root-cause analysis, code reviews, testing, deployments, observability, and production support. The role collaborates with technical and business stakeholders and requires advanced conversational English and Spanish.
Top Skills: Automated TestingBig DataCi/CdCloud Data PlatformsData IntegrationEtl/EltGitJavaObservabilityRelease PipelinesSQLTalend
YesterdaySaved
Remote
MEX
Senior level
Senior level
Information Technology • Logistics • Professional Services • Consulting
Designs, develops, and maintains scalable batch and real-time data pipelines and curated datasets for enterprise credit risk, lending, and analytics. Builds solutions with Databricks and cloud-native platforms, implements data quality, monitoring, testing, and observability, and supports platform modernization. Partners with technical and business teams, troubleshoots production issues, contributes to architecture, improves reliability and efficiency, and mentors junior engineers.
Top Skills: Apache AirflowSparkAvroAzure Data FactoryAzure Data ServicesCi/CdDatabricksDelta LakeGitGitHadoopJavaKafkaOrcParquetPysparkPythonSQL
All Filters
JobType
New Jobs
Job Category
Experience
Industry
Company Name
Company Size

Sign up now Access later

Create Free Account