Top Data Engineer Jobs in Mexico City

6 Days AgoSaved
In-Office or Remote
2 Locations
Senior level
Senior level
Automotive
Designs, builds, deploys, and optimizes scalable GCP data platforms and pipelines. The role manages ingestion, transformation, governance, security, orchestration, infrastructure as code, automation, monitoring, and cost efficiency. It integrates data systems and services, translates business requirements into reliable data solutions, collaborates with technical stakeholders, documents processes, and promotes engineering best practices.
Top Skills: Apache AirflowAstronomerBigQueryCi/CdCloud FunctionsColumnar DatabasesDataflowDataformDataprocDbtFinopsGoogle Cloud Platform (Gcp)Infrastructure As Code (Iac)JavaMicroservicesMySQLNosql DatabasesPostgresPub/SubPythonService-Oriented Architecture (Soa)SQLTektonTerraform
An Hour AgoSaved
Remote
MX
Entry level
Entry level
Software • Quantum Computing • Metaverse • Infrastructure as a Service (IaaS)
Design, develop, deploy, and operate scalable cloud-based data, analytics, automation, and tooling solutions. Integrate Azure AI and machine learning capabilities, implement monitoring and telemetry, analyze usage data, and collaborate with cross-functional teams to deliver customer-focused systems. Own solutions across design, implementation, and operations while contributing to Microsoft’s data platform innovation.
Top Skills: AgileAi/Ml ModelsAWSAzure Cognitive ServicesAzure Machine LearningAzure OpenaiCC#C++DevOpsGCPGitJavaJavaScriptAzurePython
3 Hours AgoSaved
Remote
México
Entry level
Entry level
Artificial Intelligence • Information Technology • Machine Learning
Build and optimize batch and streaming data pipelines, semantic models, and transformations on Databricks using Python and Spark. Define data contracts, quality rules, and data SLOs; guide domain teams on data modeling and product boundaries; mentor engineers and analysts; and improve reusable platform templates, tooling, and engineering practices.
Top Skills: SparkAWSAzureDatabricksDbtFeature StoreGCPLookerMcpMosaic Ai Agent FrameworkPythonTableau
Reposted 3 Hours AgoSaved
Remote
México
Senior level
Senior level
Artificial Intelligence • Cloud • Information Technology • Machine Learning • Natural Language Processing • Consulting
Build and maintain scalable data pipelines, ETL/ELT processes, data models, metric definitions, and data infrastructure. Develop production solutions with Python, SQL, Spark, dbt, Fivetran, Airflow, Databricks, Azure, and cloud storage. Collaborate with data scientists, engineers, analysts, and stakeholders to deliver reliable, accessible data and self-service tools. Troubleshoot pipeline issues, improve data quality and scalability, and use AI-assisted development tools to accelerate engineering work.
Top Skills: Amazon S3Apache AirflowSparkClaude CodeCodexData LakesData ModelingData PipelinesDatabricksDbtEltETLFivetranAzurePythonSQL
Reposted 22 Hours AgoSaved
Remote
MX
Mid level
Mid level
Cloud • Information Technology • Software • Consulting
Support migration of archived Concur data to AWS S3 using AWS Glue Catalog and Amazon Athena. Upload and organize data in S3, create Glue Catalog tables from DDL, configure Athena, develop and run SQL queries for data retrieval and validation, interpret relational schemas, join and validate datasets, and perform basic file handling, extraction, and structured data analysis.
Top Skills: Amazon AthenaAmazon S3Aws Glue CatalogSQL
14 Days AgoSaved
In-Office or Remote
14 Locations
Senior level
Senior level
Artificial Intelligence • Information Technology • Machine Learning • Software
Lead enterprise data deployments across Latin America, integrating customer data into ClickHouse or warehouse-native Snowflake and BigQuery environments. Build orchestration pipelines, canonical data models, semantic layers, and AI-assisted discovery workflows. Partner directly with customer data teams, resolve complex integration challenges, optimize performance, and establish the region’s deployment playbook. Professional English and Spanish fluency are required, with regular customer travel and potential future relocation within Latin America.
Top Skills: AirbyteBigQueryClickhouseCubeDagsterDbtMetricqlPythonSnowflakeSnowparkSQL
7 Days AgoSaved
Remote
MX
30-30 Hourly
Senior level
30-30 Hourly
Senior level
Software
Build and operate the people-data platform supporting analytics and AI use cases. Design scalable data models, production pipelines, integrations, quality controls, lineage, governance, privacy, and access controls. Monitor reliability, manage incidents, document technical assets, and define infrastructure improvements. Own the people data lake’s AI-readiness roadmap and collaborate with the People Analytics Product Owner, BI Developer, enterprise data, HR technology, security, legal, and privacy teams.
Top Skills: APIsAutomated TestingAzure Blob StorageAzure CloudAzure Data FactoryAzure Data LakeAzure SynapseDatabricksEtl/EltPythonSQLVersion ControlWorkday
8 Days AgoSaved
Remote
6 Locations
Senior level
Senior level
Cloud • Information Technology • Software • Consulting • Web3
Build and operate production data infrastructure, including batch and streaming pipelines, warehouses and lakehouses, retrieval systems, and data-quality controls. Design for schema evolution, reliability, cost, security, compliance, and observability using cloud platforms and modern orchestration tools. Partner directly with clients, explain technical trade-offs, and deploy containerized systems with CI/CD and infrastructure as code. Support AI applications through embedding, indexing, and retrieval pipelines.
Top Skills: AirflowAmazon EmrAWSAws GlueAzureAzure Ai SearchAzure Sql DatabaseBicepBigQueryCi/CdClaude CodeCursorDagsterDatabricksDbtDockerFaissFlinkGitGithub ActionsGithub CopilotGoogle ColabJupyterKafkaPgvectorPineconePrefectPythonQdrantRedshiftSnowflakeSparkSQLTerraform
8 Days AgoSaved
Remote
México
Mid level
Mid level
Artificial Intelligence • Cloud • Information Technology • Machine Learning • Natural Language Processing • Consulting
Build and enhance data engineering frameworks, scalable data solutions, centralized ETL logic, metric definitions, and concise data models. Develop self-service data tools supporting data science, internal applications, and organizational decision-making. Collaborate with cross-functional stakeholders, use AI coding agents such as Claude Code or Codex, and write production-level Python and SQL.
Top Skills: Amazon S3Apache AirflowSparkClaude CodeCodexDbtFivetranPythonSQL
13 Days AgoSaved
Remote
6 Locations
Entry level
Entry level
Information Technology • Software • Consulting
Design and deliver cloud-native data platforms, scalable big data pipelines, lakehouses, warehouses, transformation frameworks, BI solutions, and ML-enabled data workflows. Lead client-facing pre-sales engagements, technical workshops, proof-of-concepts, executive conversations, and enterprise data modernization initiatives. Build reusable architectural IP, influence senior stakeholders, and translate ambiguous business challenges into production-ready solutions using AWS and modern data technologies.
Top Skills: Amazon AthenaAmazon BedrockAmazon EmrAmazon KinesisAmazon QuicksightAmazon RedshiftApache AirflowApache HadoopApache HudiApache IcebergApache KafkaAWSAws GlueAws Lake FormationCi/CdCollibraDatabricksDbtDelta LakeGitGlue CatalogHdfsHiveLookerMlflowPower BIPysparkPythonSagemakerScalaSnowflakeSparkSQLTableauUnity CatalogWeights & BiasesYarn
15 Days AgoSaved
Remote
México
Entry level
Entry level
Information Technology • Logistics • Professional Services • Consulting
Designs and maintains scalable Azure-based data pipelines and infrastructure for batch, real-time, analytics, and reporting workloads. Builds Databricks Spark applications, Lakehouse architectures, Delta Tables, APIs, and streaming ingestion frameworks. Responsibilities include data governance, security, compliance, monitoring, troubleshooting, performance optimization, DataOps, and CI/CD automation. Collaborates with data scientists, analysts, application teams, and business stakeholders to deliver reliable data solutions.
Top Skills: Amazon KinesisApache AirflowApache KafkaSparkAuto LoaderAzureAzure Api ManagementAzure Data FactoryAzure Data Lake StorageAzure DatabricksAzure DevopsAzure Event HubsAzure FunctionsAzure Synapse AnalyticsChange Data CaptureCi/CdDatabricks LakehouseDataopsDelta LakeDelta Live TablesDelta TablesDockerHadoopInfrastructure As CodeJavaKubernetesMicroservicesMySQLOpenshiftOraclePl/SqlPysparkPythonRest ApisScalaSQLSQL Server
21 Days AgoSaved
Remote
México
Mid level
Mid level
Information Technology • Software • Business Intelligence
Designs, builds, deploys, and optimizes cloud-based data pipelines on GCP. Integrates data from databases, APIs, and streaming platforms using ETL and ELT processes; manages BigQuery and NoSQL databases; performs data modeling and query optimization; and uses Airflow and modern transformation tools. Collaborates with data scientists, analysts, clients, and business stakeholders in distributed multinational teams.
Top Skills: AlteryxApache AirflowSparkBigQueryConfluenceDatabricksDataformDbtEltETLGitGoogle Cloud Platform (Gcp)JIRALookerAzureNoSQLPower BIPythonSQLTableauTalend
New

Cut your apply time in half.

Use ourAI Assistantto automatically fill your job applications.

Use For Free
Application Tracker Preview
21 Days AgoSaved
Remote
MEX
Entry level
Entry level
Information Technology • Logistics • Professional Services • Consulting
Designs and maintains scalable cloud-native ETL/ELT pipelines, data models, Lakehouse architectures, APIs, and real-time streaming solutions. Uses Azure, Databricks, Spark, Python, SQL, and distributed data technologies to support analytics and business intelligence. Ensures data quality, governance, security, performance, and reliability while collaborating with analysts, data scientists, application teams, and stakeholders. Supports DataOps, CI/CD, monitoring, troubleshooting, and automated deployment of enterprise data platforms.
Top Skills: Amazon KinesisApache AirflowApache KafkaSparkAuto LoaderAzureAzure Api ManagementAzure Data FactoryAzure Data Lake StorageAzure DatabricksAzure DevopsAzure Event HubsAzure FunctionsAzure Synapse AnalyticsChange Data CaptureCi/CdDatabricks LakehouseDataopsDelta LakeDelta Live TablesDelta TablesDockerHadoopInfrastructure As CodeJavaKubernetesMicroservicesMySQLOpenshiftOraclePl/SqlPysparkPythonRestful ApisScalaSQLSQL Server
21 Days AgoSaved
Remote
MEX
Expert/Leader
Expert/Leader
Information Technology • Logistics • Professional Services • Consulting
Leads the design and modernization of enterprise data platforms, pipelines, and lakehouse architectures supporting credit risk, decisioning, analytics, and AI initiatives. Architects scalable ETL/ELT and streaming solutions, establishes data quality, governance, lineage, security, and observability standards, and optimizes distributed workloads. Partners with cross-functional stakeholders, drives technical roadmaps and architectural decisions, mentors engineers, supports compliance and audits, and leads complex initiatives across multiple teams.
Top Skills: AirflowSparkAvroAWSAzureCi/CdDatabricksDelta LakeDevOpsGCPGitHadoopInfrastructure As CodeJavaKafkaOrcParquetPysparkPythonSQL
Reposted 8 Days AgoSaved
In-Office
2 Locations
Senior level
Senior level
Food
Design and implement scalable batch and streaming data pipelines using Spark and Scala on Microsoft Fabric/Synapse. Build JVM REST microservices, optimize Spark performance and data models, enforce engineering rigor (functional programming, SOLID), and maintain CI/CD pipelines with GitHub Actions, Maven/SBT. Collaborate in Agile teams to deliver reliable, high-performance data solutions.
Top Skills: SparkArm TemplatesAWSAzureAzure Data Factory (Adf)Azure FunctionsAzure Synapse AnalyticsDelta LakeFunctional ProgrammingGCPGitGithub ActionsHadoopJavaJvmMavenMicrosoft FabricRest ApisSbtScala
21 Days AgoSaved
In-Office or Remote
2 Locations
Mid level
Mid level
Insurance
Build and maintain scalable batch and streaming ETL/ELT pipelines using Python, SQL, PySpark, Databricks, and Delta Lake. Design medallion data models, ingest diverse sources, optimize Spark and SQL performance, implement data quality monitoring, orchestrate workflows, and apply testing, version control, and CI/CD practices. Partner with analysts, data scientists, and stakeholders to create reliable datasets, while documenting lineage and design decisions.
Top Skills: Apache AirflowAuto LoaderAWSAzureCi/CdDatabricksDatabricks SqlDatabricks WorkflowsDbtDelta LakeDelta Live TablesEvent HubsGCPGitKafkaLakeflow Declarative PipelinesPysparkPythonSQLStructured StreamingTerraformUnity Catalog
4 Hours AgoSaved
Remote
MX
Senior level
Senior level
Artificial Intelligence • Financial Services
Designs and operates scalable data architectures, pipelines, ETL processes, databases, warehouses, and lakes. Improves data quality, governance, observability, security, and reliability while partnering with stakeholders to develop effective data solutions. Provides technical leadership, mentors engineers, leads code reviews and cross-functional initiatives, resolves production issues, and drives improvements to engineering platforms and practices.
Top Skills: AirflowAWSAzureAzure Sql Data WarehouseDatabricksGCPHadoopJavaJenkinsKafkaNosql DatabasesPysparkPythonRedshiftRelational DatabasesS3ScalaSparkSQL
Reposted YesterdaySaved
Remote
México
Junior
Junior
Software
Build and maintain data models, dashboards, reports, and dbt models supporting cross-functional decision-making. Process large datasets, establish reporting sources of truth, bridge technical and business teams, and document knowledge bases and business definitions for internal AI systems. The role also involves identifying cost-saving opportunities, communicating with stakeholders, and using AI to develop solutions and fix bugs.
Top Skills: AIDbtPower BISQLTableau
YesterdaySaved
Remote
México
Senior level
Senior level
Artificial Intelligence • Cloud • Information Technology • Machine Learning • Natural Language Processing • Consulting
Design, build, operate, and scale change data capture pipelines using Debezium, Kafka or Azure Event Hubs, and Databricks. Manage streaming data across hundreds of databases and thousands of tables, ensuring reliability, low latency, schema consistency, and data quality. Troubleshoot replication lag, schema drift, backpressure, and recovery issues. Collaborate with backend and data teams on schemas and SLAs while documenting architecture and operational procedures.
Top Skills: Azure Data FactoryAzure Event HubsAzure SqlCi/CdDatabricksDebeziumInfrastructure As CodeKafkaLakeflow CdcSQL
13 Days AgoSaved
Remote or Hybrid
Mexico City, Cuauhtémoc, Mexico City, MEX
115K-150K Annually
Senior level
115K-150K Annually
Senior level
Healthtech • Biotech
Design, build, deploy, and operate scalable full-stack, cloud-native software and infrastructure for clinicogenomic data products. Responsibilities include distributed systems, Kubernetes, serverless services, event-driven architecture, CI/CD, infrastructure as code, technical direction, engineering standards, AI-assisted development, cross-functional collaboration, and mentoring.
Top Skills: AWSAws Api GatewayAws CdkAws LambdaAzureCi/CdGCPGoInfrastructure As CodeJavaJavaScriptKubernetesPulumiPythonTerraformTypescript
5 Days AgoSaved
Remote
México
Senior level
Senior level
Information Technology • Consulting
Leads the design, development, deployment, and maintenance of data pipelines and analytics platforms using Microsoft Fabric and Azure. Builds Lakehouse and Warehouse solutions, develops transformations with SQL, Python, and PySpark, and ensures data quality, security, performance, and governance. Collaborates with architects, developers, analysts, and stakeholders in Agile teams, contributes to CI/CD and documentation, leads technical discussions, mentors junior engineers, and provides technical guidance across client engagements.
Top Skills: SparkAzure Data FactoryAzure DevopsAzure Sql DatabaseCi/CdData WarehouseDatabricksEtl/EltGitInfrastructure As CodeLakehouseAzureMicrosoft FabricPysparkPythonSQLSQL ServerSynapse AnalyticsT-Sql
29 Days AgoSaved
Remote
México
Senior level
Senior level
Professional Services • Software
Designs and operates production data platforms supporting warehouses, streaming and batch pipelines, semantic layers, RAG systems, embeddings, vector stores, and AI feature pipelines. Responsibilities include dimensional modeling, data contracts, lineage, observability, testing, governance, cost optimization, CI/CD, cloud infrastructure, and API delivery. Partners with data scientists and stakeholders, mentors engineers, and ensures reliable, secure, and maintainable data systems.
Top Skills: AirflowAksApache IcebergAWSAzureChromaCi/CdDagsterDbtDelta LakeEksFastapiFeature StoresFlinkGCPGkeGoGreat ExpectationsGrpcHudiJavaKafkaKinesisKubernetesMilvusMlflowOcrPgvectorPineconePrefectPythonRustScalaSodaSpark Structured StreamingSQLTypescriptWeaviateWeights & Biases
7 Days AgoSaved
In-Office or Remote
5 Locations
Senior level
Senior level
Information Technology
Designs, builds, and maintains large-scale data pipelines using Databricks, PySpark, and Apache Spark. Develops ETL/ELT processes, lakehouse architectures, and optimized data workflows while ensuring quality, reliability, performance, and observability. Collaborates with architects, analysts, and stakeholders to deliver scalable data solutions. Uses orchestration, version control, CI/CD, and automation practices to support modern analytics platforms.
Top Skills: Apache AirflowSparkAWSAzureAzure Data FactoryCi/CdDatabricksDelta LakeGitJavaMicrosoft FabricPysparkPythonSQL
8 Days AgoSaved
In-Office or Remote
14 Locations
Senior level
Senior level
Other
Designs, builds, and maintains scalable data pipelines, data warehouse architecture, feature stores, model-training workflows, and real-time inference services. Develops semantic search, vector database, RAG, LLM orchestration, dashboarding, and AI-powered BI solutions. Partners with stakeholders, documents workflows, supports model deployment, and applies data engineering and MLOps best practices using Python, Spark, Hadoop, Kafka, SQL, and cloud databases.
Top Skills: Amazon RedshiftSparkCi/CdDatastax AstradbGitHadoopHugging Face TransformersKafkaLangchainLinuxLlama-4LlamaindexMlopsPl/SqlPythonRagSnowflakeSQLUnixVector DatabasesWindows
Reposted One Month AgoSaved
In-Office
Mexico City, Cuauhtémoc, Mexico City, MEX
Senior level
Senior level
News + Entertainment
Own and build batch and real-time data pipelines to support marketing and fandom initiatives. Collaborate with stakeholders to model data, deliver high-quality datasets, enable analytics and ML, source data from APIs, and partner with data scientists and analytics engineers to drive business insights.
Top Skills: BigQueryFlinkGreenplumHiveIcebergKafkaPythonRedshiftSnowflakeSparkSQLTeradataVertica
All Filters
JobType
New Jobs
Job Category
Experience
Industry
Company Name
Company Size

Sign up now Access later

Create Free Account