Top Data Engineer Jobs in Sao Paulo

Reposted YesterdaySaved
In-Office
São Bernardo do Campo, São Paulo, BRA
Mid level
Mid level
Software
Build and operate batch and streaming pipelines on Databricks using PySpark and Delta Lake. Modernize legacy data warehouses into governed Lakehouse architectures, optimize Spark jobs and Delta tables, implement data quality and governance controls, and troubleshoot production incidents. Collaborate with DevOps, platform, and analytics engineers on observability, security, and compliance while delivering well-tested Python and SQL solutions.
Top Skills: AirflowSparkAuto LoaderAzure Data FactoryAzure DevopsClaudeCloudwatchDatabricksDatabricks WorkflowsDbtDelta LakeDynatraceEvent HubsGithub CopilotKafkaPostgresPysparkPythonSQLStructured StreamingTerraformUnity Catalog
Reposted 2 Days AgoSaved
In-Office
São Paulo, BRA
Mid level
Mid level
Software
Build and operate batch and streaming pipelines on Databricks using PySpark and Delta Lake. Modernize legacy warehouses into governed Lakehouse architectures, optimize Spark jobs and Delta tables, implement data quality and governance controls, and troubleshoot production incidents. Collaborate with DevOps, platform, and analytics teams on security, observability, and compliance while writing tested Python and SQL. Use AI development tools to accelerate coding, testing, documentation, and prototyping.
Top Skills: AirflowSparkAuto LoaderAzure Data FactoryAzure DevopsClaudeCloudwatchDatabricksDatabricks WorkflowsDbtDelta LakeDynatraceEvent HubsGithub CopilotKafkaPostgresPysparkPythonSQLStructured StreamingTerraformUnity Catalog
Reposted 3 Days AgoSaved
In-Office
São Bernardo do Campo, São Paulo, BRA
Mid level
Mid level
Software
Designs, maintains, and supports ETL pipelines and data workflows across development, staging, and production environments. Builds data models, datasets, validation queries, DDL, and deployment scripts for institutional analytics. Customizes reusable pipelines for clients, monitors production jobs, troubleshoots failures, supports modernization initiatives, and collaborates with analysts, client teams, and stakeholders to translate requirements into data solutions.
Top Skills: Amazon RedshiftApache AirflowBigQueryBitbucketDockerDocker ComposeGitGitGitlabPower BIPythonSnowflakeSparkSQLSynapseTableauTableau Server Rest Api
Reposted An Hour AgoSaved
Remote
11 Locations
65K-90K Annually
Mid level
65K-90K Annually
Mid level
Analytics
Own end-to-end AI-automated data platform migration projects (1-4 concurrent): scope, plan, execute, configure Datafold Migration Agent, run customer check-ins, manage stakeholders, partner with engineering, and improve delivery playbooks.
Top Skills: AIDatabricksDatafold Migration AgentDbtETLIncremental ProcessingOrchestration ToolsSnowflakeStored ProceduresStreaming
YesterdaySaved
Remote
2 Locations
Entry level
Entry level
Information Technology • Consulting
Owns end-to-end data engineering for ingestion, medallion architecture, curated datasets, and Power BI semantic layers. Builds scalable ETL/ELT pipelines, integrates ERP and CRM systems, implements data quality and governance controls, optimizes performance and cloud costs, and provides technical leadership. Requires Microsoft Fabric or Azure expertise, Azure Data Factory, advanced Power BI and DAX skills, data contract design, and experience with Bronze/Silver/Gold architectures.
Top Skills: AzureAzure Data FactoryBicepCi/CdCRMData LakesData WarehousesDaxErpFusion PlmInfor IonInfor SytelineMicrosoft FabricOnelakeOnestreamPower BITerraformWorkday
Entry level
Artificial Intelligence • Machine Learning • Software • Financial Services
Designs, builds, and maintains production data pipelines and infrastructure supporting analytics, machine learning, and financial decision-making. Responsibilities include data ingestion, modeling, quality validation, workflow orchestration, database optimization, monitoring, troubleshooting, and managing scalable AWS or on-premises systems. The role collaborates with data scientists and analysts to ensure reliable access to large-scale and time-series financial data.
Top Skills: AWSJavaMongoDBMySQLPostgresPythonSQL
8 Days AgoSaved
Remote
2 Locations
Mid level
Mid level
Information Technology
Designs and develops scalable data pipelines, automated ingestion processes, and data warehousing solutions. Builds dimensional and physical data models, transforms raw data using SQL and Python, and implements data quality validation and monitoring. Collaborates with analysts, data scientists, and business stakeholders to translate requirements into technical specifications. Applies data governance, access controls, security measures, and regulatory compliance practices.
Top Skills: BigQueryDatabricksDbtPythonRedshiftSnowflakeSQL
9 Days AgoSaved
Remote
Brazil
Mid level
Mid level
Information Technology • Consulting
Support Data & AI Governance initiatives through data analysis, curation, stewardship, documentation, cataloging, lineage, quality improvement, and policy maintenance. Use Python for analysis and automation while collaborating with technical and business stakeholders to improve data organization, accessibility, usability, and compliance with governance standards.
Top Skills: Python
9 Days AgoSaved
Remote
Brazil
Mid level
Mid level
Information Technology • Consulting
Design, develop, and maintain scalable cloud data pipelines and infrastructure for analytics and machine learning solutions. Collaborate with cross-functional teams to ensure data integrity, usability, and quality. Build CI/CD pipelines for ETL processes, document workflows and architectures, and provide technical troubleshooting. The role requires Databricks, Spark, Python, SQL, dimensional modeling, and Data Lake or Data Warehouse experience, with Azure, AWS, Power BI, and Unity Catalog as beneficial skills.
Top Skills: SparkAWSAzureCi/CdData LakeData WarehouseDatabricksETLPower BIPythonSQLUnity Catalog
Reposted 18 Days AgoSaved
In-Office
São Paulo, BRA
Mid level
Mid level
Artificial Intelligence • Consulting
Design, build and own end-to-end data solutions and infrastructure. Implement scalable pipelines using cloud and big-data technologies, optimize performance, create REST APIs and CI/CD, coach teammates, and communicate technical solutions to business stakeholders.
Top Skills: AnsibleAWSBeamBigQueryCompute EngineDataflowDockerFlinkGCPHdfsJavaKubernetesLinuxMachine LearningNifiPythonReactRedisReduxRest ApiScalaSpark
Reposted 19 Days AgoSaved
Remote
12 Locations
Senior level
Senior level
Information Technology • Software
We're seeking a Data Engineer with 5+ years of experience to build data processing solutions using Python, design data models, and process large data streams while ensuring high performance across various data technologies.
Top Skills: AirflowAWSAzureCassandraFlinkFlumeGCPHadoopHbaseHiveInformaticaKafkaKinesisLookerMongoDBPandasPower BIPysparkPythonRabbitMQRedshiftS3SnowflakeSnsSparkSql Server Integration ServicesSqsTableauTalend
21 Days AgoSaved
In-Office
São Paulo, BRA
Junior
Junior
Greentech
Build and operate reliable data infrastructure for scientific and operational use. Responsibilities include developing ingestion pipelines, orchestrated workflows, warehouse data models, integrations, testing, monitoring, alerting, documentation, and data quality investigations. The role collaborates with scientists and engineers to productionize datasets for geospatial and statistical analysis and improve data lineage, permissions, schemas, and self-service access.
Top Skills: AirflowAWSContainerizationContinuous Integration And DeploymentDagsterData WarehousesDbtGCPGdalGeopandasInfrastructure As CodePostgisPrefectPythonRasterioRelational DatabasesSQLStacZarr
New

Cut your apply time in half.

Use ourAI Assistantto automatically fill your job applications.

Use For Free
Application Tracker Preview
13 Days AgoSaved
Remote
São Paulo, BRA
Mid level
Mid level
Information Technology • Software
Build and maintain data models, ETL/ELT pipelines, and reliable streaming integrations for high-volume telecom and customer data. Develop transformations in dbt, orchestrate workflows with Airflow, write optimized SQL, and implement data quality, monitoring, documentation, and reconciliation processes. The role partners with the Central Data team to deliver accurate, timely, traceable datasets for reporting and self-service analytics.
Top Skills: Amazon EmrAmazon RedshiftApache AirflowApache HiveApache IcebergApache KafkaSparkAws AthenaAws GlueAws S3BigQueryDbtExasolPythonSnowflakeSQL
YesterdaySaved
In-Office
São Bernardo do Campo, São Paulo, BRA
Senior level
Senior level
Software
Lead the replacement of TapClicks with a custom ETL pipeline in AWS. Document legacy ingestion logic, evaluate Fivetran versus custom solutions, ingest Amazon DSP and other digital advertising data, and build a standalone system supporting client dashboards. Collaborate with data science teams to align outputs with attribution and scoring models, while also contributing to application and backend development.
Top Skills: Amazon Advertising ApiAmazon EcsAmazon RdsAmazon S3AWSDsp PlatformsETLFivetranMlopsPostgresSQLTapclicks
YesterdaySaved
In-Office
São Paulo, BRA
Senior level
Senior level
Software
Lead the replacement of TapClicks with a standalone ETL pipeline in AWS. Document legacy ingestion logic, evaluate Fivetran versus custom development, and ingest data from Amazon DSP and other programmatic advertising platforms. Design schemas and transformations, support backend application work, and coordinate with data science teams to align pipeline outputs with attribution, scoring, and dashboard requirements.
Top Skills: Amazon Advertising ApiAmazon DspAWSEcsEltETLFivetranMlopsPostgresRdsS3SQL
14 Days AgoSaved
Remote
Brazil
Senior level
Senior level
Information Technology • Software
Builds the data and knowledge infrastructure powering AI agents, including production RAG pipelines, embeddings, vector search, reranking, ingestion, ETL/ELT, document processing, and knowledge architecture. Integrates Salesforce, ServiceNow, enterprise repositories, PostgreSQL, vector stores, and Snowflake while ensuring data quality, security, lineage, scalability, and retrieval performance. Develops multi-tenant indexing and feedback systems and collaborates with GenAI engineers in a remote Brazil-based role requiring fluent English and occasional travel.
Top Skills: AirflowBgeCi/CdCohereConfluenceDagsterDbtE5ElasticsearchFull-Text SearchGitGraphQLHaystackJsonbLangchainLlamaindexNeo4JOcrOpenai EmbeddingsOpensearchPandasPdfplumberPgvectorPineconePostgresPrefectPymupdfPythonQdrantRestSalesforceServicenowSharepointSnowflakeSnowflake Cortex AiSQLWeaviate
Reposted 2 Days AgoSaved
In-Office
São Paulo, BRA
Senior level
Senior level
Software
Design and implement Salesforce Data Cloud solutions, including data ingestion, model mapping, identity resolution, harmonization, calculated insights, segmentation, and activation workflows. Build integrations using APIs, middleware, and ETL patterns; support data quality, governance, security, and compliance. Partner with business and architecture teams, troubleshoot platform issues, document technical designs, and maintain scalable customer data solutions.
Top Skills: ApexCloud Data WarehousesCrm AnalyticsETLLightning ComponentsMarketing CloudMiddlewareRest ApisSalesforceSalesforce Data CloudSoap ApisSOQLSQL
16 Days AgoSaved
Remote
Sítio Israel, Ibiúna, São Paulo, BRA
Mid level
Mid level
Big Data • Security • Cybersecurity
Build and operate scalable batch and streaming data pipelines powering AI systems. Own ingestion, storage, transformation, serving, monitoring, validation, and continuous improvement across the data lifecycle. Partner with AI engineers and researchers to translate model requirements into reliable production infrastructure, using data lakes, warehouses, relational, graph, and vector databases. Define data strategy and ensure high-quality, observable, timely data for AI models and agents.
Top Skills: SparkAWSDatadogDbtDockerDuckdbGoKubernetesNeo4JPgvectorPostgresPythonTemporalTrino
Reposted 17 Days AgoSaved
Remote
3 Locations
66K-82K Annually
Mid level
66K-82K Annually
Mid level
Artificial Intelligence • Digital Media • Information Technology • Internet of Things • Software
Build and maintain scalable batch and streaming data pipelines, develop and optimize data models in a cloud warehouse, implement ELT/ETL and orchestration (Airflow/Dagster), ensure data quality and monitoring, handle sensitive data with HIPAA-aligned practices, and evolve the platform's data architecture while partnering with analytics, product, and engineering teams.
Top Skills: AirflowAWSAzureBigQueryDagsterDbtGCPPythonRedshiftSnowflakeSQL
17 Days AgoSaved
Remote
Brazil
Mid level
Mid level
Agency • Marketing Tech
Build and maintain end-to-end data pipelines, models, data marts, and semantic layers for marketing, client, BI, product, and AI consumers. Integrate fragmented advertising and customer data across multiple platforms and tenants, manage dbt projects, ensure data quality, optimize performance, and support client-specific modeling. Use AI-assisted development workflows, collaborate cross-functionally, and deliver reliable, AI-ready data products.
Top Skills: Amazon AdsCi/CdClaude CodeCursorDbtGa4GCPGitGithub CopilotGoogle Ads ApiInfrastructure As CodeJinjaKlaviyoLinkedin AdsMeta Ads ApiMicrosoft AdvertisingPythonShopifySnowflakeSQLTiktok Ads
6 Days AgoSaved
In-Office
São Paulo, BRA
Senior level
Senior level
Information Technology • Consulting
Build and register domain-grounded data agents over governed Databricks and Snowflake datasets. Lead lakehouse migrations to open formats such as Apache Iceberg, curate catalogue metadata, and support data-contract workflows. Partner with business stakeholders and subject-matter experts to onboard commercial, finance, R&D, and real-world data domains. Ensure data quality, observability, lineage, governed access, masking, and row- and column-level security while enabling reliable AI-agent access to enterprise data.
Top Skills: Amazon S3Apache AirflowApache IcebergAWSCi/CdCollibraDatabricksDatabricks WorkflowsDbtEmbeddingsGitHorizonMcpPysparkPythonSnowflakeSnowflake CortexSnowflake GenieSparkSQLUnity CatalogVector DatabasesYaml
Reposted 19 Days AgoSaved
Remote or Hybrid
11 Locations
Mid level
Mid level
Artificial Intelligence • Big Data • Machine Learning • Analytics • Business Intelligence • Consulting • Generative AI
The Data Engineer role involves developing and maintaining data pipelines, ensuring data quality, and collaborating with teams to deliver solutions tailored for financial services.
Top Skills: AWSAzure SynapseBigQueryDbtEltETLFivetranFivetranInformaticaKafkaLookerMatillionMatillionMySQLOraclePostgresPower BIPyramid AnalyticsRedshiftRiverySigmaSisenseSnowflakeSparkSQL ServerSsisTableauTalendThoughtspot
Reposted 19 Days AgoSaved
Remote
10 Locations
24K-42K Annually
Junior
24K-42K Annually
Junior
Artificial Intelligence • Greentech • Robotics
Own image and video dataset lifecycle and quality, build and maintain data pipelines and versioning, create internal tools for labeling and curation, coordinate with labeling teams, and improve training-data collection using model-assisted and active-learning approaches.
Top Skills: CvatDjangoDvcFastapiLabel StudioNextjsPythonReactS3SQLSupervisely
Reposted 29 Days AgoSaved
In-Office
São Paulo, BRA
Internship
Internship
Artificial Intelligence • Consulting
Work with cross-functional teams to design, build, and deploy end-to-end data solutions. Own conception and implementation of data infrastructure, pipelines, ML model integration, REST APIs, and CI/CD. Collaborate, train, and present findings; stay current on cloud and big-data technologies.
Top Skills: AnsibleApache BeamApache FlinkApache NifiSparkAWSCi/CdDockerGCPGoogle BigqueryGoogle Compute EngineGoogle DataflowHdfsJavaKubernetesPythonReactRedisReduxRest ApisScalaUnix/Linux Cli
9 Days AgoSaved
In-Office
São Paulo, BRA
Senior level
Senior level
Food
Manage eDiscovery data-source insights, preservation operations, records, and metrics. Automate Microsoft Purview workflows using PowerShell and Microsoft Graph, build sync pipelines from Microsoft 365, Google Workspace, mobile devices, and endpoints into RelativityOne, and generate collection and processing reports. The role also supports process engineering, audit readiness, scalable data collection, and high-stakes digital forensics and litigation projects.
Top Skills: GoGoogle WorkspaceMicrosoft 365Microsoft GraphMicrosoft Purview EdiscoveryOauth2PowershellPythonRelativityoneRest ApiTypescript
All Filters
JobType
New Jobs
Job Category
Experience
Industry
Company Name
Company Size

Sign up now Access later

Create Free Account