Maximum of 25 job preferences reached.
Top Remote Data Engineer Jobs in San Antonio, TX
Mobile • Software
Lead design and delivery of scalable production data pipelines and data products (Spark/Python/SQL) on Databricks; integrate diverse federal health systems, define reusable ingestion patterns, ensure data quality/governance, mentor engineers, and troubleshoot/optimize production workflows.
Top Skills:
SparkAPIsCommand-LineDatabricksEltETLGitPythonSQL
Consulting
Design, build, and operate scalable batch and streaming data pipelines and a lakehouse on AWS. Implement ETL/ELT, CDC, metadata/cataloging, lineage, governance, data quality, IaC, and CI/CD for reliable, observable, and secure data platforms to support analytics and ML.
Top Skills:
Amazon S3AWSCdcCi/CdDelta LakeEncryptionEtl/EltIamInfrastructure As CodeJavaLakehousePythonScalaSecrets ManagementSQL
Insurance • Software
Build and maintain scalable ETL/data pipelines and internal frameworks to support ML and product teams. Drive architecture, observability, data quality, and engineering best practices across AI/ML workflows and prompt-engineering pipelines.
Top Skills:
AirflowAWSCi/CdDagsterDbtGCPPrefectPythonSQLTerraform
Edtech
Lead design, build, and operate scalable Databricks/Delta Lake data pipelines and Kimball dimensional models using dbt. Manage governance with Unity Catalog, optimize performance and cost, operationalize ML with MLflow, and coordinate/quality-check offshore vendor engineering. Mentor engineers, uphold standards, and deliver trusted data products that power analytics, reporting, and MLOps workflows.
Top Skills:
Azure Event HubsDatabricksDatabricks Dashboards (Genie)DbtDelta LakeDelta LakehouseDockerKafkaLakeflow ConnectMlflowMlopsPower BIPysparkPythonSalesforceSparkSQLStructured StreamingTableauUnity Catalog
Aerospace • Information Technology • Professional Services • Security • Software
Lead SAP S/4HANA data migration activities: extract, map, transform, load, validate, and reconcile master and transactional data using BODS, Migration Cockpit, LSMW/BDC and HANA SQL. Support mock loads, cutover execution, defect remediation, and collaborate with functional owners to cleanse and correct source data.
Top Skills:
BdcETLFlat-File LoadsLsmwS/4Hana Migration CockpitSap Data Services (Bods)Sap EccSap HanaSap Hana SqlSap S/4HanaSQL
Healthtech • Professional Services • Software
Design, build, and maintain scalable data pipelines, data models, ETL/ELT, and BI solutions. Collaborate with data scientists and stakeholders to improve data quality, accessibility, and PHI security. Develop and document workflows, optimize processing/storage, support data warehouse and visualization, and adopt AI tools to drive data-informed outcomes.
Top Skills:
Ai ToolsBigQueryCdcCloud SqlDatastreamDbtGCPHvrSnowflakeSQLSQL ServerSsisTableau
Food • Healthtech • Telehealth
The Senior Data Engineer will construct and optimize data pipelines, work with cross-functional teams, ensure data quality and governance, and mentor junior engineers while leading advanced analytics initiatives.
Top Skills:
Apache AirflowBigQueryDagsterJavaLuigiMySQLNode.jsPostgresPythonRedshiftScalaSnowflakeSQLTypescript
Information Technology • Database • Consulting
Design, build, and operate scalable data transformation pipelines and dimensional models using dbt, SQL, and AWS. Implement infrastructure-as-code (AWS CDK), CI/CD, and query performance optimizations while ensuring data quality, governance, and collaboration with cross-functional and offshore teams.
Top Skills:
Amazon MwaaAmazon RedshiftAws CdkAws CodecommitAws CodepipelineAws GlueAws LambdaAws Step FunctionsCi/CdDbtJinjaPysparkPythonSQLVersion Control
Information Technology • Database • Consulting
Lead Data Engineer builds and maintains scalable data pipelines and lakehouse/warehouse platforms (Databricks) to support analytics, BI, reporting, and AI. Responsibilities include data modeling, ETL/ELT, medallion architecture, orchestration and monitoring (Airflow/Cron), performance tuning, CDC/incremental loads, data quality, and leading/mentoring a team while collaborating with stakeholders and maintaining documentation.
Top Skills:
Apache AirflowAWSAzureCronData WarehousingDatabricksEltETLGCPHadoopHbaseHiveLakehouseMedallion ArchitecturePigPower BIPysparkPythonSparkSQLTableau
Healthtech
Design and build cloud-based data applications and pipelines to analyze clinical data, enrich and provision datasets, and support clinical/operational processes. Mentor junior engineers, drive R&D for repeatable templates, collaborate with Product/Platform/Architecture, and prioritize source control, documentation, and simple solutions using modern data, ETL, and LLM tooling.
Top Skills:
AdfAirflowBigQueryDbtEhrEmbeddingsEmrFastapiFhirFivetranFlaskFlinkGlueHadoopHl7InformaticaKafkaLangchainLinuxLlamaindexLlmNifiNoSQLOlapPythonRagRedshiftSnowflakeSparkSQLSynapseVector Dbs
Cloud • Information Technology • Other • Productivity • Software
Architect, build, and maintain Snowflake-based ELT pipelines and production data models. Implement observability, data quality, and governance; optimize cost and freshness. Partner with Analytics, Product, and Engineering, use dbt/CI-CD and AI-assisted tooling, and participate in reviews and architecture planning to support analytics, BI, and ML.
Top Skills:
Apache AirflowClaude CodeCursorDbtFivetranGitGithub ActionsGithub CopilotLlmsPrefectPythonSnowflakeSQL
Agency • Information Technology
Design, build, and maintain scalable ETL/ELT pipelines into Snowflake, enforce data quality, optimize data models and query performance for Power BI reporting, produce documentation, and troubleshoot data issues while collaborating with stakeholders.
Top Skills:
AWSAzureEltETLGCPOraclePower BISnowflakeSQL
New
Cut your apply time in half.
Use ourAI Assistantto automatically fill your job applications.
Use For Free
Agency • Information Technology
Design, implement, and maintain PySpark-based data reconciliation solutions for financial systems. Build matching algorithms, integrate with rules engines, process large distributed datasets, identify and resolve discrepancies, and collaborate with analysts and architects to improve data quality and governance.
Top Skills:
SparkDroolsHadoopHbaseHiveKafkaKinesisPyspark
Agency • Information Technology
Design, implement, and maintain data quality processes using cloud platforms (Azure/AWS/GCP). Collaborate with stakeholders to define standards, perform data profiling and cleansing, document issues and remediation, generate quality metrics, and support cross-functional teams to improve data architectures and integration.
Top Skills:
AWSAzureGCPJavaPythonSQL
Agency • Information Technology
Design, implement, and maintain PySpark applications to automate large-scale financial data reconciliations. Build transformation and matching algorithms, integrate with rules engines, analyze data gaps, and collaborate with analysts and architects to ensure data quality and system resilience.
Top Skills:
SparkDroolsHadoopHbaseHiveKafkaKinesisNoSQLPysparkPythonSQL
Agency • Information Technology
Design, build, and optimize PySpark applications and ETL pipelines to process large-scale datasets from SQL/NoSQL sources, data lakes, and streaming platforms. Ensure data quality, error handling, performance tuning, and collaborate with analysts, scientists, and architects to deliver scalable data solutions.
Top Skills:
Apache AirflowSparkData LakeETLLuigiNoSQLPysparkPythonSQLStreaming Platforms
Agency • Information Technology
Design, develop, test, and deploy high-performance Spark/Scala data processing applications and ETL pipelines using the Cloudera Hadoop ecosystem. Optimize Spark and platform performance, ensure data integrity and security, collaborate with data scientists and analysts, troubleshoot issues, and implement version control and CI/CD for Spark applications.
Top Skills:
SparkCdhCloudera HadoopFlumeGitGitlabHbaseHdfsHiveImpalaJenkinsKafkaNifiNoSQLOoziePostgresScalaSpark SqlSQLSqoop
Reposted 10 Days AgoSaved
Agency • Information Technology
Design and implement data federation and lakehouse architectures, build scalable ETL/ELT pipelines with Python and Spark, optimize performance across federated queries, manage Delta/Iceberg/Hudi tables, and enforce governance, security, and access controls for analytics and AI teams.
Top Skills:
AdlsApache IcebergSparkAws EmrAzure DatabricksData FederationData LakehouseDbtDelta LakeDremioGCPGlueHudiKubernetesPulumiPysparkPythonS3SQLStarburstTerraformTrino (Presto)
Agency • Information Technology
Design and develop Big Data applications using Java, Spark, and MapReduce on Cloudera/Hadoop ecosystems. Work with Hive, Impala, YARN, Kafka, and large datasets; write tests (JUnit), perform data analysis, and script in Unix/Python. Lead and manage global technology teams and follow industry best practices.
Top Skills:
SparkClouderaHadoopHiveImpalaJavaJunitKafkaMapreducePythonUnix ShellYarn
Agency • Information Technology
Design and implement big data solutions (Spark, Hive, Java, CDP). Analyze and consolidate disparate data sources, produce functional specifications, review data models, gather stakeholder requirements, validate implementations, support production deployments, investigate data quality and data lineage, and collaborate with technology leads to ensure data completeness and accuracy.
Top Skills:
CdpData LineageData TracingDatabasesExcelHiveJavaPowerPointSparkSQLVisioWord
Agency • Information Technology
Design, build, test, and maintain high-performance Python data applications and backend services. Lead technical design, review code, mentor junior engineers, integrate with databases and cloud services, optimize performance, and ensure security, scalability, and reliability.
Top Skills:
DjangoFastapiFlaskGitLinux/UnixMicroservicesMongoDBMySQLNoSQLPostgresPythonRestful ApisSQL
Reposted 10 Days AgoSaved
Agency • Information Technology
Design, develop, maintain, and optimize enterprise Tableau dashboards and BI solutions. Integrate and cleanse data from Oracle/Sybase/Big Data, tune SQL, automate Tableau extracts, manage Tableau Server on Linux, support testing, UAT, deployments, and troubleshoot data integrity and performance issues.
Top Skills:
Big DataLinuxOracleSQLSybaseTableau ApisTableau DesktopTableau PrepTableau Server
Agency • Information Technology
Design, develop, and maintain PySpark applications and ETL pipelines to process, transform, and integrate large-scale datasets from SQL, NoSQL, data lakes, and streaming sources. Optimize Spark job performance, implement robust error handling, and collaborate with data analysts, scientists, and architects using orchestration tools like Airflow or Luigi.
Top Skills:
Apache AirflowSparkData LakeLuigiNoSQLPysparkPythonSQL
Automotive
Own the 2-3 year architecture and reliability of high-volume catalog, pricing, and inventory ingestion. Lead migrations from legacy batch to modern systems, define data SLAs, build observability and orchestration, mentor senior engineers, and resolve cross-team, high-severity data problems.
Top Skills:
Ai Coding ToolsAws EksBigQueryDatabricksEc2FlinkKafkaKinesisKubernetesPysparkPythonRdsRedpandaSnowflakeSparkSqs
Mobile • Other • Software • Analytics
Design, build, and maintain scalable data systems and ETLs on GCP; connect production data to business systems; generate analytics and ML-driven insights for GTM teams; work cross-functionally and introduce modern tools (Airflow, DBT, Presto, Hightouch).
Top Skills:
AirflowAirflowData LakeData WarehouseDbtETLGCPGoogle AnalyticsHightouchIntercomMarketoMetabasePrestoPythonSalesforce
Let Your Resume Do The Work
Upload your resume to be matched with jobs you're a great fit for.
Success! We'll use this to further personalize your experience.
Top Companies in San Antonio, TX Hiring Remote Data Engineers
See AllPopular San Antonio, TX Remote Job Searches
Remote Jobs in San Antonio
Remote Content Jobs in San Antonio
Remote Customer Success Jobs in San Antonio
Remote IT Jobs in San Antonio
Remote Cyber Security Jobs in San Antonio
Remote Tech Support Jobs in San Antonio
Remote Data & Analytics Jobs in San Antonio
Remote Analysis Reporting Jobs in San Antonio
Remote Analytics Jobs in San Antonio
Remote Business Intelligence Jobs in San Antonio
Remote Data Engineer Jobs in San Antonio
Remote Data Science Jobs in San Antonio
Remote Machine Learning Jobs in San Antonio
Remote Data Management Jobs in San Antonio
Remote UX Designer Jobs in San Antonio
Remote Software Engineer Jobs in San Antonio
Remote Android Developer Jobs in San Antonio
Remote C# Jobs in San Antonio
Remote C++ Jobs in San Antonio
Remote DevOps Jobs in San Antonio
Remote Front End Developer Jobs in San Antonio
Remote Golang Jobs in San Antonio
Remote Hardware Engineer Jobs in San Antonio
Remote iOS Developer Jobs in San Antonio
Remote Java Developer Jobs in San Antonio
Remote Javascript Jobs in San Antonio
Remote Linux Jobs in San Antonio
Remote Engineering Manager Jobs in San Antonio
Remote .NET Developer Jobs in San Antonio
Remote PHP Developer Jobs in San Antonio
Remote Python Jobs in San Antonio
Remote QA Jobs in San Antonio
Remote Ruby Jobs in San Antonio
Remote Salesforce Developer Jobs in San Antonio
Remote Scala Jobs in San Antonio
Remote Finance Jobs in San Antonio
Remote HR Jobs in San Antonio
Remote Internships in San Antonio
Remote Legal Jobs in San Antonio
Remote Marketing Jobs in San Antonio
Remote Operations Jobs in San Antonio
Remote Office Manager Jobs in San Antonio
Remote Operations Manager Jobs in San Antonio
Remote Product Manager Jobs in San Antonio
Remote Project Manager Jobs in San Antonio
Remote Sales Jobs in San Antonio
Remote Account Executive (AE) Jobs in San Antonio
Remote Account Manager (AM) Jobs in San Antonio
Remote Sales Leadership Jobs in San Antonio
Remote Sales Development Representative Jobs in San Antonio
Remote Sales Engineer Jobs in San Antonio
Remote Sales Operations Jobs in San Antonio
All Filters
Total selected ()
No Results
No Results





















