Senior Data Platform Engineer

Posted 15 Hours Ago
Be an Early Applicant
Hiring Remotely in India
Remote
Senior level
Security • Software • Cybersecurity
The Role
Designs and operates scalable data and machine learning platforms using Databricks, Python, SQL, Spark, and PySpark. Builds batch, streaming, and production ML pipelines; manages Delta Lake, Unity Catalog, Workflows, MLflow, governance, data quality, observability, and model deployment. Partners with data scientists and business teams, optimizes performance and costs, contributes to platform architecture, mentors engineers, and documents operational practices.
Summary Generated by Built In

Who we are

DigiCert is a global leader in intelligent trust. We protect the digital world by ensuring the security, privacy, and authenticity of every interaction. Our AI-powered DigiCert ONE platform unifies PKI, DNS, and certificate lifecycle management, to secure infrastructure, software, devices, messages, AI content and agents. Learn why more than 100,000 organizations, including 90% of the Fortune 500, choose DigiCert to stop today’s threats and prepare for a quantum-safe future at www.digicert.com


Job summary

We are looking for a Senior Data Platform Engineer to design, build, and operate the foundational data and ML infrastructure that powers analytics, reporting, and machine learning across DigiCert. This role sits at the intersection of data engineering and ML platform work — you will own the systems, pipelines, and tooling that data scientists, analysts, and engineers rely on every day. You bring deep expertise in Databricks, a strong engineering mindset, and hands-on experience building and maintaining ML pipelines in production. You will partner closely with data science, analytics, product, and business teams to deliver a platform that is reliable, governed, and built to scale.


What you will do

  • Design, build, and maintain scalable data and ML pipelines using Python and SQL, processing large-scale datasets across batch and streaming workloads
  • Own and evolve core platform infrastructure on Databricks — including Delta Lake table architecture, Unity Catalog governance, Databricks Workflows orchestration, and compute optimization
  • Build and maintain end-to-end ML pipelines: feature engineering, model training pipelines, experiment tracking (MLflow), and model deployment/serving infrastructure
  • Collaborate with data scientists to operationalize models — bridging the gap between experimentation and production-grade ML systems
  • Define and enforce data platform standards: ingestion patterns, data modeling conventions, medallion architecture (Bronze/Silver/Gold), and pipeline reliability practices
  • Implement data quality, observability, and monitoring frameworks to ensure platform health and data trustworthiness
  • Optimize pipelines for performance, cost, and reliability at scale using Spark and PySpark
  • Evaluate, integrate, and govern new platform tooling and data sources within the Databricks ecosystem
  • Contribute to architectural decisions and help drive the long-term data platform roadmap
  • Participate in code reviews, technical design discussions, and engineering standards
  • Mentor junior engineers and elevate overall platform and data engineering practices
  • Document platform architecture, pipeline design, and operational runbooks

What you will have

  • 6+ years of experience in data engineering, data platform, or ML engineering roles
  • Strong proficiency in Python and SQL, with a track record of building production-grade data pipelines using both
  • Hands-on Databricks expertise: Delta Lake, Unity Catalog, Databricks Workflows, PySpark, and the Databricks ecosystem broadly
  • Experience building and maintaining ML pipelines in production — feature engineering, training pipelines, experiment tracking, and model deployment
  • Familiarity with MLflow or comparable experiment tracking and model registry tools
  • Experience working on cloud data platforms (AWS, Azure, or GCP)
  • Strong understanding of data modeling, dimensional design, and analytics-friendly data architecture
  • Experience with batch and incremental/CDC pipeline patterns
  • Proficiency with Git, version control, and CI/CD practices for data and ML workflows
  • Strong engineering judgment — you think about reliability, maintainability, and cost, not just correctness
  • Clear communication and comfort working with both technical and non-technical stakeholders

Nice to have

  • Experience with streaming or near real-time pipelines (Kafka, Kinesis, Spark Structured Streaming)
  • Familiarity with feature store platforms (Databricks Feature Store, Feast, or Tecton)
  • Experience with LLM pipelines, RAG architectures, or AI/BI tooling (Genie, AI Functions)
  • Knowledge of data quality and observability tooling (Great Expectations, Monte Carlo, etc.)
  • Exposure to dbt or similar SQL-based transformation frameworks
  • Infrastructure-as-code experience (Terraform, Databricks Asset Bundles)
  • Experience working in Agile or Scrum environments
  • Prior experience mentoring engineers or shaping platform standards

Benefits

  • Generous time off policies
  • Top shelf benefits
  • Education, wellness and lifestyle support

To protect candidate information and maintain a secure hiring process, all applications must be submitted through our careers portal. Resumes or CVs sent directly via email will not be reviewed or considered.


#LI-SD1


Skills Required

  • 6+ years of experience in data engineering, data platform, or ML engineering roles
  • Strong proficiency in Python and SQL
  • Production-grade data pipeline development experience
  • Hands-on Databricks experience, including Delta Lake, Unity Catalog, Databricks Workflows, and PySpark
  • Experience building and maintaining production ML pipelines, including feature engineering, model training, experiment tracking, and model deployment
  • Familiarity with MLflow or comparable experiment tracking and model registry tools
  • Experience with cloud data platforms such as AWS, Azure, or GCP
  • Strong understanding of data modeling, dimensional design, and analytics-friendly data architecture
  • Experience with batch and incremental or CDC pipeline patterns
  • Proficiency with Git, version control, and CI/CD practices for data and ML workflows
  • Strong engineering judgment focused on reliability, maintainability, and cost
  • Clear communication with technical and non-technical stakeholders
  • Experience with streaming or near-real-time pipelines using Kafka, Kinesis, or Spark Structured Streaming
  • Familiarity with feature stores such as Databricks Feature Store, Feast, or Tecton
  • Experience with LLM pipelines, RAG architectures, or AI/BI tooling
  • Knowledge of data quality and observability tools such as Great Expectations or Monte Carlo
  • Exposure to dbt or similar SQL-based transformation frameworks
  • Infrastructure-as-code experience with Terraform or Databricks Asset Bundles
  • Experience working in Agile or Scrum environments
  • Prior experience mentoring engineers or shaping platform standards

DigiCert Compensation & Benefits Highlights

The following summarizes recurring compensation and benefits themes identified from responses generated by popular LLMs to common candidate questions about DigiCert and has not been reviewed or approved by DigiCert.

  • Leave & Time Off Breadth Vacation/PTO and sick leave are characterized as strong, and some accounts mention a sabbatical program.
  • Retirement Support The package includes a 401(k) with company matching, with recent confirmations of this benefit.
  • Flexible Benefits Hybrid and work-from-home options are referenced consistently, indicating practical flexibility in how and where work is done.

DigiCert Insights

Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Lehi, Utah
1,372 Employees
Year Founded: 2003

What We Do

DigiCert is the digital trust provider of choice for leading companies around the globe, enabling individuals, businesses, governments, and consortia to engage online with confidence, knowing their digital footprint is secure.

Similar Jobs

CVS Health Logo CVS Health

Senior Platform Engineer

Fitness • Healthtech • Retail • Pharmaceutical
In-Office or Remote
47 Locations
119959 Employees
83K-222K Annually

Ampd Energy Logo Ampd Energy

Senior Engineer

Energy • Renewable Energy
In-Office or Remote
2 Locations
88 Employees

IDEXX Logo IDEXX

Senior Data Engineer

Healthtech • Pet • Biotech
In-Office or Remote
5 Locations
6764 Employees
110K-130K Annually

ServiceTitan Logo ServiceTitan

Senior Software Engineer

Artificial Intelligence • Cloud • Fintech • Machine Learning • Mobile • Software
Remote or Hybrid
India
2760 Employees

Similar Companies Hiring

Hanover Park Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
42 Employees
Kepler  Thumbnail
Fintech • Software
New York, New York
6 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account