Senior Data Platform Engineer

Posted 7 Hours Ago
Be an Early Applicant
Hiring Remotely in IN
Remote
Senior level
Software
The Role
Own the production event-streaming platform supporting the k0rdent-ai control plane. Design and operate Kafka infrastructure, CDC and connector pipelines, schema governance, stream processing, Kubernetes deployments, and infrastructure automation. Build self-service onboarding, observability, security, multi-tenant isolation, delivery guarantees, and disaster recovery. Partner with backend and SRE teams, participate in on-call, document architecture, and mentor engineers.
Summary Generated by Built In
Company Description

Mirantis, an IREN company, is the Kubernetes-native AI infrastructure company, enabling organizations to build and operate scalable, secure, and sovereign infrastructure for modern AI, machine learning, and data-intensive applications. By combining open source innovation with deep expertise in Kubernetes orchestration, Mirantis empowers platform engineering teams to deliver composable, production-ready developer platforms across any environment—on-premises, in the cloud, at the edge, or in sovereign data centers. As enterprises navigate the growing complexity of AI-driven workloads, Mirantis delivers the automation, GPU orchestration, and policy-driven control needed to manage infrastructure with confidence and agility. Committed to open standards and freedom from lock-in, Mirantis ensures that customers retain full control of their infrastructure strategy.  https://www.mirantis.com/

Job Description

We are looking for an experienced data platform engineer to own the event-streaming backbone behind the k0rdent-ai platform — the multi-tenant control plane for enterprise GPU infrastructure. Every cluster provisioned, every GPU-hour consumed, and every tenant action produces events that must move reliably from our services into analytics, billing, observability, and downstream systems.  You will design and operate that pipeline end to end: the streaming platform itself, the connectors and change-data-capture flows that feed it, the schemas that keep producers and consumers compatible, and the automation that lets product teams onboard themselves without waiting on you. Working within an agile framework alongside engineering teams across the US, Europe, and India, you will make event data a dependable product surface rather than a best-effort side channel.

Main Responsibilities:

  • Design, deploy, and operate the production streaming platform on Kubernetes — cluster lifecycle, topics, partitioning and retention strategy, capacity planning, and upgrades.

  • Build and maintain change-data-capture and connector pipelines moving data between operational databases, the streaming layer, and analytical stores.

  • Own schema governance — a schema registry, compatibility rules, and versioning discipline so producer changes never silently break consumers.

  • Develop custom connectors, transformations, and stream-processing logic in Java or Go where off-the-shelf components fall short.

  • Build self-service onboarding so product teams can provision topics, schemas, and access through infrastructure as code rather than tickets.

  • Automate the platform with Terraform and GitOps — no manually configured clusters, no undocumented topic definitions.

  • Instrument the pipeline for production: metrics, lag and throughput monitoring, alerting, and dashboards that both the platform team and application owners can act on.

  • Enforce security and multi-tenant isolation across the streaming layer — mTLS, authentication, topic-level RBAC, encryption, and audit trails.

  • Guarantee delivery semantics that the business can rely on — ordering, idempotency, exactly-once where required, replay, and disaster recovery.

  • Partner with backend and SRE teams on event schema design, producer patterns, and failure modes; participate in on-call for the platform you own.

  • Mentor engineers on streaming architecture and raise the team's bar through design review and documentation.

Qualifications

Required Skills/Abilities: 

  •  10+ years in software, data, or platform engineering, including deep hands-on ownership of a production event-streaming platform.

  • Expert-level Apache Kafka (or Confluent Platform/Cloud) — administration, tuning, troubleshooting, and capacity management at production scale.

  • Strong Kafka Connect and CDC experience — Debezium, JDBC, and custom connector or transformation development.

  • Proven schema-registry practice and compatibility management across many independent producers and consumers.

  • Solid programming ability in Java, Python, or Go — enough to build connectors, transformations, and platform tooling, not only configure them.

  • Strong SQL and relational database knowledge (PostgreSQL, MySQL, or equivalent), including replication and CDC mechanics.

  • Kubernetes fluency — deploying, operating, and debugging stateful workloads.

  • Infrastructure as code discipline (Terraform or equivalent) and CI/CD pipeline ownership.

  • Clear written English for design docs, runbooks, and asynchronous review across global time zones.

Must Have

  • Streaming: Apache Kafka or Confluent, Kafka Connect, Schema Registry, and stream processing.

  • Data Movement: CDC tooling (Debezium or equivalent), connector development, and pipelines into analytical stores.

  • Kubernetes Native: Kubernetes, Docker, and Helm-packaged workloads in production.

  • Automation: Terraform, CI/CD (Jenkins, GitHub Actions, or equivalent), and GitOps practice.

  • Data Stores: PostgreSQL at scale, plus at least one cloud data warehouse or search store (Snowflake, BigQuery, Elasticsearch, or equivalent).

  • Cloud: AWS or GCP — networking, IAM, managed Kubernetes.

  • Security & Observability: mTLS, OIDC/SSO, RBAC, and Prometheus/Grafana monitoring.

Nice to Have

  • Confluent certification, or contributions to Kafka-ecosystem open source.

  • Stream processing with Flink, Kafka Streams, or ksqlDB.

  • Event-driven architecture patterns — outbox, saga, event sourcing, CloudEvents.

  • Exposure to GPU infrastructure, AI/ML data pipelines, or telemetry at high cardinality.

  • Usage metering, billing, or chargeback pipelines built on event data.

  • Alternative brokers (NATS JetStream, Pulsar) and a view on the tradeoffs.

  • Go, for working directly in our backend services' producer code.

  • Data quality, lineage, or catalog tooling.

  • Compliance exposure — SOC 2, ISO 27001, or similar audit support.

Education and Experience:

  • Bachelor’s degree in Computer Science & Engineering or related field or 10 years related experience.

Additional Information

What does Mirantis offer you?

  • Work with an established Silicon Valley leader in the cloud infrastructure industry;
  • Work with exceptionally passionate, talented and engaging colleagues, helping Fortune 500 and Global 2000 customers implement next-generation cloud technologies;
  • Be a part of cutting-edge, open-source innovation;
  • Thrive in the high-energy environment of a young company where openness, collaboration, risk-taking, and continuous growth are valued;
  • Professional development and training;
  • Attend conferences and working groups;
  • Company outings, happy hours, hackathons, and tech talks;
  • Receive a competitive compensation package with a strong benefits plan.

We are a Leader for Container Management in G2 (#2 after AWS)!

Skills Required

  • 10+ years in software, data, or platform engineering, including hands-on ownership of a production event-streaming platform
  • Expert-level Apache Kafka or Confluent administration, tuning, troubleshooting, and capacity management
  • Kafka Connect and CDC experience with Debezium, JDBC, and custom connector or transformation development
  • Schema Registry experience and compatibility management across producers and consumers
  • Programming ability in Java, Python, or Go
  • Strong SQL and relational database knowledge, including replication and CDC mechanics
  • Kubernetes experience deploying, operating, and debugging stateful workloads
  • Infrastructure as code experience with Terraform or equivalent
  • CI/CD pipeline ownership and GitOps practice
  • Production experience with Kubernetes, Docker, and Helm-packaged workloads
  • PostgreSQL at scale plus a cloud data warehouse or search store such as Snowflake, BigQuery, or Elasticsearch
  • AWS or GCP experience, including networking, IAM, and managed Kubernetes
  • Security and observability experience with mTLS, OIDC/SSO, RBAC, Prometheus, and Grafana
  • Clear written English for design documents, runbooks, and asynchronous reviews
  • Bachelor’s degree in Computer Science, Engineering, or a related field, or 10 years of related experience
  • Confluent certification or Kafka ecosystem open-source contributions
  • Stream processing experience with Flink, Kafka Streams, or ksqlDB
  • Experience with event-driven architecture patterns such as outbox, saga, event sourcing, or CloudEvents
  • Exposure to GPU infrastructure, AI/ML data pipelines, or high-cardinality telemetry
  • Experience building usage metering, billing, or chargeback pipelines from event data
  • Experience with NATS JetStream or Pulsar
  • Data quality, lineage, or catalog tooling experience
  • SOC 2, ISO 27001, or similar compliance exposure
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Campbell, CA
729 Employees
Year Founded: 1999

What We Do

We are dedicated to helping organizations increase developer productivity and ship code faster on public and private clouds. We provide a ZeroOps experience to remove the stress of managing cloud native infrastructure by combining software and automation tools with our cloud native expertise to deliver the industry's leading secure cloud platforms. Our capabilities allow us to provide a secure and reliable cloud native platform that includes validated FIPS-140-2 Encryption and DISA STIG ready capabilities. Who do we serve? We serve a wide range of industries, building on our extensive customer experience to provide distinct value in specific verticals including Financial Services, Government & Education, Healthcare, Manufacturing, and Telecommunications. Mirantis serves many of the world’s leading enterprises, including Adobe, DocuSign, Inmarsat, PayPal, Reliance Jio, Societe Generale, Splunk, and S&P Global. Learn more at www.mirantis.com.

Similar Jobs

Kenvue Logo Kenvue

Platform Engineer

eCommerce • Healthtech • Software
In-Office or Remote
17 Locations
17919 Employees

Netskope Logo Netskope

Senior Devops Engineer

Cloud • Security • Software • Cybersecurity
Remote
India
1479 Employees

DigiCert Logo DigiCert

Platform Engineer

Security • Software • Cybersecurity
Remote
India
1372 Employees

Ampd Energy Logo Ampd Energy

Senior Engineer

Energy • Renewable Energy
In-Office or Remote
2 Locations
88 Employees

Similar Companies Hiring

Kepler  Thumbnail
Artificial Intelligence • Fintech • Software
New York, New York
9 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees
Revel.io Thumbnail
Aerospace • Hardware • Robotics • Software
US
50 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account