This is a remote position.
Veltris is a global technology consulting and digital engineering company delivering enterprise solutions across Cloud, AI, Data Engineering, Digital Transformation, and Product Engineering. We help organizations modernize their technology platforms by building highly scalable, secure, and cloud-native infrastructure that powers mission-critical business applications.
We are looking for an experienced Senior Platform Engineer to design, build, and operate enterprise-grade platform services across on-premises and hybrid cloud environments. The ideal candidate will have deep expertise in infrastructure automation, Kubernetes platform engineering, GitOps, distributed infrastructure, and observability.
- Design, deploy, and operate enterprise-grade platform services across on-premises and hybrid cloud environments.
- Build and maintain cloud-like platform capabilities including compute, storage, networking, data processing, and observability services.
- Develop Infrastructure as Code (IaC) using Terraform to automate infrastructure provisioning and lifecycle management.
- Design and implement GitOps-based deployment strategies using ArgoCD and GitHub Actions.
- Build, administer, and optimize Kubernetes clusters for production workloads.
- Deploy and maintain large-scale distributed data platforms including Hadoop/EMR-based infrastructure.
- Design and manage distributed object storage platforms such as MinIO or S3-compatible storage solutions.
- Deploy, configure, and optimize distributed caching platforms such as Redis and Memcached.
- Implement enterprise observability solutions using OpenTelemetry, Prometheus, Grafana, and Loki.
- Automate CI/CD pipelines and improve deployment reliability across multiple environments.
- Ensure platform scalability, availability, resiliency, and operational excellence.
- Collaborate with development, DevOps, infrastructure, and security teams to deliver highly available platform services.
- Troubleshoot production infrastructure issues and perform performance tuning across Kubernetes clusters and underlying infrastructure.
- Contribute to platform capacity planning, scalability forecasting, and infrastructure optimization.
Experience
- 8+ years of experience in Platform Engineering, DevOps, Site Reliability Engineering (SRE), or Infrastructure Engineering.
- Hands-on experience designing and operating enterprise platform infrastructure.
- Strong production experience in hybrid cloud and on-premises environments.
- Terraform
- Infrastructure Automation
- Advanced Kubernetes administration
- Cluster lifecycle management
- Production Kubernetes operations
- Container orchestration
- ArgoCD
- GitHub Actions
- GitOps
- CI/CD Pipeline Automation
- Hadoop
- EMR On-Prem
- Distributed Data Platforms
- MinIO
- S3-compatible Object Storage
- Distributed Storage Architecture
- Redis
- Memcached
- OpenTelemetry (OTel)
- Prometheus
- Grafana
- Loki
- Metrics, Logging & Monitoring
- Nexus Repository or equivalent Artifact Repository
- Hypervisor administration and troubleshooting
- Linux/Operating System performance tuning
- Backup and Disaster Recovery
- Capacity Planning
- Scalability Forecasting
- Infrastructure Performance Analysis
- Platform Reliability Engineering
- Experience with enterprise-scale platform modernization initiatives.
- Experience supporting highly available distributed systems.
- Knowledge of cloud-native infrastructure design patterns.
- Exposure to storage, networking, and virtualization technologies.
- Experience working in Agile and DevOps environments.
- Strong analytical and troubleshooting skills.
- Excellent communication and collaboration abilities.
- Ability to work independently in a remote environment.
- Strong ownership mindset with a focus on operational excellence.
- Ability to mentor junior engineers and drive platform best practices.
Skills Required
- 8+ years in Platform Engineering, DevOps, SRE, or Infrastructure Engineering
- Design and operate enterprise platform infrastructure in hybrid cloud and on-premises environments
- Hands-on production experience with hybrid cloud and on-premises environments
- Terraform
- Infrastructure automation
- Advanced Kubernetes administration, cluster lifecycle management, production Kubernetes operations
- ArgoCD
- GitHub Actions
- GitOps-based deployment strategies
- CI/CD pipeline automation
- Hadoop
- EMR (on-prem)
- MinIO or S3-compatible object storage
- Redis
- Memcached
- OpenTelemetry (OTel)
- Prometheus
- Grafana
- Loki
- Troubleshooting and performance tuning for Kubernetes clusters and underlying infrastructure
- Nexus Repository or equivalent artifact repository
- Hypervisor administration and troubleshooting
- Linux/operating system performance tuning
- Backup and disaster recovery
- Capacity planning, scalability forecasting, infrastructure performance analysis
- Platform reliability engineering and supporting highly available distributed systems
- Experience with enterprise-scale platform modernization and cloud-native design patterns
- Exposure to storage, networking, and virtualization technologies
- Experience working in Agile and DevOps environments
- Strong analytical, communication, collaboration, and mentoring skills
What We Do
Veltris – Innovate, Accelerate, Transform to enable technology-driven Enterprise, Business, and Industry transformations. We are a next generation technology services company specializing in developing products, platforms and solutions in Data & Artificial Intelligence (“Data/AI”), Engineering R&D (“ER&D”) and Digital Product Engineering Services (“PES”).







