Senior Infrastructure Engineer

Posted 5 Days Ago
Be an Early Applicant
Singapore, SGP
In-Office
Senior level
Artificial Intelligence • Automotive
The Role
Designs, operates, and scales production hybrid infrastructure across on-premises, cloud, and customer environments. Responsibilities include Kubernetes platform management, secure networking, observability, deployment automation, GitOps, infrastructure as code, high availability, incident response, disaster recovery, and internal developer platform capabilities. The role supports autonomous driving services, data systems, streaming platforms, and internal tooling while maintaining strict reliability and security standards.
Summary Generated by Built In
A world empowered by autonomy. We build robotic vehicles to improve logistics safety, forge a greener Earth, and enhance human lives.

We are a closely-knit team aspiring to change the world through disruptive technology. We are innovators. We are tinkerers. We are problem-solvers. And we have a fair amount of magic dust up our sleeves. We have a plan for fleet-level deployment of autonomous vehicles, and we are looking for the best-of-the-best to join us in making this a reality.


About Venti Technologies

Based in the U.S. and Asia, Venti Technologies is the leader in safe-speed autonomous logistics systems, developing the future of goods transportation. Using rigorous mathematics, deep learning, and theoretically-grounded algorithms, Venti has a proprietary collection of autonomy technologies including a suite of powerful logistics algorithms. Venti’s proven value proposition of saving costs, increasing vehicle utilization, and improving safety is recognized by customers and driving growth. Launched in 2018, Venti brings together an unsurpassed team internationally. The company has autonomous systems deployed in Asia for industrial and logistics sites and a growing pipeline. Venti has offices in Cambridge (Massachusetts, USA), Suzhou (China), and Singapore – our Asian headquarters.

 

Role responsibilities
  • Provide and operate a production-ready hybrid infrastructure platform to host all services and applications—online, offline, data, streaming, platform, and internal tooling—for the autonomous driving business.

  • Design, deploy, operate, and scale production Kubernetes clusters across on-premises, cloud, and hybrid environments; own capacity planning, autoscaling, upgrades, and multi-tenancy.

  • Ensure high availability and meet strict SLO/SLA requirements for central and customer-site deployments; participate in incident management and post-incident reviews.

  • Design and implement secure networking for on-premises, cloud, hybrid, and customer environments, including VPN, switches, routers, firewalls, and potential 5G adoption for high-throughput data/video streaming, following cybersecurity best practices.

  • Define and operate observability, logging, monitoring, and alerting across the infrastructure and application stack; establish SLO/SLI-based alerting, dashboards, and distributed tracing to reduce MTTR.

  • Build multi-environment application deployment and release automation, including blue/green and canary strategies, GitOps, Helm/Kustomize, and service mesh for traffic management, mTLS, and observability.

  • Develop internal platform capabilities, reusable IaC modules, golden paths, and self-service provisioning to improve developer velocity and infrastructure reliability.

  • Establish production infrastructure operational processes: change management, capacity planning, security patching, disaster recovery, and operational runbooks.

Required experience
  • Bachelor’s or Master’s degree in Computer Science or a related field.

  • 5+ years of experience implementing and operating on-premises and cloud compute, storage, and networking infrastructure.

  • 3+ years of experience designing and operating production Kubernetes/OpenShift clusters at scale, including autoscaling, cluster scaling, networking/CNI, storage, security, and upgrades.

  • Excellent Linux administration and scripting skills; strong hands-on experience with Bash, Python, and Terraform.

  • Experience with hybrid infrastructure across onprem, Azure, AWS, or GCP.

  • Experience with Infrastructure as Code, GitOps, and configuration management: Terraform, Ansible, Packer, ArgoCD/Flux, Helm/Kustomize.

  • Experience running Docker and Kubernetes/OpenShift in production at scale.

  • Experience designing and operating observability stacks: Prometheus/Grafana, ELK/Loki/OpenSearch, OpenTelemetry; ability to define SLO/SLI and actionable alerting.

  • Strong understanding of infrastructure security: network segmentation, mTLS, RBAC, secrets management, image signing, vulnerability management, and CIS benchmarks.

  • Hands-on experience with on-premises networking: VPN, switches, routers, firewalls, and troubleshooting network bottlenecks.

  • Excellent communication and collaboration skills; ability to work with cross-functional teams and customer-facing deployment environments.

Bonus experience

  • Experience operating OpenStack onprem/hybrid cluster.

  • Experience designing internal developer platforms or platform engineering capabilities, including self-service infrastructure, golden paths, and IaC module design.

  • Production experience with service mesh at multi-cluster scale.

  • Experience with high-throughput video streaming infrastructure, 5G connectivity, or autonomous driving/robotics/edge environments.

We also offer world-class benefits, fantastic culture, flexible working arrangements, and a great international working environment. Come and join us!  Come change the world!

Skills Required

  • Bachelor’s or Master’s degree in Computer Science or a related field
  • 5+ years implementing and operating on-premises and cloud compute, storage, and networking infrastructure
  • 3+ years designing and operating production Kubernetes or OpenShift clusters at scale
  • Excellent Linux administration and scripting skills
  • Strong hands-on experience with Bash, Python, and Terraform
  • Experience with hybrid infrastructure across on-premises, Azure, AWS, or GCP
  • Experience with Infrastructure as Code, GitOps, and configuration management
  • Experience running Docker and Kubernetes or OpenShift in production at scale
  • Experience designing and operating observability stacks including Prometheus, Grafana, ELK, Loki, OpenSearch, or OpenTelemetry
  • Strong understanding of infrastructure security, including network segmentation, mTLS, RBAC, secrets management, image signing, vulnerability management, and CIS benchmarks
  • Hands-on experience with on-premises networking, including VPNs, switches, routers, firewalls, and network troubleshooting
  • Excellent communication and collaboration skills with cross-functional and customer-facing teams
  • Experience operating OpenStack on-premises or in hybrid clusters
  • Experience designing internal developer platforms, self-service infrastructure, golden paths, and reusable IaC modules
  • Production experience with service mesh at multi-cluster scale
  • Experience with high-throughput video streaming, 5G connectivity, autonomous driving, robotics, or edge environments
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Suzhou
112 Employees
Year Founded: 2018

What We Do

Venti Technologies mission is to create a world powered by autonomy, which is greener, safer and more efficient. Venti is a global leader in Al-powered autonomous vehicle solutions for moving goods in logistics hubs. We create software which powers autonomous vehicles that are already deployed operationally at our customers. Founded in 2018 by a team with strong MIT roots, the company is pioneering the future of transportation, enabling best-in-class safety and operational efficiency for our customers. We continue to innovate to create world-class software products. For more information, please visit www.ventittech.ai

Similar Jobs

Airwallex Logo Airwallex

Software Engineer

Artificial Intelligence • Fintech • Payments • Business Intelligence • Financial Services • Generative AI
In-Office
Singapore, SGP
2300 Employees

Doodle Labs Logo Doodle Labs

Infrastructure Engineer

Aerospace • Hardware • Internet of Things • Robotics • Wearables • App development • Automation
In-Office
Singapore, SGP
50 Employees
Hybrid
Singapore, SGP
289097 Employees

Hitachi Logo Hitachi

Site Reliability Engineer

Fintech • Information Technology • Logistics
In-Office
2 Locations
33676 Employees

Similar Companies Hiring

Hanover Park Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
42 Employees
Kepler  Thumbnail
Artificial Intelligence • Fintech • Software
New York, New York
9 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account