Senior Infrastructure Engineer

Venti

Singapore

On-site

SGD 120,000 - 180,000

Full time

2 days ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

Venti Technologies is seeking a Senior Platform/Infrastructure Engineer to architect and operate a production-grade hybrid infrastructure for autonomous driving services. You will manage Kubernetes/OpenShift clusters, ensure high availability, and lead secure networking across on‑prem, cloud and customer sites.

You will build scalable deployment pipelines with GitOps tools, implement observability stacks, and drive IaC practices.

Qualifications

  • Bachelor’s or Master’s degree in Computer Science or a related field.
  • 5+ years of experience implementing and operating on-premises and cloud compute, storage, and networking infrastructure.
  • 3+ years of experience designing and operating production Kubernetes/OpenShift clusters at scale.
  • Excellent Linux administration and scripting skills; strong hands-on with Bash, Python, and Terraform.
  • Experience with hybrid infrastructure across onprem, Azure, AWS, or GCP.
  • Experience with Infrastructure as Code, GitOps, and configuration management: Terraform, Ansible, Packer, ArgoCD/Flux, Helm/Kustomize.
  • Experience running Docker and Kubernetes/OpenShift in production at scale.
  • Experience designing and operating observability stacks: Prometheus/Grafana, ELK/Loki/OpenSearch, OpenTelemetry; ability to define SLO/SLI and actionable alerting.
  • Strong understanding of infrastructure security: network segmentation, mTLS, RBAC, secrets management, image signing, vulnerability management, CIS benchmarks.

Responsibilities

  • Provide and operate a production-ready hybrid infrastructure platform to host all services and applications for the autonomous driving business.
  • Design, deploy, operate, and scale production Kubernetes clusters across on-premises, cloud, and hybrid environments; manage capacity planning, autoscaling, upgrades, and multi-tenancy.
  • Ensure high availability and meet strict SLO/SLA requirements for central and customer-site deployments; participate in incident management and post-incident reviews.
  • Design and implement secure networking for on-premises, cloud, hybrid, and customer environments, including VPN, switches, routers, firewalls, and potential 5G adoption for high-throughput data/video streaming.
  • Define and operate observability, logging, monitoring, and alerting across the stack; establish SLO/SLI-based alerting, dashboards, and distributed tracing to reduce MTTR.
  • Build multi-environment application deployment and release automation, including blue/green and canary strategies, GitOps, Helm/Kustomize, and service mesh for traffic management, mTLS, and observability.
  • Develop internal platform capabilities, reusable IaC modules, golden paths, and self-service provisioning to improve developer velocity and infrastructure reliability.
  • Establish production infrastructure operational processes: change management, capacity planning, security patching, disaster recovery, and operational runbooks.

Skills

Kubernetes
OpenShift
Bash
Python
Terraform
Ansible
Packer
ArgoCD
Flux
Helm
Kustomize
GitOps
Prometheus
Grafana
OpenTelemetry

Education

Bachelor’s or Master’s degree in Computer Science or related field

Tools

Terraform
Ansible
Packer
ArgoCD
Flux
Helm
Kustomize

Job description

A world empowered by autonomy. We build robotic vehicles to improve logistics safety, forge a greener Earth, and enhance human lives. We are a closely-knit team aspiring to change the world through disruptive technology. We are innovators. We are tinkerers. We are problem-solvers. And we have a fair amount of magic dust up our sleeves. We have a plan for fleet-level deployment of autonomous vehicles, and we are looking for the best-of-the-best to join us in making this a reality.

About Venti Technologies

Based in the U.S. and Asia, Venti Technologies is the leader in safe-speed autonomous logistics systems, developing the future of goods transportation. Using rigorous mathematics, deep learning, and theoretically-grounded algorithms, Venti has a proprietary collection of autonomy technologies including a suite of powerful logistics algorithms. Venti’s proven value proposition of saving costs, increasing vehicle utilization, and improving safety is recognized by customers and driving growth. Launched in 2018, Venti brings together an unsurpassed team internationally. The company has autonomous systems deployed in Asia for industrial and logistics sites and a growing pipeline. Venti has offices in Cambridge (Massachusetts, USA), Suzhou (China), and Singapore – our Asian headquarters.

Role Responsibilities
  • Provide and operate a production-ready hybrid infrastructure platform to host all services and applications—online, offline, data, streaming, platform, and internal tooling—for the autonomous driving business.
  • Design, deploy, operate, and scale production Kubernetes clusters across on-premises, cloud, and hybrid environments; own capacity planning, autoscaling, upgrades, and multi‑tenancy.
  • Ensure high availability and meet strict SLO/SLA requirements for central and customer-site deployments; participate in incident management and post‑incident reviews.
  • Design and implement secure networking for on‑premises, cloud, hybrid, and customer environments, including VPN, switches, routers, firewalls, and potential 5G adoption for high‑throughput data/video streaming, following cybersecurity best practices.
  • Define and operate observability, logging, monitoring, and alerting across the infrastructure and application stack; establish SLO/SLI‑based alerting, dashboards, and distributed tracing to reduce MTTR.
  • Build multi‑environment application deployment and release automation, including blue/green and canary strategies, GitOps, Helm/Kustomize, and service mesh for traffic management, mTLS, and observability.
  • Develop internal platform capabilities, reusable IaC modules, golden paths, and self‑service provisioning to improve developer velocity and infrastructure reliability.
  • Establish production infrastructure operational processes: change management, capacity planning, security patching, disaster recovery, and operational runbooks.
Required Experience
  • Bachelor’s or Master’s degree in Computer Science or a related field.
  • 5+ years of experience implementing and operating on-premises and cloud compute, storage, and networking infrastructure.
  • 3+ years of experience designing and operating production Kubernetes/OpenShift clusters at scale, including autoscaling, cluster scaling, networking/CNI, storage, security, and upgrades.
  • Excellent Linux administration and scripting skills; strong hands‑on experience with Bash, Python, and Terraform.
  • Experience with hybrid infrastructure across onprem, Azure, AWS, or GCP.
  • Experience with Infrastructure as Code, GitOps, and configuration management: Terraform, Ansible, Packer, ArgoCD/Flux, Helm/Kustomize.
  • Experience running Docker and Kubernetes/OpenShift in production at scale.
  • Experience designing and operating observability stacks: Prometheus/Grafana, ELK/Loki/OpenSearch, OpenTelemetry; ability to define SLO/SLI and actionable alerting.
  • Strong understanding of infrastructure security: network segmentation, mTLS, RBAC, secrets management, image signing, vulnerability management, and CIS benchmarks.
  • Hands‑on experience with on‑premises networking: VPN, switches, routers, firewalls, and troubleshooting network bottlenecks.
  • Excellent communication and collaboration skills; ability to work with cross‑functional teams and customer‑facing deployment environments.
Bonus Experience
  • Experience operating OpenStack onprem/hybrid cluster.
  • Experience designing internal developer platforms or platform engineering capabilities, including self‑service infrastructure, golden paths, and IaC module design.
  • Production experience with service mesh at multi‑cluster scale.
  • Experience with high‑throughput video streaming infrastructure, 5G connectivity, or autonomous driving/robotics/edge environments.

We also offer world-class benefits, fantastic culture, flexible working arrangements, and a great international working environment. Come and join us! Come change the world!

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Infrastructure Engineer
Senior Infrastructure Engineer

Venti Technologies • Singapore

On-site
SGD 120,000 - 200,000
World-class benefits
Fantastic culture
Flexible working arrangements
+1
Senior Software Engineer C++
Senior Software Engineer C++

Venti Technologies • Singapore

On-site
SGD 70,000 - 90,000
C++ Software Engineer Intern
C++ Software Engineer Intern

Venti • Singapore

On-site
SGD 13,000 - 20,000
Stipend
Autonomous Vehicle Service Engineer Intern
Autonomous Vehicle Service Engineer Intern

Venti • Singapore

On-site
World-class benefits
Flexible working arrangements
International working environment
Autonomous Vehicle Hardware Intern
Autonomous Vehicle Hardware Intern

Venti Technologies • Singapore

On-site
SGD 11,000 - 20,000
World-class benefits
Flexible working arrangements
Great international working environment
+1
System Quality Engineer
System Quality Engineer

Venti • Singapore

On-site
SGD 60,000 - 80,000
Principal Robotics Engineer (Robotics Solutions & Integration)
Principal Robotics Engineer (Robotics Solutions & Integration)

Venti • Singapore

On-site
SGD 180,000 - 260,000
Senior Recruiter
Senior Recruiter

Venti Technologies • Singapore

On-site
SGD 80,000 - 120,000
World-class benefits
Flexible working arrangements
Great international working environment
Control Engineer
Control Engineer

Linuxconfig • Singapore

On-site
SGD 80,000 - 120,000
Flexible working arrangements
International working environment
World-class benefits
Staff Research Scientist, Localization and Mapping
Staff Research Scientist, Localization and Mapping

Venti Technologies • Singapore

On-site
SGD 90,000 - 120,000