SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.
Venti Technologies is building safe-speed autonomous logistics systems for goods transportation, with deployed autonomous systems across Asia and a growing customer pipeline. The company combines rigorous mathematics, deep learning, and proprietary autonomy algorithms to deliver cost savings, increased vehicle utilization, and improved safety.
As Senior Infrastructure Engineer, you will own the design, deployment, and operation of a production-ready hybrid infrastructure platform supporting all services, applications, and internal tooling for the autonomous driving business. Key responsibilities include:
- Design, deploy, and operate production Kubernetes clusters across on-premises, cloud, and hybrid environments, managing capacity planning, autoscaling, upgrades, and multi-tenancy.
- Ensure high availability and strict SLO/SLA compliance for central and customer-site deployments; lead incident management and post-incident reviews.
- Design secure networking across on-premises, cloud, hybrid, and customer environments, including VPN, switches, routers, firewalls, and potential 5G adoption for high-throughput data and video streaming.
- Define and operate observability, logging, monitoring, and alerting across infrastructure and application stacks; establish SLO/SLI-based alerting, dashboards, and distributed tracing.
- Build multi-environment application deployment and release automation using blue/green, canary strategies, GitOps, Helm/Kustomize, and service mesh for traffic management and mTLS.
- Develop internal platform capabilities, reusable Infrastructure-as-Code modules, golden paths, and self-service provisioning to improve developer velocity.
- Establish production infrastructure operational processes including change management, capacity planning, security patching, disaster recovery, and runbooks.
Required: Bachelor's or Master's in Computer Science or related field; 5+ years implementing and operating on-premises and cloud infrastructure; 3+ years designing and operating production Kubernetes/OpenShift clusters at scale; strong Linux administration and scripting (Bash, Python, Terraform); hybrid infrastructure experience (Azure, AWS, GCP); Infrastructure as Code and GitOps expertise; production Docker/Kubernetes experience; observability stack design (Prometheus/Grafana, ELK/Loki/OpenSearch, OpenTelemetry); infrastructure security knowledge; on-premises networking hands-on experience; excellent cross-functional collaboration skills.
Bonus: OpenStack experience, internal developer platform design, multi-cluster service mesh, high-throughput video streaming, 5G, or autonomous driving/robotics/edge environment experience.