📢 New: get today's jobs on our WhatsApp Channel
Jobiglo

No results.

Senior DevOps Engineer – Kubernetes & AI Platform

Halian | Managed Services, Recruitment and Contract Staffing · Abou Dabi

Senior 🇬🇧 English
Kubernetes Terraform Ansible Helm GitOps Python Bash Go GitLab CI GitHub Actions Jenkins Prometheus Grafana ELK OpenTelemetry RBAC Image governance Secrets management Network segmentation Disaster recovery

Job description

About the role

We are looking for a Senior DevOps Engineer to lead the design, deployment, and operation of enterprise‑grade Kubernetes platforms on Nutanix infrastructure. You will work closely with AI engineers and platform architects to deliver scalable, secure, and high‑performance environments for machine‑learning workloads.

Key responsibilities

  • Deploy, administer, and optimise Kubernetes clusters on Nutanix AHV, ensuring resilience and scalability.
  • Design infrastructure for AI/ML workloads, including GPU‑enabled nodes, model deployment frameworks and large‑scale data pipelines.
  • Automate provisioning and configuration using Terraform, Ansible, Helm and GitOps practices.
  • Build and maintain CI/CD pipelines for containerised applications and ML workflows.
  • Implement monitoring, logging, tracing and alerting with Prometheus, Grafana, ELK and OpenTelemetry.
  • Enforce security controls such as RBAC, image governance, secrets management, network segmentation and workload protection.
  • Optimise cluster performance, storage, networking and compute resources for CPU and GPU workloads.
  • Collaborate with development teams to onboard applications and define operational best practices.
  • Define and test backup, recovery and disaster‑recovery procedures for Kubernetes platforms.

Required profile

  • Minimum 5 years of experience in DevOps, Platform Engineering, Cloud Operations or SRE, with at least 3 years of production Kubernetes administration.
  • Strong hands‑on experience with Nutanix technologies (AHV, Prism, Karbon, Files, Objects).
  • Proven experience supporting AI/ML platforms such as Kubeflow, MLflow, KServe or Ray.
  • Advanced knowledge of Infrastructure‑as‑Code and automation tools (Terraform, Ansible, Helm, GitOps).
  • Proficiency in scripting languages (Python, Bash, Go or equivalent).
  • Experience managing GPU‑enabled Kubernetes clusters and NVIDIA accelerators.
  • Solid understanding of Kubernetes networking, storage, security and policy frameworks.
  • Hands‑on experience with CI/CD tools (GitLab CI, GitHub Actions, Jenkins).
  • Experience implementing observability solutions in cloud‑native environments.
  • Bachelor’s degree in Computer Science, IT, Engineering or related field, or equivalent practical experience.

Required skills

  • Kubernetes
  • Nutanix AHV, Prism, Karbon, Files, Objects
  • GPU orchestration, NVIDIA technologies
  • Terraform, Ansible, Helm, GitOps
  • Python, Bash, Go
  • CI/CD (GitLab CI, GitHub Actions, Jenkins)
  • Observability tools (Prometheus, Grafana, ELK, OpenTelemetry)
  • Security controls (RBAC, image governance, secrets management, network segmentation)
  • Backup and disaster‑recovery for Kubernetes

Questions fréquentes

Le salaire n'est pas communiqué publiquement par le recruteur. Vous pouvez postuler et négocier directement avec Halian | Managed Services, Recruitment and Contract Staffing.
Cliquez sur "Postuler maintenant" en haut de la page. Vous pouvez importer votre CV en 1 clic — Jobiglo extrait automatiquement vos informations et postule pour vous.

Why are you reporting this job?

Thank you for your report. We will review this job.

Apply in 30 seconds

Enter your email to apply. An account will be created automatically.

By continuing, you accept our terms of use.

Already have an account? Login

💬 Chat with us on Telegram Chat on WhatsApp

Published 1 month ago

Expires 3 weeks from now

37 views · 0 interested

Boost your chances

Upload your CV — we will match you with relevant openings.

Analyzing your CV...

Halian | Managed Services, Recruitment and Contract Staffing

Abou Dabi