Lead HPC Engineer – AI Infrastructure
Osmii · Abu Dhabi
وصف الوظيفة
About the role
We are currently partnered with a premier global hyperscaler that is rapidly expanding its AI digital infrastructure footprint in Abu Dhabi, UAE. We are looking for a Lead HPC Engineer to architect, optimise, and scale ultra‑low latency, custom hardware clusters powering next‑generation AI, machine learning, and scientific workloads.
Key responsibilities
- Architect & deploy large‑scale GPU/accelerator clusters for enterprise AI training.
- Optimise ultra‑low latency fabrics using RoCEv2, RDMA, and custom network topologies.
- Tune parallel file systems (Lustre, GPFS) and Linux kernel stacks for maximum throughput.
- Drive automation across Slurm, Kubernetes, Python, and C/C++ environments.
Required profile
- 5+ years in HPC, hyperscale infrastructure, or supercomputing environments.
- Deep expertise in low‑latency networking, parallel storage, and workload managers.
- Strong background in performance profiling, hardware optimisation, and automation.
Required skills
- GPU/accelerator clusters
- RoCEv2
- RDMA
- Custom network topologies
- Lustre
- GPFS
- Linux kernel
- Slurm
- Kubernetes
- Python
- C/C++
Questions fréquentes
لماذا تبلغ عن هذا العرض؟
اكتشف المزيد
الرواتب والأدلة وعمليات البحث في الإمارات العربية المتحدة.
الرواتب حسب المهنة
قدم طلبك في 30 ثانية
أدخل بريدك الإلكتروني للتقديم. سيتم إنشاء حساب تلقائياً.
بالمتابعة، أنت توافق على شروط الاستخدام.
لديك حساب بالفعل؟ تسجيل الدخول
عزز فرصك
حمّل سيرتك الذاتية وسنقترح عليك الوظائف التي تناسب ملفك.
جاري تحليل سيرتك الذاتية...
Osmii
Abu Dhabi