📢 جديد: تابع عروض اليوم على قناتنا في واتساب
Jobiglo

لا توجد نتائج.

Senior LLMOps / AI Platform Engineer

Hire Rightt - Executive Search & HR Advisory · Doubaï

جديد
Senior 25,000 - 30,000 AED/شهر 🇬🇧 English
Python FastAPI vLLM Hugging Face Kubernetes Docker Helm AWS NVIDIA GPU inference performance optimization LangChain LangGraph LangSmith Langfuse RAG embeddings Qdrant Milvus OpenTelemetry Prometheus Grafana PostgreSQL Redis GitHub Actions CI/CD

وصف الوظيفة

About the role

The Senior LLMOps / AI Platform Engineer will design, build, and operate production‑grade large language model (LLM) and generative AI infrastructure for a financial services firm.

Key responsibilities

  • Deploy and operate self‑hosted LLMs using vLLM, SGLang and Ollama.
  • Optimize LLM inference for latency, throughput, concurrency, GPU memory, KV cache and cost.
  • Manage GPU workloads across multiple NVIDIA GPUs.
  • Deploy AI services on Kubernetes / AWS EKS with Docker and Helm.
  • Implement reliability mechanisms, health checks, monitoring, automated recovery and model refresh strategies.
  • Set up observability using Langfuse/LangSmith, OpenTelemetry, Prometheus and Grafana.
  • Deploy and optimise Retrieval‑Augmented Generation (RAG) systems, embedding models and vector databases such as Qdrant and Milvus.
  • Support AI agents and workflows built with LangChain and LangGraph.
  • Build and maintain CI/CD pipelines for AI services and infrastructure.
  • Troubleshoot production issues across LLMs, GPUs, Kubernetes, networking and AI applications.

Required profile

  • Strong experience with Python and FastAPI.
  • Hands‑on expertise in deploying self‑hosted LLMs and using Hugging Face models.
  • Deep knowledge of Kubernetes, Docker, Helm and AWS cloud services.
  • Proven ability to optimise NVIDIA GPU inference performance.
  • Experience with LangChain, LangGraph and LangSmith tooling.

Required skills

  • Python, FastAPI
  • vLLM, Hugging Face, self‑hosted LLM deployment
  • Kubernetes, Docker, Helm, AWS
  • NVIDIA GPU inference and performance tuning
  • LangChain, LangGraph, LangSmith, Langfuse
  • RAG, embeddings, vector databases (Qdrant, Milvus)
  • Observability: OpenTelemetry, Prometheus, Grafana
  • PostgreSQL, Redis
  • GitHub Actions, CI/CD pipelines

Questions fréquentes

Le salaire proposé pour ce poste est de 25-30k AED par mois. Le détail figure dans l'annonce.
Cliquez sur "Postuler maintenant" en haut de la page. Vous pouvez importer votre CV en 1 clic — Jobiglo extrait automatiquement vos informations et postule pour vous.

لماذا تبلغ عن هذا العرض؟

شكراً لإبلاغك. سنراجع هذا العرض.

اكتشف المزيد

الرواتب والأدلة وعمليات البحث في الإمارات العربية المتحدة.

قدم طلبك في 30 ثانية

أدخل بريدك الإلكتروني للتقديم. سيتم إنشاء حساب تلقائياً.

قدم الآن →

بالمتابعة، أنت توافق على شروط الاستخدام.

لديك حساب بالفعل؟ تسجيل الدخول

لديك سؤال حول هذا العرض؟

اطرحه هنا: ستصلك تفاصيل العرض كاملة عبر البريد الإلكتروني، فوراً.

💬 راسلنا على تيليجرام

منشور منذ 10 ساعات

ينتهي شهر من الآن

9 مشاهدات · 0 مهتم

عزز فرصك

حمّل سيرتك الذاتية وسنقترح عليك الوظائف التي تناسب ملفك.

جاري تحليل سيرتك الذاتية...

Hire Rightt - Executive Search & HR Advisory

Doubaï