Site Reliability Engineer – Production Sandbox Reliability
Deel · Dubai
وصف الوظيفة
About the role
We are seeking a Site Reliability Engineer to join our Production Sandbox Reliability team in Dubai. You will act as the operational guardian for enterprise customer sandboxes, ensuring they remain stable, up‑to‑date, and well‑monitored across a globally distributed environment.
Key responsibilities
- Monitor sandbox health, guaranteeing high availability, uptime, and stability.
- Roll out regular updates to microservices within each sandbox.
- Triage alerts, investigate incidents, and escalate with clear context.
- Enhance observability by expanding metrics, logs, and tracing using Datadog and the Grafana stack (Mimir, Loki, Tempo).
- Participate in post‑incident reviews and drive actionable improvements.
- Communicate proactively with engineering and customer‑facing teams during incidents, maintenance, and upgrades.
- Join a follow‑the‑sun on‑call rotation covering APAC, EMEA, and LATAM time zones.
Required profile
- Minimum 4 years of experience in SRE, DevOps, or Infrastructure Engineering.
- Strong verbal and written communication skills for both technical and non‑technical audiences.
- Self‑starter mindset focused on operational excellence and continuous improvement.
Required skills
- Node.js or Go programming.
- AWS services (EKS, S3, RDS).
- Kubernetes with Helm and ArgoCD.
- Observability tools: Datadog, Grafana, Mimir, Loki, Tempo, Zabbix.
Questions fréquentes
لماذا تبلغ عن هذا العرض؟
اكتشف المزيد
الرواتب والأدلة وعمليات البحث في الإمارات العربية المتحدة.
الرواتب حسب المهنة
قدم طلبك في 30 ثانية
أدخل بريدك الإلكتروني للتقديم. سيتم إنشاء حساب تلقائياً.
بالمتابعة، أنت توافق على شروط الاستخدام.
لديك حساب بالفعل؟ تسجيل الدخول
منشور منذ 4 أسابيع
ينتهي 4 أسابيع من الآن
41 مشاهدات · 0 مهتم
عزز فرصك
حمّل سيرتك الذاتية وسنقترح عليك الوظائف التي تناسب ملفك.
جاري تحليل سيرتك الذاتية...
Deel
Dubai