Jobiglo

No results.

Senior Site Reliability Engineer (SRE)

Laine

New
Senior 🇬🇧 English
GCP Cloud Monitoring Cloud Build AlloyDB Cloud Run OpenTelemetry Langfuse Prisma Google Cloud Platform Sentry PostgreSQL Prometheus Node.js Grafana Typescript Docker Nestjs Vertex AI Supabase Redis Graphite

Job description

About the role

Laine.ai builds AI‑powered workflows for the legal industry, where strict data security, high availability, and deterministic latency are essential. As a Senior Site Reliability Engineer you will own the cloud infrastructure, observability, and deployment pipelines, working closely with core engineering to ensure reliable, low‑latency AI services.

Key responsibilities

  • Design, provision and maintain secure, reproducible GCP infrastructure (Cloud Run, Cloud SQL/AlloyDB, VPCs) using Infrastructure as Code.
  • Optimize PostgreSQL/AlloyDB clusters, manage connection pooling, tune queries and implement zero‑downtime migrations with Prisma.
  • Monitor and improve uptime, rate limits and latency for AI model inference on Vertex AI and external LLM APIs, including tracing with Langfuse.
  • Build centralized telemetry and alerting pipelines using OpenTelemetry, Sentry, Prometheus/Grafana and GCP Cloud Monitoring, and define SLOs/SLIs.
  • Maintain fast, reliable CI/CD pipelines with Cloud Build and containerized workflows.
  • Enforce least‑privilege IAM policies, manage secrets, and ensure compliance with SOC 2 and ISO 27001.
  • Lead incident response, post‑mortems and continuous improvement of on‑call processes.

Required profile

  • 4+ years of professional SRE or DevOps experience in a high‑growth SaaS environment.
  • Deep expertise with Google Cloud Platform, especially Cloud Run, networking and IAM.
  • Strong production experience with PostgreSQL or AlloyDB, including performance tuning and backup strategies.
  • Solid software engineering background with TypeScript/Node.js, NestJS and Docker.
  • Proven track record managing infrastructure as code and multi‑environment deployments.
  • Experience building observability stacks and defining reliable alerting.
  • Security‑first mindset with experience securing data layers and audit logging.

Required skills

  • Google Cloud Platform (Cloud Run, Cloud Build, Cloud Monitoring)
  • AlloyDB / PostgreSQL
  • TypeScript, Node.js, NestJS
  • Prisma ORM
  • Docker
  • OpenTelemetry, Sentry, Prometheus, Grafana
  • Langfuse
  • Vertex AI

Questions fréquentes

Le salaire n'est pas communiqué publiquement par le recruteur. Vous pouvez postuler et négocier directement avec Laine.
Cliquez sur "Postuler maintenant" en haut de la page. Vous pouvez importer votre CV en 1 clic — Jobiglo extrait automatiquement vos informations et postule pour vous.

Why are you reporting this job?

Thank you for your report. We will review this job.

Apply in 30 seconds

Enter your email to apply. An account will be created automatically.

By continuing, you accept our terms of use.

Already have an account? Login

💬 Chat with us on Telegram Chat on WhatsApp

Published 6 hours ago

Expires 1 month from now

1 views · 0 interested

Boost your chances

Upload your CV — we will match you with relevant openings.

Analyzing your CV...

Laine