Incident Manager – AI Lab – Remote
Ethos
Job description
About the role
We are partnering with a leading foundational AI lab to find experienced incident managers who can help train their latest language model on professional document, spreadsheet, and slide‑deck tasks. The role focuses on creating, evaluating, and refining AI‑generated outputs across core reliability workflows.
Key responsibilities
- Create, evaluate, and refine AI‑generated documents, spreadsheets, and slide decks for core workflows such as postmortems, runbooks, and reliability review decks.
- Develop and maintain blameless postmortems, root‑cause analyses, on‑call runbooks, severity and escalation write‑ups, SLO/error‑budget reports, and remediation trackers.
Required profile
- Minimum 4 years experience as a Site Reliability Engineer, Incident Commander, or production/on‑call engineering lead.
- Proven ownership of writing or reviewing blameless postmortems and root‑cause analyses.
Required skills
- Root‑Cause Analyses
- Blameless Postmortems
What we offer
- $80 per hour.
- Flexible commitment of 5–20 hours per week (or more), fully remote work on your own schedule.
Questions fréquentes
Why are you reporting this job?
Apply in 30 seconds
Enter your email to apply. An account will be created automatically.
By continuing, you accept our terms of use.
Already have an account? Login
Published 1 day ago
Expires 1 month from now
7 views · 0 interested
Boost your chances
Upload your CV — we will match you with relevant openings.
Analyzing your CV...
Ethos