Senior Site Reliability Engineer – AI Training Expert

Askethos — Canada · Posted ~1 hour ago

Senior Contract Remote $80/hour, up to $1600/weeK

Skills

Site reliability engineering Production incident management Incident response Blameless postmortems Root-cause analysis On-call operations Runbook development SLOs Error budgets

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

A fully remote expert role for an experienced Site Reliability Engineer who will help improve AI-generated professional materials. The work draws on real-world production operations expertise, including incident reviews, root-cause analysis, on-call runbooks, reliability metrics, remediation tracking, and technical review presentations.

Highlights

Fully remote expert opportunity with flexible hours and an attractive hourly rate. Work can be performed on your own schedule while applying deep production reliability expertise to sophisticated AI evaluation and training workflows.

Description

About This Opportunity We're working with a leading foundational AI lab to find experienced senior site reliability engineers who can help train their latest language model on professional document, spreadsheet, and slide deck tasks. We're looking for senior site reliability engineers with 4+ years running or reviewing production incidents to create, evaluate, and refine AI-generated documents, spreadsheets, and slide decks across core workflows: blameless postmortems and root-cause analyses, on-call runbooks, severity and escalation write-ups, SLO and error-budget reports, remediation action-item trackers, and reliability review decks. Compensation: $80/hour Commitment: Flexible, 5-20 hours per week (or more if desired) Location: Fully remote, work on your own schedule Start date: ASAP Qualifications 4+ years as a Site Reliability Engineer, Incident Commander, or production/on-call engineering lead Direct ownership of writing or reviewing blameless postmortems and root-cause analyses Expert-level document, spreadsheet, and slide craftsmanship, with excellent written communication and attention to detail About Ethos Ethos is a new expert network built by a McKinsey/SoftBank/DeepMind team and backed by world-leading investors like General Catalyst. We connect experts with investors and consultancies for paid expert calls, speaking engagements, and advisory opportunities. Key Requirements 4+ years as a Site Reliability Engineer, Incident Commander, or production/on-call engineering leadDirect ownership of writing or reviewing blameless postmortems and root-cause analysesExpert-level document, spreadsheet, and slide craftsmanship, with excellent written communication and attention to detail