Production Engineer

Askethos — Canada · Posted ~3 hours ago

Senior Contract Remote $80/hour, up to $1600/weeK

Skills

production engineering Site Reliability Engineering incident management incident response blameless postmortems root-cause analysis on-call operations SLOs error budgets technical documentation SRE Error budgets Incident management

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

Experienced production engineers are invited to help evaluate and refine AI-generated operational documentation. Work includes blameless postmortems, root-cause analyses, on-call runbooks, severity and escalation reports, SLO and error-budget materials, and reliability reviews. The engagement is fully remote, flexible, starts immediately, and requires at least four years of relevant incident or production experience.

Highlights

Flexible fully remote expert opportunity paying $80/hour, with 5–20 hours per week or more, an immediate start, and work that leverages production reliability expertise.

Description

About This Opportunity We're working with a leading foundational AI lab to find experienced production engineers who can help train their latest language model on professional document, spreadsheet, and slide deck tasks. We're looking for production engineers with 4+ years running or reviewing production incidents to create, evaluate, and refine AI-generated documents, spreadsheets, and slide decks across core workflows: blameless postmortems and root-cause analyses, on-call runbooks, severity and escalation write-ups, SLO and error-budget reports, remediation action-item trackers, and reliability review decks. Compensation: $80/hour Commitment: Flexible, 5-20 hours per week (or more if desired) Location: Fully remote, work on your own schedule Start date: ASAP Qualifications 4+ years as a Site Reliability Engineer, Incident Commander, or production/on-call engineering lead Direct ownership of writing or reviewing blameless postmortems and root-cause analyses Expert-level document, spreadsheet, and slide craftsmanship, with excellent written communication and attention to detail About Ethos Ethos is a new expert network built by a McKinsey/SoftBank/DeepMind team and backed by world-leading investors like General Catalyst. We connect experts with investors and consultancies for paid expert calls, speaking engagements, and advisory opportunities. Key Requirements 4+ years as a Site Reliability Engineer, Incident Commander, or production/on-call engineering leadDirect ownership of writing or reviewing blameless postmortems and root-cause analysesExpert-level document, spreadsheet, and slide craftsmanship, with excellent written communication and attention to detail