Staff Platform Engineer

Myndstack — United States · Posted ~2 hours ago

Lead Full-time Remote

Skills

multi-region compute GPU infrastructure autoscaling infrastructure as code CI/CD SRE SLOs error budgets incident management cloud infrastructure Kubernetes GPU

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

A small infrastructure team is looking for a Staff Platform Engineer to own the compute and data foundations behind high-scale systems. You will design and operate multi-region GPU infrastructure, handle autoscaling for bursty workloads, own infrastructure-as-code and deployment pipelines, and establish strong reliability practices including SLOs, error budgets, on-call processes, and post-incident reviews. The role emphasizes genuine ownership and architectural problem-solving rather than ticket-driven work.

Highlights

Fully remote staff-level infrastructure role with substantial ownership over compute, data, reliability, deployment, and production operations. The role offers meaningful architectural autonomy, work on large-scale GPU infrastructure, and direct collaboration with engineering teams.

Description

Home/Careers/ Careers Infrastructure Staff platform engineer Own the compute and data planes underneath every stack we ship — the layer clients never see and can never afford to have fail. Team Infrastructure Location Remote (IST ±4) Type Full-time Package Competitive equity About The Role You'll work on the substrate: multi-region compute scheduling, autoscaling under bursty inference load, and the storage and networking seams that everything above depends on. This is a small team with real ownership. You will not be handed a ticket queue — you'll be handed a constraint and trusted to design your way out of it. What you'll do ▸Design and operate multi-region GPU compute with predictable cost and 99.99% availability▸Own infrastructure-as-code, CI, and the deployment path from merge to production▸Set the reliability bar: SLOs, error budgets, on-call practice, and post-incident review▸Work directly with client engineering teams during embedded engagements What we're looking for ▸6+ years building and operating production distributed systems▸Deep Kubernetes and cloud-provider experience, and the scars to go with it▸Fluency in a systems language — Go, Rust, or equivalent▸You've been on-call for something that mattered, and improved it Nice to have —GPU scheduling or inference-serving experience—Experience in a consulting or embedded-engineering model Not quite your role? See the other three openings. Apply Staff platform engineer No cover letter needed. Links and a few honest lines beat a formatted CV. Leave this empty Name Email Links Why this role