Site Reliability / Infrastructure Engineer

Medaltv — United States · Posted ~4 hours ago

Skills

Site Reliability Engineering Infrastructure engineering Incident response Production operations Scaling On-call operations Postmortems

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

A Site Reliability and Infrastructure Engineering role supporting systems that operate at massive scale. You will own on-call operations, incident response, postmortems, infrastructure scaling, and reliability improvements while partnering closely with engineering teams to ensure infrastructure keeps pace with rapid growth and demanding workloads.

Highlights

Infrastructure role operating at very large scale, with direct ownership of reliability, incident response, on-call operations, postmortems, and infrastructure scaling. The work provides exposure to demanding media ingestion pipelines and high-volume systems while collaborating directly with engineering teams.

Description

About The Company General Intuition is the frontier lab for acting in space and time. We build large action models and world models that can perceive, predict, and act across virtual and physical environments. General Intuition builds on the strength of Medal, the world's largest and fastest-growing platform for gaming clips, where millions of gamers capture, share, and discover new games every year. We recently raised $320M at a $2.3B valuation led by Khosla Ventures with participation from General Catalyst, Eric Schmidt, and Jeff Bezos, to discover the next generation of real-world intelligence. The Role Medal's infrastructure handles billions of clips, video ingestion pipelines, and social features at a massive scale most engineers never get to touch. The work centers on reliability, incident response, scaling, and making sure our infrastructure keeps up with our growth. You'll own the on-call rotation, drive postmortems, and work directly with engineering teams to meet their infra needs. The right person probably came through startups and scale-ups, has been in the room when things broke at 2am, has scaled databases under pressure, and knows the difference between a durable fix and a patch that buys you a week. What We're Looking For Infrastructure-as-code: Strong fluency in Terraform, with real experience owning infrastructure-as-code at scaleElasticsearch depth: Hands-on experience running ES for user-facing features, not just as a log sinkGCP depth: You know it maybe a little too well: Kubernetes, VPC, IAM, Cloud Logging, and the managed services ecosystemDatabase scaling: Deep, hands-on experience scaling and sharding relational databases (MySQL, Postgres) in productionIncident response instincts: You can work a P0 calmly, communicate clearly under pressure, and run a postmortem that prevents recurrenceCI/CD: You've worked with GitHub Actions in a production environmentCommunication (crucial!): You flag issues clearly and rapidly during incidents and lead/write actionable postmortemsExperience at startups: You are comfortable in an environment of rapid growth where scaling up is a priorityGreat judgment: You know the difference between a durable, sustainable fix and a patch that buys you a week Our Stack Electron, React, Redux, Styled Components & other modern web-based technologies C# and C++ for native Windows recording & more Swift for iOS, Kotlin for Android Java, Redis, RabbitMQ, Kubernetes for backend Terraform, Salt, GitHub Actions, CircleCI for IaC and CI/CD Compensation Range: $180K - $275K