Site Reliability Engineer

Astreya — United States · Posted ~3 hours ago

Senior Full-time

Skills

cloud infrastructure system architecture CI/CD platform engineering technical leadership Kubernetes APIs

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

An organization is seeking a senior reliability engineer to shape platform strategy, design scalable infrastructure, improve delivery automation, and guide engineering teams on mission-critical services.

Highlights

Technical leadership position focused on platform architecture, automation, reliability, and mentoring engineers on critical systems.

Description

Role Summary Our Platform Services team is hiring a technical leader to shape the architecture and direction of Platform Services for FedRAMP High and IL5. This role will influence technical decisions across centralized platform APIs, reusable CI components, continuous delivery workflows, service onboarding, federal promotion automation, and paved-road platform capabilities. This is a system-level technical leadership role. You should be able to take a broad perspective, guide other engineers, build consensus across teams, and make strong contributions under time pressure on mission-critical systems. You will lead design direction, coordinate complex work, and raise the technical bar through architecture reviews, mentorship, and clear communication. What You’ll Do Set the technical strategy for Platform Services work supporting FedRAMPHigh and IL5. Own key architectural decisions for centralized platform APIs, CIcomponents, continuous delivery workflows, automated service onboarding, and federal promotion gates. Lead the move from manual federal promotion processes toward reliable,automated, gated delivery workflows. Define patterns for service onboarding, deployment validation, operationalreadiness, audit-friendly change control, and rollback safety. Partner with other teams to align dependencies, remove blockers, anddrive clear technical decisions across the federal build-out. Lead design reviews for services moving into federal environments, withattention to isolation, observability, security controls, and operational ownership. Mentor senior and mid-level engineers; raise the quality of designs, code reviews, runbooks, and incident analysis.Represent Platform Services in program planning, execution commit work, risk reviews, and senior engineering discussions.Communicate difficult technical concepts clearly to stakeholders and help the team reach durable decisions. You Are an Ideal Candidate If You Have 8-12+ years of experience in SRE, platform engineering, infrastructure, distributed systems, or cloud operations.Demonstrated technical leadership across multi-team infrastructure or platform programs.Strong software engineering skills in Python, Go, Ruby, or similar languages, plus experience with infrastructure-as-code and production automation.Deep experience operating large-scale cloud or hybrid-cloud systems with strong reliability, security, and observability requirements.Strong understanding of CI/CD, deployment orchestration, Kubernetes, AWS, and production incident response.Ability to translate compliance and security requirements into practical engineering systems and operating models.Excellent written communication, architecture review, mentoring, and cross-functional leadership skills. Preferred Experience building or operating FedRAMP High, DoD IL5, GovCloud, FIPS, or other regulated environments.Experience with platform APIs, progressive delivery, CI components, continuous delivery, artifact promotion, or deployment control planes.Experience leading automation of previously manual operational workflows.