Summary
β¨ AIβGenerated
As a Senior Infrastructure Engineer on a platform team, you will architect, build, and operate the cloud infrastructure and developer platform supporting a suite of digital products. You will own Kubernetes, infrastructure-as-code, CI/CD and GitOps, networking, DNS, observability, and cloud security while creating reliable self-service foundations for engineering teams.
Highlights
Senior platform role with end-to-end ownership of critical cloud infrastructure and developer tooling. The position offers broad technical responsibility across Kubernetes, infrastructure automation, delivery pipelines, networking, observability, and cloud security.
Description
About Level
Level is a learning technology company dedicated to helping students build real academic and life skills with confidence and joy.
We combine proven curriculum principles with world class interactive design to make meaningful practice something students want to come back to, not something they struggle through.
We support what teachers, schools, and parents are already doing by increasing student engagement with high quality, standards aligned practice that reinforces classroom learning.
As an Senior Infrastructure Engineer on the Platform team, you will architect, build, and operate the cloud infrastructure and developer platform that every Level product runs on.
You will own critical infrastructure end-to-end β the Kubernetes platform, infrastructure-as-code, CI/CD and GitOps delivery, networking (ingress and egress), DNS, observability, and cloud security posture β and provide the reliable, self-service foundations the rest of engineering builds on.
You will work on a small, senior-leaning team where infrastructure decisions have direct, visible impact on reliability, performance, cost, and developer velocity.
What You'll Do
Cloud Infrastructure & IaC β Design, build, and operate secure, highly available AWS infrastructure using Terraform/OpenTofu with a GitOps workflow (Atlantis).
Own capacity planning, DR, and cost optimization for the systems you run.Kubernetes & Platform Operations β Operate and evolve EKS: autoscaling (Karpenter), upgrades, core add-ons, and Helm-based delivery (ArgoCD).CI/CD & Developer Enablement β Build and maintain GitHub Actions pipelines that let platform and product teams ship fast and safely, with self-service tooling where it makes sense.Networking, Ingress & DNS β Own ingress/egress (Traefik), service mesh and mTLS (Linkerd/Envoy), load balancing, edge TLS, and DNS (Route 53, Terraform-managed).Observability & Reliability β Build observability with OpenTelemetry and SigNoz; use telemetry to drive reliability, performance, and cost decisions.
Serve as an escalation point for complex incidents, leading troubleshooting and post-mortems.Security & Compliance β Apply cloud security best practices across identity, secrets, and network boundaries, with particular care for student data and K-12 privacy.
Operate posture/vulnerability tooling (Security Hub, GuardDuty, Inspector, Snyk) and org guardrails (Control Tower, SCPs).Ownership & Mentorship β Set standards, mentor engineers, and leave the platform better than you found it.
What You Need
5+ years operating large-scale cloud infrastructure (AWS strongly preferred)Deep IaC experience (Terraform/OpenTofu; CloudFormation/Pulumi/CDK also relevant)Strong Docker/Kubernetes (EKS) production experienceScripting/automation proficiency (Python, Go, or Bash)Solid cloud networking fundamentals (VPC, DNS, load balancing, ingress, firewalls/WAF, VPNs) and security best practicesProven CI/CD and GitOps experience (GitHub Actions or similar)Observability experience (metrics/logs/traces) used to drive real decisionsTrack record leading infrastructure projects independently, end to endStrong communication across technical and non-technical audiences
Nice to Have
ArgoCD, Atlantis, Linkerd/Envoy, Traefik, Karpenter, HelmOpenTelemetry, SigNoz (our stack), Grafana, Datadog, or PrometheusBackstage or other internal developer platform experienceAI/ML infra experience (GPU scheduling, model/agent hosting, inference gateways)Rust service CI/CD, CloudFront/CDN experienceAWS Solutions Architect / DevOps Engineer β Professional certificationDistributed-systems background, OSS infrastructure contributions