Site Reliability Engineer

New York Technology Partners — United States · Posted ~3 hours ago

Mid Full-time Onsite

Skills

AWS Infrastructure Incident Management On-call Operations

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

A platform engineering role focused on maintaining reliable infrastructure, leading incident response, participating in on-call rotations, and improving operational excellence.

Highlights

Develop reliability expertise while working on cloud infrastructure, incident response, and collaborative engineering practices.

Description

Candidates must be comfortable working onsite 5x a week We are looking for a high-energy, enthusiastic Platform Engineer with a desire to refine their craft. The ideal candidate will be eager to contribute, demonstrate exceptional technical aptitude, and bring a positive attitude to our collaborative environment in the heart of Chicago. In this role, you can expect to: • Drive Incident Management: Lead and cultivate a culture of incident management within the team. • Participate in On-Call Duties: Contribute to incident command and on-call rotations. • Build Technical Skills: Develop your technical expertise while working closely with a team of skilled engineers. • Collaborate Cross-Functionally: Engage in learning, teaching, and collaborating with different teams across the organization. You may be a good fit for our team if you have: • Professional AWS Experience: Proven experience in implementing and maintaining scalable and reliable infrastructure on Amazon Web Services (AWS). • Diverse Technical Interests: Enjoy working on a range of scopes including software engineering, cloud infrastructure, DevSecOps, and SRE. • Efficiency Improvement Skills: A track record of driving efficiency improvements in software at scale. • Cross-Functional Collaboration: Experience working cross-functionally to promote and implement engineering culture changes. • SaaS or Managed Software Experience: Hands-on experience with Software as a Service (SaaS) or other managed software offerings. • Public Cloud Expertise: Expertise in one or more major public cloud platforms. Qualifications: • Educational Background: Degree in Computer Science, Engineering, or a related field is preferred. • Linux Systems and CI/CD: Deep knowledge of Linux systems and CI/CD tools such as Jenkins. • Development and Automation: Experience in developing applications, automation tools, and the necessary infrastructure for large environments. Experience developing and maintaining infrastructure as code (IaC) templates using Terraform. • Security Proficiency: Well-versed in vulnerability management and security products. • Troubleshooting Skills: Expertise in troubleshooting using monitoring tools. • IT Operations Knowledge: Understanding of IT Operations best practices in always-up environments. • Strong written and verbal communication abilities. • Containerization Experience: Familiarity with containerization technologies such as Docker and Kubernetes. Our Ideal Candidate: • Develops and Automates: Has created applications and automation tools to build, deploy, monitor, and integrate data sources for our systems and applications. • Collaborates and Engages: Enjoys working closely with multiple teams, fostering a collaborative environment.