Site Reliability Engineer / Production Engineer

Hunter Bond — United States · Posted ~8 hours ago

Mid Full-time Hybrid

Skills

Linux distributed systems production infrastructure automation reliability engineering performance optimization monitoring observability

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

Join a high-performing infrastructure team as a hands-on reliability engineer focused on keeping critical, high-scale systems dependable and fast. You will automate operations, strengthen observability, resolve complex production issues, and collaborate closely with software engineers to build resilient services.

Highlights

Exceptional compensation, strong bonus potential, excellent benefits, and the opportunity to work on highly available systems at significant scale with experienced engineering teams.

Description

Job Title: Site Reliability Engineer / Production Engineer Sector: Fintech Location: NYC (Hybrid) Salary: Up to $400k USD Base + market-leading bonus + outstanding benefits My client is looking for a hands-on Site Reliability Engineer to join a high-performing infrastructure engineering team responsible for the reliability, automation and performance of critical trading platforms. You'll work across Linux, distributed systems and production infrastructure, partnering with software engineers to improve resilience, automate operations and solve complex technical challenges at scale. Role Build and improve highly available production infrastructureAutomate operational processes and eliminate manual tasksWork on complex production, reliability and performance issuesImprove monitoring, observability and platform reliabilityWork closely with engineering teams to deliver resilient applicationsSkills / Experience 1 - 20 years of experience in an SRE/Production Engineering role (Hiring at all levels)Strong Linux systems experienceGood Python or Go development skills for automation and toolingExperience supporting distributed systems in productionKnowledge of Kubernetes, networking and CI/CD in an on-prem environmentStrong understanding of monitoring and observability platformsExcellent problem-solving skills with an engineering-first mindsetWhy Apply? Work on large-scale, business-critical trading infrastructureJoin an engineering-led environment with genuine technical ownershipSolve complex reliability and performance challenges every dayOutstanding compensation and career progressionSells Flexible hours/work optionsThe chance to work with some of the most cutting-edge technology in the IndustryTechnologists only report to technologistsTechies treated as top commodity so ‘spoiled’Small team size in a growing company so an ability to make a significant impactConstantly exciting greenfield projects in an ever-evolving environmentNo red tapeBeautiful offices