Site Reliability Engineer - Linux

Fintalpartners — United States · Posted ~3 hours ago

Senior Full-time

Skills

Linux Site reliability engineering Python Automation Production operations Troubleshooting Observability Incident response Networking Routing

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

A Site Reliability Engineer is sought to build, automate, and maintain highly available Linux production environments supporting low-latency financial systems. You will develop Python automation, troubleshoot production issues, improve observability and deployment processes, support networking, and strengthen reliability, scalability, and performance.

Highlights

Support highly available, performance-critical Linux infrastructure for real-time financial systems. The role combines automation, production troubleshooting, observability, networking, and reliability engineering while working closely with software developers and infrastructure teams.

Description

We're partnering with a leading proprietary trading firm that's looking to hire a Site Reliability Engineer to support and optimize the infrastructure powering its low-latency trading platforms. This is an opportunity to work in an environment where reliability, automation, and performance are mission-critical, collaborating closely with software engineers, traders, and infrastructure teams to keep production systems running at peak efficiency. What You'll Do Build, automate, and maintain highly available Linux-based production environments.Develop automation tools and operational workflows using Python.Monitor, troubleshoot, and resolve production issues across mission-critical trading infrastructure.Improve system reliability, scalability, and performance through automation and continuous improvement.Work closely with developers to enhance deployment processes, observability, and incident response.Support networking infrastructure, including troubleshooting connectivity, routing, DNS, and latency-related issues.Participate in an on-call rotation for production support. What We're Looking For Strong experience administering Linux production environments.Solid Python scripting skills.Solid understanding of networking fundamentals, including TCP/IP, DNS, routing, switching, and network troubleshooting.Experience supporting high-availability, distributed systems in a production environment.Familiarity with monitoring and observability tools such as Prometheus, Grafana, Splunk, or similar.Strong troubleshooting skills and the ability to perform under pressure in a fast-paced environment.Experience in financial services or trading environments is advantageous but not required. Why Join? Work on some of the most performance-sensitive infrastructure in the industry.Collaborate with world-class engineers solving complex technical challenges.Competitive compensation, annual bonus, and outstanding benefits.Significant opportunities for career progression and technical ownership.Modern technology stack with a strong engineering-first culture.