Site Reliability / DevOps Engineer

Insight Global — United States · Posted ~2 hours ago

Contract Hybrid $60-$70/hour

Skills

DevOps site reliability engineering Linux Windows Server automation scripting monitoring alerting distributed systems network troubleshooting database troubleshooting system performance high-availability systems

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

A Site Reliability/DevOps Engineer is sought for a 12-month hybrid engagement supporting large-scale enterprise systems. The role covers infrastructure administration, automation, monitoring and alerting, Linux and Windows environments, distributed systems, troubleshooting, and reliability improvements.

Highlights

12-month hybrid contract with strong exposure to enterprise infrastructure, automation, monitoring, high-availability distributed systems, and reliability engineering.

Description

Title: DevOps Engineer Pay Rate: $60-70/h Duration: 12 month contract Location: Hybrid (Westlake, TX preferred; Austin, TX secondary) Onsite: Mondays, Wednesdays, ThursdaysRemote: Tuesdays and FridaysCandidate will work remotely until a seat becomes available onsiteInterview: 2 rounds Day-to-Day: Administer, support, monitor, and deploy enterprise systemsDevelop automation scripts to improve operational efficiencyBuild application monitoring dashboards and alerting solutionsProactively identify and resolve infrastructure and application issuesSupport Linux and Windows server environmentsMonitor high-availability, large-scale distributed systemsTroubleshoot networking, database, and system performance issuesCollaborate with teams to improve system reliability and operational processesParticipate in SDLC practices and process improvement initiativesDrive issue resolution within a complex trading ecosystem Job Description: Insight Global is seeking a Site Reliability Engineer (SRE)/DevOps for a top financial services client. This candidate will be responsible for supporting and maintaining enterprise-level infrastructure within a large-scale, high-availability environment. The ideal candidate will leverage strong systems administration, automation, monitoring, and troubleshooting skills to ensure platform reliability and performance. This role requires hands-on experience across Windows and Linux environments, monitoring tools, scripting languages, and networking technologies. The candidate will proactively build monitoring solutions, automate operational processes, and investigate complex issues across a distributed trading ecosystem while maintaining a strong sense of ownership and customer focus. Must-Haves: • 6-8 years of experience with enterprise level administration and support • 6-8 years of experience in writing automation scripts, building application dashboards for proactive monitoring, setting up alerts for early determination of the issues in Grafana, Influx, Datadog, Moog, ThousandEyes etc.. • 6-8 years of experience practicing SDLC (Software Development Lifecycle) practice, process improvements • Hands on enterprise systems administration, monitoring, and deployment activities • Experience with Windows 2016, 2019, 2022 hosted via Virtual Machine • Knowledge of IP networking including DNS, DHCP, firewalls, IP routing, etc. • Familiarity with large scale distributed systems and high-availability architectures • Linux and Windows system administration, troubleshooting, and tuning • Development experience in one or more or programming languages such as .Net, C#, Powershell, Yaml, Java, Python, Bash • Knowledge of one or more of SQL, Oracle, MongoDB, Postgres databases • Bachelor's degree in computer science or related discipline Plusses: • Financial services industry experience • Agile methodologies