Site Reliability Engineer

Insight Global — United States · Posted ~4 hours ago

Mid Full-time

Skills

site reliability engineering monitoring logging tracing distributed systems data pipelines Python Grafana ELK Splunk Dynatrace Terraform

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

A technology services organization is looking for a Site Reliability Engineer to design observability solutions, maintain production reliability, and support scalable data platforms and distributed systems.

Highlights

Opportunity to build reliable large-scale systems, improve observability, and work with modern infrastructure and data technologies.

Description

Insight Global is hiring for a Secret Site Reliability Engineer to support a federal client in Charleston, SC. If you have experience in designing, building and maintaining telemetry, logging, and tracing systems, this role will be a great next step. Key Responsibilities: Build and maintain monitoring, logging, tracing, metrics, alerting, and dashboarding solutions for data platforms and pipelines.Implement observability tools such as Grafana, Elastic/ELK, Splunk, and Dynatrace to improve platform visibility and reliability.Monitor system health, performance, availability, latency, throughput, and storage utilization across distributed environments.Troubleshoot complex application, infrastructure, networking, and data-flow issues in production systems.Design, develop, and support batch and real-time data pipelines to ensure reliable data processing and delivery.