Senior SRE Forward Deployed Engineer

E Solutions Global — United States · Posted ~2 hours ago

Senior Full-time Onsite

Skills

SRE observability event correlation infrastructure support AI engineering ServiceNow AI agents

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

A senior hands-on engineering position responsible for improving reliability, reducing operational noise, analyzing production issues, and guiding architecture decisions.

Highlights

Senior technical leadership role focused on reliability improvements, operational excellence, and AI-enabled infrastructure solutions.

Description

SRE Forward-Deployed Engineer (AIOPS) Location: Irving, TX / Charlotte, NC (100% Onsite) Experience Level : 10+ years Skills Required: SRE, Observability, Event Correlation, Alert Noise Reduction, Infrastructure & Application Support, ServiceNow, Confluence, AI/Agentic Engineering, and hands-on experience with Cursor/Devin, coupled with strong solution architecture and stakeholder management capabilities. Role Overview We are looking for a senior, hands-on Forward-Deployed Engineer (FDE) to act as a technical leader and trusted advisor, driving observability, SRE, infrastructure, and AI-enabled operational excellence initiatives. The ideal candidate will deliver immediate solutions while defining long-term architecture, best practices, and roadmaps. Key Responsibilities • Drive SRE and observability initiatives to improve reliability and operational efficiency. • Analyze production events, perform root cause analysis, and reduce alert noise through event correlation. • Partner with infrastructure, application support, and operations teams to enhance service health and resilience. • Design and implement automation, monitoring, and AI-driven operational solutions. • Define architecture, standards, and strategic roadmaps for observability and intelligent operations. • Mentor teams and act as a technical leader across engineering and operations functions. Required Skills • Strong experience in Site Reliability Engineering (SRE) and production operations. • Expertise in observability, monitoring, logging, tracing, and event correlation. • Deep knowledge of enterprise infrastructure, cloud platforms, and production services. • Hands-on experience in application support and operational excellence. • Proficiency with ServiceNow and Confluence. • Experience with AI-enabled engineering, agentic workflows, and automation. • Hands-on experience with tools such as Cursor and Devin. • Ability to balance hands-on execution with strategic architecture and roadmap planning. Preferred • Experience with AIOps, DevOps, CI/CD, and infrastructure automation. • Strong stakeholder management and communication skills.