Senior Staff Site Reliability Engineer, AIOps

Palo Alto Networks — United States · Posted ~1 day ago

Lead Visa History ✓

Skills

Site reliability engineering AIOps Artificial intelligence Collaboration

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

A Senior Staff Site Reliability Engineer specializing in AIOps is sought to apply AI-driven approaches to large-scale reliability and operational challenges. You will work on meaningful engineering problems, collaborate with highly skilled technical teams, and help integrate AI into reliability practices while driving innovation and technical impact.

Highlights

Senior technical opportunity focused on applying AI to reliability engineering and solving real-world technology problems. The role emphasizes collaboration, innovation, meaningful technical impact, and working alongside highly experienced engineering professionals.

Description

Our Mission At Palo Alto Networks®, we’re united by a shared mission—to protect our digital way of life. We thrive at the intersection of innovation and impact, solving real-world problems with cutting-edge technology and bold thinking. Here, everyone has a voice, and every idea counts. If you’re ready to do the most meaningful work of your career alongside people who are just as passionate as you are, you’re in the right place. Who We Are In order to be the cybersecurity partner of choice, we must trailblaze the path and shape the future of our industry. This is something our employees work at each day and is defined by our values: Disruption, Collaboration, Execution, Integrity, and Inclusion. We weave AI into the fabric of everything we do and use it to augment the impact every individual can have. If you are passionate about solving real-world problems and ideating beside the best and the brightest, we invite you to join us! We believe collaboration thrives in person. That’s why most of our teams work from the office full time, with flexibility when it’s needed. This model supports real-time problem-solving, stronger relationships, and the kind of precision that drives great outcomes. Job Summary Your Career Palo Alto Networks runs a large hybrid infrastructure and is one of the largest GCP customers. As a Site Reliability Engineer, you will be part of a team supporting the services running on this infrastructure. This includes automation, architecture, performance, metrics, troubleshooting, security, and reliability. Our stack includes Kubernetes, Docker, GCP, AWS, Ansible, Terraform, Vault, Gitlab, Spinnaker, Tensorflow, Datadog, Elasticsearch, Kafka, Hadoop, MySQL, Percona, MongoDB, Python, and Go. We don’t expect you to know all these, but we do expect you to learn the ones needed for this role. Your Impact Contribute to the success of SRE and DevOpsDevelop expertise in new technologiesWork with developers, researchers, data scientists, and security expertsDesign, build and operate reliable, secure Cloud infrastructureEnsure that applications are production-ready, scalable, and reliableDevelop tools and automation frameworksAutomate robust deployment of robust servicesOrchestrate end-to-end monitoring and alertingParticipate with SRE and Dev teams in the on-call rotationLead root cause analysis of critical business and production issuesMentor and champion SRE cultureParticipate in design reviews Qualifications Your Experience BS or MS in Computer Science, a related field, or equivalent professional experienceExpertise in configuration management with a framework such as Ansible, Terraform, HelmExperience in Production Engineering, DevOps, or Site ReliabilityExpertise in private or public cloudStrong Linux administration, internals, and network troubleshootingProficiency with programming languages like Python, Golang, and shell scripting to automate tasksFamiliarity with CI/CD pipelines, GitLab and GitHub preferredAbility to diagnose and troubleshoot complex distributed systems handling high volume transactionsExcellent written and verbal communication, able to collaborate and rally supportSelf-disciplined, self-managed, self-motivated and strong sense of ownership, urgency, and drivePassion for infrastructure and monitoring as codeReady to understand and dissect new technology stacks quickly Compensation Disclosure The compensation offered for this position will depend on qualifications, experience, and work location. For candidates who receive an offer at the posted level, the starting base salary (for non-sales roles) or base salary + commission target (for sales/com-missioned roles) is expected to be the annual range listed below. The offered compensation may also include restricted stock units and a bonus. A description of our employee benefits may be found here. $151,600.00 - $245,300.00/yr Our Commitment We’re trailblazers that dream big, take risks, and challenge cybersecurity’s status quo. It’s simple: we can’t accomplish our mission without diverse teams innovating, together. We are committed to providing reasonable accommodations for all qualified individuals with a disability. If you require assistance or accommodation due to a disability or special need, please contact us at accommodations@paloaltonetworks.com. Palo Alto Networks is an equal opportunity employer. We celebrate diversity in our workplace, and all qualified applicants will receive consideration for employment without regard to age, ancestry, color, family or medical care leave, gender identity or expression, genetic information, marital status, medical condition, national origin, physical or mental disability, political affiliation, protected veteran status, race, religion, sex (including pregnancy), sexual orientation, or other legally protected characteristics. All your information will be kept confidential according to EEO guidelines. Is role eligible for Immigration Sponsorship?: Yes