Site Reliability Engineer III

Jpmorganchase — United States · Posted ~3 hours ago

Senior Full-time Visa History ✓

Skills

SRE cloud infrastructure CI/CD application monitoring reliability engineering cloud monitoring automation

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

A large technology environment is seeking an SRE professional to improve availability, automate operations, manage cloud infrastructure, and enhance reliability of critical applications.

Highlights

Experienced reliability engineering role focused on mission-critical systems, automation, scalability, and improving application operations.

Description

Job Description There's nothing more exciting than being at the center of a rapidly growing field in technology and applying your skillsets to drive innovation and modernize the world's most complex and mission-critical systems. As a Site Reliability Engineer III at JPMorgan Chase within the Asset and Wealth Management team, you will solve complex and broad business problems with simple and straightforward solutions. Through code and cloud infrastructure, you will configure, maintain, monitor, and optimize applications and their associated infrastructure to independently decompose and iteratively improve on existing solutions. You are a significant contributor to your team by sharing your knowledge of end-to-end operations, availability, reliability, and scalability of your application or platform. Job Responsibilities Supports and collaborates with other engineers to design, develop, and implement deployment and reliability approaches using automated CI/CD pipelinesImplements infrastructure, configuration, and network as code for applications and platforms in your remitContributes to observability improvements including white and black box monitoring, service level objective alerting, and telemetry collectionAssists in identifying and resolving complex problems by using service level indicators and objectives to proactively address issues before they impact customersParticipates in incident response, triage, and post-incident analysis; helps document findings and remediation actions to prevent recurrenceIdentifies opportunities to eliminate or automate remediation of recurring issues to reduce toil and improve overall operational stabilityUses enterprise-authorized AI capabilities to accelerate incident triage, troubleshooting, and post-incident analysis, validating outputs and handling operational data according to sensitivity and security requirementsProactively recognizes roadblocks and identifies improvements to solve operational problems, including exploring new technologies where appropriateDocuments and shares knowledge within your organization via internal forums and communities of practiceSupports adoption of site reliability engineering best practices within your team Required Qualifications, Capabilities, And Skills Formal training or certification on site reliability engineering concepts and 3+ years applied experience Foundational understanding of SRE culture and principles, including Service Level Indicators (SLIs) and Service Level Objectives (SLOs)General observability and monitoring understanding with working experience using industry-standard tooling (e.g., Grafana, Dynatrace, Prometheus, Datadog, Splunk, CloudWatch)Proficiency in at least one scripting or programming language such as Python, Bash, or similar for tool development and operational supportExperience with incident and response management, including on-call participation and structured post-incident reviewKnowledge of CI/CD pipelines and best practices using tools such as Jenkins, GitLab CI, or similarSkills in automating repetitive tasks and managing configurations at scale using tools such as Ansible, Terraform, or similarFamiliarity with container technologies and container orchestration (e.g., Docker, Kubernetes)Working knowledge of using enterprise-authorized AI capabilities to support SRE workflows, with strong validation habits and awareness of data sensitivityAbility to validate AI-assisted operational recommendations before applying changes, escalating when uncertain and following data sensitivity requirements Preferred Qualifications, Capabilities, And Skills Experience operating in cloud environments (AWS and/or Azure), including understanding of resiliency, scalability, and observability patternsFamiliarity with Kubernetes ingress, networking, and certificate deployment patternsExperience improving infrastructure-as-code patterns (e.g., Terraform modules, reusable configurations)Understanding of controls-focused operations in regulated environments, including change management discipline and audit supportExperience with service mesh, load balancing, and DNS troubleshooting (e.g., ALB/NLB, Route 53)Drive to self-educate and evaluate emerging technologies in the SRE and cloud-native spaceStrong communication skills with the ability to collaborate across different levels and stakeholder groups ABOUT US JPMorganChase, one of the oldest financial institutions, offers innovative financial solutions to millions of consumers, small businesses and many of the world's most prominent corporate, institutional and government clients under the J.P. Morgan and Chase brands. Our history spans over 200 years and today we are a leader in investment banking, consumer and small business banking, commercial banking, financial transaction processing and asset management. We offer a competitive total rewards package including base salary determined based on the role, experience, skill set and location. Those in eligible roles may receive commission-based pay and/or discretionary incentive compensation, paid in the form of cash and/or forfeitable equity, awarded in recognition of individual achievements and contributions. We also offer a range of benefits and programs to meet employee needs, based on eligibility. These benefits include comprehensive health care coverage, on-site health and wellness centers, a retirement savings plan, backup childcare, tuition reimbursement, mental health support, financial coaching and more. Additional details about total compensation and benefits will be provided during the hiring process. We recognize that our people are our strength and the diverse talents they bring to our global workforce are directly linked to our success. We are an equal opportunity employer and place a high value on diversity and inclusion at our company. We do not discriminate on the basis of any protected attribute, including race, religion, color, national origin, gender, sexual orientation, gender identity, gender expression, age, marital or veteran status, pregnancy or disability, or any other basis protected under applicable law. We also make reasonable accommodations for applicants' and employees' religious practices and beliefs, as well as mental health or physical disability needs. Visit our FAQs for more information about requesting an accommodation. JPMorgan Chase & Co. is an Equal Opportunity Employer, including Disability/Veterans About The Team J.P. Morgan Asset & Wealth Management delivers industry-leading investment management and private banking solutions. Asset Management provides individuals, advisors and institutions with strategies and expertise that span the full spectrum of asset classes through our global network of investment professionals. Wealth Management helps individuals, families and foundations take a more intentional approach to their wealth or finances to better define, focus and realize their goals.