DevOps Engineer / Site Reliability Engineer

Apriden — United States · Posted ~10 hours ago

Mid Full-time

Skills

DevOps AWS CI/CD Cloud infrastructure Troubleshooting Java Python

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

A technology team is hiring a DevOps engineer to improve application reliability, automate deployments, manage cloud infrastructure, and collaborate on production troubleshooting.

Highlights

Role focused on cloud operations, automation, reliability improvements, and solving complex production issues.

Description

DevOps Engineer / SRE (Java/Python, AWS, CI/CD) Reston, VA Job Description Overview We are seeking a highly skilled DevOps Engineer/Site Reliability Engineer (SRE) to join our team. The ideal candidate will be responsible for identifying application faults, bugs, and system issues, conducting root cause analysis, and collaborating with development teams to implement effective solutions. The candidate should have strong experience in DevOps practices, cloud infrastructure, automation, and troubleshooting production environments. Key Responsibilities • Identify, analyze, and troubleshoot application bugs, system failures, and performance issues. • Perform root cause analysis and work closely with development teams to drive issue resolution. • Monitor application and infrastructure health, ensuring high availability and reliability. • Build, maintain, and optimize CI/CD pipelines for automated deployments. • Manage and support cloud infrastructure on AWS. • Implement Infrastructure as Code (IaC) using Terraform. • Configure and maintain Jenkins and other DevOps automation tools. • Investigate repository-related issues, deployment failures, and source code management problems. • Collaborate with cross-functional teams including Development, QA, Operations, and Security. • Automate operational processes using Java or Python scripting. • Ensure best practices in system monitoring, logging, release management, and incident response. Required Skills & Qualifications • 4+ years of experience in DevOps, SRE, Production Support, or related roles. • Strong programming experience in Java or Python (experience with both is preferred). • Hands-on experience with AWS Cloud Services. • Experience with Terraform for Infrastructure as Code. • Strong knowledge of Jenkins and CI/CD pipeline implementation. • Expertise in troubleshooting application defects, production incidents, and performance bottlenecks. • Understanding of source code repositories and repository-related issues (Git, Bitbucket, GitHub, etc.). • Experience with Linux/Unix administration and shell scripting. • Strong analytical, debugging, and problem-solving skills. Preferred Qualifications • Experience with Docker and Kubernetes. • Familiarity with monitoring tools such as CloudWatch, Prometheus, Grafana, or Splunk. • Knowledge of Agile and DevOps methodologies. • AWS or DevOps-related certifications are a plus.