Summary
✨ AI‑Generated
A technology team is hiring a DevOps engineer to improve application reliability, automate deployments, manage cloud infrastructure, and collaborate on production troubleshooting.
Highlights
Role focused on cloud operations, automation, reliability improvements, and solving complex production issues.
Description
DevOps Engineer / SRE (Java/Python, AWS, CI/CD)
Reston, VA
Job Description
Overview
We are seeking a highly skilled DevOps Engineer/Site Reliability Engineer (SRE) to join our team.
The ideal candidate will be responsible for identifying application faults, bugs, and system issues, conducting root cause analysis, and collaborating with development teams to implement effective solutions.
The candidate should have strong experience in DevOps practices, cloud infrastructure, automation, and troubleshooting production environments.
Key Responsibilities
• Identify, analyze, and troubleshoot application bugs, system failures, and performance issues.
• Perform root cause analysis and work closely with development teams to drive issue resolution.
• Monitor application and infrastructure health, ensuring high availability and reliability.
• Build, maintain, and optimize CI/CD pipelines for automated deployments.
• Manage and support cloud infrastructure on AWS.
• Implement Infrastructure as Code (IaC) using Terraform.
• Configure and maintain Jenkins and other DevOps automation tools.
• Investigate repository-related issues, deployment failures, and source code management problems.
• Collaborate with cross-functional teams including Development, QA, Operations, and Security.
• Automate operational processes using Java or Python scripting.
• Ensure best practices in system monitoring, logging, release management, and incident response.
Required Skills & Qualifications
• 4+ years of experience in DevOps, SRE, Production Support, or related roles.
• Strong programming experience in Java or Python (experience with both is preferred).
• Hands-on experience with AWS Cloud Services.
• Experience with Terraform for Infrastructure as Code.
• Strong knowledge of Jenkins and CI/CD pipeline implementation.
• Expertise in troubleshooting application defects, production incidents, and performance bottlenecks.
• Understanding of source code repositories and repository-related issues (Git, Bitbucket, GitHub, etc.).
• Experience with Linux/Unix administration and shell scripting.
• Strong analytical, debugging, and problem-solving skills.
Preferred Qualifications
• Experience with Docker and Kubernetes.
• Familiarity with monitoring tools such as CloudWatch, Prometheus, Grafana, or Splunk.
• Knowledge of Agile and DevOps methodologies.
• AWS or DevOps-related certifications are a plus.