DevOps Engineer / Site Reliability Engineer

Smart It Frame Llc — United States · Posted ~4 hours ago

Mid Full-time Onsite

Skills

AWS CI/CD Python Java Cloud infrastructure Troubleshooting

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

A DevOps and reliability engineering role responsible for maintaining cloud environments, improving deployment pipelines, resolving system issues, and ensuring stable production operations.

Highlights

Opportunity to improve production reliability, automate deployments, and work on cloud infrastructure and operational excellence.

Description

Role - DevOps Engineer / SRE (Java/Python, AWS, CI/CD) Location - Reston, VA Duration - Fulltime Job Description Overview We are seeking a highly skilled DevOps Engineer/Site Reliability Engineer (SRE) to join our team. The ideal candidate will be responsible for identifying application faults, bugs, and system issues, conducting root cause analysis, and collaborating with development teams to implement effective solutions. The candidate should have strong experience in DevOps practices, cloud infrastructure, automation, and troubleshooting production environments. Key Responsibilities Identify, analyse, and troubleshoot application bugs, system failures, and performance issues. Perform root cause analysis and work closely with development teams to drive issue resolution. Monitor application and infrastructure health, ensuring high availability and reliability. Build, maintain, and optimize CI/CD pipelines for automated deployments. Manage and support cloud infrastructure on AWS. Implement Infrastructure as Code (IaC) using Terraform. Configure and maintain Jenkins and other DevOps automation tools. Required Skills & Qualifications 4+ years of experience in DevOps, SRE, Production Support, or related roles. Strong programming experience in Java or Python (experience with both is preferred). Hands-on experience with AWS Cloud Services. Experience with Terraform for Infrastructure as Code. Expertise in troubleshooting application defects, production incidents, and performance bottlenecks. Experience with Linux/Unix administration and shell scripting.