DevOps Engineer

Stelvioinc — United States · Posted ~1 hour ago

Mid Full-time Onsite

Skills

DevOps Cloud Infrastructure Kubernetes CI/CD Infrastructure as Code Monitoring Automation Microservices Event-Driven Architecture Production Support Cloud Reliability AWS Azure Google Cloud Event Streaming Distributed SQL

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

A DevOps Engineer is sought to automate and improve infrastructure, development, and deployment processes in a cloud-native, event-driven microservices environment. Responsibilities include managing cloud platforms, Kubernetes, CI/CD, infrastructure as code, monitoring, automation, reliability, security, performance, cost optimization, and participation in a 24/7 production on-call rotation.

Highlights

Full-time onsite DevOps role focused on scalable cloud infrastructure, Kubernetes, CI/CD, infrastructure as code, observability, automation, and reliability for business-critical production systems.

Description

DevOps Engineer Frisco, TX | On-site | Full-Time We are working with a client that develops and operates complex technology platforms where scalability, reliability, and availability are critical. They are looking for a DevOps Engineer to help automate and improve the infrastructure, development, and deployment processes supporting a cloud-native, event-driven microservices environment. You’ll work across cloud infrastructure, Kubernetes, CI/CD, Infrastructure as Code, monitoring, and automation, partnering closely with software development, QA, and IT teams. The role also participates in a 24/7 on-call rotation supporting business-critical production systems. Responsibilities Design, deploy, and maintain scalable infrastructure across AWS, Azure, or Google CloudManage cloud environments with a focus on availability, reliability, security, performance, and costOperate Kubernetes platforms, including stateful workloads, distributed SQL databases, and event-streaming technologies such as Kafka or NATSBuild and maintain containerized CI/CD pipelines using tools such as Dagger, GitHub Actions, or GitLab CIAutomate build, testing, and deployment processes across applications and servicesImplement Infrastructure as Code using Terraform or Ansible, alongside Helm and GitOps workflowsManage Kubernetes deployments using tools such as Argo CDAutomate infrastructure provisioning, configuration, and scaling to create consistent and repeatable environmentsManage containerized applications using Docker and KubernetesImplement and maintain monitoring, logging, tracing, and alerting using technologies such as Prometheus, Grafana, OpenTelemetry, and ELKMonitor production systems, troubleshoot performance issues, and respond to incidents as part of the on-call rotationWork closely with developers, system administrators, and QA engineers to improve deployment and operational processesIntegrate security practices into infrastructure and deployment workflows, including secrets management, access controls, and policy-as-codeSupport compliance with relevant standards including SOC 2, ISO 27001, and PCI DSS Qualifications Bachelor’s degree in Computer Science, Information Technology, Software Engineering, or a related field, or equivalent practical experience3+ years of experience within DevOps, Site Reliability Engineering (SRE), or a similar roleStrong hands-on experience with at least one major cloud platform: AWS, Azure, or Google CloudExperience designing, building, and maintaining CI/CD pipelinesHands-on experience with CI/CD technologies such as Dagger, GitHub Actions, GitLab CI, or JenkinsStrong Infrastructure as Code experience using technologies such as Terraform, Ansible, and HelmExperience with GitOps deployment approaches and tools such as Argo CDStrong knowledge of Docker and KubernetesAutomation and scripting experience using Python, Bash, or PowerShellExperience with monitoring and observability technologies such as Prometheus, Grafana, OpenTelemetry, or ELKStrong knowledge of Git, GitHub, or GitLabStrong troubleshooting skills across infrastructure, deployments, and production systemsAbility to collaborate effectively across development, QA, operations, and infrastructure teams Benefits Paid vacation, sick leave, and bereavement leaveHealth and dental plansRetirement plansEmployee and Family Assistance Program (EFAP)Employee referral program