DevOps & Site Reliability Engineer
Voltagrid โ United States ยท Posted ~1 day ago
๐ Log in to save this job, tailor your resume & track your apply process โ 7 days free, no card needed.
Log in to add to target listDescription
Position Title: DEVOPS & SRE ENGINEER
Location: HOUSTON, TX
FLSA Class: EXEMPT
Responsible to: Directo of Software Engineering
Position Summary: DevOps / Site Reliability Engineer to implement and evolve the infrastructure, deployment pipelines, and reliability posture of our systems.
You'll work closely with engineering teams to build scalable, observable, and resilient infrastructure while driving a culture of operational excellence.
Essential Duties And Responsibilities
Design, build, and maintain cloud infrastructureManage and optimize Kubernetes clusters and containerized workloads in productionDevelop and maintain infrastructureascode using Terraform (or equivalent tooling)Build and improve CI/CD pipelines to enable fast, safe, and reliable deploymentsImplement and maintain monitoring, alerting, and observability systems (Prometheus, Grafana, Datadog, or similar)Define and track SLIs/SLOs, participate in incident response, root cause analysis, and blameless postmortemsIdentify and eliminate toil through automation and selfservice toolingConfigure and maintain onprem baremetal servers and Linuxbased infrastructureConfigure, maintain, and optimize virtualized assetsCollaborate with development teams on system design, capacity planning, and performance optimizationParticipate in oncall rotations and ensure production readiness of new services
Other Requirements
4+ years of experience in DevOps, SRE, or infrastructure engineering rolesStrong experience with at least one major cloud provider (AWS, GCP, or Azure AWS preferred)Deep hands-on experience with Kubernetes and Docker in production environmentsProficiency with infrastructureascode tools, particularly TerraformExperience building and maintaining CI/CD pipelines (GitHub Actions, GitLab CI, Jenkins, or similar)Solid understanding of monitoring and observability (metrics, logs, traces)Strong scripting skills (Bash, Python, or Go)Experience with incident management, SLObased reliability practices, and capacity planningStrong Linux systems administration skills (Ubuntu, RHEL/CentOS, or similar)Experience with virtualization platforms including VM provisioning, storage, networking, and cluster managementSolid understanding of networking, DNS, load balancing, and security fundamentals
Nice To Have
Contributions to internal developer platforms or platform engineering initiativesProxmox VE experienceCertifications in cloud platforms (AWS SA, CKA, etc.)
The above statements are intended to describe the general nature and level of work being performed by employees assigned to this classification.
All personnel may be required to perform duties outside of their normal responsibilities from time to time, as needed.
VoltaGrid is an Equal Opportunity Employer that does not discriminate on the basis of actual or perceived race, creed, color, religion, alienage or national origin, ancestry, citizenship status, age, disability or handicap, sex, marital status, veteran status, sexual orientation, genetic information, arrest record, or any other characteristic protected by applicable federal, state or local laws.
Our management team is dedicated to this policy with respect to recruitment, hiring, placement, promotion, transfer, training, compensation, benefits, employee activities, and general treatment during employment.
We have 68,012 jobs that might be an even better fit for you
DontApply's real value goes far beyond a single job link or company name. Just upload your resume โ in under a minute we'll analyze all 68,012 jobs and tell you exactly which ones you should apply to right now.
Upload My Resume