Senior DevOps / SRE Engineer

Intellecttinc — United States · Posted ~2 hours ago

Senior Hybrid

Skills

Azure AKS Scalability Monitoring Deployment Linux Microservices NoSQL databases Distributed systems Troubleshooting Python Terraform Cassandra MongoDB PostgreSQL Databricks Kafka Event Hub NATS Java Datadog Splunk

🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.

Log in to add to target list

Summary ✨ AI‑Generated

Join a senior engineering team responsible for the reliability and operation of cloud-based microservices. Work with Kubernetes, distributed systems, NoSQL storage, messaging platforms, infrastructure as code, production monitoring, and observability tooling. The role requires strong Linux expertise and advanced troubleshooting skills, with regular office presence.

Highlights

Senior SRE/DevOps role with deep exposure to Azure cloud, Kubernetes, distributed systems, microservices, databases, observability, production monitoring, and troubleshooting in a technically demanding environment.

Description

SRE / Sr. DevOps Engineer Seattle based Client Hybrid working is needed -3 days office in Seattle-WA -Azure Cloud, AKS – Scalability, monitoring, deployment, check logs, ensure node and pod health. -Databases include - Cassandra, Mongo, PostGres, NoSQL -Databricks Notebooks – experience with Databricks to know how a notebook is created and run - run queries against the database and finding discrepancies and perform fixes. -Experience with using Kafka, Event Hub, NATS or any messaging broker. -JAVA Based microservices, responsible for deployment, scripting language is python. Should have an understanding around terraform. -Emphasis on Logs and Monitoring (datadog and splunk) We are seeking an experienced, self-motivated Senior Engineer who is technically very strong with strong Linux background, with deep knowledge in micro services, backend storage design, NoSQL database, distributed systems and very good troubleshooting skills. Typical activities include production monitoring, creating monitoring dashboards, setting up alerts, triaging alerts coupled with the ability to drive efforts and solution improvements effectively across various IT and business functions. In this role, person will be responsible for setting up monitoring dashboards, alerts, maintaining production systems, deploying code in Production, monitoring alerts, resolving issues, and leading production troubleshooting calls. Working with Product Owners and other developers to implement highly scalable reactive application platform solutions in Cloud based Linux environments. Summary of Experience • Requires 10+ years experience in the IT industry • Requires 10+ years of software and DevOps development engineering • Experience in working with cloud environment Azure preferred. • Experience with Kubernetes, Azure Kubernetes (AKS) preferred. • Experience with using Kafka, Event Hub, NATS or any messaging broker. • Experience with Cassandra, PostgresSQL, Mongo, Elastic Search, Cosmos DB • Experience on Azure DevOps, Jenkins/ Python / Terraform / Ansible • Experience with Databricks • Experience with DataDog, Splunk or other logging and APM tools. • Experience in working with Linux environment. • Experience building complex, scalable, high-performance software systems that have been successfully delivered to customers • Demonstrated knowledge of best practices for the design and implementation of large-scale systems as well as experience in taking such systems from design to production • Experience building and operating mission critical, highly available (24x7) systems