Principal DevOps Engineer - Azure
Tenex Ai — United States · Posted ~2 hours ago
🔓 Log in to save this job, tailor your resume & track your apply process — 7 days free, no card needed.
Log in to add to target listDescription
Company Overview
TENEX is an AI-native, automation-first, built-for-scale Managed Detection and Response (MDR) provider.
We are a force multiplier for defenders, helping organizations enhance their cybersecurity posture through advanced threat detection, rapid response, and continuous protection.
Our team is composed of industry experts with deep experience in cybersecurity, automation, and AI-driven solutions.
Backed by leading investors, we are rapidly growing and seeking top talent to join our mission of revolutionizing the AI-Native MDR landscape.
We’re a fast-growing startup backed by industry experts and top-tier investors led by Crosspoint Capital Partners and also backed by Shield Capital, DTCP (formerly Deutsche Telekom Capital Partners), Deepwork Capital, and the Florida Opportunity Fund.
Seed round led by Andreessen Horowitz (a16z).
As an early employee, you’ll play a meaningful role in defining and building our culture.
Get in on the ground floor.
We’re a small but well-funded team that just raised a substantial round – joining now comes with limited risk and unlimited upside.
As a Principal DevOps Engineer, you will be a key technical leader responsible for the architecture, evolution, and operation of our Azure infrastructure, CI/CD pipelines, and Site Reliability Engineering (SRE) practices.
You will keep the platform highly available, secure, and performant as it scales to handle petabytes of security data and billions of daily events.
You'll work closely with Software Engineering, AI/ML, and Security Operations teams to define the technical vision and architecture for our production systems, driving automation and operational excellence to minimize toil and accelerate product delivery.
This role requires deep, hands-on Azure platform expertise, a strong software engineering foundation, and fluency in DevSecOps principles.
Culture is one of the most important things at TENEX.AI.
Explore our culture deck at culture.tenex.ai to witness how we embody it, prioritizing the irreplaceable collaboration and community of in-person work.
Location: This role will require Monday through Thursday onsite in our Kansas City office (preferred), with San Jose or Sarasota, FL also considered.
WFH Friday.
Candidates must live in or be willing to relocate to one of these three cities.
Job Responsibilities
Own the architecture of our Azure platform as it scales to petabytes of security data and billions of daily events.
Own our Azure governance and environment model, including subscription and management group structure, Azure Policy, network topology, and identity.
Lead Site Reliability Engineering (SRE) initiatives, defining and driving adherence to critical Service Level Objectives (SLOs) and Service Level Indicators (SLIs), and managing on-call rotations.
Drive operational excellence by implementing advanced monitoring, observability (logs, metrics, tracing), automated provisioning, and disaster recovery strategies.
Establish and enforce DevSecOps practices, standardizing CI/CD pipelines, infrastructure-as-code (IaC), security testing, and deployment mechanisms for rapid, secure, and reliable software delivery.
Automate deployment, scaling, and management of microservices and event-driven systems using containerization and orchestration technologies (Docker, Kubernetes, AKS).
Maintain workload portability across the platform so that infrastructure decisions remain reversible.
Partner with engineering teams to optimize application performance, resource utilization, and cloud cost efficiency.
Mentor and influence engineering teams on best practices in Azure architecture, reliability, and security-first development.
Collaborate with Product Management and Security Operations to translate new product requirements and operational needs into scalable and cost-effective platform solutions.
Evaluate and drive the adoption of new infrastructure technologies and engineering methodologies to maintain a competitive advantage.
Required Skills & Qualifications
8+ years of progressive experience in DevOps, Site Reliability Engineering (SRE), or Platform Engineering roles.
Deep, hands-on expertise building production Azure platforms, including AKS, Entra ID and workload identity federation, VNet design and Private Link, Key Vault, Azure Policy, and subscription or landing zone architecture.
This is a platform engineering role rather than a Microsoft 365, Intune, or Windows administration role.
Experience standing up or moving production workloads across cloud environments, with the ability to describe the design, the data path, the cutover, and what broke.
Experience building secure and compliant (e.g., SOC 2, ISO 27001) environments.
Deep understanding of microservices architecture, containerization (Docker, Kubernetes), and event-driven systems.
Extensive experience with Infrastructure-as-Code tools (e.g., Terraform, Bicep) and CI/CD best practices.
Production experience writing and shipping software in Go or Python, beyond scripting and configuration.
Experience with monitoring and observability tools (Prometheus, Grafana, Azure Monitor, ELK stack, or similar).
Familiarity with real-time data pipelines and stream processing (e.g., Kafka, Event Hubs, Service Bus, Pub/Sub).
Proven track record of architecting, building, and operating highly scalable, distributed, and secure enterprise-grade SaaS platforms.
Nice-to-have
Working knowledge of more than one major cloud provider, deep enough to judge where the provider models differ rather than assume they match.
Prior experience in cybersecurity (SIEM, EDR, SOAR, or MDR) or an MSSP environment.
Experience with large-scale data warehousing/lakehouse technologies (e.g., Azure Data Explorer, Microsoft Fabric, Snowflake, BigQuery).
Background leading technical initiatives in high-growth startups or enterprise SaaS.
Familiarity with the underlying infrastructure to support AI/ML model deployment and monitoring (MLOps).
Education & Certifications
10-12 years of experience, Bachelor's or Master's degree in Computer Science, Engineering, and or years of relative experienceRelevant certifications (Azure Solutions Architect Expert, Kubernetes, or security-related credentials) are a plus.
Certifications complement production depth and do not substitute for it.
We have 123,872 jobs that might be an even better fit for you
DontApply's real value goes far beyond a single job link or company name. Just upload your resume — in under a minute we'll analyze all 123,872 jobs and tell you exactly which ones you should apply to right now.
Upload My Resume