Principal DevOps Engineer - Azure

TENEX.AI

San Jose (CA)

Hybrid

USD 170,000 - 210,000

Full time

9 days ago

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

TENEX.AI is seeking a Principal DevOps Engineer to lead the architecture, evolution, and operation of our Azure infrastructure, CI/CD pipelines, and SRE practices. You will ensure high availability, security, and performance as the platform scales to immense volumes of data.

You will collaborate with Software Engineering, AI/ML, and Security Operations to define a technical vision, drive automation, and minimize toil. Strong Azure, Go/Python, and IaC experience are essential.

Qualifications

  • 8+ years in DevOps, SRE, or Platform Engineering.
  • Hands-on Azure production platforms (AKS, Entra ID, VNet, Key Vault, Policy).
  • Experience across cloud environments and secure, compliant ecosystems (SOC 2, ISO 27001).
  • Deep microservices, containers, and event-driven systems.
  • Proficient with IaC (Terraform, Bicep) and CI/CD.
  • Experience shipping Go or Python beyond scripting.
  • Strong monitoring/observability (Prometheus, Grafana, Azure Monitor).
  • Knowledge of real-time data pipelines (Kafka, Event Hubs).
  • Proven ability to architect, build, and operate scalable SaaS platforms.

Responsibilities

  • Own Azure platform architecture at scale with petabytes of data.
  • Define governance, topology, and identity in Azure.
  • Lead SRE initiatives, SLOs/SLIs, on-call rotations.
  • Implement monitoring, observability, DR strategies.
  • Enforce DevSecOps with IaC, CI/CD, security testing.
  • Automate deployment, scaling, and management of microservices.
  • Collaborate with engineering to optimize performance and cost.
  • Mentor teams on Azure reliability and security practices.
  • Translate product and ops needs into scalable platform solutions.

Skills

Azure architecture
SRE/DevOps
Terraform/Bicep
Go/Python
Docker/Kubernetes
Monitoring/Observability
Security/Compliance
CI/CD

Education

Bachelor's or Master's in CS/Engineering

Tools

AKS
Entra ID
VNet/Private Link
Key Vault
Terraform

Job description

Company Overview

TENEX is an AI-native, automation-first, built-for-scale Managed Detection and Response (MDR) provider. We are a force multiplier for defenders, helping organizations enhance their cybersecurity posture through advanced threat detection, rapid response, and continuous protection. Our team is composed of industry experts with deep experience in cybersecurity, automation, and AI-driven solutions. Backed by leading investors, we are rapidly growing and seeking top talent to join our mission of revolutionizing the AI-Native MDR landscape.

We’re a fast-growing startup backed by industry experts and top-tier investors led by Crosspoint Capital Partners and also backed by Shield Capital, DTCP (formerly Deutsche Telekom Capital Partners), Deepwork Capital, and the Florida Opportunity Fund. Seed round led by Andreessen Horowitz (a16z). As an early employee, you’ll play a meaningful role in defining and building our culture. Get in on the ground floor. We’re a small but well‑funded team that just raised a substantial round – joining now comes with limited risk and unlimited upside.

As a Principal DevOps Engineer, you will be a key technical leader responsible for the architecture, evolution, and operation of our Azure infrastructure, CI/CD pipelines, and Site Reliability Engineering (SRE) practices. You will keep the platform highly available, secure, and performant as it scales to handle petabytes of security data and billions of daily events.

You'll work closely with Software Engineering, AI/ML, and Security Operations teams to define the technical vision and architecture for our production systems, driving automation and operational excellence to minimize toil and accelerate product delivery. This role requires deep, hands‑on Azure platform expertise, a strong software engineering foundation, and fluency in DevSecOps principles.

Culture is one of the most important things at TENEX.AI. Explore our culture deck at culture.tenex.ai to witness how we embody it, prioritizing the irreplaceable collaboration and community of in-person work.

Location: This role will require Monday through Thursday onsite in our Kansas City office (preferred), with San Jose or Sarasota, FL also considered. WFH Friday. Candidates must live in or be willing to relocate to one of these three cities.

Job Responsibilities
  • Own the architecture of our Azure platform as it scales to petabytes of security data and billions of daily events.
  • Own our Azure governance and environment model, including subscription and management group structure, Azure Policy, network topology, and identity.
  • Lead Site Reliability Engineering (SRE) initiatives, defining and driving adherence to critical Service Level Objectives (SLOs) and Service Level Indicators (SLIs), and managing on‑call rotations.
  • Drive operational excellence by implementing advanced monitoring, observability (logs, metrics, tracing), automated provisioning, and disaster recovery strategies.
  • Establish and enforce DevSecOps practices, standardizing CI/CD pipelines, infrastructure‑as‑code (IaC), security testing, and deployment mechanisms for rapid, secure, and reliable software delivery.
  • Automate deployment, scaling, and management of microservices and event‑driven systems using containerization and orchestration technologies (Docker, Kubernetes, AKS).
  • Maintain workload portability across the platform so that infrastructure decisions remain reversible.
  • Partner with engineering teams to optimize application performance, resource utilization, and cloud cost efficiency.
  • Mentor and influence engineering teams on best practices in Azure architecture, reliability, and security‑first development.
  • Collaborate with Product Management and Security Operations to translate new product requirements and operational needs into scalable and cost‑effective platform solutions.
  • Evaluate and drive the adoption of new infrastructure technologies and engineering methodologies to maintain a competitive advantage.
Required Skills & Qualifications
  • 8+ years of progressive experience in DevOps, Site Reliability Engineering (SRE), or Platform Engineering roles.
  • Deep, hands‑on expertise building production Azure platforms, including AKS, Entra ID and workload identity federation, VNet design and Private Link, Key Vault, Azure Policy, and subscription or landing zone architecture. This is a platform engineering role rather than a Microsoft 365, Intune, or Windows administration role.
  • Experience standing up or moving production workloads across cloud environments, with the ability to describe the design, the data path, the cutover, and what broke.
  • Experience building secure and compliant (e.g., SOC 2, ISO 27001) environments.
  • Deep understanding of microservices architecture, containerization (Docker, Kubernetes), and event‑driven systems.
  • Extensive experience with Infrastructure‑as‑Code tools (e.g., Terraform, Bicep) and CI/CD best practices.
  • Production experience writing and shipping software in Go or Python, beyond scripting and configuration.
  • Experience with monitoring and observability tools (Prometheus, Grafana, Azure Monitor, ELK stack, or similar).
  • Familiarity with real‑time data pipelines and stream processing (e.g., Kafka, Event Hubs, Service Bus, Pub/Sub).
  • Proven track record of architecting, building, and operating highly scalable, distributed, and secure enterprise‑grade SaaS platforms.
Nice-to-have
  • Working knowledge of more than one major cloud provider, deep enough to judge where the provider models differ rather than assume they match.
  • Prior experience in cybersecurity (SIEM, EDR, SOAR, or MDR) or an MSSP environment.
  • Experience with large‑scale data warehousing/lakehouse technologies (e.g., Azure Data Explorer, Microsoft Fabric, Snowflake, BigQuery).
  • Background leading technical initiatives in high‑growth startups or enterprise SaaS.
  • Familiarity with the underlying infrastructure to support AI/ML model deployment and monitoring (MLOps).
Education & Certifications
  • 10-12 years of experience, Bachelor's or Master's degree in Computer Science, Engineering, and or years of relative experience
  • Relevant certifications (Azure Solutions Architect Expert, Kubernetes, or security‑related credentials) are a plus. Certifications complement production depth and do not substitute for it.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Principal DevOps Engineer
Principal DevOps Engineer

TENEX.AI • San Jose (CA)

On-site
USD 180,000 - 260,000
Principal DevOps Engineer
Principal DevOps Engineer

TENEX.AI • Kansas City (MO), Northern (KY)

Hybrid
USD 150,000 - 200,000
Senior AI/ML Engineer
Senior AI/ML Engineer

TENEX.AI • United States

Hybrid
USD 120,000 - 150,000
Competitive salary and benefits package
Opportunities for growth and development
Cutting-edge AI-driven technologies
Software Engineer
Software Engineer

tenex • San Jose (CA)

Hybrid
USD 120,000 - 180,000
Software Engineer II
Software Engineer II

TENEX.AI • San Jose (CA)

On-site
USD 100,000 - 130,000
Competitive salary and benefits package
Opportunity for growth in AI and cybersecurity
AI/ML Engineer
AI/ML Engineer

TENEX.AI • United States

Hybrid
USD 100,000 - 150,000
Competitive salary
Benefits package
Growth and development opportunities
Director of Forward Deployed Engineering
Director of Forward Deployed Engineering

Tenex • Sarasota (FL), Scottsdale (AZ), Kansas City (MO)

On-site
USD 180,000 - 240,000
Staff Site Reliability Engineer
Staff Site Reliability Engineer

TENEX.AI • Sarasota (FL)

On-site
USD 150,000 - 230,000
Software Engineer II
Software Engineer II

Tenex • San Jose (CA)

Hybrid
USD 140,000 - 200,000
Software Engineer
Software Engineer

TENEX.AI • Kansas City (MO)

Hybrid
USD 80,000 - 110,000
Competitive salary and benefits package
Growth and development opportunities in AI
Collaborative work culture