DevOps & Cloud Platform Engineer

Aira Technologies

San Francisco (CA)

On-site

USD 180,000 - 240,000

Full time

14 days+
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

Aira Technologies in San Francisco builds AI agents for cellular networks. You will build and operate the cloud platform our agent systems run on—the infrastructure, pipelines, and Kubernetes environments that take prototypes to production under live network traffic.

You own deployment, scaling, and uptime, collaborating with ML researchers and engineers to ensure secure, observable, and cost-efficient operation across Azure, Helm, Jenkins, and Terraform.

Qualifications

  • 5+ years running production cloud infrastructure with on-call/incident response.
  • Strong Azure experience: AKS, ACR, Entra ID, workload identity, Private Link, VNets, NSGs.
  • Hands-on Kubernetes with Helm, autoscaling, limits, upgrades, and kubectl.
  • Observability with Prometheus/Grafana and meaningful SLOs/alerts.
  • CI/CD ownership; Jenkins used in production; GitOps a plus.
  • Security in pipelines: SAST/DAST/SCA and RBAC design.
  • Secure networking: restricted egress, proxies, VPNs, and hybrid connectivity.

Responsibilities

  • Design and run AKS clusters and Istio service mesh; multi-tenant Kubernetes in Azure.
  • Provision Azure infra with Terraform and Bicep; manage ACR, Key Vault, VNets, and access.
  • Maintain Jenkins pipelines and Helm packaging for consistent deployments.
  • Keep platform secure: CVE remediation, secret management, and image scanning.
  • Manage capacity, cost, and performance: node sizing, HPA tuning, and monitoring.

Skills

Azure AKS
Kubernetes
Terraform
Bicep
Helm
Jenkins
RBAC
Security
Cost optimization

Tools

Kubectl
Prometheus
Grafana
Istio
ACR
Key Vault
Entra ID
Private Link

Job description

We build AI agents that run cellular networks.

Mobile operators spent billions on 5G but can't extract value - vendor lock-in blocks innovation, monolithic systems don't scale, operational complexity multiplies. We deploy multi-agent systems with fine-tuned LLMs that autonomously optimize network performance, generate operational tools, and execute operator intent. Applied AI at planetary scale: millions of towers, petabytes of telemetry, real-time inference affecting billions of users.

Founded by Qualcomm and Intel veterans. Team from Stanford, Berkeley, Cornell and other leading institutions. Backed by AT&T Ventures, Intel Capital, In-Q-Tel, NeoTribe, Acrew, and more.

The Role

Build and operate the cloud platform our agent systems run on - the infrastructure, pipelines, and Kubernetes environments that take prototypes to production under live network traffic. You own how it gets deployed, scaled, and kept running, working alongside ML researchers, application engineers, and wireless engineers.

What You'll Do
  • Design and run our AKS clusters and Istio service mesh: node pools, autoscaling, tenant isolation, upgrades, mTLS, and traffic policy. These run in customer-managed Azure environments.
  • Provision Azure infrastructure with Terraform and Bicep, including ACR, Key Vault, Entra ID, VNets, NSGs, Private Link, workload identity, and storage. Deploy into environments with private registries, restricted outbound access, proxies, VPNs, and bastion hosts.
  • Maintain the Jenkins pipelines that build and ship Aira releases, and package Naavik with Helm so it installs the same way for every operator, including per-tenant values and upgrade paths.
  • Keep the platform secure: image and dependency scanning, CVE remediation, secrets in Key Vault, workload identity, and least-privilege RBAC.
  • Manage capacity and cost: node fleet sizing, HPA tuning, and the monthly Azure bill.
What We're Looking For
  • 5+ years running production cloud infrastructure, including on-call and incident response.
  • Strong Azure experience: AKS, ACR, Entra ID, workload identity, Private Link, VNets, NSGs, routing, Key Vault, and quota management.
  • Hands-on Kubernetes: Helm, autoscaling, resource limits, cluster upgrades, and troubleshooting with kubectl.
  • Experience building observability with Prometheus, Grafana, or similar, including SLOs and alerts that are actually useful.
  • Terraform or Bicep in production, and ownership of a CI/CD pipeline. We use Jenkins. GitOps experience is a plus.
  • Security in the pipeline: SAST, DAST, SCA, CVE remediation, secrets management, and RBAC design.
  • Secure networking: restricted egress, segmentation, proxies, and VPN or hybrid connectivity.
Why Aira
  • Direct access to founders, partners, and researchers; high autonomy and visible impact
  • Own high-impact infrastructure that supports real-world AI systems at massive scale.
  • Partner with tier-1 mobile operators globally

We're building the intelligent network. Help us ship it

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Software Engineer - Data Platform
Senior Software Engineer - Data Platform

Aira Technologies • San Francisco (CA)

On-site
USD 180,000 - 240,000
Direct access to founders
Multi‑agent systems in production
Global tier‑1 operators
+1
AI-Driven Cloud Platform Engineer for Cellular Networks
AI-Driven Cloud Platform Engineer for Cellular Networks

Aira Technologies • San Francisco (CA)

On-site
USD 180,000 - 240,000
Full-Stack AI Engineer — Platform Architect
Full-Stack AI Engineer — Platform Architect

TechDigital Group • New York (NY)

On-site
USD 180,000 - 280,000
AI Engineer - Backend
AI Engineer - Backend

AGI, Inc. • San Francisco (CA)

On-site
USD 120,000 - 150,000
Competitive cash and meaningful equity
Top-tier relocation and immigration support
Software engineer, infrastructure
Software engineer, infrastructure

Altara • San Francisco (CA), Northern (KY)

Hybrid
USD 140,000 - 200,000
Senior AI Platform Engineer (#5640)
Senior AI Platform Engineer (#5640)

Talanto • Union (NJ), Northern (KY)

Hybrid
USD 140,000 - 210,000
Flexible working format
Competitive salary
Professional development
Cloud Platform Engineer (Agentic AI)
Cloud Platform Engineer (Agentic AI)

Luxoft • United States

Remote
EUR 60,000 - 90,000
Network Engineer
Network Engineer

OpenAI • San Francisco (CA)

Hybrid
USD 293,000 - 385,000
Relocation assistance
Hybrid work model (3 days in office)
AI Engineer - Forward Deployed
AI Engineer - Forward Deployed

Moring • Atlanta (GA)

On-site
USD 120,000 - 160,000
Agentic Infrastructure Engineer — Frontier AI Lab
Agentic Infrastructure Engineer — Frontier AI Lab

Aionia Group • New York (NY)

On-site
USD 250,000 - 500,000
Competitive equity
On-site collaboration with a small team
High compensation package