Staff DevOps Engineer

nextiva

Deutschland

Hybrid

EUR 103.000 - 155.000

Vollzeit

Vor 6 Tagen
Sei unter den ersten Bewerbenden
Bewerbungsgenerator

Mach aus dieser Rolle ein Vorstellungsgespräch — ein Lebenslauf und ein Anschreiben, die darauf ausgerichtet sind, was dieser Arbeitgeber sucht.

Schaffe es an den ATS-Filtern vorbei

Benefits dieser Stelle

Major medical insurance
Vision and dental coverage
Life insurance
Paid time off and holidays
Christmas bonus and savings match

Zusammenfassung

nextiva is seeking a staff-level DevOps leader to design and evolve a Kubernetes-based platform engineering practice for a globally available microservices environment. This role is an individual-contributor technical leadership position, not a people manager.

You will own multi-cluster architecture, establish GitOps workflows with Argo CD/Flux, and build a self-service developer platform with secure service mesh, GPU/ML scheduling, and observability.

Qualifikationen

  • Bachelor's degree in Computer Science or related field, or equivalent work experience.
  • 8+ years of DevOps, platform, or infrastructure engineering experience.
  • 5+ years hands-on Kubernetes experience in production, including cluster architecture, upgrades, and multi-tenant environments.
  • Strong experience operating cloud infrastructure across AWS and GCP.
  • Strong GitOps experience with Argo CD or Flux, and Infrastructure-as-Code experience with Terraform or Pulumi.
  • Deep understanding of container networking, ingress, and service mesh (e.g., Istio, Linkerd, Cilium), plus middleware experience with Nginx, Kafka, and Redis at scale.
  • Experience with GPU scheduling and resource quotas for ML/AI workloads on GKE and EKS, and operating observability platforms such as Prometheus, Grafana, Datadog, or OpenTelemetry.

Aufgaben

  • Own the architecture and roadmap for the multi-cluster Kubernetes platform, including scaling, upgrades, and multi-tenancy.
  • Establish GitOps-based deployment workflows using tools such as Argo CD or Flux.
  • Design and evolve an internal developer platform that gives engineering teams self-service access to compute, environments, and observability.
  • Architect service mesh, networking, and ingress strategies for reliable, secure service-to-service communication.
  • Define platform standards for GPU and ML workload scheduling and partner with AI/ML teams on training and inference infrastructure needs.
  • Drive capacity planning and cost optimization across Kubernetes and cloud infrastructure, and own platform reliability through SLOs, incident response, and postmortems.
  • Mentor experienced engineers and represent platform engineering in cross-org technical decisions.

Kenntnisse

Kubernetes
AWS
GCP
GitOps
Terraform
Pulumi
Istio
Service mesh
Nginx
Kafka
Datadog
Grafana
OpenTelemetry

Ausbildung

Bachelor's degree in Computer Science or related field

Tools

Argo CD
Flux
Terraform
Pulumi
Kubernetes
Prometheus
Grafana
Datadog
OpenTelemetry
Istio

Jobbeschreibung

Role overview

Staff-level DevOps opening to lead the design and evolution of a Kubernetes-based platform engineering practice for a globally available, redundant microservices environment. The role combines deep technical ownership of multi-cluster Kubernetes infrastructure with broad influence across engineering, including paved-road tooling for AI/ML workloads. It is an individual-contributor technical leadership role, not a people-management position.

Responsibilities
  • Own the architecture and roadmap for the multi-cluster Kubernetes platform, including scaling, upgrades, and multi-tenancy.
  • Establish GitOps-based deployment workflows using tools such as Argo CD or Flux.
  • Design and evolve an internal developer platform that gives engineering teams self-service access to compute, environments, and observability.
  • Architect service mesh, networking, and ingress strategies for reliable, secure service-to-service communication.
  • Define platform standards for GPU and ML workload scheduling and partner with AI/ML teams on training and inference infrastructure needs.
  • Drive capacity planning and cost optimization across Kubernetes and cloud infrastructure, and own platform reliability through SLOs, incident response, and postmortems.
  • Mentor experienced engineers and represent platform engineering in cross-org technical decisions.
Requirements
  • Bachelor's degree in Computer Science or a related field, or equivalent work experience.
  • 8+ years of DevOps, platform, or infrastructure engineering experience.
  • 5+ years of hands-on Kubernetes experience in production, including cluster architecture, upgrades, and multi-tenant environments.
  • Strong experience operating cloud infrastructure across AWS and GCP.
  • Strong GitOps experience with Argo CD or Flux, and Infrastructure-as-Code experience with Terraform or Pulumi.
  • Deep understanding of container networking, ingress, and service mesh (e.g., Istio, Linkerd, Cilium), plus middleware experience with Nginx, Kafka, and Redis at scale.
  • Experience with GPU scheduling and resource quotas for ML/AI workloads on GKE and EKS, and operating observability platforms such as Prometheus, Grafana, Datadog, or OpenTelemetry.
Nice to have
  • Experience designing internal developer platforms or paved-road tooling.
  • Strong Linux, networking, storage, and security fundamentals, plus excellent cross-team communication skills.
Benefits and work setup
  • Remote-eligible across Mexico; team members within roughly 80 km of the Guadalajara office are expected onsite to support collaboration.
  • Package includes major medical insurance for the employee and dependents, vision and dental coverage, life insurance, paid time off (10 personal days before the first anniversary, then vacation plus additional personal days), a 30-day Christmas bonus, 50% vacation premium, company-matched food vouchers, and a 13% matched savings fund.
  • Employee Assistance Program and ongoing learning and development opportunities.
Hol dir deinen kostenlosen, vertraulichen Lebenslauf-Check.
oder ziehe deine Datei hierhin.
Similar jobs

Ähnliche Jobs, die dir auch gefallen könnten

Platform Engineer (Kubernetes) - Remote Work
Platform Engineer (Kubernetes) - Remote Work

BairesDev • Berlin

Remote
EUR 70.000 - 110.000
100% remote work
Competitive USD salary
Home office setup provided
+3
Kubernetes Engineer
Kubernetes Engineer

Franklin Fitch • Düsseldorf

Vor Ort
EUR 90.000 - 120.000
Sr DevOps Platform Engineer - Bilingual
Sr DevOps Platform Engineer - Bilingual

Embedded Shishya • Deutschland

Remote
EUR 110.000 - 150.000
Senior DevOps Engineer
Senior DevOps Engineer

Annapurna • Berlin

Hybrid
EUR 90.000 - 120.000
Remote work
Stock options
30 vacation days
+3
Staff Software Engineer
Staff Software Engineer

vCluster Labs • Berlin

Vor Ort
EUR 70.000 - 100.000
Competitive Salary
Platinum-Level Insurance
Flexible Working Schedule
+1
Senior DevOps Engineer
Senior DevOps Engineer

orioninnovation • Deutschland

Hybrid
EUR 70.000 - 110.000
Flexible work options (remote/hybrid/3
Senior DevOps Engineer
Senior DevOps Engineer

Trust In SODA • Deutschland

Hybrid
USD 104.000 - 139.000
Hybrid working with regular in-person collaboration
On-call compensation
Competitive base salary
+1
Senior DevOps Engineer Configuration Management
Senior DevOps Engineer Configuration Management

Utimaco • Aachen

Vor Ort
EUR 70.000 - 100.000
Open corporate culture
Company pension
Flexible working model
+2
Infrastructure Operations Engineer
Infrastructure Operations Engineer

lightningai • Deutschland

Hybrid
EUR 138.000 - 172.000
Discretionary bonus
Equity
401(k) matching
+1
Senior DevOps Engineer
Senior DevOps Engineer

European Tech Recruit • Deutschland

Vor Ort
EUR 60.000 - 75.000