Principal DevOps Engineer for Industrial AI Cloud (m/f/d)

T-Systems Iberia

Sevilla la Nueva

Hybrid

EUR 90,000 - 130,000

Full time

33 hours ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Benefits offered by this job

Health insurance
Meal vouchers
Telemedicine
Language classes
Hybrid work model

Job summary

T-Systems Iberia is seeking a DevOps Engineer Principal to guide enterprise customers in onboarding and adoption of an AI GPU platform. You will understand customer needs, support design, run PoCs, and ensure smooth integration of LLMs, GPU compute, and AI pipelines.

You will coordinate across infra, networking, automation, security, and AI services, mentoring teams and driving automation improvements for reliable, scalable solutions.

Qualifications

  • 3+ years in design and delivery of IaaS, PaaS & SaaS systems.
  • Experience with GPU-based infrastructure.
  • Solid knowledge of Kubernetes container tech.
  • Expert in scripting (Python/Bash).
  • Expert knowledge of automation deployments (Ansible/SaltStack/Terraform/Helm).
  • Expert knowledge of CI/CD in Kubernetes and repository management.
  • Automation with Git (GitHub/GitLab) and CI/CD tools like GitHub Actions or GitLab CI/CD.
  • Strong Linux OS experience.
  • Experience with monitoring/visualization tools (Grafana/Prometheus).
  • English proficiency at B2 level; German an advantage.

Responsibilities

  • Consult customers on GPU infrastructure and platform usage.
  • Lead onboarding and training, mentoring customer specialists.
  • Design and implement PoCs, including environment setup and pipelines.
  • Capture requirements and translate into technical specs.
  • Assist with performance optimization, troubleshooting, and validation.
  • Act as technical lead, coordinating across infra, networking, automation, security, and AI services.
  • Propose and develop automation concepts to improve services and processes.
  • Ensure reliability, scalability and security across the customer lifecycle.
  • Support monitoring, observability and capacity planning for AI workloads.

Skills

GPU infrastructure
Kubernetes
CI/CD in Kubernetes
Python/Bash scripting
Automation tools (Ansible SaltStack,-T
Git workflows (GitHub/GitLab)
Linux OS
Monitoring (Grafana Prometheus)
English (B2 level)

Tools

Docker
Grafana
Prometheus
Terraform
Helm
GitHub
GitLab
Ansible
SaltStack

Job description

T‑Systems is part of the Deutsche Telekom Group, with around 30.000 employees worldwide. We create technology with purpose to generate a positive impact on society. We are looking for curious talent, eager to learn, take on challenges, and contribute ideas that transform our customers’ experience.

We trust people: we offer autonomy, continuous support, and a collaborative environment where you can grow without limits. We are one global team, guided by respect, integrity, and a passion for doing better every day.

Job Description

NVIDIA and Deutsche Telekom are jointly developing industrial AI cloud for Europe. This AI factory in Germany will host 10,000 GPUs across NVIDIA DGX B200 systems and RTX Pro Servers. Deutsche Telekom provides secure, sovereign and fast infrastructure, including data centers, operations, security, and AI solutions. As DevOps Engineer Principal you will guide enterprise customers through onboarding, training, and early adoption of the AI platform. Your responsibility includes understanding customer requirements, supporting solution design, executing Proofs of Concept (PoCs), and ensuring smooth integration of customer workloads (LLMs, GPU compute, AI pipelines). You act as a trusted technical advisor, helping customers efficiently use their GPU clusters and AI toolchains.

What will you do?
  • Consult customers on all technical aspects related to GPU infrastructure and platform usage.
  • Lead onboarding and training, mentoring customer specialists on optimal usage of their GPU clusters and AI environments.
  • Design and implement PoCs, including environment setup, data processing pipelines, and deployment workflows.
  • Conduct requirement engineering, translating business needs into technical specifications.
  • Assist customers with performance optimization, troubleshooting, fine-tuning, and validation of delivered solutions.
  • Act as the key technical point of contact, coordinating cross-functional teams across infrastructure, networking, automation, security, and AI services.
  • Propose and develop automation concepts to improve services, processes, and operating models.
  • Ensure best practices in reliability, scalability, and security are applied across the customer lifecycle.
  • Support monitoring, observability, and capacity planning for AI workloads and GPU utilization.
Qualifications
  • Have 3+ years of experience in the design and delivery of systems based on IaaS, PaaS and SaaS.
  • Have experience with GPU based infrastructure.
  • Possess solid knowledge of Kubernetes container-based technologies.
  • Have expert knowledge of scripting languages (Python/Bash).
  • Possess expert knowledge of Automation tools and automation deployments (Ansible/Salt-Stack/Terraform/Helm).
  • Have expert knowledge of CI/CD in a Kubernetes environment and repository management.
  • Are expert in automation with Git (GitHub, GitLab) and CI/CD tools like GitHub Actions or GitLab CI/CD.
  • Have strong experience with Linux OS.
  • Have experience with monitoring and visualization tools (Grafana, Prometheus,…)
  • Speak English at B2 level (German is an advantage).
Other skills
  • Good communication skills, analytical thinking, team cooperation, presentation skills, negotiation skills.
  • Project Management- Basic.
  • Leadership skills- Basic
  • Quality management- Intermediate.
Additional Information

What do we offer you?

Work environment & flexibility
  • International, dynamic and collaborative environment.
  • T-Social: social initiatives (sports, community, health, …).
  • Hybrid work model (remote/on-site).
  • Flexible working hours.
Growth & development
  • Customized training: access to Coursera to learn whatever you want, whenever you want.
  • Weekly language classes (English, Spanish & German).
  • International Mentoring Sessions & Experience Days.
Compensation & benefits
  • Flexible compensation plan (health insurance, meal vouchers, childcare, transport).
  • Telemedicine.
  • Life and accident insurance.
  • Social fund.
Wellbeing & time off
  • 26+ working days of vacation per year.
  • Free access to specialist services (medical, legal, wellness).
  • 100% salary coverage during medical leave.

And many more advantages of being part of T-Systems!

T-Systems Iberia will only process the CVs of candidates who meet the requirements specified for each offer.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Principal DevOps Engineer for Industrial AI Cloud (m/f/d)
Principal DevOps Engineer for Industrial AI Cloud (m/f/d)

SmartRecruiters, Inc. • Spain

Remote
EUR 70,000 - 110,000
Coursera training
Language classes (English, Spanish, 2)
Life insurance
+4
Senior DevOps Engineer for Industrial AI Cloud (m/f/d)
Senior DevOps Engineer for Industrial AI Cloud (m/f/d)

SmartRecruiters, Inc. • Spain

Remote
EUR 70,000 - 110,000
Life and accident insurance
Social fund
Weekly language classes
Senior DevOps Engineer for Industrial AI Cloud (m/f/d)
Senior DevOps Engineer for Industrial AI Cloud (m/f/d)

SmartRecruiters, Inc. • Granada

On-site
EUR 55,000 - 75,000
Coursera access
Language classes
Insurance (life/accident)
+2
Senior DevOps Engineer for Industrial AI Cloud (m/f/d)
Senior DevOps Engineer for Industrial AI Cloud (m/f/d)

Iglubit • Sevilla

Hybrid
EUR 65,000 - 95,000
Health insurance
Meal vouchers
Telemedicine
+2
Principal DevOps Engineer (m/w/d) für Industrial AI Cloud
Principal DevOps Engineer (m/w/d) für Industrial AI Cloud

Deutsche Telekom • Madrid

Hybrid
EUR 110,000 - 160,000
Deutschkenntnisse
Principal DevOps Engineer (m/w/d) für Industrial AI Cloud
Principal DevOps Engineer (m/w/d) für Industrial AI Cloud

Deutsche Telekom • Bilbao

Hybrid
EUR 90,000 - 130,000
Deutschkenntnisse
Principal DevOps Engineer (m/w/d) für Industrial AI Cloud
Principal DevOps Engineer (m/w/d) für Industrial AI Cloud

Deutsche Telekom • Granada

Hybrid
EUR 90,000 - 150,000
Krankenversicherung
Essensgutscheine
Kinderbetreuung
+6
AI SWE / OpenStack Service Engineer (m/f/d)
AI SWE / OpenStack Service Engineer (m/f/d)

SmartRecruiters, Inc. • Spain

Hybrid
EUR 70,000 - 90,000
Hybrid work model
Flexible working hours
Language classes (English & German)
+1
Cloud Architect (m/f/d)
Cloud Architect (m/f/d)

T-Systems Iberia • Granada

On-site
EUR 60,000 - 85,000
Flexible schedule
Continuous training
Hybrid work model
+3
Principal DevOps Engineer (m/w/d) für Industrial AI Cloud
Principal DevOps Engineer (m/w/d) für Industrial AI Cloud

Deutsche Telekom • La Coruña

Hybrid
EUR 90,000 - 140,000
Hybrides Arbeitsmodell
Flexible Arbeitszeiten
Coursera Kurse
+6