DevOps Engineer (AI Inference)

Gcore

Poland

On-site

PLN 180,000 - 240,000

Full time

7 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Competitive compensation
Flexible working hours
Hybrid or remote options
Work from anywhere up to 45 days/year
Private medical insurance
Extra vacation and sick leave
Language courses
Modern offices with amenities
Team sports and social activities

Job summary

Gcore is seeking a DevOps Engineer to join our AI Inference Operations Team. You will design, deploy, and maintain infrastructure enabling scalable AI inference workloads on-premises.

You will collaborate with ML engineers and platform teams to optimize architecture, automate deployments with Terraform/Ansible, and ensure robust monitoring using Grafana and VictoriaMetrics.

Qualifications

  • Design, deploy, and maintain infrastructure for AI inference workloads, including GPU scheduling and data access patterns.
  • Build and manage monitoring and observability tools for AI inference platforms with dashboards and alerts.
  • Collaborate with ML engineers and platform teams to design architecture for AI workloads and test performance at scale.

Responsibilities

  • Design, develop, and maintain on-prem infrastructure for AI inference workloads.
  • Build and manage monitoring/observability tools and runbooks for system health.
  • Collaborate with ML engineers to optimize system architecture and deployment pipelines.

Skills

Kubernetes
Linux
Networking
Python
Go
Bash
Terraform
Ansible
Grafana
Git

Tools

Kubernetes operators
Argo CD
Helm
Helmfile
Slurm

Job description

This position is available only under an employment (labor) agreement.

The world’s digital experiences run on something invisible: the infrastructure and software that keep them fast, reliable, and secure. AtGcore,you’llhelp design and deliver that foundation for an AI-driven world.

We’rea global provider of infrastructure and software solutions forAI, cloud, network, and security,powering everything from real-time communication and streaming to enterprise AI and secure web applications. With210+ edge locations, 50+ cloud regions, and thousands of GPUs, your work here can reach users and businesses across the globe.

You’llcollaborate with leading technology partners such asIntel, NVIDIA, Dell, and Equinix, and work on platforms that power digital products used around the world. Our vision is simple: to connect the world to AI, anywhere, anytime.

Want to work on technology that goes beyonda single productor industry?Join a global team of550+professionalsbuilding infrastructure and software that supports the entire digital ecosystem.

We are looking fora talented DevOps Engineerto join our AI Inference Operations Team.

Job Description

As a DevOps Engineer, you will be responsible for designing, deploying, and maintaining infrastructure and services that enable scalable and secure AI inference workloads on-premises.

Qualifications

What You Will Do

  • Design, develop, and maintain infrastructure for AI inference workloads, including GPU scheduling, model deployment pipelines, and data access patterns in on-prem environments
  • Build and manage monitoring and observability tools for AI inference platforms, including dashboards, alerts, and runbooks for model health and system performance
  • Collaborate with ML engineers and platform teams to design system architecture for AI workloads, integrate inference runtimes, and test performance at scale

What We're Looking For

  • Strong understanding of Kubernetes architecture, including CNI, CSI, operators, ingress/gateway, and control plane components.
  • Hands-on experience operating and troubleshooting production Kubernetes clusters.
  • Strong Linux and networking troubleshooting skills, including DNS, routing, firewalling, TLS, MTU, connectivity and performance issues.
  • Ability to develop automation and operational tooling using Python, Go, or Bash.
  • Experience with Terraform, Ansible, or similar IaC/configuration management tools.
  • Experience with VictoriaMetrics/Grafana or similar monitoring, alerting, and troubleshooting tools.
  • Strong experience with Git-based workflows and CI/CD pipelines.

Preferred Qualifications

  • Familiarity with Cluster API or similar Kubernetes cluster lifecycle management technologies.
  • Hands-on operation or administration of Slurm clusters.
  • Knowledge of Argo CD, GitOps workflows, Helm, or Helmfile.
  • Background working with managed platforms, PaaS, or cloud services.
  • Exposure to bare metal, GPU, HPC, or other high-performance computing environments.

Nice to Have

  • Familiarity with the NVIDIA GPU stack, RDMA/InfiniBand, or high-performance networking.
  • Knowledge of OpenStack or similar cloud infrastructure platforms.
  • Hands-on experience developing Kubernetes operators or controllers.
Additional Information

At Gcore, we want you to do your best work and enjoy the journey. Our benefits are designed to support your growth, well-being, and life beyond work:

  • Competitive compensation
  • Flexible working hours and hybrid or remote options, depending on your role
  • Work from anywhere in the world for up to45 days per year
  • Private medical insurance for you and your family*
  • Extra paid vacation and sick leave days*
  • Support for life’s important moments and celebrations
  • Language courses to help you connect and grow
  • Modern, welcoming offices with snacks, drinks, and entertainment*
  • Team sports and social activities*

*Benefits may vary depending on your location.

Equal Opportunity Employer

We provide equal opportunity to all applicants without regard to race, color, religion, sex, sexual orientation, age, gender identity, gender expression, national origin, disability, or any other legally protected characteristics.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Software Python Engineer (GPU Cloud)
Software Python Engineer (GPU Cloud)

Gcore • Poland

Hybrid
PLN 180,000 - 280,000
Remote work options
Private medical insurance
Sick leave days
+2
AI Infrastructure & Platform Operations Engineer (remote in the EU)
AI Infrastructure & Platform Operations Engineer (remote in the EU)

Mirantis • Poznań

Remote
PLN 180,000 - 240,000
AI Infrastructure Engineer (GPU) - Remote EMEA
AI Infrastructure Engineer (GPU) - Remote EMEA

Pragmatike • Poland

On-site
PLN 364,000 - 474,000
Ownership of critical infrastructure
High-growth startup environment
Inclusive recruitment process
Software Engineer (Golang)
Software Engineer (Golang)

Gcore • Poland

Hybrid
PLN 180,000 - 300,000
Competitive compensation
Hybrid or remote options
Private medical insurance
+1
DevOps
DevOps

Complexio • Warszawa

On-site
PLN 180,000 - 280,000
DevOps
DevOps

Complexio Limited • Warszawa

On-site
PLN 100,000 - 130,000
Professional growth opportunities
Collaborative team environment
Continuous learning in AI and industry
Senior AI Infrastructure & Platform Operations Engineer (remote in the EU)
Senior AI Infrastructure & Platform Operations Engineer (remote in the EU)

Mirantis • Poznań

On-site
PLN 180,000 - 300,000
Junior Data Engineer (Data & AI Factory)
Junior Data Engineer (Data & AI Factory)

Devoteam • Kraków

On-site
PLN 60,000 - 80,000
Up to 26 days of vacation per year
MultiSport card
Career Management and training
+3
Mid Data Engineer (Data & AI Factory)
Mid Data Engineer (Data & AI Factory)

Devoteam • Warszawa

On-site
PLN 60,000 - 90,000
Up to 26 days of vacation
MultiSport card
Cafeteria points
+3
Sr. Software Engineer, Inference
Sr. Software Engineer, Inference

CoreWeave • Warszawa

On-site
PLN 321,000 - 471,000
Family-level Medical Insurance
Generous Pension Contribution
Tuition Reimbursement