Site Reliability Engineer

FPT Software

Madrid

Presencial

EUR 65.000 - 90.000

Jornada completa

Hace 2 días
Sé de los primeros/as/es en solicitar esta vacante
Generador de candidaturas

Consigue una respuesta de este empleador — un currículum y una carta de presentación adaptados exactamente a lo que busca para contratar.

Supera los filtros ATS

Descripción de la vacante

FPT Software in Madrid, Spain, seeks a Cloud Operations Engineer to own end-to-end operation, maintenance support and architecture governance of overseas public cloud infrastructures, covering containers, VMs, storage and networks.

You will design, implement and maintain automated CI/CD pipelines, oversee production releases of overseas containerized apps, participate in on-call rotations, and collaborate with China-based teams to standardize governance and optimize resource usage.

Formación

  • Solid mastery of Linux operating system principles and core network protocols (TCP/IP, HTTP).
  • In-depth understanding of Kubernetes and Operatos in production environments.
  • Proficiency with Prometheus and Grafana to build/maintain monitoring systems.
  • Familiar with CI/CD toolchains like ArgoCD for automated pipelines.
  • Knowledge of Nginx, APISIX, Envoy; RocketMQ, RabbitMQ, Kafka for cloud-native gateways and messaging.
  • Proficient in Python or Shell scripting for automation.
  • Understanding of GDPR/data privacy requirements for overseas operations.
  • Strong English and Chinese communication; Chinese-speaking candidates with right to work in Spain preferred.

Responsabilidades

  • Own end-to-end operation, maintenance support and architecture governance of overseas cloud infrastructures, including containers, VMs, storage and networks.
  • Design, implement and maintain automated CI/CD pipelines to support continuous integration, delivery and standardized release workflows for overseas business applications.
  • Undertake daily on-call rotation responsibilities; troubleshoot and resolve defects, bottlenecks and performance issues to optimize resource performance and service stability.
  • Manage daily operation and production release of overseas containerized and cloud-native applications, ensuring reliable online services.
  • Collaborate with domestic teams to standardize governance and migration initiatives across platforms.
  • Analyze usage of overseas cloud and application resources to optimize resource scheduling and cost-effectiveness.

Conocimientos

Linux
Kubernetes
Prometheus
Grafana
ArgoCD
Nginx
APISIX
Envoy
RocketMQ
RabbitMQ
Kafka
Python
Shell
GDPR
TCP/IP
HTTP
English
Chinese

Descripción del empleo

  • Own end-to-end operation, maintenance support and architecture governance of overseas public cloud infrastructures, covering core cloud resources including containers, cloud virtual machines, storage and networks. Also oversee the operational stability and architectural standardization of overseas databases and middleware systems.
  • Design, implement and maintain automated CI/CD pipelines to support continuous integration, continuous delivery and standardized release workflows for overseas business applications.
  • Undertake daily on-call rotation responsibilities. Proactively troubleshoot and resolve functional defects, resource bottlenecks and performance anomalies of cloud infrastructure and business applications, and drive continuous optimization of overall resource performance and service stability.
  • Manage daily operation, maintenance and production release of overseas containerized and cloudnative applications, ensuring reliable and smooth online iteration of business services.
  • Collaborate closely with domestic technical teams to formulate and implement unified containerized application management specifications across the platform, and steadily promote standardized governance, architecture optimization and business migration initiatives for overseas applications.
  • Continuously analyze the usage status of overseas cloud and application resources, drive resource scheduling optimization and efficiency improvement, and maximize overall resource utilization and costeffectiveness of overseas business environments.

We are looking for you who

  • Solid mastery of Linux operating system principles and core network protocols, including TCP/IP and HTTP, with proficient hands‑on operational capabilities.
  • In‑depth understanding of the Kubernetes ecosystem and the working mechanisms of core components; proficient in operating and maintaining Kubernetes Operators in production environments.
  • Proficient in the principles and practical usage of mainstream observability tools including Prometheus and Grafana, capable of building and maintaining complete monitoring systems for cloud‑native environments.
  • Familiar with mainstream CI/CD toolchains represented by ArgoCD, with solid capabilities to build, configure and maintain automated continuous integration and delivery pipelines.
  • Familiar with the architecture, operational principles, deployment and daily O&M of common cloudnative gateway systems including Nginx, APISIX and Envoy, as well as mainstream message queue middleware such as RocketMQ, RabbitMQ and Kafka.
  • Proficient in at least one mainstream scripting language (Python / Shell) to support daily automation operation, batch processing and operational tool development.
  • Have a clear understanding of overseas data compliance specifications and privacy protection regulations such as GDPR, able to carry out cloud operation and maintenance work in compliance with regional regulatory requirements.
  • Given that these positions will collaborate closely with engineering and operations teams located in China, strong Chinese and English communication skills are highly desirable. To facilitate efficient communication, collaboration, and alignment with China‑based stakeholders, Chinese‑speaking candidates with the legal right to work in Spain are strongly preferred. However, qualified international candidates meeting the required competencies will also be considered.
Consigue la evaluación confidencial y gratuita de tu currículum.
o arrastra y suelta tu archivo aquí
Similar jobs

Puestos de trabajo similares que vale la pena comparar

Sre Monitoring Engineer
Sre Monitoring Engineer

Optiwisers • Valladolid

Híbrido
EUR 62.000 - 103.000
Site Reliability Engineer
Site Reliability Engineer

Optiwisers • Madrid

Híbrido
EUR 83.000 - 111.000
Site Reliability Engineer
Site Reliability Engineer

Optiwisers • Vitoria

Híbrido
EUR 70.000 - 110.000
Site Reliability Engineer (Sre)
Site Reliability Engineer (Sre)

Optiwisers • Castro-Urdiales

Híbrido
EUR 67.000 - 123.000
Site Reliability Engineer (Sre)
Site Reliability Engineer (Sre)

Optiwisers • Chantada

Híbrido
EUR 60.000 - 90.000
Hybrid work model
Site Reliability Engineer
Site Reliability Engineer

Optiwisers • Meis

Híbrido
EUR 65.000 - 111.000
Site Reliability Engineer
Site Reliability Engineer

ThunderSoft • Madrid

Presencial
EUR 52.000 - 76.000
Sre Monitoring Engineer
Sre Monitoring Engineer

Optiwisers • Chantada

Híbrido
EUR 83.000 - 124.000
Global Cloud SRE: Kubernetes, CI/CD & Observability
Global Cloud SRE: Kubernetes, CI/CD & Observability

FPT Software • Madrid

Presencial
EUR 65.000 - 90.000
Site Reliability Engineer
Site Reliability Engineer

Optiwisers • Valencia

Híbrido
EUR 83.000 - 124.000