Senior Site Reliability Engineer (remote within EMEA)

FYUL

Warszawa

Hybrid

PLN 240,000 - 420,000

Full time

14 days+
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Benefits offered by this job

Private health insurance
Flexible hours
Remote work

Job summary

FYUL is hiring a Senior SRE II to join Platform Infrastructure. You will architect and drive large-scale automation, set engineering standards, and mentor junior SREs.

You will split time between hands-on platform work (Kubernetes, AWS, GCP, CI/CD, observability) and technical leadership, guiding design choices and cost/reliability trade-offs. Your daily tasks include designing highly available infrastructure across AWS accounts, operating EKS clusters, evolving core services, and driving

Qualifications

  • Solid Linux systems administration background and Python scripting.
  • Strong AWS knowledge: EKS, IAM, VPC networking, RDS, S3; multi-account AWS experience is a plus.
  • Hands-on production Kubernetes (EKS) experience, Helm chart development, CNI networking (Cilium), pod networking/IPAM, container security (ECR).
  • Terraform and Terragrunt for multi-environment management; GitOps with ArgoCD.

Responsibilities

  • Architect and manage highly available, secure, and scalable infra across multiple AWS accounts and environments using infrastructure as code.
  • Design and operate Amazon EKS clusters and networking for containerized workloads.
  • Own and evolve core platform services: cloud networking, Kubernetes, databases, and messaging systems.
  • Drive large-scale automation and standardise Terraform/Terragrunt and GitOps across teams.
  • Lead on-call and incident response, write runbooks and postmortems.
  • Drive cost optimization and FinOps initiatives across the platform.

Skills

Linux
Python
AWS EKS
IAM
VPC networking
RDS
S3
Kubernetes
Terraform
Terragrunt
GitOps
ArgoCD
PostgreSQL
MySQL
MongoDB
Jenkins
GitHub Actions
Helm
Grafana
Prometheus
Loki
Tempo
Mimir
Incident management
FinOps
12-Factor App
Multi-account AWS
Cilium

Tools

Terraform
Terragrunt
ArgoCD
Jenkins
GitHub Actions
Helm
Atlantis
Cilium
ECR

Job description

About the team:

Platform Infrastructure builds, operates, and continuously evolves FYUL's container platform and cloud foundation. We foster a DevOps culture through self-service tooling, enabling product engineering teams to ship reliable, secure, and cost-efficient services as the business scales. The team owns our AWS cloud accounts, Kubernetes platform, cloud networking, observability stack, core databases, CI/CD pipelines, and infrastructure-as-code, and acts as the go-to partner for engineering teams on cloud and DevOps topics.

About the role:

We're hiring a Senior SRE II to join Platform Infrastructure as one of the team's senior individual contributors. At this level, you're the go-to person for our most complex infrastructure problems: you architect and drive large-scale automation and reliability initiatives, set standards other engineers follow, and mentor Associate and mid-level SREs. You'll split your time between hands-on platform work - Kubernetes, AWS, GCP, CI/CD, observability - and technical leadership: proposing designs, reviewing others' work, and helping the team make good build-vs-buy and cost/reliability trade-offs.

Your daily tasks will include:
  • Infrastructure & reliability: Architect and manage highly available, secure, and scalable infrastructure across multiple AWS accounts and environments using infrastructure as code.
  • Design and operate our Amazon EKS clusters, including networking policies, persistent storage, and scaling strategies for containerized workloads.
  • Own and evolve core platform services: cloud networking, Kubernetes, and the databases and messaging systems engineering teams depend on.
  • Automation & infrastructure as code: Drive large-scale automation projects and set standards for using Terraform / Terragrunt and GitOps (ArgoCD) across teams.
  • Lead adoption of automation to reduce manual operational work and keep environments consistent and repeatable.
  • Observability & incident response: Be the go-to person for solving complex, cross-service infrastructure problems.
  • Drive initiatives that improve reliability and observability (Grafana, Prometheus, Loki, Tempo, Mimir) so systems scale with minimal manual intervention.
  • Participate in on-call rotation, lead incident response for production issues, and write clear runbooks, ADRs, and postmortems.
  • Security & cost efficiency: Lead security efforts within the team - IAM, encryption, secure logging - and mentor others on secure infrastructure practices.
  • Audit infrastructure spend regularly and drive cost optimization across the platform (rightsizing, autoscaling, FinOps practices).
  • Collaboration & mentorship: Mentor mid-level SREs, provide detailed feedback, and support onboarding of new team members.
  • Communicate complex technical concepts clearly to both engineers and non-technical stakeholders.
  • Partner with product engineering squads to understand their needs and represent Platform Infrastructure in cross-team initiatives.
Your qualifications:

These reflect the technical bar we hold Senior SRE II's to internally, based on our SRE competency framework and current stack.

  • Core technical experience:
  • Solid Linux systems administration background and comfort scripting in Python.
  • Strong AWS knowledge: EKS, IAM (roles, policies, IRSA), VPC networking, RDS, S3, SQS, and familiarity with the Well-Architected Framework; experience in multi-account AWS environments is a strong plus.
  • Hands-on experience operating and troubleshooting Kubernetes (EKS) at production scale, including Helm chart development, CNI networking (we run Cilium), pod networking/IPAM concepts, and container security (ECR, image scanning).
  • Proficiency with Terraform (modules, state management) and ideally Terragrunt for multi-environment management; GitOps experience with ArgoCD.
  • Experience with Postgres, MySQL and/or MongoDB in production scale, including Aurora.
  • CI/CD experience with Jenkins (Jenkinsfile, shared libraries) and/or GitHub Actions, and familiarity with deployment strategies such as blue-green and canary.
  • Experience with the Grafana observability stack (Grafana, Prometheus, Loki, Tempo, Mimir) - metrics design, dashboarding, alerting, log aggregation, and distributed tracing. Not only using but also maintaining it.
  • Practical incident management experience: on-call rotations, structured incident response, and writing runbooks/postmortems.
  • Working knowledge of 12-Factor App principles and cost optimization / FinOps awareness.
  • How you work:
  • A methodical, data-driven approach to troubleshooting rather than guessing.
  • Strong written communication - you write runbooks, ADRs, and postmortems that others can actually follow.
  • Comfortable driving initiatives with ambiguous ownership, and taking accountability for outcomes rather than waiting to be asked.
  • Track record of mentoring less senior engineers and giving direct, constructive feedback.
  • Several years of hands-on production infrastructure/SRE experience, with demonstrated ownership of initiatives at a senior individual-contributor level (leading design work, setting standards, being the escalation point for hard problems).
  • Nice to have:
  • GCP Experience.
  • Experience with Kafka / AWS MSK.
  • Prior experience in regulated or compliance-sensitive environments (security best practices, access reviews).
  • Experience contributing to a platform/DevEx roadmap that other engineering teams consume as a self-service product.
Our tech stack:
  • Languages: PHP (Symfony), Node.js (TypeScript), Angular (TypeScript).
  • Data: PostgreSQL, Redis, MongoDB.
  • Infra: AWS, Kubernetes, Terraform, Helm, Atlantis
  • Engineering Tools: Postman, Git, GitHub Copilot, PhpStorm, Grafana, Kibana, Prometheus.
  • Remote work Tools: Jira, Miro, Google Workspace, Slack.
  • Development Practices: Pair Programming, Code Reviews, Continuous Integration/Deployment.
What we offer:
  • A global, inclusive team that's as supportive as it is ambitious and serious about getting things done
  • An opportunity to work remotely or in a modern and welcoming office in Riga
  • Flexible working hours (start your day as late as 11 AM)
  • Private health insurance
  • 2 extra paid days off to focus on your mental or physical well-being
  • 1 extra paid day off to celebrate a Birthday or any other celebration of your choice
  • Internal and external learning opportunities
  • Access to mentorship, internal meetups, and hackathons, both on-site and online
  • Free and healthy lunch if you work from the Rīga office
  • Design and order your own merch using our platforms with an employee discount
  • Exciting team-building events and parties you'll never forget!
FYUL is the engine that powers on-demand commerce at global scale.

Formed in 2024 through the merger of Printful, Printify, and Snow Commerce, we bring together tech, talent, and infrastructure to help people turn ideas into beautiful products.

From solo creators to entertainment giants, FYUL powers merch that connects with millions, backed by advanced tech, premium production, and global reach.

We're a fast-growing global company working toward powering great brands, great experiences, and great people.

We are an equal-opportunity workplace. We're committed to diversity and inclusion and make hiring decisions based solely on qualifications, merit, and work experience.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

IT Systems Automation Engineering Manager (remote within EMEA)
IT Systems Automation Engineering Manager (remote within EMEA)

FYUL • Warszawa

Hybrid
PLN 390,000 - 607,000
Remote work flexibility
Private health insurance
Learning opportunities and mentorship
+1
Backend Software Engineer | Node.js | AI Builder | Remote
Backend Software Engineer | Node.js | AI Builder | Remote

Hostinger • Warszawa

On-site
PLN 180,000 - 240,000
360 Growth
Freedom & responsibility
Wellness
+1
Principal Infrastructure Engineer
Principal Infrastructure Engineer

Sezzle • Warszawa

On-site
PLN 564,000 - 938,000
Principal Infrastructure Engineer New Poland, Remote
Principal Infrastructure Engineer New Poland, Remote

Sezzle • Poland

Remote
PLN 520,000 - 866,000
Full-Stack Developer | Node.js | AI Builder | Remote
Full-Stack Developer | Node.js | AI Builder | Remote

JobCubby • Warszawa

Hybrid
PLN 180,000 - 320,000
Growth opportunities
Home office budget
Wellness program
+1
DevOps Engineer (Cloud) - Freelance (Part-time)
DevOps Engineer (Cloud) - Freelance (Part-time)

Lingaro • Poland

Hybrid
PLN 180,000 - 240,000
Office as an option
Workation
Great Place to Work certified
+7
DevOps Engineer
DevOps Engineer

Joy Studios • Warszawa

On-site
PLN 260,000 - 360,000
DevOps & Data Quality Platform Engineer (Remote)
DevOps & Data Quality Platform Engineer (Remote)

Lingarogroup • Poland

Hybrid
USD 140,000 - 200,000
Stable employment
Office or remote work option
Workation policy
+4
IT Support Engineer L2
IT Support Engineer L2

Fundraise Up • Warszawa

Hybrid
PLN 180,000 - 240,000
31 days off
Telemedicine plan
Home office setup assistance
+5
Senior Frontend Product Engineer (React + TypeScript) - B2B
Senior Frontend Product Engineer (React + TypeScript) - B2B

Fresha • Warszawa

On-site
PLN 279,000 - 446,000
RSUs
Private healthcare
Competitive salary