Senior SRE/DevOps (Platform Tribe)

Playson

United States

Remote

USD 140,000 - 210,000

Full time

14 days+
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Benefits offered by this job

Competitive Salary
Quarterly Bonuses
Unlimited Paid Vacation
Unlimited Paid Sick Leave
Flexible Schedule
Remote Work
Medical Insurance
Financial Support for Life Events
Professional Development
International exposure
B2B contracts

Job summary

Playson in the United States is seeking a Senior SRE/DevOps Engineer to own reliability and performance for a high-traffic, microservices-based platform. You’ll operate at scale, handle incidents, on-call rotations, and drive automation across Kubernetes, CI/CD, and infrastructure as code.

Join a lean Platform Tribe and partner with engineering to minimize user impact during deployments, improve observability, and introduce new tools for scalability.

Qualifications

  • Hands-on Kubernetes in high-load environments.
  • GitOps experience with FluxCD or ArgoCD.
  • Incident response and postmortems in production.
  • AWS, Terraform, Docker, and CI/CD expertise.
  • Observability with Datadog/Prometheus/Grafana and ELK/CloudWatch.
  • Networking fundamentals and scripting (Python/Go/Node).

Responsibilities

  • Own system reliability by monitoring health and responding to incidents in real time.
  • Participate in 24/7 on-call rotations for production stability.
  • Investigate incidents, perform RCAs, and implement long-term fixes.
  • Improve monitoring and observability across Kubernetes (EKS).
  • Deploy and manage infrastructure with Terraform, Helm, Flux/ArgoCD.
  • Automate and improve resilience to reduce manual work.
  • Maintain CI/CD pipelines and IaC practices.
  • Collaborate with engineering to minimize user impact during deployments.
  • Introduce new tools to boost scalability and performance.
  • Handle environment-specific requests for day-to-day platform ops.

Skills

Kubernetes
GitOps
AWS
Terraform
Docker
CI/CD
Observability
Scripting
Git
Incident response
On-call
Networking
PagerDuty

Tools

FluxCD
ArgoCD
Terraform (IaC)
Docker
Kubernetes

Job description

About the Role

Playson is a globally recognised iGaming supplier delivering a high-performance, microservices-based platform designed to process billions of financial transactions every day. Operating at high scale (5–7k RPS), our cross-regional infrastructure is built to deliver near-zero latency and a seamless player experience under constant load.

We’re looking for a Senior SRE / DevOps Engineer to join our Platform Tribe - a lean & senior team where ownership is high and expectations are even higher. This is a deeply hands-on role at the core of a high-traffic system, where you’ll be directly responsible for maintaining reliability, performance, and stability in a fast-paced environment.

You’ll be working on real-time production challenges, handling incidents, managing alerts, and being part of a critical on-call rotation. This role requires resilience, strong decision-making under pressure, and a proactive mindset to continuously improve systems operating at scale.

If you thrive in high-load environments, enjoy solving complex production issues, and want to have a direct impact on systems used by millions - this is the place for you.

Key Responsibilities

  • Own system reliability by actively monitoring platform health, managing alerts, and responding to incidents in real time

  • Participate in 24/7 on-call rotations, taking full ownership of production stability in a high-traffic (5–7k RPS) environment

  • Investigate incidents, perform root cause analysis, and implement long-term fixes to prevent recurrence

  • Build and continuously improve monitoring, alerting, and observability across the Kubernetes (EKS) ecosystem

  • Deploy, manage, and optimise infrastructure using Terraform, Helm, and GitOps tools (Flux/ArgoCD)

  • Drive automation and proactively improve system resilience, reducing manual intervention and recurring issues

  • Maintain and evolve CI/CD pipelines and infrastructure-as-code practices

  • Collaborate closely with engineering teams to support deployments and minimise user impact in a live environment

  • Introduce and integrate new tools and technologies to enhance scalability, reliability, and performance

  • Handle environment-specific requests and ensure smooth day-to-day platform operations under constant load

Requirements

  • Strong hands-on experience with Kubernetes (deployment, scaling, troubleshooting) in high-load environments

  • Experience with GitOps tools such as FluxCD or ArgoCD

  • Proven experience in incident response, root cause analysis, and postmortems in production systems

  • Solid experience with AWS, Terraform, Docker, and CI/CD pipelines

  • Experience with monitoring and observability tools such as Datadog, Prometheus, Grafana, and logging stacks like ELK or CloudWatch

  • Strong understanding of networking concepts and protocols

  • Proficiency in at least one scripting language (e.g. Python, Go, Node.js)

  • Experience working with version control systems (Git)

  • Familiarity with incident management tools like PagerDuty, Opsgenie, or similar

  • Ability to operate effectively in a fast-paced, high-pressure environment with strong ownership and accountability

  • Proactive, resilient mindset with a focus on continuous improvement and system stability

What We Offer

  • Competitive Salary: We offer a competitive salary, subject to annual performance reviews

  • Quarterly Bonuses: Benefit from a transparent and systematic quarterly bonus system

  • Unlimited Paid Vacation: Enjoy unlimited paid vacation leave, including Ukrainian bank holidays

  • Unlimited Paid Sick Leave: Take unlimited paid sick leave whenever necessary

  • Flexible Schedule: We offer a flexible work schedule to accommodate your needs

  • Remote Work: Choose to work remotely, providing greater flexibility and comfort

  • Medical Insurance: Receive comprehensive medical insurance for both you and a significant other

  • Financial Support for Life Events: We provide financial support during special life events

  • Professional Development: Get reimbursement for professional development courses and training

  • International exposure : Attend industry expos, team gatherings & global meet-ups

  • B2B contracts

Recruitment Process

  1. HR Interview (30-45 min)

  2. Interview with a Product Owner (60 min)

  3. Technical interview (90 min)

  4. Final Interview with C-level (60 min)

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Backend Engineer (Platform Tribe)
Senior Backend Engineer (Platform Tribe)

Playson • United States

Remote
USD 120,000 - 160,000
Competitive Salary
Quarterly Bonuses
Unlimited Paid Vacation
+8
Site Reliability Engineer
Site Reliability Engineer

sportygroup • United States

Remote
USD 150,000 - 210,000
Remote-first company
Competitive salary
Weekend Site Reliability Engineer
Weekend Site Reliability Engineer

sportygroup • United States

Remote
USD 140,000 - 170,000
Remote-first company
Competitive salary with quarterlyBonu
28 days paid annual leave
+4
Weekend Site Reliability Engineer
Weekend Site Reliability Engineer

Sporty Group • United States

On-site
USD 120,000 - 180,000
Remote first
Bonuses (quarterly)
28 days leave
+4
Site Reliability Engineer (SRE)
Site Reliability Engineer (SRE)

Social Discovery Ventures • United States

Remote
USD 120,000 - 180,000
REMOTE OPPORTUNITY
Vacation 28 days/year
Wellness days 7/year
+4
Senior SRE/DevOps Engineer – Remote, High-Load Kubernetes
Senior SRE/DevOps Engineer – Remote, High-Load Kubernetes

Playson • United States

Remote
USD 140,000 - 210,000
Competitive Salary
Quarterly Bonuses
Unlimited Paid Vacation
+8
Site Reliability Engineer
Site Reliability Engineer

Sporty Group • United States

Remote
USD 140,000 - 190,000
Remote-first company
Quarterly performance bonuses
28 days paid annual leave
+4
Senior DevOps Engineer
Senior DevOps Engineer

Linuxconfig • Northern (KY)

On-site
USD 120,000 - 160,000
Fully remote work
Paid vacation
Private medical insurance
+1
Senior Engineer - DevOps
Senior Engineer - DevOps

Hard Rock Digital • United States

On-site
USD 160,000 - 210,000
Hybrid/Remote work options
Competitive compensation
Senior DevOps Engineer/Site Reliability Engineer-East Coast
Senior DevOps Engineer/Site Reliability Engineer-East Coast

Stellar Cyber • United States

Remote
USD 165,000 - 215,000
Pre-IPO Stock Options
Medical, Dental & Vision care
401(k)
+6