Senior Software Engineer (Application Operations)

Reward Gateway

Greater London

Hybrid

GBP 65,000 - 95,000

Full time

9 days ago

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

On-call allowance £500/week
Hybrid work in London
Enhanced rates for public holidays

Job summary

Reward Gateway is looking for a hands-on Application Operations Engineer to improve the reliability of business-critical PHP applications. You’ll work across PHP, MySQL and AWS, investigate incidents, perform root cause analysis, and build automation to reduce operational effort.

The role emphasizes production operations, observability and continuous improvement, with on-call duties and a hybrid London-based pattern.

Qualifications

  • Strong PHP experience with troubleshooting and safe fixes.
  • Operational support across cloud environments.
  • Experience supporting production applications in medium to large organisations.

Responsibilities

  • Investigate, troubleshoot and resolve production issues across PHP apps, APIs, MySQL, and AWS.
  • Participate in on-call rota and incident response.
  • Perform root cause analysis and implement long-term fixes.
  • Analyze logs, metrics, and traces with Datadog and Kibana.
  • Write production-grade PHP code and automation to reduce manual work.

Skills

PHP
MySQL
AWS
Datadog
Kibana
Incident management
On-call

Tools

CI/CD

Job description

Reward Gateway, part of Edenred, helps organisations engage, motivate and retain their people through employee benefits, recognition and wellbeing solutions. Our platforms support millions of users globally and play a critical role in our clients' day-to-day operations.

As part of our Platform Engineering & Technical Operations team, you'll play a key role in ensuring our business-critical applications remain stable, reliable and performant for millions of users worldwide.

Your Role in Our Mission

We're looking for a hands-on engineer who enjoys solving complex production problems and improving the reliability of business-critical applications. Working across our PHP, MySQL and AWS estate, you'll investigate incidents, perform root cause analysis, build automation and implement lasting engineering solutions that reduce operational effort and improve customer experience.

This role offers a unique opportunity to combine software engineering, operational excellence and reliability engineering while helping shape the future of Application Operations at Reward Gateway.

PHP applications running in AWS

MySQL databases

Kibana log analysis

Modern engineering and CI/CD practices

Working Pattern & On-Call

Standard hours: Monday to Friday, 9am to 6pm.

1-in-4 on-call rotation.

£500 on-call allowance per week (in addition to base salary).

Additional hourly payments for call-outs.

Enhanced rates for public holiday coverage.

Historically low call-out volumes, supported by a strong focus on automation, operational maturity and continuous improvement.

This role offers a hybrid work model to be present in our London office twice a week.

Why Join Us?

Application Operations is a growing engineering capability at Reward Gateway, focused on reliability, automation and continuous improvement.

You’ll have the opportunity to:

  • Solve complex production challenges across customer-facing platforms.
  • Influence the reliability and performance of services used by millions of people worldwide.
  • Build automation and tooling that removes repetitive work and delivers measurable business value.
  • Develop expertise across software engineering, cloud operations, observability and platform reliability.
  • Work closely with Engineering, Product, Platform and SRE teams to drive meaningful technical improvements.
  • Join a collaborative team that values ownership, learning and continuous improvement.
  • This removes a lot of the repetition around "improving reliability", "automation", "operational excellence" and "production incidents" while still selling the role.
What You'll Be Doing
  • Investigate, troubleshoot and resolve production issues across PHP applications, APIs, MySQL databases and AWS-hosted services.
  • Participate in the on-call rota, support incident response activities and contribute to post-incident reviews.
  • Perform root cause analysis, identify recurring issues and implement long-term fixes.
  • Analyse logs, metrics and traces using Datadog, Kibana and supporting observability tools.
  • Assess customer and business impact during incidents and communicate progress to technical and non-technical stakeholders.
  • Write production-quality PHP code, automation and operational tooling to reduce manual effort and improve reliability.
  • Develop and maintain runbooks, playbooks and operational documentation.
  • Partner with Engineering, Platform and SRE teams to improve application performance, resilience and supportability.
  • Enhance monitoring, alerting and operational visibility across services.
  • Contribute to continuous improvement initiatives that reduce operational toil and increase service reliability.

A job description is available on request.

Key Skills & Experience Required
  • Experience supporting customer-facing production applications in medium to large scale organisations, including incident management, root cause analysis and operational support across complex cloud-based environments.
  • Strong PHP experience, with the ability to investigate issues, troubleshoot code and implement safe fixes where required.
  • Strong MySQL operational knowledge, including query analysis, performance troubleshooting, slow query investigation and database issue diagnosis.
  • Proven experience investigating and resolving complex production issues across applications, APIs, integrations and databases.
  • Experience managing or contributing to the resolution of high-priority production incidents, including root cause analysis and post-incident improvement activities.
  • Hands-on experience using Datadog (or equivalent observability platforms), together with log analysis tools such as Kibana, to investigate and diagnose production issues.
  • Experience assessing customer and business impact during incidents and using this information to prioritise remediation activities.
  • Strong communication and documentation skills, including maintaining runbooks, operational documentation and providing clear stakeholder updates during incidents.
  • Familiarity with ITSM tooling and incident/problem management processes, with a focus on reducing operational toil through automation and continuous improvement.
Interview Process
  • Screening call with a member of the Talent Acquisition Team
  • First stage interview with Application Operations Leadership and peer (practical scenario or technical assessment relevant to the L2.5 operating model)
  • Final stage interview with Director or VP

At Reward Gateway | Edenred we are committed to ensuring an inclusive and accessible recruitment process for all candidates. If you have any specific requirements or need reasonable adjustments at any stage of the recruitment journey, please let your Talent Acquisition Partner know. Your needs are important to us, and we want to ensure an equitable experience for every candidate.

Be Comfortable. Be You.

At Reward Gateway | Edenred, we want everyone to feel comfortable bringing their passion, creativity and individuality to work. We value diverse backgrounds, perspectives and experiences because we believe diversity drives innovation. Join us and help make the world a better place to work.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Staff QA Engineer
Staff QA Engineer

Reward Gateway • City Of London

Hybrid
GBP 75,000 - 80,000
Flexible holiday up to 40 days
Wellbeing allowance GBP 400
Private Medical Insurance
+4
Product Analytics Lead
Product Analytics Lead

Reward Gateway • Greater London

On-site
GBP 60,000 - 70,000
Flexible holiday plan up to 40 days per year
£400 a year Wellbeing Allowance
Private Medical Insurance
+3
Sales Consultant
Sales Consultant

Reward Gateway • Greater London

Hybrid
GBP 40,000 - 60,000
Flexible holiday plan of up to 40 days
£400 a year Wellbeing Allowance
Private Medical Insurance
+2
Staff QA Engineer
Staff QA Engineer

JobCubby • Greater London

Hybrid
GBP 75,000 - 80,000
Flexible holidays up to 40 days per yr
Wellbeing allowance £400/yr
Private medical insurance
+3
Senior Technical Programme Manager (Senior TPM)
Senior Technical Programme Manager (Senior TPM)

Rewardgateway • Greater London

On-site
GBP 85,000 - 137,000
VP of Enterprise Architecture
VP of Enterprise Architecture

Reward Gateway • Greater London

On-site
GBP 180,000 - 200,000
Staff Quality Assurance Engineer
Staff Quality Assurance Engineer

Reward Gateway • Greater London

Hybrid
GBP 90,000 - 120,000
Flexible holiday plan up to 40 days
Wellbeing allowance £400/year
Private Medical Insurance
+3
Staff QA Engineer
Staff QA Engineer

Reward Gateway • Greater London

Hybrid
GBP 75,000 - 80,000
40 days annual holiday
Wellbeing allowance £400/year
Private medical insurance
+3
Senior Software Engineer (Application Operations)
Senior Software Engineer (Application Operations)

Rewardgateway • Greater London

Hybrid
GBP 70,000 - 100,000
Hybrid work model
On-call rotation compensation
Senior Software Engineer (Python / TypeScript / JavaScript)
Senior Software Engineer (Python / TypeScript / JavaScript)

United States Digital Space LLC • Greater London

Hybrid
GBP 90,000 - 130,000
L&D allowance
Hybrid work structure
25 days holiday + public holidays
+1