As a preferred supplier to one of our biggest Financial Client, I am seeking for a SRE for a position in Amsterdam, Netherlands. As a Reliability Engineer, you will work across two DevOps teams: PBDB X-files and PBDB Platform & Generic. Together, these teams provide the foundation for our international private banking applications in Germany, France, and Belgium. In this role, you will help ensure the reliability and continuity of both blocks by investigating incidents, identifying root causes, improving monitoring and observability, supporting releases, and driving structural improvements across the application landscape. This gives you a unique opportunity to combine operational excellence with continuous improvement in a complex international banking environment.
Your main responsibilities include:
- Investigating customer-impacting issues and acting as a point of contact for support and operational stakeholders.
- Analyzing incidents and problems, identifying root causes, and coordinating resolution activities across teams and domains.
- Taking ownership of reliability topics that extend beyond the scope of a single application or team.
- Monitoring application performance, availability, and operational health, and translating insights into improvement initiatives.
- Creating and maintaining operational documentation and knowledge articles.
- Supporting testing, release validation, and production verification activities.
- Working closely with Software Engineers, QA Engineers, Product Owners, and other stakeholders to improve service reliability.
- Contributing to incident, problem, change, and crisis management processes.
- Participating in refinements, risk assessments, and operational readiness reviews.
- Identifying automation opportunities to reduce operational effort and improve service resilience.
- Driving improvements in observability, monitoring, alerting, and operational excellence.
Your profile:
- A proactive, analytical, and pragmatic mindset.
- At least 2 years of experience in Reliability Engineering, Operations, Site Reliability Engineering (SRE), Application Support, or a similar role within an agile environment.
- Strong troubleshooting and root cause analysis skills.
- Experience with incident management, problem management, change management, and crisis management.
- Experience working in DevOps teams and collaborating closely with software engineers.
- Knowledge of application architectures, APIs, and distributed systems.
- Experience with monitoring and observability platforms such as Splunk, Azure Monitor, AppDynamics, or similar tooling.
- The ability to translate operational insights into structural improvements and automation opportunities.
- Strong communication skills, both verbal and written, in English.
- A continuous improvement mindset and willingness to challenge existing processes.
- ITIL4 certification is considered a plus.
- Experience within the financial services or banking sector is considered a plus.