This position works as an SRE Engineer within the Surveillance product family. The role maintains an end to end view of the application and infrastructure ecosystem and manages service delivery to consumers in line with agreed Service Level Agreements.
The engineer automates and implements technical solutions to address business needs. With full ownership of production platforms, the role focuses on maintaining environment stability, resolving production issues quickly, driving continuous service improvements, and coordinating with key stakeholders.
What we offer
- Industry leading leave policies
- Gender neutral parental leave
- Full reimbursement through childcare assistance benefits
- Sponsorship for relevant industry certifications and education
- Employee Assistance Program for employees and their families
- Comprehensive hospitalization insurance for employees and dependents
- Accident and term life insurance coverage
- Complimentary health screening for employees aged 35 and above
Key Responsibilities
- Develop strong technical expertise for a suite of Deutsche Bank applications, including business workflows, architecture, and infrastructure configuration
- Handle incidents and service requests raised by application users and upscale unresolved issues to Level 3 support
- Perform real time monitoring to maintain application availability and meet service level targets, while improving monitoring capabilities
- Build and maintain effective relationships with stakeholders across business, development, infrastructure teams, and third party vendors
- Conduct post incident reviews and feed insights into incident, problem, and change management processes
- Drive service improvements and operational efficiencies that enhance platform stability and customer satisfaction
Skills and Experience
Applicants should demonstrate strong knowledge in several of the following technologies and areas:
- Proficiency in Python or Java
- Experience working with Google Cloud Platform
- Understanding of GCP services such as GKE, Secret Management, BigQuery, Airflow monitoring, and Dataproc
- Experience creating monitoring dashboards using tools like Looker or Tableau
- Familiarity with container technologies including Docker and Kubernetes
- Experience with monitoring and alerting tools such as Prometheus, Grafana, Datadog, New Relic, or Splunk
- Experience working with CI/CD pipelines using tools like Jenkins, GitLab CI, GitHub Actions, or Azure DevOps
- Understanding of networking concepts including TCP/IP, DNS, HTTP, and load balancing
- Knowledge of database technologies including SQL and NoSQL
- Support production platforms through monitoring, issue remediation, and scheduled maintenance activities
- Collaborate with infrastructure and delivery services teams to improve availability and resilience
- Participate in disaster recovery planning and testing
- Contribute to incident investigations, root cause analysis, and knowledge documentation
- Develop runbooks and knowledge articles based on production incidents
- Support the implementation of operational tools and best practices for effective platform support
Support and Development
- Training and development programs to support career growth
- Guidance and mentoring from experienced professionals
- A culture focused on continuous learning and professional progression
- Flexible benefits that can be tailored to individual needs
Deutsche Bank promotes a culture where employees are encouraged to excel together. The organization values responsible actions, commercial thinking, initiative, and collaboration.
The company celebrates the achievements of its people and promotes a positive, fair, and inclusive workplace for all employees.
If an employer asks you to pay any kind of fee, please notify us immediately. Talentd does not charge any fee from applicants and we do not allow other companies to do so.
- Bachelor's degree in Computer Science, Information Technology, or related field