A leading technology firm is seeking an experienced Site Reliability Engineering (SRE) / DevOps Manager to lead teams and drive reliability roadmaps. This role involves optimizing platform reliability and performance, developing strategies aligned with business goals, overseeing incident management, and ensuring scalable and secure infrastructure in a cloud environment. The ideal candidate will have a strong background in cloud-native technologies and excellent leadership skills.
Qualifications
Minimum of 3 years' management experience in SRE or DevOps roles.
Strong background in cloud-native environments and technologies.
Experience with incident management and reliability improvement frameworks.
Excellent leadership and team management skills.
Strong problem-solving abilities and a proactive approach to challenges.
Responsibilities
Lead and manage SRE and DevOps teams to optimize platform reliability and performance.
Develop and implement reliability roadmaps, ensuring alignment with business goals.
Collaborate with cross-functional teams to drive cloud-native strategies and solutions.
Oversee incident management processes and implement proactive measures to minimize downtime.
Ensure the scalability and security of our infrastructure in a cloud environment.
Skills
Management experience in SRE or DevOps
Cloud-native environments
Incident management
Leadership skills
Problem-solving abilities
Job description
A leading technology firm is seeking an experienced Site Reliability Engineering (SRE) / DevOps Manager to lead teams and drive reliability roadmaps. This role involves optimizing platform reliability and performance, developing strategies aligned with business goals, overseeing incident management, and ensuring scalable and secure infrastructure in a cloud environment. The ideal candidate will have a strong background in cloud-native technologies and excellent leadership skills.