Get more replies from employers
Send a job-specific resume in minutes.
Netskope is seeking an engineer to ensure reliable operation of production environments for our Data Infrastructure and products. You will lead monitoring, incident response, and automation efforts across a large-scale distributed cloud system, collaborating on capacity planning and CI/CD support.
The role offers hands-on cloud tech exposure, including AWS and GCP, Docker, and Kubernetes. You will contribute to system uptime and user experience, with on-call responsibilities and opportunities to
Please note, this team is hiring across all levels and candidates are individually assessed and appropriately leveled based upon their skills and experience.
We're looking for an engineer to ensure the reliable operation of production environments for our Data Infrastructure and products, running at scale on large-volume distributed cloud systems. You'll focus on maximizing system reliability, automating routine tasks, and improving production efficiency.
This role offers hands-on exposure to modern cloud technologies — Docker, Kubernetes, networking, and platforms like AWS and GCP — while you contribute directly to system uptime and user experience for a large-scale distributed application.
You will lead production monitoring and incident response, drive automation to reduce manual work, and collaborate with Engineering on root cause analysis, CI/CD support, and capacity planning.
In this role, you will be responsible for seamless operation of production environments for our Data Infrastructure and products within large-scale, high-volume distributed cloud systems. You will concentrate on maximizing system reliability, automating routine tasks, and maintaining the efficiency of our production systems.
This role provides practical cloud exposure to help you deepen your expertise in modern technology stacks, such as Docker, Kubernetes, Networking, and major public cloud providers like AWS and GCP. You will have the chance to contribute to enhancing user experience and system uptime for a large-scale distributed application.
#LI-JB3