Sr. DevOps Engineer, Data Platform

Netskope

San Francisco (CA)

On-site

USD 140,000 - 185,000

Full time

4 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Netskope is seeking an engineer to ensure reliable operation of production environments for our Data Infrastructure and products. You will lead monitoring, incident response, and automation efforts across a large-scale distributed cloud system, collaborating on capacity planning and CI/CD support.

The role offers hands-on cloud tech exposure, including AWS and GCP, Docker, and Kubernetes. You will contribute to system uptime and user experience, with on-call responsibilities and opportunities to

Qualifications

  • 5+ years of industry experience in a technical role.
  • Proficiency in Python scripting for automation and tooling.
  • Hands-on with containerization/orchestration (Docker, Kubernetes).
  • Strong understanding of public cloud infrastructure (AWS; GCP also considered).
  • Experience with observability/monitoring tools (Grafana, Prometheus) a plus.
  • IaC using Terraform and CI/CD workflows with GitHub Actions a plus.
  • Willingness to participate in on-call rotation and support production incidents.

Responsibilities

  • Monitor system performance using observability dashboards and write runbooks for incident response.
  • Lead incident response as first point of contact for production alerts and perform root cause analysis.
  • Identify repetitive tasks and drive automation to reduce manual work.
  • Support CI/CD by creating, testing, and maintaining automated deployment pipelines.
  • Track capacity and resource usage to forecast scaling needs.

Skills

Scripting (Python)
Incident response
Observability mindset
Linux/Unix proficiency
Cloud concepts (WAN/Networking)

Education

BSCS or equivalent
MSCS or equivalent (strongly preferred)

Tools

Docker
Kubernetes
AWS
GCP
Grafana
Prometheus
Terraform
GitHub Actions

Job description

About the Role:

Please note, this team is hiring across all levels and candidates are individually assessed and appropriately leveled based upon their skills and experience.

We're looking for an engineer to ensure the reliable operation of production environments for our Data Infrastructure and products, running at scale on large-volume distributed cloud systems. You'll focus on maximizing system reliability, automating routine tasks, and improving production efficiency.

This role offers hands-on exposure to modern cloud technologies — Docker, Kubernetes, networking, and platforms like AWS and GCP — while you contribute directly to system uptime and user experience for a large-scale distributed application.

You will lead production monitoring and incident response, drive automation to reduce manual work, and collaborate with Engineering on root cause analysis, CI/CD support, and capacity planning.

What's in it for you:

In this role, you will be responsible for seamless operation of production environments for our Data Infrastructure and products within large-scale, high-volume distributed cloud systems. You will concentrate on maximizing system reliability, automating routine tasks, and maintaining the efficiency of our production systems.

This role provides practical cloud exposure to help you deepen your expertise in modern technology stacks, such as Docker, Kubernetes, Networking, and major public cloud providers like AWS and GCP. You will have the chance to contribute to enhancing user experience and system uptime for a large-scale distributed application.

What you will be doing:
  • Monitoring: Use observability dashboards to monitor system performance, error rates, and resource utilization. Define new dashboards and alerts as and when required and write technical runbooks for Incident response.
  • Incident Response: Act as the initial point of contact for production alerts, mitigate/solve the ongoing issues, escalade highly complex problems to the Engineering team, and conduct thorough root cause analysis for incidents.
  • Automation: Identify repetitive manual tasks and streamline them through automation.
  • Support CI/CD: Help create, test, and maintain automated application deployment pipelines.
  • Capacity Management: Assist in tracking system resource usage (CPU, memory, storage) and traffic volume to help forecast and scale infrastructure needs
Required skills and experience:
  • 5+ years of overall industry experience in a relevant technical role
  • Proficiency in at least one scripting language, preferably Python
  • Hands-on experience with containerization and orchestration technologies such as Docker, Kubernetes, and related cluster concepts
  • Strong understanding of public cloud infrastructure, with a preference for AWS (GCP experience also considered)
  • Working knowledge of Linux/Unix command-line environments and core web protocols including HTTP, gRPC, DNS, and TCP/IP
  • Experience with observability and monitoring tools such as Grafana and Prometheus is a plus
  • Familiarity with Infrastructure as Code (IaC) using Terraform, along with workflow automation via GitHub Actions, is a plus
  • Willingness to participate in an on-call rotation and respond to production incidents with flexibility to support critical systems
Education
  • BSCS or equivalent required, MSCS or equivalent strongly preferred

#LI-JB3

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

DevOps Engineer
DevOps Engineer

Stelvio Inc. • Frisco (TX)

On-site
USD 110,000 - 160,000
Paid vacation
Sick leave
Bereavement leave
+2
Devops Developer
Devops Developer

Itgiantsolutions • United States

On-site
USD 90,000 - 120,000
Sr. DevOps Engineer
Sr. DevOps Engineer

TalentOla • Boston (MA)

Hybrid
USD 130,000 - 160,000
DevOps Engineer
DevOps Engineer

Evlo AI • Washington

On-site
USD 120,000 - 160,000
DevOps Engineer
DevOps Engineer

Careernet • Coopersburg

Hybrid
USD 120,000 - 140,000
DevOps Engineer
DevOps Engineer

BrothersTech • United States

On-site
USD 120,000 - 160,000
Sr. Lead DevOps Engineer
Sr. Lead DevOps Engineer

Solomon Page • New York (NY)

On-site
USD 200,000 - 225,000
DevOps Engineer
DevOps Engineer

Eurobase People • Germany (OH)

Hybrid
USD 90,000 - 130,000
Access to modern cloud technologies
Ownership of impactful projects
Culture that values learning and growth
DevOps Engineer
DevOps Engineer

VTG Defense • McLean (VA)

On-site
USD 100,000 - 140,000
DevOps Engineer
DevOps Engineer

VT Group (VTG) • Chantilly (VA)

On-site
USD 120,000 - 150,000