Platform Infrastructure Engineer (SRE Core)

Menlo Security

United States

On-site

USD 80,000 - 120,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Menlo Security is seeking a Platform Infrastructure Engineer to design, deploy, and manage our global infrastructure on GCP and AWS. You will build IaC with Terraform and Spacelift, manage networking, security, and multi-region deployments, and maintain observability with Grafana/Prometheus.

You’ll collaborate across teams to align on capacity planning and disaster recovery, while contributing to a secure, scalable platform.

Qualifications

  • Bachelor’s degree in Computer Science or related field or equivalent experience.
  • Proficiency in Python, Bash and Go for scripting and tooling.
  • Strong Kubernetes expertise and Terraform for IaC.

Responsibilities

  • Implement, deploy, and maintain VM and Kubernetes infrastructure on GCP and AWS across multiple regions.
  • Build and maintain IaC using Terraform modules and Spacelift, provisioning networking, compute, storage, and security.
  • Implement observability using Grafana Cloud, Prometheus, and OTel collectors with dashboards and alerts.
  • Manage certificate lifecycle, DNS automation, ingress, and service mesh networking.
  • Collaborate with Engineering, Product, Compliance, and Security on capacity planning and architecture decisions.
  • Automate toil with scripts and CI/CD pipelines; participate in 24x7 on-call rotation.

Skills

Python
Bash
Go
Kubernetes
Terraform
Spacelift
LLM-based coding
Networking
GCP
AWS
GitOps

Education

Bachelor’s degree in Computer Science or related field

Tools

Terraform
Spacelift
Grafana Cloud
Prometheus
Kubernetes tooling
GKE
TLS/HTTPs networks

Job description

Menlo Security’s mission is enabling the world to connect, communicate and collaborate securely without compromise. COVID-19 has made our mission all the more real. We support customers across various enterprises including Fortune 500 companies, 9/10 of the largest global banks and the Department of Defense.

The world has fundamentally changed. We are growing from 400 employees into the next phase of our journey, and we need passionate talent filled with empathy and agility. The right candidate for the job is ethical, hyper-organized, fanatical about seeing things through to completion, service-oriented, and humble enough to take feedback and coaching yet confident enough to provide feedback and coaching.

Menlo is well-funded for growth and our investors are second to none. They include Vista Equity Partners (“Vista”), General Catalyst, JPMC, American Express, HSBC, and Ericsson Ventures.

Summary

Platform Infrastructure Engineering builds and operates Menlo Security’s Infrastructure Platform, enabling our customers to connect to the Internet without compromise. As a Platform Infrastructure Engineer, you’ll join a globally distributed team of experienced engineers building and managing the company’s core infrastructure services on a cloud-native platform built on Google Kubernetes Engine and VMs spanning multiple regions and environments. The team manages infrastructure as code with Terraform and Spacelift, deploys with Helm, and emphasizes security-first design, comprehensive observability, and multi-region resilience. The team also uses AI-assisted development and code-review tools, including Gemini Code Assist, as part of the standard engineering workflow, and this role is expected to use LLM-based tooling to build and troubleshoot infrastructure code efficiently.

Outcomes & KPIs
Key Outcome(s) Owned:
  • Reliable, secure, and scalable infrastructure across GCP and AWS supporting Menlo’s platform globally.

  • Reduced operational toil and incident recurrence through automation and Infrastructure as Code practices.

  • Comprehensive, end-to-end observability framework providing deep platform visibility, proactive health monitoring, and accelerated incident detection and resolution.

Success Metrics / KPIs:
  • Infrastructure uptime/availability across regions (e.g., 99.9%+)

  • Mean time to detect (MTTD) and mean time to resolve (MTTR) for incidents

  • Percentage of infrastructure changes deployed via IaC (Terraform) vs. manual changes

  • On‑call incident volume and reduction in repeat/preventable incidents

  • Lead time for provisioning new infrastructure

What You’ll Do
  • Implement, deploy, and maintain VM and Kubernetes infrastructure on GCP and AWS across dozens of clusters spanning development, staging, and production environments in multiple regions

  • Build and maintain Infrastructure as Code using Terraform modules and Spacelift (or equivalent TACOS), provisioning networking, compute, storage, and security components, and implementing multi‑layer configuration management workflows

  • Implement and maintain observability solutions using Grafana Cloud, Prometheus/Mimir, and OTel collectors, designing dashboards and alerting rules across all platform components

  • Manage certificate lifecycle, DNS automation, ingress controllers, and service mesh networking with Cilium

  • Partner with peers and across Engineering, Product, Compliance, and Security teams to align on requirements and consult on capacity planning, disaster recovery, and architectural decisions

  • Identify and eliminate toil through automation — writing scripts, building CI/CD pipelines, and using AI-assisted coding tools to move faster

  • Participate in a 24×7 on‑call rotation as part of a globally distributed team, responding to incidents and driving post‑incident reviews

Functional Competencies
Required:
  • Bachelor’s degree in Computer Science, a related technical field, or equivalent practical experience

  • Proficiency in common programming and scripting languages, particularly Python, Bash, and Go

  • Understanding of network topologies, communication protocols (e.g., TCP/IP, HTTP/S, UDP, TLS), and enterprise‑grade connectivity solutions

  • Kubernetes expertise, including cluster administration, RBAC, networking, workload management, and troubleshooting in production environments

  • Proven experience with Terraform for infrastructure provisioning and management

  • Knowledge of Google Cloud Platform services including GKE, VPC networking, Cloud DNS, Artifact Registry, Secret Manager, IAM, Gemini Code Assist, and Workload Identity

  • Clear understanding of how to use LLM-based code‑assist tools to effectively build and troubleshoot software

Preferred / Nice to Have:
  • Experience with GitOps methodologies and tools
Our Compensation and Benefits

At Menlo Security, Base Salary is one part of our competitive total compensation and benefits package and is determined using a salary range. The base salary range for this role is 112,000 CAD – 168,000 CAD.

In accordance with Canadian law, the range provided is Menlo Security’s reasonable estimate of the base compensation for this role. The actual amount may be higher or lower, based on non‑discriminatory factors such as experience, knowledge, skills, abilities, and location. All employees may be eligible to become Menlo Security shareholders through eligibility for stock‑based compensation grants, which are awarded to employees based on company and individual performance.

Menlo Security does not accept unsolicited resumes from search firm recruiters. Fees will not be paid in the event a candidate submitted by a recruiter without an agreement in place is hired; such resumes will be deemed the sole property of Menlo Security.

Menlo Security is an equal opportunity employer. All aspects of employment will be based on merit, competence, performance, and business needs. We do not discriminate on the basis of race, color, religion, marital status, age, national origin, ancestry, physical or mental disability, medical condition, pregnancy, genetic information, gender, sexual orientation, gender identity or expression, veteran status, or any other status protected under federal, state, or local law.

All qualified applicants will receive consideration for employment without regard to race, sex, color, religion, sexual orientation, gender identity, national origin, protected veteran status, or on the basis of disability

Why Menlo?

Our culture is collaborative, inclusive, and fun! We have five core values: Stay Aligned, Get It Done, Customer Empathy, Think Creatively and Help Each Other Out. We believe in open communication, supporting new ideas, and sharing a mutual mindset of what we’re aiming to achieve together. There are tremendous opportunities to take initiative, implement new ideas, and have a hand in building a legacy.

TO ALL AGENCIES: Please, no phone calls or emails to any employee of Menlo Security outside of the Talent organization. Menlo Security’s policy is to only accept resumes from agencies via Ashby (ATS). Agencies must have a valid services agreement executed and must have been assigned by the Talent team to a specific requisition. Any resume submitted outside of this process will be deemed the sole property of Menlo Security. In the event a candidate submitted outside of this policy is hired, no fee or payment will be paid.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Business Development Representative - Public Sector
Senior Business Development Representative - Public Sector

Menlo Security • United States

On-site
USD 58,000 - 75,000
Senior HRIS Analyst
Senior HRIS Analyst

Menlo Security Inc. • United States

On-site
USD 90,000 - 130,000
Product Designer
Product Designer

Menlo Security • Mountain View (CA)

Hybrid
USD 126,000 - 190,000
Senior Business Development Representative - Public Sector
Senior Business Development Representative - Public Sector

Menlo Security Inc. • United States

On-site
USD 58,000 - 75,000
Stock-based compensation grants
Security Engineer
Security Engineer

Menlo Research • San Francisco (CA)

On-site
USD 120,000 - 160,000
Manager, Software Engineering (Cortex Platform)
Manager, Software Engineering (Cortex Platform)

Palo Alto Networks • United States

Hybrid
USD 170,000 - 210,000
Hybrid work model
Competitive compensation
Employee benefits
Manager, Software Engineering (Cortex Platform)
Manager, Software Engineering (Cortex Platform)

Palo Alto Networks, Inc. • Santa Clara (CA)

Hybrid
USD 190,000 - 320,000
Senior Platform Security Engineer
Senior Platform Security Engineer

Socket.dev • New York (NY)

Hybrid
USD 185,000 - 232,000
Comprehensive health, vision, dental, and disability coverage
Equity options
Flexible work culture focused on ownership and impact
Manager, Security Engineering, Cloud & AppSec
Manager, Security Engineering, Cloud & AppSec

Horizon3 AI • United States

On-site
USD 149,000 - 185,000
Health, vision & dental insurance
Flexible vacation policy
Generous parental leave
+2
Senior Software Developer
Senior Software Developer

Menlo Innovations • Ann Arbor (MI)

On-site
USD 80,000 - 120,000
Health insurance (73%-87% subsidized)
Life and disability insurance
20 days paid time off annually
+3