Job Summary
System Engineer II - Infrastructure Support (GCP / Kubernetes / Windows)
Providing day-to-day support for GCP infrastructure and Kubernetes-based infrastructure, with a focus on Google Cloud Compute Engine, Windows environments, and automated delivery through Terraform and GitHub.
Supports the operation, stability, and ongoing improvement of cloud-based infrastructure services, ensuring reliable performance in a production and compliance-driven environment.
Responsibilities
- Support day-to-day operation of GCP-based infrastructure, including Compute Engine workloads and Kubernetes platform and cluster configurations
- Work with leadership and technical teams to support the strategic direction of the Infrastructure Engineering function
- Support automated build and deployment practices using infrastructure as code and CI/CD pipelines
- Maintain installation, configuration, and operational procedures, and contribute to improving runbooks and documentation
- Participate in infrastructure support and engineering activities across BAU operations and project delivery
- Provide incident support, troubleshooting, and root cause analysis for production issues
- Work closely with networking, security, and application teams to resolve issues and support service delivery
- Support compliance, audit, and security requirements across managed systems
- Contribute to improving the availability, performance, and resilience of infrastructure services
What Are We Looking For in This Role
Minimum Qualifications
- Bachelor's degree in Computer Science or related field, or equivalent experience.
- 2-5 years of hands-on experience supporting cloud-based infrastructure in a production environment, with exposure to GCP (Compute Engine), Windows systems, and Kubernetes-based platforms.
- General understanding of cloud networking and access controls, with experience working alongside networking and security teams.
- Good communication skills.
Desired Skills and Capabilities
- Hands-on experience supporting Google Cloud Platform (GCP) environments from an infrastructure support perspective, with focus on Google Compute Engine (GCE) and Windows-based workloads
- Good working understanding of cloud networking and access controls (including service accounts), with experience working alongside dedicated networking and security teams
- Working knowledge of Kubernetes (GKE preferred) from an infrastructure support perspective, including cluster configuration, namespace configuration, ingress/service routing, and security configuration, along with troubleshooting of infrastructure-related issues
- Solid Windows systems engineering background, including server build, patching, and operational support in a cloud-hosted environment
- Experience with infrastructure as code, preferably Terraform, for provisioning and maintaining cloud resources
- Familiarity with GitHub or similar source control platforms, supporting CI/CD-driven workflows and deployments
- Solid understanding of production support practices, including incident support, troubleshooting, root cause analysis, and working from runbooks
- Experience supporting services in a regulated or compliance-driven environment (e.g. PCI), with awareness of security, audit, and change requirements
- Working knowledge of Active Directory, including user and group management and basic policy administration
- Basic working knowledge of Linux systems, sufficient to support Kubernetes-based workloads
- Understanding of security and access control concepts, including IAM, service accounts, and patch and vulnerability management
- Ability to contribute to both operational support and project delivery activities, working across application, infrastructure, and security teams in a distributed environment
- Strong collaboration and communication skills
Key Qualifications
- 2-5 years experience supporting cloud infrastructure in production environments
- Hands-on experience with GCP (Google Compute Engine) and Windows-based workloads
- Working knowledge of Kubernetes (GKE preferred), supporting application deployments and troubleshooting
- Experience with Terraform and infrastructure as code, and GitHub-based workflows
- Strong production support skills, including incident response, troubleshooting, and root cause analysis
- Working knowledge of Active Directory and core access management concepts
- Basic understanding of Linux systems in support of Kubernetes workloads
- Familiarity with security, access control, and compliance requirements (e.g. PCI)
- Experience working alongside networking and security teams in a shared responsibility model