- Own foundational infrastructure and the internal developer platform used by engineering teams
- Build golden-path templates, self-service tooling, local development environments, and CI/CD pipelines
- Manage GCP, AWS, and/or Azure infrastructure using Terraform and Kubernetes operators/Crossplane
- Operate Kubernetes fleets, including cluster lifecycle, upgrades, autoscaling, node management, and multi-cluster patterns
- Define service packaging and deployment with Helm
- Design service mesh, ingress, mTLS, routing, rate limiting, canary releases, circuit breaking, and RPC traffic management
- Own Argo CD release and deployment workflows, progressive delivery, rollback safety, and environment promotion
- Operate observability using OpenTelemetry, Prometheus, Grafana, distributed tracing, logging, alerting, and dashboards
- Build zero-trust access, VPN, BeyondCorp, short-lived credentials, and on-premises connectivity
- Partner with Security on secrets management, policy-as-code, and secure-by-default patterns meeting HIPAA and SOC 2 requirements
- Make architectural decisions, write production code, and establish engineering patterns for other teams
- Report security events or risks to the information security office and attest to security requirements upon hire and annually
Requirements
- 6+ years of software engineering experience in infrastructure, platform, or site reliability engineering roles
- Experience building internal developer platforms with a strong product mindset
- Experience with public cloud platforms: GCP, AWS, and/or Azure
- Experience with on-premises or hybrid environments and data center/cloud migrations
- Experience with cloud-native technologies
- Experience with Infrastructure-as-Code using Terraform or Pulumi
- Experience with controller-based infrastructure management, including Crossplane and Kubernetes operators
- Experience with service mesh technologies and software-defined networking
- Experience with GitOps and progressive delivery workflows, including Argo CD, Flux, or Kargo
- Experience with observability tools including Prometheus, Grafana, and OpenTelemetry
- Experience in regulated industries with HIPAA and SOC 2 obligations
- Ability to work in the United States; sponsorship question included in application
Core Competencies
Demonstrates expertise in managing cloud infrastructure across GCP, AWS, and Azure, utilizing Infrastructure-as-Code with Terraform and Kubernetes. Proficient in building internal developer platforms, implementing observability tools, and ensuring compliance with HIPAA and SOC 2 standards.
Highest-signal resume keywords
- Cloud Infrastructure Management
- Infrastructure-as-Code with Terraform
- Kubernetes Operations
- Observability Tools (Prometheus, Grafana)
- Regulated Industry Compliance (HIPAA, SOC 2)
ATS Optimization Keywords
Hard Skills
- Software Engineering
- Infrastructure Management
- Cloud-Native Technologies
- Service Mesh Technologies
- GitOps Workflows
- CI/CD Pipelines
- Controller-Based Infrastructure ManagementService Packaging with Helm
- Progressive Delivery Workflows
- Data Center/Cloud Migrations
Industry Keywords
- HIPAA
- SOC 2
- Security Management
- Policy-as-Code
- Zero-Trust Access
Tools & Technologies
- Terraform
- Kubernetes
- Argo CD
- OpenTelemetry
- Grafana
- Prometheus
- Crossplane
- VPN
- BeyondCorp
- MTLS