US/Europe/Middle East based | Early-stage, large-scale GPU Cloud Service Provider
We’re working with a high-growth, early-stage GPU Cloud / NeoCloud provider building large-scale, multi-tenant GPU infrastructure for AI workloads. The business is pre-/early-revenue and security is being designed into the platform from day one, not bolted on later.
This is a hands-on role for someone who understands real infrastructure, not theoretical security.
What You’ll Do
- Own end-to-end security architecture for a large-scale AI/GPU platform
- Design and implement zero-trust security across:
- GPU nodes, DPUs (BlueField), networking, storage
- Kubernetes, KubeVirt, VMs, and bare metal
- Define and enforce tenant isolation (compute, memory, network, storage)
- Secure GPU virtualization (MIG, SR-IOV, passthrough, multi-tenant inference)
- Lead firmware & hardware security (BIOS, BMC, NIC, GPU, secure boot, TPM, attestation)
- Design high-performance network security (IB, RoCE, Ethernet) without killing latency
- Own Kubernetes and platform security (RBAC, runtime, secrets, image signing, supply chain)
- Build secure customer-facing APIs (auth, rate limiting, abuse prevention)
- Lead security compliance and audits (SOC 2, ISO 27001, NIST, FedRAMP)
- Support incident response, threat modelling, and red teaming
- Be a trusted partner to engineers, leadership, and customers
What We’re Looking For
Must-have
- 10+ years in security with deep infrastructure experience
- Strong background in:
- Data centre or cloud infrastructure
- Networking at scale
- Virtualisation and containers
- Experience with GPU clouds, HPC, AI platforms, or hyperscale systems
- Comfortable reviewing architecture and calling out what’s broken
- Pragmatic mindset — security that enables delivery, not blocks it
Nice to have
- High-performance storage systems
- Experience with demanding enterprise or AI customers