We are looking for a Senior DevOps / Platform Engineer with ~6+ years of experience to take complete 0-to-1 ownership of our cloud infrastructure. In this role, you will architect, build, and scale production-ready cloud environments from the ground up, automate end-to-end workflows, and drive reliability, security, and developer productivity using modern tools and AI automation.
Responsibilities:
- 0-to-1 Ownership: Architect, build, and maintain scalable, production-ready cloud infrastructure from scratch with single-point technical ownership.
- Infrastructure-as-Code: Automate all infrastructure provisioning strictly using modular and reusable Terraform.
- Cloud Workload Management: Manage, scale, and optimise cloud infrastructure primarily on GCP (AWS background is also acceptable).
- Containers and Orchestration: Design, deploy, and manage production Kubernetes (GKE/EKS) clusters and containerised applications.
- CI/CD Pipelines: Build, automate, and streamline developer deployment pipelines using GitHub Actions, GitLab CI, or Jenkins.
- Observability and SRE: Implement monitoring, logging, and alerting stacks (Prometheus, Grafana, Datadog, ELKand lead incident response and root-cause analysis (RCA).
- Security and Optimisation: Oversee IAM policies, secrets management, and zero-trust concepts while actively optimising cloud spend.
- AI Productivity Integration: Leverage modern AI tools (GitHub Copilot, Claude, LLM scripting) to automate scripts, debug systems, and scale engineering efficiency.
Requirements:
- Experience: ~5 years of hands-on experience in DevOps, Platform Engineering, or SRE roles.
- IaC Mastery: Deep, hands-on expertise with Terraform for full infrastructure automation.
- Cloud Expertise: Strong background with GCP (preferred) or AWS.
- Containerization: Direct experience managing Kubernetes (GKE/EKS) and Docker in production.
- Scripting: Proficiency in Python, Go, or Bash.
- Product Ownership: Proven ability to drive projects independently from concept through production rollout.
- AI Tooling Comfort: High comfort level using AI-assisted tools/workflows to accelerate scripting and developer tooling.
- Core Engineering: Strong Linux administration, cloud security, and networking fundamentals.
Good to Have:
- Experience in high-scale startup, healthcare, or fintech environments.
- Exposure to multi-cluster Kubernetes, service meshes, or zero-trust architectures.
- Hands-on experience building Internal Developer Platforms (IDPs) or custom automation tools.