We are looking for a highly skilled Cloud & SecOps Engineer to architect, automate, optimize, and secure cloud-native platforms across multiple cloud providers. The ideal candidate combines deep expertise in Kubernetes, cloud infrastructure, automation, and cost optimization with a strong engineering mindset focused on scalability, resilience, operational excellence, and security. This role is ideal for someone who treats cloud infrastructure as a product, continuously seeks efficiency gains, and understands that every architectural decision impacts reliability, developer productivity, and cloud spend.
KEY RESPONSIBILITIES
- Design, deploy, and manage production-grade infrastructure across AWS, Azure, and GCP with cloud-agnostic operational strategies.
- Architect highly available, fault-tolerant, and disaster-resilient systems while leading cloud migration and modernization initiatives.
- Evaluate infrastructure architecture trade-offs balancing performance, reliability, scalability, and cost.
- Build, operate, and optimize large-scale Kubernetes platforms, including multi-cluster and multi-region deployments.
- Manage Kubernetes ecosystem tools such as Helm, ArgoCD/FluxCD, Service Meshes (Istio/Linkerd), Ingress Controllers, and Operators.
- Develop Internal Developer Platforms (IDP), self-service capabilities, and automation solutions to enhance developer productivity.
- Implement Infrastructure as Code (Terraform/OpenTofu, Pulumi, CloudFormation) and create reusable infrastructure modules and templates.
- Automate infrastructure provisioning, deployments, backup, recovery, and operational workflows to eliminate manual dependencies.
- Design and maintain enterprise-grade CI/CD pipelines with GitOps practices and advanced deployment strategies (Blue-Green, Canary, Progressive Delivery, Automated Rollbacks).
- Drive FinOps initiatives including cloud cost governance, rightsizing, resource optimization, Reserved Instances/Savings Plans, storage lifecycle management, and cluster consolidation.
- Establish cost visibility, governance frameworks, budgets, dashboards, and cost-aware architectural practices to improve infrastructure efficiency.
- Implement and manage observability platforms using Prometheus, Grafana, OpenTelemetry, and ELK/OpenSearch.
- Define and monitor SLIs, SLOs, and error budgets while leading incident response, root cause analysis, and reliability improvement initiatives.
- Implement DevSecOps practices, secure cloud and Kubernetes environments, and manage IAM, RBAC, secrets, network policies, and vulnerability remediation.
- Support security, compliance, and SecOps initiatives including container security, supply chain security, runtime threat detection, CSPM, incident response, and compliance standards such as SOC2, ISO27001, PCI-DSS, and HIPAA.
EXPERIENCE
- 5+ years of SecOps, Platform Engineering, SRE, or Cloud Infrastructure experience.
- 3+ years of hands-on Kubernetes production experience.
- Experience operating infrastructure across at least two major cloud providers.
- Strong Linux administration and troubleshooting skills.
- Experience supporting mission-critical production systems.
PREFERRED QUALIFICATIONS
- Kubernetes Certifications (CKA, CKAD, CKS)
- AWS, Azure, or GCP Certifications
- Edge computing experience
- Serverless architecture experience
- Security certifications (optional): CISSP, CCSP, Security+, AWS Security Specialty