Holidu is one of the world’s fastest-growing holiday rental technology companies. Our mission is to make booking and hosting holiday homes free of doubt and full of joy, by helping hosts generate more bookings with less work and helping guests find a holiday home they truly enjoy. Our team of 700 colleagues from 60+ nations shares a passion for tech, an ambition for constant improvement, and a relentless drive to bring the best experience to more than 40k holiday rental hosts and 4 million annual guests.
Your future team
The DevOps team at Holidu is a central team across the entire tech organization, responsible for creating and maintaining the infrastructure that powers all of our products and services. Our team is a group of professionals who are passionate about automating processes, implementing modern DevOps toolsets, enhancing system reliability, and optimising workflows.
In this role, you will contribute to the continuous improvement of our DevOps processes, collaborate with cross-functional teams, and apply best practices for scalable, reliable, and secure systems. The ideal candidate has a solid technical foundation, a strong hands-on approach, and the ability to deliver results with minimal supervision.
Our Tech Stack
- Cloud: AWS (EC2, S3, RDS, EKS, Elasticache, Lambda)
- Container Orchestration: Kubernetes with Helm
- Infrastructure as Code: Terraform + Terragrunt, [Pulumi/ CDK](Optional)
- Monitoring & Observability: Prometheus, Grafana, Elastic Stack, OpenTelemetry
- CI/CD: Jenkins, GitHub Actions, ArgoCD, ArgoRollouts
- Scripting: Python, Go, Bash
- Version Control: GitHub
- Collaboration: Jira (Agile)
- Automation: N8N, AI-assisted tooling (Agentic ADK)
Your role in this journey
- Infrastructure as Code: Implement and maintain infrastructure definitions using Terraform, Pulumi, or similar tools. Ensure IaC standards are followed and contribute improvements to existing modules and patterns.
- Cloud Operations: Manage and monitor AWS services, ensuring system performance, availability, and adherence to best practices. Troubleshoot production issues and participate in capacity planning.
- Kubernetes & Containers: Maintain and troubleshoot Kubernetes clusters — deploying workloads, managing configurations, scaling services, and resolving incidents to support high-availability applications. Experience in upgrading EKS is a plus.
- CI/CD Pipelines: Maintain and improve CI/CD pipelines to ensure smooth, automated software delivery. Identify bottlenecks and implement enhancements across Jenkins, GitHub Actions, ArgoRollouts and ArgoCD.
- Monitoring & Alerting: Maintain and extend our monitoring stack (Prometheus, Grafana). Build dashboards, configure alerts, and improve observability to ensure comprehensive visibility into system health and performance.
- Cost Awareness: Participate in cost reviews, identify wasteful resource usage, and implement cost-saving measures in collaboration with senior engineers and product teams.
- Automation & Scripting: Write and maintain scripts (Python, Bash, or Go) to automate operational tasks and improve workflows. Leverage AI-powered automation tools like N8N to reduce manual intervention where applicable.
- Collaboration: Work effectively with development, security, and operations teams to support smooth deployments, share knowledge, and align on operational standards.
Your backpack is filled with
Required
- 4+ years of experience in a DevOps, SRE, or cloud engineering role with hands-on production experience.
- Solid working experience with AWS services (EC2, EKS, S3, RDS, Lambda) and cloud infrastructure management.
- Hands-on experience with Docker and Kubernetes in production environments — deploying, scaling, and troubleshooting containerized workloads.
- Practical experience with at least one Infrastructure as Code tool (Terraform, Pulumi, or AWS CDK).
- Experience maintaining and improving CI/CD pipelines using tools like Jenkins, GitHub Actions, or ArgoCD.
- Proficiency in scripting with Python, Bash, or Go for operational automation.
- Working knowledge of monitoring and observability tools such as Prometheus, Grafana, or similar platforms.
- Familiarity with logging and log aggregation systems (Elastic Stack,Open Telemetry, or similar).
- Solid understanding of Linux administration, networking fundamentals, and system security basics.
- Strong communication skills with the ability to collaborate across teams and explain technical decisions clearly.
Nice to Have
- Experience with Helm charts and Kubernetes package management.
- Familiarity with GitOps workflows (e.g., Github Actions, ArgoCD, Flux).
- Experience with designing AWS services based architectures is a plus.
- Experience with AI automation or low-code/no-code platforms such as N8N is a plus.
- Familiarity with prompt engineering and using AI tools to augment DevOps workflows.
- Exposure to cost optimization strategies …