Get more replies from employers
Send a job-specific resume in minutes.
Zigsaw is seeking an experienced DevOps Lead in India to design, manage, and optimize scalable cloud and on-premise infrastructure. You will lead robust CI/CD pipelines, monitor reliability, enforce security, and drive SRE maturity across teams.
The role requires 5+ years in DevOps with hands-on leadership, experience in production SaaS environments, and a proactive mindset to build resilient systems for rapid business growth.
Design, manage, and optimize cloud and on-premise infrastructure.
Ensure high availability, scalability, redundancy, and disaster recovery planning.
Manage Linux-based production environments.
Optimize infrastructure cost without compromising reliability.
Handle scaling strategies for increasing device load and customer growth.
Build and maintain robust CI/CD pipelines.
Reduce deployment risks and deployment time.
Automate build, deployment, rollback, and environment provisioning processes.
Standardize deployment practices across teams and products.
Establish strong monitoring, alerting, and observability systems.
Implement proactive incident detection and root cause analysis.
Reduce downtime and improve platform stability.
Drive SRE-oriented operational maturity.
Implement infrastructure security best practices.
Manage access control, secrets management, SSL, firewall policies, backups, and vulnerability handling.
Ensure infrastructure hardening and operational compliance.
Manage Dockerized environments and orchestration platforms.
Improve deployment consistency and environment portability.
Support microservices architecture where applicable.
Work closely with backend and database teams on: Performance tuning Query optimization support Load balancing Caching strategies Replication and failover systems
Lead and mentor DevOps engineers.
Create operational SOPs and infrastructure standards.
Build accountability, documentation culture, and ownership within the team.
Coordinate with Development, QA, Support, and Product teams.
Handle production incidents with urgency and ownership.
Build escalation systems and incident response frameworks.
Conduct postmortem analysis and preventive planning.
Linux Server Administration AWS / GCP / Azure Docker Kubernetes Jenkins / GitHub Actions / GitLab CI Nginx / Apache Load Balancers & Reverse Proxies Networking & Security Monitoring Tools (Prometheus, Grafana, ELK, Zabbix, etc.) Infrastructure Automation Shell Scripting / Python
High-availability architecture Distributed systems Scaling real-time applications Database replication and clustering Message brokers (RabbitMQ, Kafka, Redis Streams, etc.) API infrastructure SSL, DNS, VPN, CDN, WAF
Nice to Have Experience in IoT or telematics platforms Experience managing large-scale real-time tracking systems Terraform / Infrastructure as Code SRE practices Cost optimization at scale Multi-region deployment experience
This role is not for someone who only executes tickets.
We expect the person to: Think proactively instead of reactively Build systems before problems become incidents Create operational leverage through automation Reduce dependency on manual intervention Build infrastructure that supports aggressive business growth Create visibility and measurable operational KPIs
The DevOps TL will be evaluated on: Platform uptime Deployment frequency & stability MTTR (Mean Time to Recovery) Infrastructure scalability Security incident reduction Alert quality and monitoring maturity Automation coverage Infrastructure cost efficiency Team efficiency and operational discipline
5+ years in DevOps / Infrastructure Engineering 2+ years leading teams or handling critical production infrastructure Experience managing production SaaS environments at scale
We are not looking for a “server administrator.”
We are looking for someone who: Understands business impact of infrastructure decisions Can scale systems under uncertainty Handles pressure calmly during outages Builds processes, not heroics Has strong ownership mindset Can challenge poor engineering practices Thinks in terms of reliability engineering, not firefighting
For most SaaS companies, DevOps becomes a support function.
For us, it is a growth constraint or growth accelerator.
A weak DevOps team creates: Slow releases Customer dissatisfaction Downtime Engineering bottlenecks Support overload Revenue risk A strong DevOps function compounds the effectiveness of every other department.
That is why this role is strategically important.