WAGNIFY Information Technologies is an IT services, consulting, and business solutions company built to simplify and strengthen how organizations operate. We combine innovation with collective intelligence to accelerate digital transformation, and to turn technology from an obstacle into a strategic advantage.
Our Mission
We understand that every business has unique needs, and we respond with tailored, innovative, and sustainable IT solutions. With deep expertise in Database Management, DevOps, AIOps, and Cloud Services, we help our clients optimize their infrastructure, boost operational efficiency, and accelerate growth, delivering smarter solutions that create real, meaningful outcomes.
Our Vision
To be a leading IT solutions provider, helping organizations use technology efficiently and securely in the digital world of tomorrow. We partner with businesses across finance, insurance, retail, energy, and technology, delivering scalable, secure, and innovative solutions that support their digital transformation. Built on innovation, reliability, and long‑term partnership, we’re committed to our clients’ lasting success.
About the Job
We are looking for an experienced Senior DevOps Engineer to join our technology team and take an active role in designing, building, and continuously improving our cloud infrastructure and software delivery processes.
In this role, you will work closely with software engineering, architecture, security, and operations teams to build scalable, secure, highly available, and observable platforms .
Beyond day‑to‑day operations, we expect our Senior DevOps Engineer to drive automation, contribute to architectural decisions, improve engineering standards, and help the organization adopt DevOps and Site Reliability Engineering best practices.
Responsibilities
- Design, build, maintain, and continuously improve scalable and highly available cloud infrastructure.
- Own and improve CI/CD pipelines, deployment strategies, and release automation processes.
- Design and manage containerized environments using Docker and Kubernetes.
- Build and maintain infrastructure using Infrastructure as Code (IaC) practices and tools such as Terraform and Ansible.
- Implement and improve monitoring, logging, alerting, and observability solutions.
- Improve system reliability, scalability, performance, security, and operational efficiency.
- Design and implement automation to reduce manual operational workload and repetitive tasks.
- Troubleshoot complex production issues, perform root cause analysis, and implement permanent corrective actions.
- Contribute to cloud architecture, capacity planning, disaster recovery, and high‑availability strategies.
- Collaborate with engineering and security teams to integrate security practices into infrastructure and delivery pipelines.
- Establish and promote DevOps, GitOps, automation, and operational best practices across engineering teams.
- Participate in technical design and architecture discussions and provide guidance on infrastructure‑related decisions.
- Identify infrastructure and cloud cost optimization opportunities.
- Mentor engineers and contribute to improving the team’s technical capabilities and engineering culture.
Qualifications
- Strong professional experience in DevOps, Platform Engineering, Site Reliability Engineering roles.
- Hands‑on experience with at least one major cloud platform such as AWS, Microsoft Azure, or Google Cloud Platform.
- Strong hands‑on experience with Kubernetes and Docker in production environments.
- Strong knowledge of Infrastructure as Code, preferably Terraform.
- Experience with configuration management and automation tools such as Ansible.
- Strong experience designing and maintaining CI/CD pipelines using tools such as Jenkins, GitLab CI/CD, GitHub Actions, or Azure DevOps.
- Experience with observability and monitoring technologies such as Prometheus, Grafana, ELK/OpenSearch, or equivalent solutions.
- Experience with Git and modern software development and deployment workflows.
- Experience with Helm, Argo CD, Flux, orequivalent solutions.
- Understanding of High Availability, Disaster Recovery, scalability, and fault‑tolerant architecture.
- Experience with production environments and incident troubleshooting.
Nice to have:
- Experience implementing DevSecOps practices.
- Knowledge of secrets management technologies such as HashiCorp Vault or cloud‑native alternatives.
- Experience with microservices and distributed systems.
- Knowledge of SRE concepts such as SLIs, SLOs, error budgets, and incident management.
- Relevant Kubernetes, or DevOps certifications.
General Skills and Attributes
- Strong analytical thinking and problem‑solving skills, particularly when dealing with complex production environments.
- Ability to take ownership of technical problems from identification through resolution.
- Strong sense of accountability and a proactive approach to identifying risks and improvement opportunities.
- Ability to balance reliability, security, performance, cost, and delivery speed when making technical decisions.
- Strong communication skills and the ability to collaborate effectively with both technical and non‑technical stakeholders.
- Comfortable working across multiple engineering teams and influencing technical decisions without relying solely on authority.
- Passion for automation, continuous improvement, reliability, and engineering excellence.
- Ability to document technical solutions and communicate architectural decisions clearly.
- Willingness to share knowledge, mentor other engineers, and contribute to the development of engineering standards.
- Curious and continuous‑learning mindset with an interest in evaluating and adopting new technologies where they provide meaningful value.