Job Description
Beghou Consulting is seeking a hands‑on Site Reliability / DevOps Engineer to design, build, automate, and operate secure, scalable, and cost‑efficient cloud infrastructure and data platforms across Azure and AWS. This role plays a critical part in enabling reliable data and application delivery by implementing Infrastructure as Code (IaC), CI/CD automation, and cloud best practices. You will ensure operational excellence, security, and disaster‑recovery readiness while supporting the organisation’s cloud modernization and data‑driven initiatives. You will work closely with data engineering, application development, and client teams to translate platform requirements into resilient infrastructure solutions. By standardising environments, improving deployment pipelines, and automating operational workflows, this role helps accelerate development velocity and reduce operational risk.
Company: Beghou
Website: Visit Website
Business Type: Enterprise
Company Type: Service
Business Model: B2B
Funding Stage: Bootstrapped
Industry: Information Technology
Salary Range: ₹ 11-14 Lacs PA
Key Responsibilities
Infrastructure as Code (IaC)
- Design, develop, and maintain Terraform modules to provision and manage cloud infrastructure.
- Follow best practices for Terraform state management, module design, and testing.
- Apply CDK for Terraform (CDKTF) or Terragrunt patterns where appropriate.
Cloud & Data Platform Operations
- Architect, provision, and operate services across Azure and AWS, including compute, storage, networking, and databases.
- Manage Databricks workspaces, clusters, and jobs to support scalable data pipelines.
- Implement tagging standards, cost‑optimization, and security controls across cloud and Databricks environments.
CI/CD & Automation
- Build, test, and maintain Azure DevOps Pipelines or GitHub Actions for infrastructure and application deployments.
- Manage source control in Azure DevOps or GitHub, enforcing branching strategies, pull‑request workflows, and policy‑as‑code checks.
Collaboration & Documentation
- Partner with client teams and data engineers to safely integrate infrastructure changes into development workflows.
- Create and maintain architecture diagrams, runbooks, and operational documentation for technical and non‑technical stakeholders.
Programming & Tooling
- Develop automation scripts and CLI tools using Python to streamline operational processes.
- Contribute to internal SDKs, shared libraries, and infrastructure utilities.
Required Qualifications
- Experience: 3+ years in a DevOps or SRE role supporting production workloads on Azure and AWS.
- IaC Expertise: Strong hands‑on experience with Terraform, including Terraform Cloud, Terragrunt, module design, and testing.
- CI/CD: Proven experience building and maintaining Azure DevOps Pipelines and managing repositories in Azure DevOps or GitHub.
- Programming: Strong Python skills with an emphasis on clean, testable, reusable code.
- Data Platforms: Experience provisioning and managing Databricks workspaces, clusters, and jobs.
- Disaster Recovery: Hands‑on experience designing DR strategies, automated backups, and performing recovery drills.
- Cloud Fundamentals: Solid understanding of Azure and AWS core services (compute, networking, storage, IAM, databases).
- Problem‑Solving: Strong troubleshooting skills with a structured, calm approach to incident response.
- Communication: Clear written and verbal communication with strong documentation skills.
Preferred Qualifications
- Advanced IaC: Experience with CDK for Terraform (CDKTF) or Terragrunt.
- Languages & Tools: Terraform, Python, PowerShell, Bicep.
- Security & Compliance: Knowledge of cloud security best practices (CIS benchmarks, Azure Policy, AWS IAM).
- Networking: Deep understanding of VNets, VPCs, load balancers, VPNs, and cross‑region connectivity.
- Serverless & Containers: Experience with Azure Functions, AWS Lambda, and container registries.