Overview
We are looking for a Site Reliability Engineer (SRE) to help build, scale, and maintain highly reliable, secure, and performant cloud-native systems. You will work closely with engineering and platform teams to improve system reliability, automate operations, and support CI/CD pipelines across multi-cloud environments.
Company
- Company: Resource Funnel
- Website: Visit Website
- Business Type: Consulting Firm
- Company Type: Product & Service
- Business Model: B2B
- Funding Stage: Bootstrapped
- Industry: Consulting
- Salary Range: ₹ 8-12 Lacs PA
Responsibilities
- Ensure high availability, reliability, and scalability of production systems.
- Design, build, and maintain infrastructure across AWS, Azure, and GCP.
- Implement and manage Infrastructure as Code (IaC) using Terraform.
- Build, maintain, and optimize CI/CD pipelines using tools like Jenkins.
- Monitor system health using tools such as Prometheus, Grafana, ELK, Datadog, or Stackdriver.
- Troubleshoot production issues, perform root cause analysis, and participate in incident response.
- Work with containers, Kubernetes, and microservices architectures.
- Collaborate with development teams to improve deployment reliability and system performance.
- Automate repetitive operational tasks using scripting and tooling.
Required Qualifications
- Bachelor’s degree in Computer Science, Engineering, or a related field (or equivalent practical experience).
- 3–7 years of experience as an SRE, DevOps Engineer, or Platform Engineer.
- Strong scripting skills in Python, Bash, Go, or similar languages.
- Solid experience with Linux system administration.
- Hands-on experience with at least one major cloud provider: AWS, Azure, or GCP.
- Proficiency with CI/CD tools such as Jenkins, GitLab CI, or GitHub Actions.
- Experience with monitoring, logging, and alerting tools.
- Strong understanding of containers, Kubernetes, and microservices.
- Experience working with databases like PostgreSQL, MySQL, Redis, or similar.
- Good understanding of networking concepts (DNS, VPC, TLS, load balancing).
- Strong communication skills and ability to work collaboratively in a team environment.
Preferred Qualifications
- Experience with Cloud Run, Cloud SQL, or serverless architectures.
- Familiarity with incident management tools such as Jira.
- Experience designing or operating distributed systems and high-throughput architectures.