Company: Resource Funnel
Website: Visit Website
Business Type: Consulting Firm
Company Type: Product & Service
Business Model: B2B
Funding Stage: Bootstrapped
Industry: Consulting
Salary Range: ₹ 8-12 Lacs PA
Job Description
Role Overview
We are looking for a Site Reliability Engineer (SRE) to help build, scale, and maintain highly reliable, secure, and performant cloud-native systems. You will work closely with engineering and platform teams to improve system reliability, automate operations, and support CI/CD pipelines across multi-cloud environments.
Key Responsibilities
- Ensure high availability, reliability, and scalability of production systems.
- Design, build, and maintain infrastructure across AWS, Azure, and GCP.
- Implement and manage Infrastructure as Code (IaC) using Terraform.
- Build, maintain, and optimize CI/CD pipelines using tools like Jenkins.
- Monitor system health using tools such as Prometheus, Grafana, ELK, Datadog, or Stackdriver.
- Troubleshoot production issues, perform root cause analysis, and participate in incident response.
- Work with containers, Kubernetes, and microservices architectures.
- Collaborate with development teams to improve deployment reliability and system performance.
- Automate repetitive operational tasks using scripting and tooling.
Required Qualifications
- Bachelor's degree in Computer Science, Engineering, or a related field (or equivalent practical experience).
- 3-7 years of experience as an SRE, DevOps Engineer, or Platform Engineer.
- Strong scripting skills in Python, Bash, Go, or similar languages.
- Solid experience with Linux system administration.
- Hands-on experience with at least one major cloud provider: AWS, Azure, or GCP.
- Proficiency with CI/CD tools such as Jenkins, GitLab CI, or GitHub Actions.
- Experience with monitoring, logging, and alerting tools.
- Strong understanding of containers, Kubernetes, and microservices.
- Experience working with databases like PostgreSQL, MySQL, Redis, or similar.
- Good understanding of networking concepts (DNS, VPC, TLS, load balancing).
- Strong communication skills and ability to work collaboratively in a team environment.
Preferred Qualifications
- Experience with Cloud Run, Cloud SQL, or serverless architectures.
- Familiarity with incident management tools such as Jira.
- Experience designing or operating distributed systems and high-throughput architectures.