Site Reliability Engineer

MNR Solutions Pvt. Ltd.

Bengaluru

On-site

INR 900,000 - 1,500,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

An established industry player is seeking a skilled Site Reliability Engineer to ensure the smooth operation of cloud infrastructure. In this role, you will leverage your expertise in AWS, GCP, and Kubernetes to maintain high availability and reliability while automating deployments and optimizing performance. You'll collaborate with cross-functional teams to enhance system efficiency and manage resources effectively. If you're passionate about cloud technologies and eager to make a significant impact, this is the perfect opportunity for you to thrive in a dynamic environment and contribute to innovative solutions.

Qualifications

  • 10-12 years of experience in Site Reliability Engineering with a focus on cloud infrastructure.
  • Strong proficiency in Python and automation tools for optimizing performance.

Responsibilities

  • Ensure system availability and reliability on AWS and GCP, focusing on uptime.
  • Automate infrastructure and deployments using CI/CD tools and Terraform.

Skills

SRE
Python
Cloud
Kubernetes
DataDog
Azure DevOps
Jenkins
Octopus

Tools

Terraform
AWS
GCP

Job description

Position: Site Reliability Engineer
Location: Chennai, Bangalore
Experience: 10-12 Years
Skills: SRE, DataDog, Azure DevOps, Jenkins, Octopus, Cloud, Python

Job Description:

  1. Ensure smooth production on AWS and GCP by maintaining availability, scalability, and reliability with Kubernetes (GKE), AWS ECS, and cloud-native services, focusing on uptime and availability.
  2. Build and automate infrastructure with Terraform, CI/CD tools, and Python scripts to reduce manual tasks and optimize error rates and throughput.
  3. Provide 24x7 on-call support to ensure system availability and quick issue resolution, minimizing MTTR.
  4. Monitor infrastructure with telemetry, tracking latency, error rates, and other SLIs to ensure seamless operations.
  5. Improve system performance by analyzing metrics from OS, containers, APIs, and apps to address issues early, focusing on response times and resource usage.
  6. Automate deployments using CI/CD, ensuring performance, compliance, and cost efficiency.
  7. Plan immutable infrastructure deployments with automated pipelines, ensuring low latency and cost optimization.
  8. Collaborate with teams (.NET, Java, APIs, Python) to optimize testing and automate deployments for reliable releases, managing error budgets.
  9. Design scalable systems, manage platforms for high demand, and monitor capacity and throughput.
  10. Automate processes for efficiency and resource management, reducing saturation. Ensure binaries and configurations work across environments, focusing on scalability.
  11. Balance feature development with system stability by managing SLOs and error budgets.
  12. Experience managing cloud infrastructure on AWS, GCP, and Kubernetes with a focus on scalability and SLO-driven performance.
  13. Proficiency with tools like DataDog, Azure DevOps, Jenkins, and Octopus for code deployment and monitoring throughput and latency.
  14. Strong background in software development, test automation, and Infra-as-code (Terraform) for efficient deployments.
  15. Expertise in Python, .NET, or Java for automating tasks and optimizing performance and latency.
  16. Familiarity with distributed storage systems, handling RPA toolsets, large datasets, focusing on cost efficiency and data throughput.
  17. Experience with Kubernetes and AWS/GCP services for resource management and resource usage.
  18. Proactive in identifying bottlenecks, troubleshooting, and improving system performance.
  19. Ability to design scalable systems to support business growth, ensuring SLO adherence.

Interested candidates should share their resume at rubi.jena@mnrsolutions.in

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer
Site Reliability Engineer

Tech Mahindra • Bengaluru

On-site
INR 1,500,000 - 2,200,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

VMC Soft Technologies, Inc • Hyderabad

Hybrid
INR 1,500,000 - 2,000,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

F-Prime Capital • Pune District

On-site
INR 1,500,000 - 2,000,000
Senior Site Reliability Engineer - Cloud Infrastructure
Senior Site Reliability Engineer - Cloud Infrastructure

WITS Innovation Lab • Chandigarh

On-site
INR 1,800,000 - 3,000,000
Site Reliability Engineer Lead
Site Reliability Engineer Lead

Hilabs • Pune District

On-site
INR 1,500,000 - 2,500,000
Site Reliability Engineer
Site Reliability Engineer

ScaleneWorks People Solutions LLP • Pune District

On-site
INR 3,500,000 - 5,500,000
Site Reliability Engineer
Site Reliability Engineer

SourcingXPress • Mumbai

On-site
INR 800,000 - 1,200,000
Site Reliability Engineering Lead
Site Reliability Engineering Lead

Infosys • Hyderabad

On-site
INR 1,200,000 - 1,800,000
Site Reliability Engineer
Site Reliability Engineer

SourcingXPress • Maharashtra

On-site
INR 700,000 - 1,800,000
Site Reliability Engineer
Site Reliability Engineer

Elgebra • Chennai District

On-site
INR 1,500,000 - 2,500,000