Site Reliability Engineer Lead- GC

PureSoftware Ltd

Bengaluru

On-site

INR 2,500,000 - 4,500,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

A technology company located in Karnataka, Bengaluru is seeking an experienced professional to focus on system reliability and Kubernetes operations. The role entails designing scalable architectures, automating deployments, managing cloud services, and implementing monitoring solutions. The ideal candidate will collaborate with various teams, ensuring high availability and security of Kubernetes environments. Key responsibilities include incident response and capacity planning to accommodate future workloads, along with maintaining comprehensive documentation.

Responsibilities

  • Collaborate with software development teams to ensure reliability is prioritized.
  • Design scalable architectures for critical applications.
  • Manage high availability of Kubernetes clusters.
  • Drive automation to streamline deployment in cloud environments.
  • Implement CI/CD pipelines for Kubernetes applications.
  • Leverage cloud services to optimize Kubernetes solutions.
  • Implement monitoring solutions for applications and infrastructure.
  • Respond to incidents minimizing downtime.
  • Conduct security audits for Kubernetes environments.
  • Maintain documentation for processes and troubleshooting.
  • Collaborate with security to apply best practices and run audits.
  • Document configurations and troubleshooting steps.

Tools

Kubernetes
AWS
GCP
Azure
Terraform
Ansible

Job description

Responsibilities
  1. System Reliability:
    • a. Collaborate with software development teams to ensure reliability is a key consideration throughout the software development life cycle.
    • b. Design and implement scalable and resilient architectures for mission-critical applications.
  2. Kubernetes Operations:
    • a. Design, implement, and manage Kubernetes clusters, ensuring high availability, fault tolerance, and scalability.
    • b. Perform upgrades, patch management, and security enhancements for Kubernetes infrastructure.
  3. Automation and Infrastructure as Code (IaC):
    • a. Drive automation efforts to streamline deployment, scaling, and management of applications on Kubernetes and/or cloud environments.
    • b. Implement CI/CD pipelines for deploying and updating Kubernetes applications.
    • c. Develop and maintain Infrastructure as Code scripts (e.g., Terraform, Ansible) for provisioning and managing cloud and container resources.
  4. Cloud Integration:
    • a. Leverage cloud services (AWS, GCP, Azure) to optimize Kubernetes infrastructure and seamlessly integrate with other cloud-native solutions.
    • b. Implement best practices for deploying and managing Kubernetes on cloud platforms.
  5. Monitoring and Alerting:
    • a. Implement effective monitoring and alerting solutions for Kubernetes clusters, applications, and underlying infrastructure.
    • b. Proactively identify and address performance bottlenecks and reliability issues.
  6. Incident Response:
    • a. Respond to and resolve incidents related to Kubernetes infrastructure and applications, ensuring minimal downtime and impact on users.
    • b. Conduct post-incident reviews and implement improvements to prevent future issues.
  7. Capacity Planning:
    • a. Perform capacity planning to ensure the Kubernetes infrastructure can accommodate current and future workloads in the cloud.
  8. Security:
    • a. Collaborate with the security team to implement and maintain security best practices for Kubernetes environments in the cloud.
    • b. Conduct regular security audits and vulnerability assessments.
  9. Collaboration and Documentation:
    • a. Work closely with development, operations, and other teams to ensure a collaborative approach to infrastructure and application reliability.
    • b. Maintain clear and comprehensive documentation for processes, configurations, and troubleshooting steps.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer
Site Reliability Engineer

Smart Ims • Bengaluru

Hybrid
INR 1,200,000 - 2,000,000
Sr. Site Reliability Engineer
Sr. Site Reliability Engineer

Visa Consolidated Support Services India • Bengaluru

On-site
INR 1,500,000 - 2,500,000
Site Reliability Engineer
Site Reliability Engineer

Tecnoprism • Bengaluru

On-site
INR 900,000 - 1,400,000
Site Reliability Engineer
Site Reliability Engineer

Insight Global • Bengaluru

On-site
INR 2,500,000 - 5,000,000
Site Reliability Engineer
Site Reliability Engineer

Tata Consultancy Services • Bengaluru, Hyderabad

Hybrid
INR 800,000 - 1,500,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

iLink Digital • Chennai

On-site
INR 1,200,000 - 1,800,000
Senior Cloud Site Reliability Engineer
Senior Cloud Site Reliability Engineer

Augusta Infotech • Bengaluru

Hybrid
INR 1,500,000 - 2,500,000
Staff Site Reliability Engineer
Staff Site Reliability Engineer

Stryker Group • Gurugram District

On-site
INR 1,200,000 - 1,800,000
Site Reliability Engineer
Site Reliability Engineer

Innodata Inc. • India

On-site
INR 2,400,000 - 4,000,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Headout • Bengaluru

On-site
INR 1,200,000 - 1,800,000