Site Reliability Engineer

JPS Tech Solutions Pvt Ltd

Arizona

On-site

USD 140,000 - 190,000

Full time

2 days ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

JPS Tech Solutions Pvt Ltd is seeking a talented Site Reliability Engineer (SRE) with a strong background in Google Cloud Platform and Red Hat OpenShift administration. The role focuses on reliability, performance, and scalability of on-premise and cloud-based systems while reducing Google Cloud costs.

You will ensure uptime, design cloud infrastructure, automate processes, monitor systems, and collaborate with development and operations teams to support growth and efficient operations.

Qualifications

  • Bachelor's degree in computer science, engineering or related field.
  • 8+ years of experience in site reliability engineering or similar role.
  • Strong knowledge of Google Cloud Platform and OpenShift administration.
  • Experience with automation tools and infrastructure as code.

Responsibilities

  • Ensure reliability and uptime of critical services and infrastructure.
  • Design, implement, and manage cloud infrastructure using Google Cloud services.
  • Develop and maintain automation scripts and tools to improve efficiency.
  • Implement monitoring solutions and respond to incidents to minimize downtime.
  • Collaborate with development and operations teams to improve reliability and performance.
  • Conduct capacity planning and performance tuning for future growth.
  • Create and maintain documentation for system configurations and processes.

Skills

Google Cloud Platform
Kubernetes
Prometheus
Grafana
Python
Bash
CI/CD
Networking
Looker
Vertex AI

Education

Bachelor's degree in Computer Science or related field

Tools

Terraform
Ansible
Puppet
Jenkins
GitLab CI
Azure Pipelines

Job description

Job Description

We are looking for a talented Site Reliability Engineer (SRE) with a strong background in Google Cloud Platform (Google Cloud Platform), and RedHat OpenShift administration. The ideal candidate will be responsible for ensuring the reliability, performance, and scalability of our on-premise and cloud-based systems along with focus on reducing costs for Google Cloud.

System Reliability:

Ensure the reliability and uptime of critical services and infrastructure.

Google Cloud Expertise:

Design, implement, and manage cloud infrastructure using Google Cloud services.

Automation:

Develop and maintain automation scripts and tools to improve system efficiency and reduce manual intervention.

Monitoring and Incident Response:

Implement monitoring solutions and respond to incidents to minimize downtime and ensure quick recovery.

Collaboration:

Work closely with development and operations teams to improve system reliability and performance.

Capacity Planning:

Conduct capacity planning and performance tuning to ensure systems can handle future growth.

Documentation:

Create and maintain comprehensive documentation for system configurations, processes, and procedures.

Qualifications:

Education: Bachelor's degree in computer science, Engineering, or a related field.

Experience: 8+ years of experience in site reliability engineering or a similar role.

Skills:

Proficiency in Google Cloud services (Compute Engine, Kubernetes Engine, Cloud Storage, BigQuery, Pub/Sub, etc.).

Familiarity with Google BI and AI/ML tools (Looker, BigQuery ML, Vertex AI, etc.)

Experience with automation tools (Terraform, Ansible, Puppet).

Familiarity with CI/CD pipelines and tools (Azure pipelines Jenkins, GitLab CI, etc.).

Strong scripting skills (Python, Bash, etc.).

Knowledge of networking concepts and protocols.

Experience with monitoring tools (Prometheus, Grafana, etc.).

Preferred Certifications:

Google Cloud Professional DevOps Engineer

Google Cloud Professional Cloud Architect

Red Hat Certified Engineer (RHCE) or similar Linux certification

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Site Reliability Engineer
Site Reliability Engineer

Compunnel, Inc. • New Jersey

On-site
USD 120,000 - 150,000
Site Reliability Engineer with GCP
Site Reliability Engineer with GCP

MSRcosmos LLC • Dallas (TX)

On-site
USD 120,000 - 180,000
Site Reliability Engineer (Only W2)
Site Reliability Engineer (Only W2)

MSRcosmos LLC • Mahwah (NJ)

On-site
USD 120,000 - 150,000
Senior SRE: Google Cloud & OpenShift Reliability
Senior SRE: Google Cloud & OpenShift Reliability

JPS Tech Solutions Pvt Ltd • Arizona

On-site
USD 140,000 - 190,000
Senior Site Reliability Engineer NEX
Senior Site Reliability Engineer NEX

NexTier Completion Solutions Inc. • Houston (TX)

On-site
USD 110,000 - 150,000
Site Reliability Engineer
Site Reliability Engineer

Harvey Nash • United States

Remote
USD 120,000 - 150,000
Senior Site Reliability Engineer – Google Distributed Cloud Edge (Edge SRE)
Senior Site Reliability Engineer – Google Distributed Cloud Edge (Edge SRE)

CoSourcing Partners Inc. • Chicago (IL)

On-site
USD 150,000 - 190,000
Senior Systems Engineer, Site Reliability Engineering, Google Cloud
Senior Systems Engineer, Site Reliability Engineering, Google Cloud

Google • Sunnyvale (CA)

On-site
USD 166,000 - 244,000
Site Reliability Engineer
Site Reliability Engineer

Ethos Group • Irving (TX)

On-site
USD 110,000 - 160,000
Systems Engineer – SRE Enablement
Systems Engineer – SRE Enablement

AutoZone • Memphis (TN)

On-site
USD 120,000 - 160,000
Hybrid work model