GCP SRE | Kubernetes & Cloud Reliability Engineer

MSRcosmos LLC

Dallas (TX)

On-site

USD 120,000 - 180,000

Full time

3 days ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

MSRcosmos LLC is seeking a Site Reliability Engineer with strong GCP and Kubernetes expertise to ensure reliability, performance, and scalability of both on-premise and cloud-based systems. The role emphasizes reducing Google Cloud costs while maintaining uptime and efficiency.

You will design and implement cloud infrastructure, automate tasks with Terraform and Ansible, monitor services with Prometheus/Grafana, and collaborate with development and operations teams to plan capacity and respond

Qualifications

  • Experience with Google Cloud Platform (GCP) and Kubernetes.
  • Experience with automation and configuration management tools.
  • Proficiency in monitoring and incident response.
  • Ability to collaborate across teams and document procedures.

Responsibilities

  • Ensure reliability and uptime of critical services.
  • Design, implement, and manage cloud infrastructure on GCP.
  • Develop automation scripts to improve efficiency.
  • Monitor and respond to incidents to minimize downtime.
  • Collaborate with development and operations teams to improve reliability and performance.
  • Perform capacity planning and performance tuning to handle growth.
  • Create and maintain documentation for configurations, processes, and procedures.

Skills

GCP fundamentals
System reliability
Monitoring concepts
Cost optimization

Tools

Kubernetes
Terraform
Ansible
Puppet
Jenkins
GitLab CI
Networking basics

Job description

MSRcosmos LLC is seeking a Site Reliability Engineer with strong GCP and Kubernetes expertise to ensure reliability, performance, and scalability of both on-premise and cloud-based systems. The role emphasizes reducing Google Cloud costs while maintaining uptime and efficiency.

You will design and implement cloud infrastructure, automate tasks with Terraform and Ansible, monitor services with Prometheus/Grafana, and collaborate with development and operations teams to plan capacity and respond

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer
Site Reliability Engineer

Compunnel, Inc. • New Jersey

On-site
USD 120,000 - 150,000
GCP & Kubernetes SRE: Reliable, Scalable Cloud Ops
GCP & Kubernetes SRE: Reliable, Scalable Cloud Ops

Compunnel, Inc. • New Jersey

On-site
USD 120,000 - 150,000
Site Reliability Engineer with GCP
Site Reliability Engineer with GCP

MSRcosmos LLC • Dallas (TX)

On-site
USD 120,000 - 180,000
Senior Cloud SRE: Kubernetes, Automation & Reliability
Senior Cloud SRE: Kubernetes, Automation & Reliability

GCS Recruitment • Moorestown Township (NJ)

On-site
USD 120,000 - 165,000
SRE & Cloud Platform Engineer (Kubernetes + IaC)
SRE & Cloud Platform Engineer (Kubernetes + IaC)

GCS Recruitment • Cherry Hill Township (NJ)

On-site
USD 110,000 - 150,000
Site Reliability Engineer
Site Reliability Engineer

JPS Tech Solutions • San Jose (CA)

On-site
USD 130,000 - 160,000
Senior Site Reliability Engineer: Cloud, Kubernetes & CI/CD
Senior Site Reliability Engineer: Cloud, Kubernetes & CI/CD

Amiri Recruiting • Mountain View (CA)

On-site
USD 130,000 - 160,000
Senior Observability & SRE Engineer — GCP/Kubernetes
Senior Observability & SRE Engineer — GCP/Kubernetes

Ontrac Solutions • United States

On-site
USD 120,000 - 180,000
Site Reliability Engineer - GCP & Automation Focus
Site Reliability Engineer - GCP & Automation Focus

Insight Global • United States

On-site
USD 100,000 - 125,000
Senior GCP Platform Engineer – SRE & CI/CD
Senior GCP Platform Engineer – SRE & CI/CD

CVS Health • Illinois

Hybrid
USD 130,000 - 261,000
Bonus program
Equity awards
Comprehensive benefits