Site Reliability Engineer - GCP & Automation Focus
Insight Global
United States
On-site
USD 89,544 - 99,187
Full time
14 days+
Get more replies from employers
Send a job-specific resume in minutes.
Start fresh or import an existing resume
Benefits offered by this job
Medical insurance
Vision insurance
401(k)
Job summary
A dynamic engineering team is seeking a Mid-to-Senior Level Site Reliability Engineer to ensure the reliability and scalability of mission-critical systems on Google Cloud Platform. In this role, you'll leverage your expertise in SRE principles and various tools to proactively identify and resolve issues, automate tasks, and enhance infrastructure processes. Collaborate with high-performing teams in a fast-paced environment while continuously improving your skills. This is an exciting opportunity to make a significant impact on the reliability and performance of critical applications.
Qualifications
5+ years of experience in Site Reliability Engineering or DevOps.
Proficiency in scripting languages like Python and Bash.
Extensive experience with HashiCorp Terraform for infrastructure-as-code.
Responsibilities
Design and manage scalable infrastructure on GCP.
Develop monitoring and alerting solutions using Datadog.
Automate operational tasks using HashiCorp Terraform.
Skills
Site Reliability Engineering
Google Cloud Platform
Python
Bash
HashiCorp Terraform
Datadog
PagerDuty
Kubernetes
Docker
Linux
Education
Bachelor's degree in Computer Science
Tools
Google Cloud Spanner
GCP Cloud Logging
ChaosSearch
Job description
A dynamic engineering team is seeking a Mid-to-Senior Level Site Reliability Engineer to ensure the reliability and scalability of mission-critical systems on Google Cloud Platform. In this role, you'll leverage your expertise in SRE principles and various tools to proactively identify and resolve issues, automate tasks, and enhance infrastructure processes. Collaborate with high-performing teams in a fast-paced environment while continuously improving your skills. This is an exciting opportunity to make a significant impact on the reliability and performance of critical applications.