GCP SRE II: Observability, Automation & Reliability

HTC Global Services

Dearborn (MI)

On-site

USD 110,000 - 140,000

Full time

24 hours ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Benefits offered by this job

Group Health (Medical, Dental, Vision)
401(k) matching
Paid Time Off
Wellness programs
Professional Development opportunities

Job summary

HTC Global Services is seeking a Site Reliability Engineer II to join our SRE team focused on observability, monitoring, and technical consulting across GCP-based data platforms. You will ensure reliability and performance of cloud and network systems through automation, monitoring, and optimization.

The role emphasizes hands-on work with cloud infrastructure, BigQuery workloads, CI/CD pipelines, and enterprise monitoring tools, while collaborating with cross-functional teams to improve incident

Qualifications

  • Bachelor's degree.
  • 4+ years of experience in IT.
  • 3+ years of development experience.
  • Practitioner-level experience with at least one coding language or framework.
  • Hands-on experience with Google Cloud Platform (GCP).
  • Experience with BigQuery.
  • Experience with Dynatrace.
  • Proficiency with monitoring and observability tools, ideally Dynatrace or comparable tools such as Datadog or New Relic.
  • Familiarity with ITSM tools such as ServiceNow, including incident, problem, and change management.

Responsibilities

  • Collaborate with infrastructure teams to automate routine tasks.
  • Monitor and manage production environments, proactively identifying and resolving issues.
  • Participate in building advanced tooling for system access monitoring, log session recording, and reliability administration across multiple geographically distributed data centers.
  • Engage with engineering teams to improve on-call efficiencies, incident management, and post-mortem analysis.
  • Perform capacity planning and optimization to support growing demands and traffic patterns.
  • Maintain monitoring and alerting systems for proactive system health checks.
  • Continuously improve system performance, stability, and security through data-driven analysis and optimization.
  • Create and maintain comprehensive documentation and diagrams to facilitate knowledge sharing.
  • Work hands-on with cloud infrastructure, BigQuery workloads, CI/CD pipelines, and enterprise monitoring tools to maintain critical systems at scale.

Skills

GCP
BigQuery
Dynatrace
Monitoring
ITSM
Coding language

Education

Bachelor's degree

Tools

Datadog
New Relic
ServiceNow

Job description

HTC Global Services is seeking a Site Reliability Engineer II to join our SRE team focused on observability, monitoring, and technical consulting across GCP-based data platforms. You will ensure reliability and performance of cloud and network systems through automation, monitoring, and optimization.

The role emphasizes hands-on work with cloud infrastructure, BigQuery workloads, CI/CD pipelines, and enterprise monitoring tools, while collaborating with cross-functional teams to improve incident

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer II — Remote, GCP & Observability
Site Reliability Engineer II — Remote, GCP & Observability

HTC Global Services, Inc. • Dearborn (MI)

Hybrid
USD 110,000 - 140,000
Hybrid work
Work-Life Balance
Career development plan
+3
Site Reliability Engineer II
Site Reliability Engineer II

HTC Global Services • Dearborn (MI)

On-site
USD 110,000 - 140,000
Group Health (Medical, Dental, Vision)
401(k) matching
Paid Time Off
+2
Senior Cloud SRE — GCP Reliability & Observability
Senior Cloud SRE — GCP Reliability & Observability

Apex Systems • Dearborn (MI)

Hybrid
USD 110,000 - 165,000
Site Reliability Engineer II
Site Reliability Engineer II

HTC Global Services, Inc. • Dearborn (MI)

Hybrid
USD 110,000 - 140,000
Hybrid work
Work-Life Balance
Career development plan
+3
Site Reliability Engineer — Cloud Observability & Automation
Site Reliability Engineer — Cloud Observability & Automation

FastTek Global • Dearborn (MI)

On-site
USD 110,000 - 150,000
Medical and Dental
Vision
PTO Program
+3
Lead SRE: Cloud Infra, Kubernetes & CI/CD
Lead SRE: Cloud Infra, Kubernetes & CI/CD

HTC Global Services • Orlando (FL)

On-site
USD 150,000 - 210,000
Health Insurance
401(k) matching
Paid Time Off
Site Reliability Engineer - GCP & Automation Focus
Site Reliability Engineer - GCP & Automation Focus

Insight Global • United States

On-site
USD 100,000 - 125,000
Senior Observability & SRE Engineer — GCP/Kubernetes
Senior Observability & SRE Engineer — GCP/Kubernetes

Ontrac Solutions • United States

On-site
USD 120,000 - 180,000
Senior GCP SRE – Reliability, Observability & Automation
Senior GCP SRE – Reliability, Observability & Automation

Bank of America • Charlotte (NC)

On-site
USD 152,000 - 192,000
Industry-leading benefits
Paid time off
Discretionary incentive eligibility
Senior Cloud SRE: Hybrid GCP & On-Prem
Senior Cloud SRE: Hybrid GCP & On-Prem

Optimum • Plano (TX)

On-site
USD 100,000 - 143,000