Site Reliability Engineer II — Remote, GCP & Observability

HTC Global Services, Inc.

Dearborn (MI)

Hybrid

USD 110,000 - 140,000

Full time

3 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Hybrid work
Work-Life Balance
Career development plan
Rewards & Recognition program
Upskilling opportunities
Career mobility

Job summary

HTC Global Services is seeking a Site Reliability Engineer II to join a remote-friendly SRE team supporting GCP-based data platforms. You will ensure the availability, reliability, and performance of cloud and network systems through automation, monitoring, and optimization.

You will collaborate across infrastructure teams to automate tasks, monitor production, and develop tooling for system access, logging, and incident management. Strong knowledge of Dynatrace, BigQuery, and CI/CD is preferred.

Qualifications

  • 4+ years of experience in IT.
  • 3+ years of development experience.
  • Practitioner-level experience with at least one coding language or framework.
  • Hands-on experience with Google Cloud Platform (GCP).
  • Experience with BigQuery.
  • Experience with Dynatrace.
  • Proficiency with monitoring and observability tools, ideally Dynatrace or comparable tools such as Datadog or New Relic.
  • Familiarity with ITSM tools such as ServiceNow, including incident, problem, and change management.

Responsibilities

  • Collaborate with infrastructure teams to automate routine tasks.
  • Monitor and manage production environments, proactively identifying and resolving issues.
  • Build tooling for system access monitoring, log session recording, and reliability administration across multiple data centers.
  • Engage with engineering teams to improve on-call efficiencies, incident management, and post-mortem analysis.
  • Perform capacity planning and optimization to support growing demands and traffic patterns.
  • Maintain monitoring and alerting systems for proactive system health checks.
  • Improve system performance, stability, and security through data-driven analysis and optimization.
  • Create and maintain comprehensive documentation and diagrams to facilitate knowledge sharing.
  • Work hands-on with cloud infrastructure, BigQuery workloads, CI/CD pipelines, and enterprise monitoring tools to maintain critical systems at scale.

Skills

Troubleshooting skills
GCP experience
Development experience
Observability
Problem-solving

Tools

Dynatrace
Datadog
New Relic
ServiceNow
BigQuery

Job description

HTC Global Services is seeking a Site Reliability Engineer II to join a remote-friendly SRE team supporting GCP-based data platforms. You will ensure the availability, reliability, and performance of cloud and network systems through automation, monitoring, and optimization.

You will collaborate across infrastructure teams to automate tasks, monitor production, and develop tooling for system access, logging, and incident management. Strong knowledge of Dynatrace, BigQuery, and CI/CD is preferred.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer II
Site Reliability Engineer II

HTC Global Services, Inc. • Dearborn (MI)

Hybrid
USD 110,000 - 140,000
Hybrid work
Work-Life Balance
Career development plan
+3
Lead SRE: Cloud Infra, Kubernetes & CI/CD
Lead SRE: Cloud Infra, Kubernetes & CI/CD

HTC Global Services • Orlando (FL)

On-site
USD 150,000 - 210,000
Health Insurance
401(k) matching
Paid Time Off
Senior AI Platform Reliability Engineer
Senior AI Platform Reliability Engineer

HTC Global Services • Orlando (FL)

On-site
USD 150,000 - 210,000
Health Insurance
401(k) matching
Paid Time Off
Lead Site Reliability Engineer, GCP & Cloud Platforms
Lead Site Reliability Engineer, GCP & Cloud Platforms

Optimum • Bethpage (NY)

On-site
USD 134,000 - 220,000
Remote GCP Cloud Adoption SRE & Automation Engineer
Remote GCP Cloud Adoption SRE & Automation Engineer

Huntington Bancshares Inc. • Columbus (OH)

Hybrid
USD 120,000 - 160,000
Senior Observability & SRE Engineer — GCP/Kubernetes
Senior Observability & SRE Engineer — GCP/Kubernetes

Ontrac Solutions • United States

On-site
USD 120,000 - 180,000
Remote SRE II: Cloud-Native Reliability & Automation
Remote SRE II: Cloud-Native Reliability & Automation

NationsBenefits, LLC • Plantation (FL)

On-site
USD 110,000 - 160,000
Unlimited PTO
Competitive compensation & benefits
Career growth opportunities
+1
Site Reliability Engineer - GCP & Automation Focus
Site Reliability Engineer - GCP & Automation Focus

Insight Global • United States

On-site
USD 100,000 - 125,000
Site Reliability Engineer
Site Reliability Engineer

Compunnel, Inc. • New Jersey

On-site
USD 120,000 - 150,000
Site Reliability Engineer (SRE) – II
Site Reliability Engineer (SRE) – II

Huntington National Bank • Columbus (OH)

Hybrid
USD 90,000 - 120,000