Site Reliability Engineer II

HTC Global Services, Inc.

Dearborn (MI)

Hybrid

USD 110,000 - 140,000

Full time

4 days ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Benefits offered by this job

Hybrid work
Work-Life Balance
Career development plan
Rewards & Recognition program
Upskilling opportunities
Career mobility

Job summary

HTC Global Services is seeking a Site Reliability Engineer II to join a remote-friendly SRE team supporting GCP-based data platforms. You will ensure the availability, reliability, and performance of cloud and network systems through automation, monitoring, and optimization.

You will collaborate across infrastructure teams to automate tasks, monitor production, and develop tooling for system access, logging, and incident management. Strong knowledge of Dynatrace, BigQuery, and CI/CD is preferred.

Qualifications

  • 4+ years of experience in IT.
  • 3+ years of development experience.
  • Practitioner-level experience with at least one coding language or framework.
  • Hands-on experience with Google Cloud Platform (GCP).
  • Experience with BigQuery.
  • Experience with Dynatrace.
  • Proficiency with monitoring and observability tools, ideally Dynatrace or comparable tools such as Datadog or New Relic.
  • Familiarity with ITSM tools such as ServiceNow, including incident, problem, and change management.

Responsibilities

  • Collaborate with infrastructure teams to automate routine tasks.
  • Monitor and manage production environments, proactively identifying and resolving issues.
  • Build tooling for system access monitoring, log session recording, and reliability administration across multiple data centers.
  • Engage with engineering teams to improve on-call efficiencies, incident management, and post-mortem analysis.
  • Perform capacity planning and optimization to support growing demands and traffic patterns.
  • Maintain monitoring and alerting systems for proactive system health checks.
  • Improve system performance, stability, and security through data-driven analysis and optimization.
  • Create and maintain comprehensive documentation and diagrams to facilitate knowledge sharing.
  • Work hands-on with cloud infrastructure, BigQuery workloads, CI/CD pipelines, and enterprise monitoring tools to maintain critical systems at scale.

Skills

Troubleshooting skills
GCP experience
Development experience
Observability
Problem-solving

Tools

Dynatrace
Datadog
New Relic
ServiceNow
BigQuery

Job description

Location: Dearborn (MI) | Employment Type: Full Time - 40 hours per week | Job Level: T3 | Work Preference: Remote Work From Home | Job Code: 244953

Job Description:
Site Reliability Engineer II
Overview / Summary

We are seeking a Site Reliability Engineer to join an SRE team focused on observability, monitoring, and technical consulting across GCP-based data platforms. This role is responsible for ensuring the availability, reliability, and performance of cloud and network systems and services through automation, monitoring, troubleshooting, and continuous optimization.

Key Responsibilities
  • Collaborate with infrastructure teams to implement critical solutions by automating routine tasks.
  • Monitor and manage production environments, proactively identifying and resolving issues.
  • Participate in building advanced tooling for system access monitoring, log session recording, and reliability administration across multiple geographically distributed data centers.
  • Engage with engineering teams to improve on-call efficiencies, incident management, and post-mortem analysis.
  • Perform capacity planning and optimization to support growing demands and traffic patterns.
  • Maintain monitoring and alerting systems for proactive system health checks.
  • Continuously improve system performance, stability, and security through data-driven analysis and optimization.
  • Create and maintain comprehensive documentation and diagrams to facilitate knowledge sharing.
  • Work hands-on with cloud infrastructure, BigQuery workloads, CI/CD pipelines, and enterprise monitoring tools to maintain critical systems at scale.
Required Qualifications
  • 4+ years of experience in IT.
  • 3+ years of development experience.
  • Practitioner-level experience with at least one coding language or framework.
  • Hands‑on experience with Google Cloud Platform (GCP).
  • Experience with BigQuery.
  • Experience with Dynatrace.
  • Proficiency with monitoring and observability tools, ideally Dynatrace or comparable tools such as Datadog or New Relic.
  • Familiarity with ITSM tools such as ServiceNow, including incident, problem, and change management.
Preferred Qualifications
  • Experience with GCP Cloud Run.
  • Experience with Python.
  • Strong troubleshooting and problem‑solving skills.
  • Familiarity with AI tools, including agents, skills, LLMs, and copilots.
  • Experience defining and tracking SLAs, SLOs, and SLIs.

What Makes HTC A Great Place To Build Your Future

HTC Global Services wants you to join our team. Come build new things with us and advance your career. At HTC Global, you’ll collaborate with experts, work alongside clients, and be part of high‑performing teams driving success together. You’ll have long‑term opportunities to grow your career and develop skills in the latest emerging technologies.

At HTC Global Services, our employees have access to a comprehensive benefits package. Benefits can include Group Health (Medical, Dental, and Vision), Paid Time Off, Paid Holidays, 401(k) matching, Group Life and Disability insurance, Professional Development opportunities, Wellness programs, and a variety of other perks.

Our success as a company is built on inclusion and diversity. HTC Global Services is committed to providing a workplace free from discrimination and harassment, where every employee is treated with dignity and respect. We celebrate differences and believe that diverse cultures, perspectives, and skills drive innovation and success. HTC is an Equal Opportunity Employer and a proud National Minority Supplier. We seek to empower each individual, fostering an environment where everyone feels valued, included, and respected.

At HTC Global Services, our culture is an embodiment of who we are – a value‑led organizationcommitted to success of our people and customers.

  • Hybrid and Workplace flexibility
  • Work-Life-Balance
  • Well‑defined career development plan
  • Rewards & Recognition program
  • L&D focuses on upskilling
  • Hands‑on experience on Emerging Technologies and Digital Transformation
  • Career Mobility programs

We do not sell or share your personal information with any third parties

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior AI Platform Reliability Engineer
Senior AI Platform Reliability Engineer

HTC Global Services • Orlando (FL)

On-site
USD 150,000 - 210,000
Health Insurance
401(k) matching
Paid Time Off
Data Engineer III
Data Engineer III

HTC Global Services, Inc. • Dearborn (MI), Northern (KY)

Hybrid
USD 110,000 - 140,000
Health benefits
PTO (Paid Time Off)
Paid Holidays
+4
Software Engineer – Full Stack
Software Engineer – Full Stack

HTC Global Services, Inc. • Dearborn (MI), Northern (KY)

Hybrid
USD 110,000 - 160,000
Hybrid & workplace flexibility
Work-Life Balance
Career development opportunities
Site Reliability Engineer #1063726
Site Reliability Engineer #1063726

FastTek Global • Dearborn (MI)

On-site
USD 120,000 - 180,000
Medical & Dental
Vision
PTO
+6
Senior Site Reliability Engineer – Multi-Cloud Architecture
Senior Site Reliability Engineer – Multi-Cloud Architecture

HTC Global Services • Madison (WI)

On-site
USD 100,000 - 130,000
Paid-Time-Off
401K matching
Life Insurance
+1
Senior Software Engineer – Full Stack / Cloud
Senior Software Engineer – Full Stack / Cloud

HTC Global Services • Dearborn (MI)

On-site
USD 120,000 - 160,000
The Digital Site Reliability Engineer (SRE) - GCP Cloud Adoption Engineer
The Digital Site Reliability Engineer (SRE) - GCP Cloud Adoption Engineer

Huntington National Bank • Columbus (OH)

Hybrid
USD 90,000 - 120,000
Senior Java Backend Engineer
Senior Java Backend Engineer

HTC Global Services, Inc. • Dearborn (MI), Northern (KY)

Hybrid
USD 110,000 - 150,000
Hybrid work environment
Work-Life Balance
Career development plan
+4
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Bank of America • Charlotte (NC)

On-site
USD 152,000 - 192,000
Industry-leading benefits
Paid time off
Discretionary incentive eligibility
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Bank of America • Plano (TX)

On-site
USD 152,000 - 192,000
Industry-leading benefits
Paid time off
Access to resources and support