Site Reliability Engineer III

US Diversity Job Search

Pennington (NJ)

On-site

USD 84,000 - 119,000

Full time

14 days+
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

BCforward is seeking a Site Reliability Engineer III to design, build, and operate tooling, automation frameworks, and observability capabilities for critical trading services at scale.

The role emphasizes reliability, incident management, CI/CD, cloud migration, and collaboration with cross-functional teams; onsite 4 days weekly for the first six months, then 3 days onsite and 2 days remote.

Qualifications

  • Strong Linux administration and troubleshooting skills.
  • Proficiency in Python and Shell scripting or similar automation languages.
  • Hands-on experience with observability platforms such as Dynatrace, Splunk, Grafana, and Prometheus.
  • Experience implementing monitoring, alerting, and operational dashboards.
  • Incident management, production support, and root cause analysis expertise.
  • Experience integrating systems via APIs and service interfaces.
  • Knowledge of cloud platforms and distributed systems architecture.
  • Experience with CI/CD pipelines and deployment automation.
  • Understanding of database technologies such as MySQL and MongoDB, and familiarity with Node.js.
  • Strong analytical, problem-solving, and debugging skills.
  • Experience in a Site Reliability Engineering organization with cross-team collaboration.
  • Hands-on experience defining and implementing SLIs, SLOs, and error budgets.

Responsibilities

  • Design, develop, and maintain SRE tooling, automation frameworks, and engineering utilities.
  • Implement and optimize observability using Dynatrace, Splunk, Grafana, and Prometheus.
  • Define and operationalize SLIs, SLOs, and error budgets to improve service reliability.
  • Build monitoring, alerting, and operational dashboards to support production excellence.
  • Lead and support incident management, root cause analysis, and continuous improvement actions.
  • Integrate systems and services through APIs and service interfaces to streamline workflows.
  • Contribute to CI/CD pipeline design and deployment automation for reliable releases.
  • Partner with application, infrastructure, and production support teams to embed SRE practices.
  • Conduct reliability engineering, resiliency assessments, and operational readiness reviews.
  • Support cloud migration and modernization programs with reliable architectures.

Skills

Linux
Python
Shell scripting
Dynatrace
Splunk
Grafana
Prometheus
APIs
CI/CD
Cloud platforms
OpenTelemetry
MySQL
MongoDB
Node.js

Tools

Kubernetes
OpenShift
Terraform
Ansible
OpenTelemetry

Job description

Job Title: Site Reliability Engineer III Location: Pennington, NJ Duration: Contract - 7 months Pay Range: $73.67/hr (W2) Job ID: 409685

About BCforward

BCforward is a leading global IT consulting and workforce solutions firm providing services and support to Fortune 500 and government clients. Founded in 1998, BCforward has grown with our customers needs into a full-service business solutions provider. With delivery centers and offices across North America and India, we take pride in building long-term relationships and delivering excellence through innovation, collaboration, and integrity.

Job Description

We are seeking a Site Reliability Engineer III to join our dynamic team within Global Markets. The successful candidate will design, build, and support tooling, automation frameworks, and observability capabilities that enable reliable operation of critical trading and business services at scale. The role will drive SRE best practices, improve operational workflows, enhance platform observability, and develop solutions that reduce manual effort while increasing service reliability and resilience.

Work Arrangement

Onsite 4 days per week for the first six months, then 3 days onsite and 2 days remote.

Responsibilities
  • Design, develop, and maintain SRE tooling, automation frameworks, and engineering utilities.
  • Implement and optimize observability using platforms such as Dynatrace, Splunk, Grafana, and Prometheus.
  • Define and operationalize SLIs, SLOs, and error budgets to improve service reliability.
  • Build monitoring, alerting, and operational dashboards to support production excellence.
  • Lead and support incident management, root cause analysis, and continuous improvement actions.
  • Integrate systems and services through APIs and service interfaces to streamline workflows.
  • Contribute to CI/CD pipeline design and deployment automation for reliable releases.
  • Partner with application, infrastructure, and production support teams to embed SRE practices.
  • Conduct reliability engineering, resiliency assessments, and operational readiness reviews.
  • Support cloud migration and modernization programs with reliable architectures.
Required Skills & Qualifications
  • Strong Linux administration and troubleshooting skills.
  • Proficiency in Python and Shell scripting or similar automation languages.
  • Hands-on experience with observability platforms such as Dynatrace, Splunk, Grafana, and Prometheus.
  • Experience implementing monitoring, alerting, and operational dashboards.
  • Incident management, production support, and root cause analysis expertise.
  • Experience integrating systems via APIs and service interfaces.
  • Knowledge of cloud platforms and distributed systems architecture.
  • Experience with CI/CD pipelines and deployment automation.
  • Understanding of database technologies such as MySQL and MongoDB, and familiarity with Node.js.
  • Strong analytical, problem-solving, and debugging skills.
  • Experience in a Site Reliability Engineering organization with cross-team collaboration.
  • Hands-on experience defining and implementing SLIs, SLOs, and error budgets.
Preferred Skills
  • Kubernetes and OpenShift experience.
  • Infrastructure as Code using Terraform or Ansible.
  • Development of self-service platforms and engineering productivity tools.
  • AIOps, intelligent alerting, and automated incident triage solutions.
  • Knowledge of observability standards including OpenTelemetry.
  • Experience in large-scale, high-availability financial services environments.
Why BCforward?

At BCforward, we believe in advancing lives and careers. When you join our team, you gain access to:

  • Competitive compensation and benefits.
  • Opportunities for growth with global clients.
  • A supportive, inclusive culture that values innovation and people.
  • Exposure to cutting-edge technologies and projects.
About Our Commitment

BCforward is an equal opportunity employer. We value diversity and are committed to creating an inclusive environment for all employees. All qualified applicants will receive consideration for employment without regard to race, color, religion, gender, sexual orientation, gender identity, national origin, age, disability, or veteran status.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

B Site Reliability Engineer Iii BC Forward Pennington, New Jersey, US $73.67-73.67
B Site Reliability Engineer Iii BC Forward Pennington, New Jersey, US $73.67-73.67

Artha Nexgen • Northern (KY)

Hybrid
USD 150,000 - 156,000
Contract 7 months
Onsite 4 days/week
Remote 2 days
Network / System Engineer II
Network / System Engineer II

BCforward • Pennington (NJ)

On-site
USD 83,000 - 90,000
Senior Site Reliability Engineer Observability & Automation
Senior Site Reliability Engineer Observability & Automation

Artha Nexgen • Northern (KY)

Hybrid
USD 150,000 - 156,000
Contract 7 months
Onsite 4 days/week
Remote 2 days
Network / System Engineer III
Network / System Engineer III

BC Forward • Jersey City (NJ)

On-site
USD 83,000 - 118,000
Site Reliability Engineer III — Hybrid Onsite/Remote
Site Reliability Engineer III — Hybrid Onsite/Remote

US Diversity Job Search • Pennington (NJ)

On-site
USD 84,000 - 119,000
Application Architect III
Application Architect III

BC Forward • Chandler (AZ)

On-site
USD 181,843,000 - 223,171,000
Application Programmer III
Application Programmer III

BC Forward • Newark (NJ)

On-site
USD 76,000 - 103,000
Competitive compensation and benefits
Opportunities for growth with global/2
Exposure to cutting-edge technologies
Senior Software Engineer III
Senior Software Engineer III

BC Forward • Kentucky

On-site
USD 96,000 - 107,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Hobbsnews • Jersey City (NJ)

On-site
USD 153,000 - 192,000
Benefits eligible
Discretionary incentive plan
Senior Site Reliability Engineer
Senior Site Reliability Engineer

United States Digital Space LLC • Charlotte (TX)

On-site
USD 153,000 - 192,000
Discretionary incentive eligible
Benefits package