HPC Data Center Production Engineer (Automation) - Banking & Finance

Hamilton Barnes

Illinois

On-site

USD 175,000 - 235,000

Full time

2 days ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Premium health, dental, and vision
Performance-based bonus
Meals and high-end office amenities
Fully funded developer tools and AI平台
Learning and conference allowance

Job summary

Hamilton Barnes, a specialist provider of sophisticated point-to-point wireless networks, seeks an HPC Data Center Production Engineer to build and scale automation and software that powers its high-performance computing data centres.

The role focuses on transforming hardware into automated, production-ready infrastructure, translating strategies into software, and building telemetry pipelines. Strong Linux fluency and Go coding are essential.

Qualifications

  • 5+ years in Production Engineering, SRE, or Infrastructure Automation in HPC/large-scale data centers.
  • Proficient in Golang and Python scripting.
  • Strong Linux administration and OS-level troubleshooting.
  • Deep understanding of data center power/cooling, cabling, and hardware interfaces (IPMI/BMC/Redfish/SNMP).
  • Experience with Grafana, Prometheus/InfluxDB, ClickHouse, MySQL for observability and data pipelines.
  • Proficient with SaltStack, Ansible, Terraform and GitHub CI/CD workflows.
  • Daily use of AI coding tools and AI analytics in software workflows.

Responsibilities

  • Design end-to-end automation workflows for servers, switches, PDUs, and sensors.
  • Develop tools for capacity planning, utilization modeling, and outage simulations.
  • Build telemetry integrations and data pipelines for centralized observability.
  • Collaborate with HPC Planning, Engineering, and Operations to automate manual pain points.
  • Own the software lifecycle of internal tooling, including maintenance and RCAs.
  • Leverage AI tooling for code generation, data analysis, and debugging.

Skills

Golang
Python scripting
Linux administration
Data center infrastructure
IPMI/Redfish/SNMP
Observability tooling
Infrastructure as Code
AI tooling

Tools

Grafana
Prometheus
InfluxDB
ClickHouse
MySQL
SaltStack
Ansible
Terraform
GitHub CI/CD

Job description

Are you looking for an exciting new opportunity?

Join a specialist provider of sophisticated point-to-point wireless networks supporting mission-critical trading operations across multiple financial markets, with a strong focus on quality, safety, and technical excellence.

The organization is currently on the lookout for an HPC Data Center Production Engineer to build, own, and scale the automation and software systems powering its high-performance computing data centres. The ideal candidate will transform raw hardware servers, switches, rack PDUs, CDUs, and liquid cooling systems into highly automated, production-ready infrastructure, translating operational strategies into software, designing outage simulations, and building advanced telemetry pipelines. Strong Linux fluency and clean Go coding ability are essential for this development-heavy, hands-on role.

Responsibilities:
  • Hardware Onboarding Automation: Design and build end-to-end automated workflows that take servers, switches, PDUs, CDUs, and environmental sensors from racked-and-cabled to production-ready.
  • Capacity & Simulation Tooling: Develop tools for power/cooling capacity planning, predictive utilization modeling, and outage simulations to test facility redundancy.
  • Observability & Telemetry Integration: Build custom telemetry integrations and metrics pipelines (IPMI/Redfish, SNMP) to normalize data from colocation facilities into centralized observability platforms.
  • Cross-Functional Architecture: Work directly with HPC Planning, Engineering, and Operations leads to convert manual pain points into maintainable, automated systems.
  • Reliability & Maintenance: Own the full software lifecycle of all internal tools, participating in scheduled maintenance windows and performing root-cause analysis on failures.
  • AI Tooling Acceleration: Leverage AI tools daily for code generation, data analysis, debugging, and predictive capacity modeling.
Skills/Must Have:
  • Experience: 5+ years in Production Engineering, Site Reliability Engineering (SRE), or Infrastructure Automation within HPC or large-scale data center environments.
  • Programming Proficiency: High proficiency in Golang alongside strong scripting capabilities in Python.
  • Linux Expertise: Mastery of Linux systems administration, OS-level troubleshooting, networking, and process management.
  • Hardware & Data Center Domain: Deep understanding of data center power/cooling infrastructure (air and liquid), structured cabling, and hardware management interfaces (IPMI, BMC, Redfish, SNMP).
  • Observability & Data Stack: Hands-on experience with Grafana, Prometheus/InfluxDB, ClickHouse, MySQL, and building custom metric exporters.
  • Infrastructure as Code: Proficiency with modern configuration management tools (SaltStack, Ansible, Terraform) and GitHub CI/CD workflows.
  • AI Integration: Daily, practical experience using LLM-based coding assistants and AI analytics tools in a professional software development workflow.
Benefits:
  • Premium health, dental, and vision coverage with top-tier benefits.
  • Generous performance-based bonus structures and long-term incentive programs.
  • Daily provided meals and high-end office amenities in a premier, collaborative workplace.
  • Fully funded access to top-tier developer tools, hardware, and AI platforms.
  • Comprehensive continuous learning and conference allowance.
Salary:
  • $175,000 - $235,000 + Performance-Based Bonus (Commensurate with experience)
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

HPC Data Center Developer
HPC Data Center Developer

P2P • New York (NY)

On-site
USD 100,000 - 130,000
HPC Developer
HPC Developer

Autonomai Recruitment • Chicago (IL)

On-site
USD 110,000 - 190,000
HPC Data Center Developer
HPC Data Center Developer

Jump Trading • Chicago (IL)

On-site
USD 150,000 - 200,000
Discretionary bonus eligibility
Medical, dental, and vision insurance
HSA, FSA, and Dependent Care options
+6
HPC Developer
HPC Developer

Autonomai Recruitment • New York (NY)

On-site
USD 130,000 - 195,000
HPC Data Center Infrastructure Planning Lead
HPC Data Center Infrastructure Planning Lead

Jump Trading • New York (NY)

On-site
USD 150,000 - 200,000
Discretionary bonus
Medical, dental, and vision insurance
HSA, FSA, and Dependent Care options
+5
Data Center Compute Engineer
Data Center Compute Engineer

Blue Signal Search • San Francisco (CA)

Hybrid
USD 150,000 - 210,000
Competitive compensation
Equity opportunity
Comprehensive benefits
+2
HPC Engineer
HPC Engineer

Autonomai Recruitment • Chicago (IL)

On-site
USD 250,000 - 500,000
HPC & Compute Engineering Lead
HPC & Compute Engineering Lead

Autonomai Recruitment • Chicago (IL)

On-site
USD 180,000 - 250,000
HPC Data Center Automation & Reliability Engineer
HPC Data Center Automation & Reliability Engineer

Hamilton Barnes • Illinois

On-site
USD 175,000 - 235,000
Premium health, dental, and vision
Performance-based bonus
Meals and high-end office amenities
+2
HPC Engineer (Elite Fintech Trading firm) $500,000 + Bonus/Benefits!
HPC Engineer (Elite Fintech Trading firm) $500,000 + Bonus/Benefits!

Hunter Bond • New York (NY)

On-site
USD 450,000 - 550,000
Exceptional bonus
Cutting-edge technology
Collaborative environment
+2