HPC Systems Engineer

Jump Trading

Illinois

On-site

USD 150,000 - 200,000

Full time

14 days+
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Benefits offered by this job

Discretionary bonus eligibility
Medical, dental, and vision insurance
HSA, FSA options
Retirement plan with employer match
Paid vacation and holidays

Job summary

Jump Trading Group is a frontier technology firm known for research-driven innovation in mathematics, physics, and computer science. Our Chicago office hosts a high-performance computing environment where a Production Engineer will ensure the reliability and performance of large-scale compute and storage systems.

The role combines architecture, engineering, and operational duties, requiring hands-on work with Linux systems, performance monitoring, and cross-team collaboration with researchers to

Qualifications

  • 5+ years in high performance computing environments.
  • Strong Linux systems administration experience.
  • Proficiency in at least one programming/scripting language (Go, Python, C).
  • Experience designing and maintaining distributed systems.
  • Familiarity with debugging tools and profiling.
  • Experience with configuration management tools (SaltStack, Ansible, Puppet).
  • Strong root cause analysis and problem-solving skills.

Responsibilities

  • Architect, engineer, and administer HPC systems.
  • Design, implement, maintain, and support compute and storage infrastructure.
  • Develop tooling to build, package, install, and upgrade software at scale.
  • Collaborate across teams to write code and testing infrastructures.
  • Provide operations support on a rotating basis and during maintenance windows.
  • Manage vendor relationships and travel as needed.

Skills

HPC experience
Linux administration
Programming languages
Distributed systems
Debugging tools
Config management
Root cause analysis
Vendor relations

Tools

SaltStack
Ansible
Puppet

Job description

Jump Trading Group is committed to world class research. We empower exceptional talents in Mathematics, Physics, and Computer Science to seek scientific boundaries, push through them, and apply cutting edge research to global financial markets. Our culture is unique. Constant innovation requires fearlessness, creativity, intellectual honesty, and a relentless competitive streak. We believe in winning together and unlocking unique individual talent by incenting collaboration and mutual respect. At Jump, research outcomes drive more than superior risk adjusted returns. We design, develop, and deploy technologies that change our world, fund start-ups across industries, and partner with leading global research organizations and universities to solve problems.

Our global High Performance Computing Team is looking to add a Production Engineer in our Chicago office. The scale of our computing environments provides unique challenges in providing good performance and reliability. Several systems including compute, scheduling, networks, and large-scale data storage must integrate seamlessly to support data pipelines and quantitative research. The ideal candidate would be a hands-on individual, highly skilled in the details and nuances of managing Linux environments with a strong software development background necessary to support uniquely customized systems at scale.

What You’ll Do:
  • Daily access to world-class, large-scale production systems as well as an unparalleled high-performance research “supercomputer”
  • Perform the functions of architect, engineer, and administrator
  • Work on a variety of production engineering projects and troubleshooting a range of IT problems
  • Design, implement, maintain, and support high performance compute and storage systems
  • Implement and support performance monitoring and fault monitoring systems
  • Monitor systems and storage performance, up to and including network components
  • Build tooling to compile, package, install, and upgrade software and operating system components at scale
  • Collaborate with team members and across teams to write code and testing infrastructures spanning both new and existing codebases in multiple programming languages
  • Develop and improve systems and user documentation
  • Participate in large, coordinated maintenance operations, including during evenings and weekends.
  • Work on global projects across a wide range of infrastructure
  • Collaborate directly with researchers to optimize their use of HPC infrastructure
  • Develop and monitor the tools used to maintain a production computing environment
  • Provide operational support on a rotating basis and as needed
  • Manage relationships with outside vendors, including traveling both domestically and internationally to meet with current and potential vendors
  • Adhere to all company cybersecurity and IT policies, including performing all work using only approved hardware and software
  • Other duties as assigned or needed
Skills You’ll Need:
  • 5+ years of professional experience in high performance computing (HPC), including parallel filesystems (e.g., Lustre, GPFS), batch systems (e.g., Slurm, Grid Engine), and high-performance network interconnects experience is a plus, but not required
  • 5+ years of experience with Linux systems administration
  • High proficiency with at least one programming/scripting language (e.g., Go, Python, C)
  • Extensive experience designing, building, and maintaining complicated, interdependent, and distributed systems
  • Extensive experience profiling and debugging application stacks (debuggers and profilers)
  • Experience with system configuration management tools (SaltStack, Ansible, Puppet, etc.)
  • A compulsion to perform root cause analysis
  • Reliable and predictable availability

If you are currently a student or recent graduate, please see our Campus postings which offer both intern and full-time opportunities.

Benefits
  • Discretionary bonus eligibility
  • Medical, dental, and vision insurance
  • HSA, FSA, and Dependent Care options
  • Employer Paid Group Term Life and AD&D Insurance
  • Voluntary Life & AD&D insurance
  • Paid vacation plus paid holidays
  • Retirement plan with employer match
  • Paid parental leave
  • Wellness Programs
Annual Base Salary Range

$150,000—$200,000 USD

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

HPC Data Center Infrastructure Planning Lead
HPC Data Center Infrastructure Planning Lead

P2P • New York (NY)

On-site
USD 125,000 - 150,000
Discretionary bonus eligibility
Medical, dental, and vision insurance
Paid vacation plus paid holidays
+2
HPC Data Center Developer
HPC Data Center Developer

P2P • New York (NY)

On-site
USD 100,000 - 130,000
CORE | Data Center Technician
CORE | Data Center Technician

Socket.dev • Chicago (IL)

On-site
USD 120,000 - 180,000
Private Medical Insurance
Vision Insurance
Dental Insurance
+4
CORE - Data Center Technician
CORE - Data Center Technician

jumptrading • United States

On-site
USD 120,000 - 180,000
Medical Insurance
Pension Plan
Paid Vacation
+3
CORE | Data Center Technician
CORE | Data Center Technician

P2P • Chicago (IL)

On-site
USD 120,000 - 180,000
Medical Insurance
Vision Insurance
Dental Insurance
+10
HPC Systems Engineer - Production & Research Infra
HPC Systems Engineer - Production & Research Infra

Jump Trading • Illinois

On-site
USD 150,000 - 200,000
Discretionary bonus eligibility
Medical, dental, and vision insurance
HSA, FSA options
+2
Quantitative Developer - Real-Time Trading Systems Engineer
Quantitative Developer - Real-Time Trading Systems Engineer

Jump Trading • New York (NY)

On-site
USD 100,000 - 150,000
Campus Systems Engineer (Full-Time)
Campus Systems Engineer (Full-Time)

P2P • Chicago (IL)

On-site
USD 200,000 - 250,000
Discretionary bonus eligibility
Medical, dental, and vision insurance
HSA, FSA, and Dependent Care options
+6
Software Engineer | Market Data Systems
Software Engineer | Market Data Systems

Jump Trading • Illinois

On-site
USD 200,000 - 250,000
Discretionary bonus
Medical, dental, and vision insurance
HSA, FSA, and Dependent Care options
+5
HPC Data Center Operational Lead
HPC Data Center Operational Lead

P2P • New York (NY)

On-site
USD 120,000 - 160,000