HPC Systems Engineer

P2P

Greater London

On-site

GBP 90,000 - 140,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Private medical, vision and dental
Travel medical insurance
Group pension scheme
Parental leave

Job summary

Jump Trading Group, based in London, is seeking an experienced Systems Engineer to design, deploy, and maintain high performance compute and storage systems. You will work closely with researchers and software teams to optimize workflows across Linux-based HPC infrastructure.

The role involves building scalable tooling, monitoring, and documentation, plus occasional travel for vendor relations and cross-site projects. A strong background in HPC, Linux, and scripting is essential for success.

Qualifications

  • 5+ years in HPC environments with Lustre/GPFS and Slurm/Grid Engine.
  • 5+ years Linux administration experience.
  • Proficiency in Go, Python, or C.
  • Experience with configuration management tools (SaltStack, Ansible, Puppet).
  • Strong debugging and profiling skills and distributed systems knowledge.

Responsibilities

  • Design, implement, maintain, and support HPC compute and storage systems.
  • Implement and support performance and fault monitoring for large-scale environments.
  • Monitor performance of compute, storage, and networks, including interconnects.
  • Build tooling to compile, package, install, and upgrade software and OS components at scale.
  • Collaborate with researchers and teams to optimize HPC usage and workflows.

Skills

HPC experience
Linux systems administration
Go
Python
C
Distributed systems design
Root cause analysis
Team collaboration

Tools

SaltStack
Ansible
Puppet

Job description

Jump Trading Group is committed to world class research. We empower exceptional talents in Mathematics, Physics, and Computer Science to seek scientific boundaries, push through them, and apply cutting edge research to global financial markets. Our culture is unique. Constant innovation requires fearlessness, creativity, intellectual honesty, and a relentless competitive streak. We believe in winning together and unlocking unique individual talent by incenting collaboration and mutual respect. At Jump, research outcomes drive more than superior risk adjusted returns. We design, develop, and deploy technologies that change our world, fund start-ups across industries, and partner with leading global research organizations and universities to solve problems.

Our global High Performance Computing Team is looking to add a Systems Engineer in our London office. The scale of our computing environments provides unique challenges in providing good performance and reliability. Several systems including compute, scheduling, networks, and large-scale data storage must integrate seamlessly to support data pipelines and quantitative research. The ideal candidate would be a hands‑on individual, highly skilled in the details and nuances of managing Linux environments with a strong software development background necessary to support uniquely customized systems at scale.

What You'll Do:
  • Design, implement, maintain, and support high performance compute and storage systems
  • Implement and support performance monitoring and fault monitoring systems
  • Monitor systems and storage performance, up to and including network components
  • Build tooling to compile, package, install, and upgrade software and operating system components at scale
  • Collaborate with team members and across teams to write code and testing infrastructures spanning both new and existing codebases in multiple programming languages
  • Develop and improve systems and user documentation
  • Participate in large, coordinated maintenance operations, including during evenings and weekends
  • Work on global projects across a wide range of infrastructure
  • Collaborate directly with researchers to optimize their use of HPC infrastructure
  • Develop and monitor the tools used to maintain a production computing environment
  • Provide operational support on a rotating basis and as needed
  • Manage relationships with outside vendors, including traveling both domestically and internationally to meet with current and potential vendors
  • Adhere to all company cybersecurity and IT policies, including performing all work using only approved hardware and software Other duties as assigned or needed
Skills You'll Need:
  • At least 5+ years of professional experience in high performance computing (HPC), including parallel filesystems (e.g., Lustre, GPFS), batch systems (e.g., Slurm, Grid Engine), and high-performance network interconnects experience is a plus, but not required
  • At least 5+ years of experience with Linux systems administration
  • High proficiency with at least one programming/scripting language (e.g., Go, Python, C)
  • Extensive experience designing, building, and maintaining complicated, interdependent, and distributed systems
  • Extensive experience profiling and debugging application stacks (debuggers and profilers)
  • Experience with system configuration management tools (SaltStack, Ansible, Puppet, etc.)
  • A compulsion to perform root cause analysis
  • Reliable and predictable availability

Benefits include:

  • Private Medical, Vision and Dental Insurance
  • Travel Medical Insurance
  • Group Pension Scheme
  • Group Life Assurance and Income Protection Schemes
  • Paid Parental Leave
  • Parking and Commuter Benefits
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior HPC Systems Engineer, London
Senior HPC Systems Engineer, London

P2P • Greater London

On-site
GBP 90,000 - 140,000
Private medical, vision and dental
Travel medical insurance
Group pension scheme
+1
Campus Systems Engineer (Full-Time)
Campus Systems Engineer (Full-Time)

P2P • Greater London

On-site
GBP 70,000 - 120,000
Private Medical
Vision Insurance
Dental Insurance
+5
DevOps/High Performance Trading System Engineer
DevOps/High Performance Trading System Engineer

Jump Trading • Greater London

On-site
GBP 50,000 - 80,000
Private Medical, Vision and Dental Insurance
Travel Medical Insurance
Group Pension Scheme
+3
HPC Operations Engineer - Banking & Finance
HPC Operations Engineer - Banking & Finance

Hamilton Barnes Associates Limited • Greater London

On-site
GBP 70,000 - 110,000
DevOps/High Performance Trading System Engineer
DevOps/High Performance Trading System Engineer

P2P • Greater London

On-site
GBP 70,000 - 110,000
Private Medical, Vision and Dental Insurance
Travel Medical Insurance
Group Pension Scheme
+6
HPC Engineer (Linux) - Sussex, Onsite
HPC Engineer (Linux) - Sussex, Onsite

Arden Resourcing • Pease Pottage

On-site
GBP 75,000 - 85,000
Annual bonus scheme
Enhanced pension contribution
Private medical and dental options
+2
HPC Systems Engineer - Up to £200k + Bonus
HPC Systems Engineer - Up to £200k + Bonus

Hunter Bond • Greater London

Hybrid
GBP 200,000
30 days holiday
All certifications and further education paid
Outstanding work-life balance
+3
HPC Engineer
HPC Engineer

LinuxRecruit • Greater London

On-site
GBP 50,000 - 70,000
Senior HPC Engineer - Hedge Fund - £150K+
Senior HPC Engineer - Hedge Fund - £150K+

Oliver Bernard • Greater London

On-site
GBP 150,000 - 190,000
High remuneration
Senior HPC Engineer
Senior HPC Engineer

CGG Services (UK) Limited • Bolney

On-site
GBP 60,000 - 80,000
Competitive salary
Annual bonus scheme
Flexible holiday program
+4