Systems Engineer

Tower Research Capital

Hong Kong

On-site

HKD 450,000 - 750,000

Full time

14 days+
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Benefits offered by this job

Generous PTO
Savings plans
Hybrid work
Free meals
Wellness reimbursements
Sports teams & events
Volunteer opportunities
Social events
Continuous learning

Job summary

Tower Research Capital seeks an experienced HPC/Infra Engineer to design, deploy, and operate large‑scale Linux compute and storage platforms across on‑prem and cloud environments. You will optimize batch workloads, build automation, and maintain HPC tools while collaborating with cross‑functional teams to ensure high availability and performance.

Responsibilities include deploying cloud infrastructure across GCP, AWS, and Azure, enhancing observability, and implementing scalable solutions for

Qualifications

  • Bachelor's degree in CS, engineering or related field.
  • Experience with Linux systems, HPC or infrastructure engineering.
  • Strong knowledge of Linux internals and networking.
  • Familiarity with cloud providers and IaC tools.
  • Scripting skills in Python/Go/Bash.
  • Experience with Docker/Kubernetes and batch schedulers.

Responsibilities

  • Design, support, and operate HPC compute/storage infra across on‑prem and cloud environments.
  • Maintain large-scale Linux systems covering compute, storage, networking, automation, and monitoring.
  • Troubleshoot OS/storage/networking/cluster scheduling issues with other teams.
  • Manage batch/workloads across CPU/GPU resources; optimize utilization.
  • Develop and operate cloud infrastructure across GCP, AWS, and Azure.
  • Build HPC management tools, access modules, and internal libraries.
  • Improve observability pipelines and system performance; drive efficiency.

Skills

Linux systems
DevOps
HPC
Cloud (GCP/AWS/Azure)
Scripting (Python/Go/Bash)
Networking (TCP/IP)
Batch schedulers (Slurm/HTCondor)
Storage systems
Monitoring/Observability
Automation

Education

Bachelor's degree in CS/Engineering

Tools

Docker
Kubernetes
Terraform
Ansible
Salt

Job description

Responsibilities
  • Designing, supporting, and operating HPC compute and storage infrastructure across on‑premises and cloud environments
  • Maintaining and improving large-scale Linux systems spanning compute, storage, networking, automation, and monitoring
  • Troubleshooting complex issues across OS, storage, networking, and cluster scheduling layers in collaboration with other infrastructure teams
  • Managing and optimizing batch and containerized workloads across diverse compute resources (CPU and GPU)
  • Developing, operating, and improving cloud infrastructure across various providers such as GCP, AWS, and Azure
  • Building and maintaining HPC management tools, user access modules, and internal libraries
  • Developing metrics and observability pipelines, analyzing system performance, and driving improvements in cluster utilization and efficiency
  • Managing code deployments, upgrades, fixes, and infrastructure lifecycle processes
  • Identifying manual or repetitive workflows and designing automation to improve reliability and user experience
  • Staying current with emerging hardware and software technologies relevant to HPC, storage, and cloud infrastructure
Qualifications
  • A bachelor’s degree or higher in computer science, engineering, or a related field
  • 1–5 years of relevant experience in Linux systems, DevOps, HPC, or infrastructure engineering
  • Strong understanding of Linux internals (process scheduling, virtual memory, filesystems, networking)
  • Experience with batch schedulers such as Slurm or HTCondor is a plus
  • Experience with distributed or networked storage systems (e.g., NFS, Weka, or object storage)
  • Experience with at least one major cloud provider (GCP, AWS, or Azure)
  • Familiarity with Infrastructure-as-Code and configuration management tools such as Ansible, Terraform, or Salt
  • Strong scripting or programming skills in Python, GO or Bash scripting
  • Hands‑on experience with container technologies such as Docker/Podman and Kubernetes
  • Solid understanding of networking fundamentals (TCP/IP, Ethernet)
  • Working knowledge of hardware and server components
  • Strong troubleshooting skills, with a bias toward automation and operational excellence
  • Clear communication skills and a strong focus on end‑user experience
Nice to Have
  • Experience managing GPU‑based compute platforms
  • Exposure to CI/CD systems and release automation
  • Experience operating hybrid on‑prem + cloud HPC environments
  • Interest in performance tuning, scalability, and systems optimization
Benefits
  • Generous paid time off policies
  • Savings plans and other financial wellness tools available in each region
  • Hybrid working opportunities
  • Free breakfast, lunch, and snacks daily
  • In‑office wellness experiences and reimbursement for select wellness expenses (e.g., gym, personal training)
  • Company‑sponsored sports teams and fitness events
  • Volunteer opportunities and charitable giving
  • Social events, happy hours, treats, and celebrations throughout the year
  • Workshops and continuous learning opportunities

Tower Research Capital is an equal opportunity employer.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

HPC Systems Engineer - Hybrid, Cloud & Automation
HPC Systems Engineer - Hybrid, Cloud & Automation

Tower Research Capital • Hong Kong

On-site
HKD 450,000 - 750,000
Generous PTO
Savings plans
Hybrid work
+6
Platform Engineer - Tribus
Platform Engineer - Tribus

Tribus • Hong Kong

On-site
HKD 480,000 - 900,000
Server Engineer (On-site Hong Kong)
Server Engineer (On-site Hong Kong)

Pragmatike • Hong Kong

On-site
HKD 420,000 - 660,000
Cloud Engineer (AWS Infrastructure & Compute Architect)
Cloud Engineer (AWS Infrastructure & Compute Architect)

Selby Jennings • Hong Kong

Hybrid
HKD 900,000 - 1,500,000
Quantitative Developer (C++)
Quantitative Developer (C++)

Tower Research Capital • Hong Kong

On-site
HKD 500,000 - 1,000,000
Hybrid work
Free meals
Wellness reimbursements
+3
GPU Server Hardware Validation Engineer
GPU Server Hardware Validation Engineer

Pragmatike • Hong Kong

On-site
HKD 420,000 - 700,000
Software Engineer, C++ - Graduate Opportunity
Software Engineer, C++ - Graduate Opportunity

Tower Research Capital LLC • Hong Kong

Hybrid
HKD 300,000 - 520,000
Generous paid time off policies
Savings plans
Hybrid working opportunities
+5
Server Engineer (HK - Hybrid)
Server Engineer (HK - Hybrid)

Pragmatike • Hong Kong

Hybrid
HKD 480,000 - 720,000
DevOps/High Performance Trading System Engineer
DevOps/High Performance Trading System Engineer

CW Talent Solutions • Hong Kong

On-site
HKD 900,000 - 1,500,000
Private medical
Travel medical insurance
Pension plan
+1
Lead System Admin - Linux Server
Lead System Admin - Linux Server

Don Nelson Recruitment Limited • Hong Kong

On-site
HKD 420,000 - 650,000