Systems Engineer

Tower Research Capital

Hong Kong

On-site

HKD 450,000 - 750,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Generous PTO
Savings plans
Hybrid work
Free meals
Wellness reimbursements
Sports teams & events
Volunteer opportunities
Social events
Continuous learning

Job summary

Tower Research Capital seeks an experienced HPC/Infra Engineer to design, deploy, and operate large‑scale Linux compute and storage platforms across on‑prem and cloud environments. You will optimize batch workloads, build automation, and maintain HPC tools while collaborating with cross‑functional teams to ensure high availability and performance.

Responsibilities include deploying cloud infrastructure across GCP, AWS, and Azure, enhancing observability, and implementing scalable solutions for

Qualifications

  • Bachelor's degree in CS, engineering or related field.
  • Experience with Linux systems, HPC or infrastructure engineering.
  • Strong knowledge of Linux internals and networking.
  • Familiarity with cloud providers and IaC tools.
  • Scripting skills in Python/Go/Bash.
  • Experience with Docker/Kubernetes and batch schedulers.

Responsibilities

  • Design, support, and operate HPC compute/storage infra across on‑prem and cloud environments.
  • Maintain large-scale Linux systems covering compute, storage, networking, automation, and monitoring.
  • Troubleshoot OS/storage/networking/cluster scheduling issues with other teams.
  • Manage batch/workloads across CPU/GPU resources; optimize utilization.
  • Develop and operate cloud infrastructure across GCP, AWS, and Azure.
  • Build HPC management tools, access modules, and internal libraries.
  • Improve observability pipelines and system performance; drive efficiency.

Skills

Linux systems
DevOps
HPC
Cloud (GCP/AWS/Azure)
Scripting (Python/Go/Bash)
Networking (TCP/IP)
Batch schedulers (Slurm/HTCondor)
Storage systems
Monitoring/Observability
Automation

Education

Bachelor's degree in CS/Engineering

Tools

Docker
Kubernetes
Terraform
Ansible
Salt

Job description

Responsibilities
  • Designing, supporting, and operating HPC compute and storage infrastructure across on‑premises and cloud environments
  • Maintaining and improving large-scale Linux systems spanning compute, storage, networking, automation, and monitoring
  • Troubleshooting complex issues across OS, storage, networking, and cluster scheduling layers in collaboration with other infrastructure teams
  • Managing and optimizing batch and containerized workloads across diverse compute resources (CPU and GPU)
  • Developing, operating, and improving cloud infrastructure across various providers such as GCP, AWS, and Azure
  • Building and maintaining HPC management tools, user access modules, and internal libraries
  • Developing metrics and observability pipelines, analyzing system performance, and driving improvements in cluster utilization and efficiency
  • Managing code deployments, upgrades, fixes, and infrastructure lifecycle processes
  • Identifying manual or repetitive workflows and designing automation to improve reliability and user experience
  • Staying current with emerging hardware and software technologies relevant to HPC, storage, and cloud infrastructure
Qualifications
  • A bachelor’s degree or higher in computer science, engineering, or a related field
  • 1–5 years of relevant experience in Linux systems, DevOps, HPC, or infrastructure engineering
  • Strong understanding of Linux internals (process scheduling, virtual memory, filesystems, networking)
  • Experience with batch schedulers such as Slurm or HTCondor is a plus
  • Experience with distributed or networked storage systems (e.g., NFS, Weka, or object storage)
  • Experience with at least one major cloud provider (GCP, AWS, or Azure)
  • Familiarity with Infrastructure-as-Code and configuration management tools such as Ansible, Terraform, or Salt
  • Strong scripting or programming skills in Python, GO or Bash scripting
  • Hands‑on experience with container technologies such as Docker/Podman and Kubernetes
  • Solid understanding of networking fundamentals (TCP/IP, Ethernet)
  • Working knowledge of hardware and server components
  • Strong troubleshooting skills, with a bias toward automation and operational excellence
  • Clear communication skills and a strong focus on end‑user experience
Nice to Have
  • Experience managing GPU‑based compute platforms
  • Exposure to CI/CD systems and release automation
  • Experience operating hybrid on‑prem + cloud HPC environments
  • Interest in performance tuning, scalability, and systems optimization
Benefits
  • Generous paid time off policies
  • Savings plans and other financial wellness tools available in each region
  • Hybrid working opportunities
  • Free breakfast, lunch, and snacks daily
  • In‑office wellness experiences and reimbursement for select wellness expenses (e.g., gym, personal training)
  • Company‑sponsored sports teams and fitness events
  • Volunteer opportunities and charitable giving
  • Social events, happy hours, treats, and celebrations throughout the year
  • Workshops and continuous learning opportunities

Tower Research Capital is an equal opportunity employer.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

HPC Systems Engineer - Hybrid, Cloud & Automation
HPC Systems Engineer - Hybrid, Cloud & Automation

Tower Research Capital • Hong Kong

On-site
HKD 450,000 - 750,000
Generous PTO
Savings plans
Hybrid work
+6
Systems Engineer, HKG
Systems Engineer, HKG

Compute, Inc. • Hong Kong

On-site
HKD 420,000 - 700,000
Health & wellness
Remote flexibility
Generous time off
+3
Software Engineer, C++
Software Engineer, C++

Tower Research Capital • Hong Kong

Hybrid
HKD 600,000 - 900,000
Generous paid time off policies
Hybrid working opportunities
Free breakfast, lunch and snacks daily
+4
Infrastructure Engineer - Eclipse Trading
Infrastructure Engineer - Eclipse Trading

Leadingnation • Hong Kong

On-site
HKD 500,000 - 700,000
Dynamic work environment
Collaborative team structure
Work-life balance
+2
[LPS] AI Platform Engineer (Technical)
[LPS] AI Platform Engineer (Technical)

LPS • Hong Kong

On-site
HKD 500,000 - 900,000
Data Center Technician
Data Center Technician

Google • Hong Kong

On-site
HKD 180,000 - 320,000
Site Reliability Engineer
Site Reliability Engineer

Hazeltree • Hong Kong

On-site
HKD 420,000 - 640,000
Health, dental, vision insurance
Retirement plan with company match
Professional development opportunities
+1
Cloud DevOps Engineer
Cloud DevOps Engineer

Quberesearchandtechnologies • Hong Kong

On-site
HKD 500,000 - 700,000
Infrastructure Linux Systems Analyst — HA, HPC & Ubuntu
Infrastructure Linux Systems Analyst — HA, HPC & Ubuntu

Swing Consulting Ltd. • Hong Kong

On-site
10-20 Days Annual Leave
5-day Work Week
Friendly and Energetic Working Environment
Core DevOps Engineer
Core DevOps Engineer

Selby Jennings • Hong Kong

On-site
HKD 900,000 - 1,300,000
Relocation support