Site Reliability Engineer - Global Hedge Fund - Hong Kong

Leadingnation

Hong Kong

On-site

HKD 900,000 - 1,700,000

Full time

14 days+
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Benefits offered by this job

Medical benefits
Brand name recognition
Internal mobility opportunities

Job summary

Leadingnation, a global hedge fund, is seeking a hands-on SRE to drive Site Reliability Engineering adoption and operational excellence. You will partner with development teams to ensure reliable deployment and operation of applications, set and monitor SLOs, and automate toil reduction.

The ideal candidate has strong Python automation skills, expertise in distributed systems and cloud, and excellent communication across infrastructure, security, and product teams.

Qualifications

  • Bachelor's or Master's degree in Computer Science, Engineering, or related field.
  • Extensive experience in Site Reliability Engineering with solid SRE principles, practices and tools.
  • Strong background in distributed systems, cloud, network architecture and software development.
  • Proficient in Python with automation and configuration management experience.
  • Solid understanding of monitoring, observability, incident management and post-incident analysis.
  • Excellent problem-solving and troubleshooting skills across complex systems.
  • Strong leadership, collaboration and communication across diverse teams.
  • Financial industry or hedge fund experience is highly preferred.

Responsibilities

  • Drive the adoption of SRE principles, methodologies, and best practices across the organization.
  • Collaborate with development teams to ensure reliable deployment and operation of applications.
  • Establish and monitor SLOs to ensure reliability and availability of critical systems.
  • Identify toil and automate improvements through code and process optimization.
  • Conduct root cause analyses for failures and incidents, implementing preventive solutions.
  • Lead incident management and post-incident reviews to improve processes.
  • Work with infrastructure, networking, and security teams to optimize performance and security.
  • Collaborate with stakeholders to define SLAs and operational requirements.
  • Stay current with industry trends in SRE to drive innovation and efficiency.

Skills

SRE principles
Distributed systems
Cloud computing
Python
Automation
Monitoring
Incident management
Leadership
Communication

Education

Bachelor's or Master's in CS/Engineering

Job description

My client, a global hedge fund, is actively seeking a hands on a highly skilled and motivated SRE to join their team. As an SRE, you will play a critical role in driving the adoption of Site Reliability Engineering practices within their organization. The ideal candidate will have a strong technical background and a passion for driving operational efficiency and continuous improvement.

The role:
  1. Drive the adoption of SRE principles, methodologies, and best practices across the organization.
  2. Collaborate closely with application development teams to ensure the successful deployment and operation of applications, including early-stage support during development.
  3. Establish and monitor key metrics, performance indicators, and service level objectives (SLOs) to ensure the reliability and availability of critical systems.
  4. Identify opportunities to eliminate toil through automation, code improvements, and process optimizations.
  5. Conduct root cause analyses for system failures and incidents, and implement engineering solutions to prevent future occurrences.
  6. Lead incident management and resolution efforts, ensuring timely and effective response to incidents, and driving post-incident reviews and process improvements.
  7. Work closely with cross-functional teams, including infrastructure, networking, and security, to optimize system performance, scalability, and security.
  8. Collaborate with stakeholders to define and refine service-level agreements (SLAs) and operational requirements.
  9. Stay abreast of industry trends and emerging technologies in Site Reliability Engineering, and leverage them to drive innovation and enhance operational efficiency.
What you offer:
  1. Bachelor's or Master's degree in Computer Science, Engineering, or a related field.
  2. Extensive experience in Site Reliability Engineering or a related field, with a strong understanding of SRE principles, practices, and tools.
  3. Strong technical background with expertise in areas such as distributed systems, cloud computing, network architecture, and software development.
  4. Strong experience in Python, also with automation and configuration management tools.
  5. Solid understanding of monitoring and observability frameworks, incident management, and post-incident analysis.
  6. Excellent problem-solving and troubleshooting skills, with the ability to analyze complex systems and identify areas for improvement.
  7. Strong leadership skills, with the ability to inspire and motivate a team, and foster a culture of collaboration, innovation, and continuous improvement.
  8. Excellent communication and interpersonal skills, with the ability to effectively communicate technical concepts to both technical and non-technical stakeholders.
  9. Experience in the financial industry or hedge fund environment highly preferred.
The sell:
  1. Extremely competitive compensation and medical benefits
  2. Brand name that will open doors for your career in the future
  3. Many opportunities for internal mobility and long term career growth
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Site Reliability Engineer - Global Financial Institution - Hong Kong
Site Reliability Engineer - Global Financial Institution - Hong Kong

NLS Executive Search • Hong Kong

On-site
HKD 600,000 - 900,000
Very competitive compensation (base +
Sr. Manager, Site Reliability & Innovation, IT
Sr. Manager, Site Reliability & Innovation, IT

CLSA • Hong Kong

On-site
HKD 900,000 - 1,200,000
Core Site Reliability Engineer
Core Site Reliability Engineer

Selby Jennings • Hong Kong

On-site
HKD 900,000 - 1,200,000
SRE Lead: Drive Reliability & Automation in Finance
SRE Lead: Drive Reliability & Automation in Finance

Leadingnation • Hong Kong

On-site
HKD 900,000 - 1,700,000
Medical benefits
Brand name recognition
Internal mobility opportunities
Senior Site Reliability Engineer (SRE)
Senior Site Reliability Engineer (SRE)

GFT TECHNOLOGIES SE • Hong Kong

On-site
HKD 700,000 - 1,200,000
Senior Full Stack Java/C# Engineer - Tier 1 Buy-side Firm - Hong Kong
Senior Full Stack Java/C# Engineer - Tier 1 Buy-side Firm - Hong Kong

Ashford Benjamin • Hong Kong

On-site
HKD 900,000 - 1,500,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Kody • Hong Kong

On-site
HKD 900,000 - 1,500,000
Competitive Package
A dynamic and innovative team
Collaborative, inclusive working env.
Senior Application Engineer
Senior Application Engineer

CW Talent Solutions • Hong Kong

On-site
HKD 900,000 - 1,300,000
Competitive base salary
Bonus potential
Medical, dental, vision coverage
+1
Site Reliability Engineer - HFT
Site Reliability Engineer - HFT

Selby Jennings • Hong Kong

On-site
HKD 480,000 - 720,000
Senior DevOps Engineer - Global Trading Firm - Hong Kong
Senior DevOps Engineer - Global Trading Firm - Hong Kong

NLS Executive Search • Hong Kong

On-site
HKD 600,000 - 1,000,000
Bonus up to 6 months