Cloud Computing Engineer

PDT Partners, LLC

New York (NY)

Hybrid

USD 195,000 - 225,000

Full time

3 days ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

PDT Partners, LLC is seeking an ambitious software engineer to join the Cloud Computing team in New York City. You will help mature and scale our cloud HPC platform, collaborating with Research and Model Implementation to deliver real tradable products.

This hybrid role requires 3 days per week in the NYC office and offers the chance to own design, automation, and operation of high-performance compute infrastructure at scale.

Qualifications

  • Bachelor's or Master's degree in Engineering or Applied Sciences.
  • 5+ years of experience building and shipping software systems.
  • 2+ years running production compute platforms and scheduling systems.
  • Fluency in Python, Go, Rust or similar languages with strong system design, testing and debugging.
  • Experience with Terraform for IaC.
  • Excellent written and verbal communication skills.
  • Experience working with researchers or other developers is a plus.
  • Hands-on experience managing inferencing and training hardware at scale.
  • Knowledge of NVIDIA GPU management, Kubernetes, Slurm, and the wider compute ecosystem.

Responsibilities

  • Design and implement cloud-based HPC systems spanning scheduling, metrics, containerization, software distribution, and cloud architecture.
  • Run the HPC plant day-to-day with 24/7 availability and ensure platform reliability.
  • Implement automation from CI/CD pipelines to production metrics and monitoring of the cloud HPC platform.
  • Collaborate with researchers and engineers to deliver high-quality cloud HPC systems that scale.
  • Handle capacity management and benchmark optimization for research and trading workloads.

Skills

Python
Go
Rust
System design
Testing
Debugging
Containerization
CI/CD
Cloud infrastructure
Performance tuning

Education

Bachelor's or Master in Engineering/Applied Sciences

Tools

Terraform
Kubernetes
Slurm
NVIDIA GPU management

Job description

The Cloud Computing team is a group of experts solving computing problems in the critical path of Research. We work directly with Research and Model Implementation teams and provide them with platform tools and computing resources to take their ideas from inception to real tradable products. We are looking for an ambitious and operationally minded software engineer to join our team as we mature and scale our cloud HPC platform to the next iteration of our firm-wide Research platform.

This is a hybrid position and will require the person to work from our New York City office at minimum 3 days a week.

PDT Partners has a 30+ year track record and a reputation for excellence. Our goal is to be the best quantitative investment manager in the world-measured by the quality of our products, not their size. PDT’s very high employee-retention rate speaks for itself. Our people are intellectually extraordinary and our community is close-knit, down-to-earth, and diverse.

Responsibilities:

We are a small flat team sitting at the cross-section of research, implementation, and platform infrastructure. Our team responsibilities span many areas. Including:

  • Design and implementation of cloud-based HPC systems. Our projects involve equal parts engineering and operations for success in our fast-moving environment. You will be expected to conceive and implement projects small and large in the intersection of HPC scheduling, metrics, containerization, software distribution, accelerated compute performance/efficiency, cloud architecture.
  • Running our HPC plant day-to-day. Our research environment is up 24/7, and we want to keep it that way. Everybody on the team contributes to the support of our platform, which thankfully is light because of our automation and quality work.
  • Implementing automation. We will always choose to work smart over working hard. You will be responsible for conception and implementation of automation from CI/CD pipelines to production metrics and monitoring of our cloud HPC platform.
  • Obsessive User Focus. All members of the team are expected to partner with researchers and engineers to deliver high-quality cloud HPC systems that are efficient and reliable. This includes leading projects to evolve it as our needs change.
  • Capacity management and benchmark optimization. Our demand for scale and performance is constant and involves challenging optimization problems for workloads critical to research and trading

Below is a list of skills and experiences we think are relevant. Even if you don’t think you’re a perfect match, we still encourage you to apply because we are committed to developing our people.

  • Bachelors orMastersdegree in an Engineering or Applied Sciences field from a rigorous academic programor equivalent professional experience.
  • 5+ years of experience building and shipping software systems
  • 2+ years running production compute platforms and scheduling systems.
  • Mastery of core software engineering concepts such as system design, testing, debugging, building reliable production systems with fluency in a language such as Python, Go, Rust etc.
  • Experience with a cloud-based infrastructure-as-code tool such as Terraform
  • Excellent written and verbal communication skills
  • Experience working with or supporting researchers and/or other developers is a plus
  • Hands-on experience managing inferencing and training hardware at scale
  • Knowledge of NVIDIA GPU management, Kubernetes, Slurm, and the wider large-scale compute ecosystem

The salary range for this role is between $195,000 and $225,000. This range is not inclusive of any potential bonus amounts . Factors that may impact the agreed upon salary within the range for a particular candidate include years of experience, level of education obtained, skill set, and other external factors.

PRIVACY STATEMENT: For information on ways PDT may collect, use, and process your personal information, please see PDT’s privacy notices .

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Cloud HPC Engineer — Research Platform (Hybrid NYC)
Senior Cloud HPC Engineer — Research Platform (Hybrid NYC)

PDT Partners, LLC • New York (NY)

Hybrid
USD 195,000 - 225,000
Platform Support Engineer
Platform Support Engineer

Quant Blueprint LLC • New York (NY)

On-site
USD 165,000 - 180,000
Social events and team-building activities
Mentorship from senior engineers
Inclusive and supportive work environment
Performance Engineer
Performance Engineer

PDT Partners • New York (NY)

On-site
USD 90,000 - 130,000
Research Engineer
Research Engineer

PDT Partners • New York (NY)

On-site
USD 190,000 - 250,000
Software Engineer
Software Engineer

PDT Partners • New York (NY)

On-site
USD 160,000 - 200,000
Junior Cloud Engineer
Junior Cloud Engineer

Jobright.ai • New York (NY)

On-site
USD 120,000 - 155,000
Medical insurance
Vision insurance
401(k)
HPC Engineer - Privacy Research Program
HPC Engineer - Privacy Research Program

Tufts University • Massachusetts

On-site
USD 89,000 - 134,000
Quant Hedge Fund - Systems Engineer - GPU expert
Quant Hedge Fund - Systems Engineer - GPU expert

Saragossa • New York (NY)

On-site
USD 300,000 - 1,100,000
HPC & Compute Engineering Lead
HPC & Compute Engineering Lead

Autonomai Recruitment • Chicago (IL)

On-site
USD 180,000 - 250,000
Senior System Software Engineer - Scientific Computing PaaS
Senior System Software Engineer - Scientific Computing PaaS

NVIDIA Gruppe • Santa Clara (CA)

On-site
USD 184,000 - 357,000