Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.
Long Ridge Partners is seeking a Machine Learning Performance Engineer (Inference) in New York to architect low-latency ML inference pipelines for high-frequency trading. You will benchmark CPU, GPU, and FPGA workloads, optimize kernels, and push models into real-time production with strict latency requirements.
You will collaborate with ML researchers, HPC, FPGA, and Datacenter teams to ensure efficient, scalable deployments across hardware platforms.
Long Ridge Partners is seeking a Machine Learning Performance Engineer (Inference) in New York to architect low-latency ML inference pipelines for high-frequency trading. You will benchmark CPU, GPU, and FPGA workloads, optimize kernels, and push models into real-time production with strict latency requirements.
You will collaborate with ML researchers, HPC, FPGA, and Datacenter teams to ensure efficient, scalable deployments across hardware platforms.