Remote Senior ML Systems Engineer - Model Serving

Atlassian

Northern (KY)

Hybrid

USD 149,000 - 235,000

Full time

3 hours ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Benefits offered by this job

Benefits and bonuses
Equity

Job summary

Atlassian is seeking a Senior ML System Engineer to design and optimize large-scale model serving systems across distributed infrastructure. You will own everything from global KV caches and auto-scaling to low-level kernel optimizations for high-concurrency production workloads.

The role requires deep expertise in C/C++ or Rust, experience with GPU inference engines, and a track record of optimizing latency and throughput in production.

Qualifications

  • 5+ years of software engineering experience, 2+ years of system performance optimization.
  • Deep low-level systems programming in C/C++ or Rust.
  • Experience with large-scale, high-concurrency production serving.
  • Experience with GPU inference engines (vLLM, SGLang, Triton, TensorRT-LLM).

Responsibilities

  • Architect and implement scalable distributed infrastructure for model serving (load balancing, auto-scaling, batch scheduling, global KV cache).
  • Optimize latency and throughput of model inference under real production workloads.
  • Build reliable, high-concurrency serving systems that serve billions of requests reliably
  • Benchmark, fine-tune, and accelerate inference engines.
  • Create robust CI/CD infrastructure for seamless model deployment and inference engine updates.
  • Partner with senior ML engineers to fine‑tune open-source LLMs and deploy

Skills

C/C++ or Rust
Distributed systems
GPU inference
Performance optimization
CI/CD for inference

Tools

vLLM
SGLang
Triton
TensorRT-LLM

Job description

Atlassian is seeking a Senior ML System Engineer to design and optimize large-scale model serving systems across distributed infrastructure. You will own everything from global KV caches and auto-scaling to low-level kernel optimizations for high-concurrency production workloads.

The role requires deep expertise in C/C++ or Rust, experience with GPU inference engines, and a track record of optimizing latency and throughput in production.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior ML Model Serving Engineer (Remote)
Senior ML Model Serving Engineer (Remote)

NEPSE Trading • Northern (KY)

Hybrid
USD 74,000 - 98,000
Senior AI Systems Engineer — Scalable Model Serving
Senior AI Systems Engineer — Scalable Model Serving

Apply • San Francisco (CA), Northern (KY)

Hybrid
USD 150,000 - 210,000
Senior ML Infra Engineer - Scalable Model Serving (Remote)
Senior ML Infra Engineer - Scalable Model Serving (Remote)

Etsy, Inc. • New York (NY)

Hybrid
USD 182,000 - 246,000
Equity package
Annual bonus
Comprehensive benefits
Lead ML Systems Engineer, GenAI Platform
Lead ML Systems Engineer, GenAI Platform

Atlassian • Northern (KY)

Hybrid
USD 166,000 - 221,000
Bonus
Equity
Benefits package
Senior Model Serving Engineer – Remote AI Infra
Senior Model Serving Engineer – Remote AI Infra

United States Digital Space LLC • United States

Remote
USD 74,000 - 98,000
GenAI ML Systems Engineer: Scalable Training & Inference
GenAI ML Systems Engineer: Scalable Training & Inference

Meta • Menlo Park (CA)

On-site
USD 180,000 - 300,000
Senior ML Systems Engineer: Scalable Search Platform Remote
Senior ML Systems Engineer: Scalable Search Platform Remote

Atlassian • Austin (TX)

On-site
USD 160,000 - 220,000
Health & wellbeing resources
Volunteer days
Community involvement
Senior ML Infrastructure Engineer - Large-Scale GPU Serving
Senior ML Infrastructure Engineer - Large-Scale GPU Serving

Selby Jennings • New York (NY)

On-site
USD 180,000 - 250,000
Remote Senior ML Systems Architect – Search at Scale
Remote Senior ML Systems Architect – Search at Scale

Atlassian • Austin (TX)

Hybrid
USD 273,000 - 356,000
Senior Inference & RL Systems Engineer (Scalable ML Infra)
Senior Inference & RL Systems Engineer (Scalable ML Infra)

Magic AI, Inc • San Francisco (CA)

On-site
USD 300,000 - 550,000
Equity compensation
401(k) matching
Health, dental and vision insurance
+4