Senior Real-Time ML Engineer (AWS, Low-Latency Inference)

TWG Global AI

Santa Monica (CA)

On-site

USD 190,000 - 290,000

Full time

14 days+
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

TWG Global AI is seeking a Senior ML Engineer for AWS and Real-Time Inference in an onsite role located in Santa Monica, CA or New York, NY. You will own the fast path: ingesting live trading data and scoring it in near real time, focusing on streaming, low-latency inference, and timely model retraining.

The role requires strong data/ML engineering experience with streaming systems (Kafka/Kinesis/MSK), proficiency in Go and Python, and production AWS exposure.

Qualifications

  • Experience with streaming systems and modern data storage formats.
  • Proficiency in Go and Python.
  • Production AWS experience.

Responsibilities

  • Streaming and storage pipelines for model training and low-latency inference.
  • Connect online inference path to run-time model and manage latency under live load.
  • Model retraining cadence with accumulating data and drift triggering.
  • Productionizing new features and detectors on the fast path with data science.

Skills

Streaming systems
Low-latency inference
Go
Python
AWS
Financial market data familiarity
Event-driven architecture

Tools

Kafka
Kinesis
MSK

Job description

TWG Global AI is seeking a Senior ML Engineer for AWS and Real-Time Inference in an onsite role located in Santa Monica, CA or New York, NY. You will own the fast path: ingesting live trading data and scoring it in near real time, focusing on streaming, low-latency inference, and timely model retraining.

The role requires strong data/ML engineering experience with streaming systems (Kafka/Kinesis/MSK), proficiency in Go and Python, and production AWS exposure.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Real-Time ML Engineer – AWS, Low-Latency Inference
Real-Time ML Engineer – AWS, Low-Latency Inference

TWG AI • Santa Monica (CA)

On-site
USD 190,000 - 290,000
Senior ML Engineer: Real-Time Inference & Streaming (AWS)
Senior ML Engineer: Real-Time Inference & Streaming (AWS)

TWG Global AI • New York (NY)

On-site
USD 190,000 - 290,000
Senior ML Engineer - Real-Time Inference & Streaming (AWS)
Senior ML Engineer - Real-Time Inference & Streaming (AWS)

TWG AI • New York (NY)

On-site
USD 190,000 - 290,000
Bonus
Medical benefits
Senior Machine Learning Engineer - AWS, Real-Time Inference, Pipelines
Senior Machine Learning Engineer - AWS, Real-Time Inference, Pipelines

TWG Global AI • Santa Monica (CA)

On-site
USD 190,000 - 290,000
Senior Machine Learning Engineer - AWS, Real-Time Inference, Pipelines
Senior Machine Learning Engineer - AWS, Real-Time Inference, Pipelines

TWG AI • Santa Monica (CA)

On-site
USD 190,000 - 290,000
Senior Machine Learning Engineer - AWS, Real-Time Inference, Pipelines
Senior Machine Learning Engineer - AWS, Real-Time Inference, Pipelines

TWG Global AI • New York (NY)

On-site
USD 190,000 - 290,000
Senior Machine Learning Engineer - AWS, Real-Time Inference, Pipelines
Senior Machine Learning Engineer - AWS, Real-Time Inference, Pipelines

TWG AI • New York (NY)

On-site
USD 190,000 - 290,000
Bonus
Medical benefits
Real-Time ML Security Systems Engineer
Real-Time ML Security Systems Engineer

Amazon • Maryland

On-site
USD 144,000 - 194,000
Low-Latency ML Inference Engineer
Low-Latency ML Inference Engineer

Career Techniques • New York (NY)

Hybrid
USD 200,000 - 300,000
Senior ML Engineer: Real-Time Trading Models
Senior ML Engineer: Real-Time Trading Models

Rapidtrade • New York (NY)

On-site
USD 180,000 - 240,000