Senior Machine Learning Engineer - AWS, Real-Time Inference, Pipelines

TWG Global AI

New York (NY)

On-site

USD 190,000 - 290,000

Full time

14 days+
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

TWG Global AI is seeking a Senior ML Engineer for AWS and Real-Time Inference to own the fast path: ingest live trading data and score it in near real time. This is a systems-heavy role focused on streaming, low-latency inference, and the retraining cadence that keeps models current with live markets.

You will develop streaming pipelines feeding model training and inference, connect the real-time detector service to a run-time model while managing latency under load, and coordinate feature

Qualifications

  • Experience with streaming systems (Kafka, Kinesis, MSK) and modern data formats.
  • Experience building low-latency, high-throughput inference services.
  • Proficiency in Go and Python; production AWS experience.

Responsibilities

  • Streaming pipelines and storage feeding model training and low-latency inference.
  • Own online inference path; connect real-time detector to run-time model while maintaining latency under load.
  • Manage model retraining cadence as data and labels accumulate, including drift-triggered retraining.
  • Productionize new features/detectors on the fast path in partnership with data science.

Skills

Streaming systems
Low-latency inference
Go programming
Python programming
Production AWS experience

Tools

Kafka
Kinesis
MSK (Managed Streaming for Kafka)
FIX protocol

Job description

The Organization

At TWG AI, we drive innovation and business transformation across a range of industries—including financial services, insurance, technology, media, and sports—by leveraging data and AI as core assets. Our AI-first, cloud-native approach delivers real-time intelligence and interactive business applications, empowering informed decision-making for both customers and employees.

We prioritize responsible data and AI practices, ensuring ethical standards and regulatory compliance. Our decentralized structure enables each business unit to operate autonomously, supported by a central AI Solutions Group, while strategic partnerships with leading data and AI vendors fuel game-changing efforts in marketing, operations, and product development.

You will collaborate with management to advance our data and analytics transformation, enhance productivity, and enable agile, data-driven decisions. By leveraging relationships with top tech startups and universities, you will help create competitive advantages and drive enterprise innovation.

At TWG, your contributions will support our goal of sustained growth and superior returns, as we deliver rare value and impact across our businesses.

The Role

As a Senior ML Engineer for AWS and Real-Time Inference, you'll own the fast path: ingesting live trading data and scoring it in near real time. It's a systems-heavy role focused on streaming, low-latency inference, and the retraining cadence that keeps models current and most directly determines whether the systems keep up with live markets.

Key Responsibilities:
  • Streaming and storage pipelines that feed both model training and low-latency inference.
  • The online inference path and its latency. A real-time detector microservice is built and unit-tested but not yet deployed — it needs to be connected to a run-time model and hold latency under live load. One known, non-trivial problem lives here: batch scoring ranks across a whole population, but single-account (or single-wallet) real-time scoring has no population to rank against, so it must threshold on calibrated raw scores.
  • The model retraining cadence as data and labels accumulate, including drift-triggered retraining.
  • Productionizing new features and detectors on the fast path, in partnership with data science.
Qualifications:
  • Strong data / ML engineering experience with streaming systems (e.g., Kafka / Kinesis / MSK) and modern data storage formats
  • Experience building low-latency, high-throughput inference services
  • Proficiency in a systems language (e.g., Go) alongside Python
  • Production AWS experience
  • Familiarity with financial market data or trading protocols a plus — the US feed is a FIX 5.0 SP2 drop-copy session with real-world quirks (nanosecond timestamps, repeating groups, dedup semantics)
  • Familiarity with chain-data infrastructure (node providers, subgraphs, event indexing) is a plus
Position Location:

This is an onsite position based out of our Santa Monica, CA or New York, NY offices.

Compensation:

The base pay for this position is $190,000-290,000. A bonus will be provided as part of the compensation package, in addition to a full range of medical, financial, and/or other benefits.

TWG is an equal opportunity employer, and all qualified applicants will receive consideration for employment without regard to race, color, religion, gender, sexual orientation, gender identity, national origin, disability, or status as a protected veteran.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Machine Learning Engineer - AWS, Real-Time Inference, Pipelines
Senior Machine Learning Engineer - AWS, Real-Time Inference, Pipelines

TWG Global AI • Santa Monica (CA)

On-site
USD 190,000 - 290,000
Senior Machine Learning Engineer - AWS, Real-Time Inference, Pipelines
Senior Machine Learning Engineer - AWS, Real-Time Inference, Pipelines

TWG AI • Santa Monica (CA)

On-site
USD 190,000 - 290,000
Senior Machine Learning Engineer - AWS, Real-Time Inference, Pipelines
Senior Machine Learning Engineer - AWS, Real-Time Inference, Pipelines

TWG AI • New York (NY)

On-site
USD 190,000 - 290,000
Bonus
Medical benefits
Real-Time ML Engineer – AWS, Low-Latency Inference
Real-Time ML Engineer – AWS, Low-Latency Inference

TWG AI • Santa Monica (CA)

On-site
USD 190,000 - 290,000
Senior Real-Time ML Engineer (AWS, Low-Latency Inference)
Senior Real-Time ML Engineer (AWS, Low-Latency Inference)

TWG Global AI • Santa Monica (CA)

On-site
USD 190,000 - 290,000
Senior ML / AWS Engineer - Pipeline Hardening, Governance & Integrations
Senior ML / AWS Engineer - Pipeline Hardening, Governance & Integrations

TWG AI • Santa Monica (CA)

On-site
USD 190,000 - 290,000
Senior ML Engineer - Real-Time Inference & Streaming (AWS)
Senior ML Engineer - Real-Time Inference & Streaming (AWS)

TWG AI • New York (NY)

On-site
USD 190,000 - 290,000
Bonus
Medical benefits
Senior ML Engineer: Real-Time Inference & Streaming (AWS)
Senior ML Engineer: Real-Time Inference & Streaming (AWS)

TWG Global AI • New York (NY)

On-site
USD 190,000 - 290,000
Senior ML / AWS Engineer - Pipeline Hardening, Governance & Integrations
Senior ML / AWS Engineer - Pipeline Hardening, Governance & Integrations

TWG Global AI • Santa Monica (CA)

On-site
USD 190,000 - 290,000
Senior Associate, AI Data Scientist
Senior Associate, AI Data Scientist

TWG Global AI • Santa Monica (CA)

On-site
USD 175,000 - 195,000