SDM, ML Data Infra, Amazon Traffic Engineering

Amazon

Vancouver

On-site

CAD 171,000 - 286,000

Full time

39 hours ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Amazon is seeking an experienced Software Development Manager to lead a team of SDEs and Data Engineers building Core Data Infrastructure for Bot Management. You will own the end-to-end data platform, from ingestion to serving, enabling ML and Science initiatives with reliable, scalable data access.

You will partner with Science leadership to translate model requirements into data investments and work with ML Platform to integrate data seamlessly with training and inference systems.

Qualifications

  • 3+ years of engineering team management experience.
  • 7+ years of working directly within engineering teams experience.
  • 3+ years of designing or architecting (design patterns, reliability and scaling) of new and existing systems experience.
  • 8+ years of leading the definition and development of multi tier web services experience.
  • Knowledge of engineering practices and patterns for the full software/hardware/networks development life cycle, including coding standards, code reviews, source control management, build processes, testing, certification, and livesite operations.
  • Experience partnering with product or program management teams.

Responsibilities

  • Data Infrastructure & Feature Engineering — Own the Feature Store, feature pipelines, and data serving layer.
  • Streaming & Real-Time Systems — Design and operate Apache Flink applications and live stream data processing.
  • Data Pipelines & Ingestion — Own batch and near real-time pipelines spanning Trails and Non-Trails sources.
  • Science & ML Platform Partnership — Serve as the primary data infrastructure partner to Applied Scientists and ML Platform.
  • People Leadership — Recruit, develop, and retain a high-performing team of SDEs and DEs.

Skills

Engineering management
Team leadership
System design
Web services

Tools

Kinesis
Kafka
Apache Flink
S3
OpenSearch
AWS Glue

Job description

Description

We are seeking an experienced Software Development Manager to lead a team of Software Development Engineers (SDEs) and Data Engineers (DEs) building the Core Data Infrastructure that underpins our ML and Science initiatives for Bot Management. You will own the end-to-end data platform—from source ingestion through transformation, feature engineering, and serving—ensuring Science and ML Platform teams have reliable, scalable, and timely access to the data they need for training, evaluation, and inference.

This is a high-impact leadership role. Our Science teams are building increasingly sophisticated models—and each requires different data formats, latencies, and serving patterns. Your team will be the backbone that makes this possible: ingesting billions of events from diverse source systems, building production-grade pipelines that transform raw signals into ML-ready feature groups, and operating the Feature Store that serves these features consistently across all model types. You will partner closely with Science leadership to translate model requirements into data infrastructure investments, and with ML Platform leadership to ensure seamless integration between your data layer and their training/inference systems. This role demands a leader who can navigate ambiguity across organizational boundaries, drive technical alignment between data engineering, software engineering and science teams, and build systems that scale with the rapid pace of model innovation.

Key job responsibilities
  • Data Infrastructure & Feature Engineering — Own the Feature Store, feature pipelines, and data serving layer. Build versioned feature groups across multiple storage backends (S3 for tabular, OpenSearch for embeddings) and production pipelines that transform disparate datasets into ML-ready features for Science teams.
  • Streaming & Real-Time Systems — Design and operate Apache Flink applications and live stream data processing for near real-time feature computation. Build event-driven architectures leveraging Kinesis and Kafka to support low-latency bot detection signals.
  • Data Pipelines & Ingestion — Own batch and near real-time pipelines spanning Trails (raw + aggregated) and Non-Trails sources (AIT, Clickstream, Customer Segmentations, OPS). Evolve pipelines from Cradle/POC to production-grade using AWS Glue. Implement data drift detection and governance frameworks.
  • Science & ML Platform Partnership — Serve as the primary data infrastructure partner to Applied Scientists and ML Platform. Define data contracts and SLAs, participate in model design reviews, and ensure the Feature Store integrates seamlessly with training and inference systems.
  • People Leadership — Recruit, develop, and retain a high-performing team of SDEs and DEs. Set goals, manage roadmaps, and foster a culture of operational excellence.
About The Team

Traffic Engineering's Bot Management organization protects Amazon's ecosystem by detecting and mitigating automated threats at scale. Our Core ML Data Infrastructure team is responsible for building and operating the foundational data infrastructure that powers bot detection, AI agent identification, and content exfiltration defense. We are building a unified, model-agnostic, production-grade ML platform that brings together training, evaluation, and inference pipelines into a cohesive system serving multiple model types across the organization.

Basic Qualifications
  • 3+ years of engineering team management experience
  • 7+ years of working directly within engineering teams experience
  • 3+ years of designing or architecting (design patterns, reliability and scaling) of new and existing systems experience
  • 8+ years of leading the definition and development of multi tier web services experience
  • Knowledge of engineering practices and patterns for the full software/hardware/networks development life cycle, including coding standards, code reviews, source control management, build processes, testing, certification, and livesite operations
  • Experience partnering with product or program management teams
Preferred Qualifications
  • Experience in communicating with users, other technical teams, and senior leadership to collect requirements, describe software product features, technical designs, and product strategy
  • Experience in recruiting, hiring, mentoring/coaching and managing teams of Software Engineers to improve their skills, and make them more effective, product software engineers

Amazon is an equal opportunity employer and does not discriminate on the basis of protected veteran status, disability, or other legally protected status.

Our inclusive culture empowers Amazonians to deliver the best results for our customers. If you have a disability and need a workplace accommodation or adjustment during the application and hiring process, including support for the interview or onboarding process, please visit https://amazon.jobs/content/en/how-we-hire/accommodations for more information. If the country/region you’re applying in isn’t listed, please contact your Recruiting Partner.

The base salary range for this position is listed below. As a total compensation company, Amazon's package may include other elements such as sign-on payments and restricted stock units (RSUs). Final compensation will be determined based on factors including experience, qualifications, and location. Amazon offers comprehensive benefits including health insurance (medical, dental, vision, prescription, basic life & AD&D insurance), Registered Retirement Savings Plan (RRSP), Deferred Profit Sharing Plan (DPSP), paid time off, and other resources to improve health and well-being. We thank all applicants for their interest, however only those interviewed will be advised as to hiring status.

CAN, BC, Vancouver - 171,400.00 - 286,200.00 CAD annually

Company - Amazon Development Centre Canada ULC

Job ID: A10496175

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

SDM, ML Data Infra, Amazon Traffic Engineering
SDM, ML Data Infra, Amazon Traffic Engineering

Socket.dev • Vancouver

On-site
CAD 171,000 - 286,000
Health benefits
RRSP
Paid time off
Machine Learning Engineer , Amazon Customer Service
Machine Learning Engineer , Amazon Customer Service

Amazon • Vancouver

On-site
CAD 115,000 - 192,000
Health insurance
RRSP
DPSP
+1
Sr. SDE, Edge AI ML Platform, Edge AI and Science
Sr. SDE, Edge AI ML Platform, Edge AI and Science

Amazon • Vancouver

On-site
CAD 140,000 - 190,000
Software Development Manager - AWS Identity & Access, AWS IAM Identity Center
Software Development Manager - AWS Identity & Access, AWS IAM Identity Center

Amazon Web Services (AWS) • Vancouver

On-site
CAD 171,000 - 287,000
Health insurance
RRSP
Paid time off
Software Development Engineer, Early Career - 2026
Software Development Engineer, Early Career - 2026

Amazon • Vancouver

On-site
CAD 90,000 - 150,000
Health insurance
RRSP
Paid time off
Sr. Applied Scientist, Workforce Solutions
Sr. Applied Scientist, Workforce Solutions

Socket.dev • Vancouver

On-site
CAD 196,000 - 327,000
Health insurance (medical, dental, and
RRSP
DPSP
+1
Data Engineer II, Alexa Audio
Data Engineer II, Alexa Audio

Amazon • Vancouver

On-site
CAD 103,000 - 173,000
Senior Security Engineer, AI Security
Senior Security Engineer, AI Security

Amazon • Vancouver

On-site
CAD 180,000 - 302,000
Software Dev Manager - AWS IAM, Control Plane - Propagation
Software Dev Manager - AWS IAM, Control Plane - Propagation

Amazon Web Services (AWS) • Vancouver

On-site
CAD 171,000 - 286,000
Applied Scientist Manager, Tax Engine
Applied Scientist Manager, Tax Engine

Amazon • Vancouver

On-site
CAD 222,800 - 372,000
Health insurance
RRSP
DPSP
+2