AI/ML Ops Engineer

Blackpoint Cyber

Canada

Hybrid

CAD 110,000 - 145,000

Full time

12 days ago

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Blackpoint Cyber in Canada seeks an AI/ML Ops Engineer to own the operational backbone of its AI/ML capability, moving models from training to production, deploying live endpoints, and ensuring scalable, reliable pipelines.

You will automate training, monitoring, and deployment across the AI lifecycle, collaborate with SOC and APG, and implement CI/CD with Terraform, Docker, Kubernetes, and AWS to deliver production‑grade infrastructure.

Qualifications

  • Proven 5+ years of hands-on ML engineering in production.
  • Strong experience deploying models end-to-end in production.
  • Excellent problem-solving and data-driven decision making.
  • Ability to collaborate with cross-functional teams including SOC and APG.

Responsibilities

  • Own ML loop end-to-end from training to deployment and retirement.
  • Build and optimize deployment pipelines and live endpoints.
  • Automate training, monitoring, and deployment pipelines.
  • Ensure production-grade, highly available systems with monitoring.

Skills

Cloud ML Infrastructure
Model Deployment
ML Pipelines
Terraform
GitHub Actions
Docker
Kubernetes
SageMaker
Bedrock
Kafka
Spark
MLFlow
Python
Bash
SQL
SparkSQL
CI/CD
ML Governance

Job description

Blackpoint Cyber is the leading provider of world‑class cybersecurity threat hunting, detection and remediation technology. Founded by former National Security Agency (NSA) cyber operations experts who applied their learningsto bring national security-grade technology solutions to commercial customers around the world, Blackpoint Cyber is in hyper‑growth mode, fueled by a recent 190m series C round.

ROLE SUMMARY

As AI/ML Ops Engineer, you will own the operational backbone of Blackpoint's AI/ML capability — taking models from training through production deployment and owning the full AI/ML loop across every pipeline the team runs at scale. You will be instrumental in building out this new function: standing up deployment pipelines, taking trained scripts and putting them on endpoints that can accept live requests, and ensuring every solution you ship is production‑grade, highly available, and well monitored. As the team's scope expands, you will take on ownership of testing across the board and automate training, monitoring, and deployment pipelines across the AI lifecycle. You will report to the Vice President of AI and Data and work closely with Engineering, the Security Operations Center (SOC), and the Adversary Pursuit Group (APG).

WHO YOU ARE
  • 5+ years of hands‑on ML Engineering experience, including having personally trained and deployed models into a production environment — you know what it takes to take an AI product from prototype to live service.
  • A well‑architected mindset — built for efficiency, performance, security, and reliability, with genuine comfort owning deployment pipelines end‑to‑end.
  • Strong analytical and problem‑solving abilities, with a focus on data‑driven decision‑making.
  • Excellent communication and interpersonal skills, with the ability to influence and collaborate with stakeholders at all levels.
WHAT YOU'LL BRING
Experienced in:
  • Cloud-based ML Infrastructure (AWS)
  • Model Development, Evaluation, & Deployment Operations (SageMaker, Bedrock)
  • Inference Streams & Event‑Driven Processing (Kafka, Spark)
  • MLOps Workflows (MLFlow, Sagemaker Pipelines)
  • Infrastructure as Code & Pipeline Automation (Terraform, AI CI/CD, GitHub Actions)
  • ML Governance (Data, Model, & Feature Versioning, Monitoring & Testing)
  • Containerized Services (Docker, Kubernetes, ECS/EKS)
  • Scripting Languages (Python, Bash)
  • Query Languages (SQL, SparkSQL)
  • GitFlow, CI/CD workflows & DevOps best practices
  • Experience with AI‑Assisted development life cycle
  • Building high‑availability, production‑grade systems with strong visibility and alerting baked in from day one
Nice to Have:
  • Transformer Neural Networks
  • Agile Scrum/Kanban
  • Anthropic, OpenAI, LiteLLM APIs and SDKs
  • Experience in Cybersecurity, IoT, or NLP fields
  • Grafana or CloudWatch (observability tooling)
HOW YOU'LL MAKE AN IMPACT
  • Own the AI/ML loop end‑to‑end at scale — across all pipelines, from model training through deployment, monitoring, and retirement.
  • Develop, optimize, and deploy ML models, drawing on direct, hands‑on experience training models yourself.
  • Design, build, and administer model‑building and serving infrastructure, taking trained scripts and standing them up as live endpoints that can accept real‑time requests.
  • Implement ML workflows as containerized Infrastructure as Code, using Terraform, GitHub Actions, Docker, and Kubernetes.
  • Build and automate standardized container pipelines for training, feature engineering, and inference channels, with CI/CD managed through GitHub.
  • Own test strategy across the full ML pipeline — model validation, integration, load, and deployment testing — as the team's testing scope continues to expand.
  • Build visibility and alerting into every deployed pipeline and hold all AI/ML solutions to a production‑grade, highly‑available, well‑monitored bar.
  • Develop ML governance utilities for oversight and administration of deployed infrastructure.
  • Implement data, feature, and model lifecycle best practices.
  • Contribute to AI architecture and design decisions, taking primary ownership of ML pipeline work.
  • Collaborate closely with cross‑functional teams, including Engineering, Blackpoint Cyber’s Security Operations Center (SOC), and the Adversary Pursuit Group (APG).

Blackpoint Cyber welcomes and encourages applications from qualified individuals of all races, colors, religions, sex, sexual orientation, gender identity or expression, national origin, age, marital status, or any other legally protected status. We are committed to equality of opportunity in all aspects of employment.

For eligible employees in the US, Blackpoint offers competitive Health, Vision, Dental, and Life Insurance plans, a robust 401k plan, Discretionary Time Off, and other minor perks. International employees receive competitive benefits in accordance with local market standards and applicable country requirements.

Blackpoint believes all employees should share in the company’s success – equity participation is available to employees globally, with program details varying by location and employment structure.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Sr. Software Engineer (AI/ML)
Sr. Software Engineer (AI/ML)

Blackpoint-Cyber • Canada

On-site
CAD 120,000 - 180,000
Software Engineer Backend
Software Engineer Backend

Blackpoint Cyber • Canada

On-site
CAD 70,000 - 110,000
Health Insurance
Vision Insurance
Dental Insurance
+3
AI / ML Engineer
AI / ML Engineer

Testlify, Inc. • Canada

On-site
CAD 100,000 - 170,000
Health Insurance
Flexible Working Style
Senior Data Engineer I - QuantumBlack, AI by McKinsey
Senior Data Engineer I - QuantumBlack, AI by McKinsey

QuantumBlack, AI by McKinsey • Toronto

On-site
CAD 100,000 - 130,000
Comprehensive benefits package
Continuous learning opportunities
Global community engagement
Senior Product Manager
Senior Product Manager

Blackpoint Cyber • Canada

Hybrid
CAD 120,000 - 180,000
Health insurance
Vision & Dental
Life Insurance
+3
AI Engineer
AI Engineer

Valsoft Corporation • Canada

On-site
CAD 100,000 - 140,000
ML Engineer New Vancouver
ML Engineer New Vancouver

Tigera, Inc. • Vancouver

On-site
CAD 160,000 - 180,000
Health benefits
Vision benefits
Dental benefits
+1
Adversarial Machine Learning Engineer - Red Teaming
Adversarial Machine Learning Engineer - Red Teaming

C-Serv • Canada

Remote
CAD 120,000 - 190,000
Fully remote Canada
Staff-level growth opportunities
Full-cycle hiring support
+1
Principal Data Engineer, LLM/AI Platforms (Remote)
Principal Data Engineer, LLM/AI Platforms (Remote)

CrowdStrike • Winnipeg

On-site
CAD 210,000 - 320,000
Market-leading compensation
Wellness programs
Generous vacation & holidays
+2
Adversarial Machine Learning Engineer - Red Teaming
Adversarial Machine Learning Engineer - Red Teaming

C-Serv Global Ltd • Canada

Hybrid
CAD 120,000 - 180,000
Fully remote Canada
Growth opportunities
Competitive compensation