AI Operations Engineer

Placements24

Upington

Hybrid

ZAR 800,000 - 1,200,000

Full time

7 days ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Competitive salary
Flexible remote work options
Professional development

Job summary

Placements24 is seeking an AI Operations Engineer in Upington to manage and optimize the operational aspects of AI infrastructure, ensuring reliability, scalability, and performance in production environments.

You will deploy, monitor, and maintain AI models, automate deployments and maintenance, and collaborate with engineering and data science teams to improve AI system efficiency.

Qualifications

  • Bachelor's degree in Computer Science, Engineering, or a related technical field.
  • 3+ years of experience in DevOps, Site Reliability Engineering (SRE), or a related operations role.
  • Experience with cloud platforms (AWS, Azure, GCP) and containerization technologies (Docker, Kubernetes).
  • Familiarity with MLOps principles and tools for managing ML lifecycles.
  • Proficiency in scripting languages (e.g., Python, Bash) and experience with CI/CD pipelines.

Responsibilities

  • Deploy, monitor, and maintain AI/ML models and associated infrastructure in production environments.
  • Automate operational tasks related to AI systems, including deployment, scaling, and maintenance.
  • Develop and implement monitoring solutions to track the performance and health of AI services.
  • Troubleshoot and resolve operational issues, ensuring minimal downtime and service disruption.
  • Collaborate with engineering and data science teams to optimize AI system performance and efficiency.

Skills

DevOps
SRE
CI/CD
Python scripting
Cloud platforms

Education

Bachelor's degree in CS/Engineering

Tools

Docker
Kubernetes
AWS
Azure
GCP

Job description

About the Role

Our client is seeking a skilled and detail-oriented AI Operations Engineer to manage and optimize the operational aspects of their AI & Emerging Technologies infrastructure in Upington . This role is critical for ensuring the reliability, scalability, and performance of AI systems in production. You will be responsible for deploying, monitoring, and maintaining AI models and platforms, working closely with engineering teams to automate processes and resolve operational issues. This is an exciting opportunity to contribute to the operational excellence of cutting-edge AI solutions in a supportive and innovative environment.

Key Responsibilities
  • Deploy, monitor, and maintain AI/ML models and associated infrastructure in production environments.
  • Automate operational tasks related to AI systems, including deployment, scaling, and maintenance.
  • Develop and implement monitoring solutions to track the performance and health of AI services.
  • Troubleshoot and resolve operational issues, ensuring minimal downtime and service disruption.
  • Collaborate with engineering and data science teams to optimize AI system performance and efficiency.
Requirements
  • Bachelor's degree in Computer Science, Engineering, or a related technical field.
  • 3+ years of experience in DevOps, Site Reliability Engineering (SRE), or a related operations role.
  • Experience with cloud platforms (AWS, Azure, GCP) and containerization technologies (Docker, Kubernetes).
  • Familiarity with MLOps principles and tools for managing ML lifecycles.
  • Proficiency in scripting languages (e.g., Python, Bash) and experience with CI/CD pipelines.
Benefits
  • Competitive salary and comprehensive benefits package.
  • Opportunity to work on advanced AI technologies in Upington .
  • Professional development and continuous learning support.
  • Collaborative and forward-thinking work culture.
  • Flexible remote work options.
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

AI Engineer - MLOps Specialist
AI Engineer - MLOps Specialist

Placements24 • LegKraal Gate

Hybrid
ZAR 900,000 - 1,300,000
Competitive salary
Professional development opportunities
Hybrid work arrangements
+2
AI Automation Engineer
AI Automation Engineer

Placements24 • Upington

Hybrid
ZAR 600,000 - 900,000
Competitive salary
Professional development opportunities
Dynamic work environment
+1
Senior AI Engineer - MLOps
Senior AI Engineer - MLOps

Placements24 • LegKraal Gate

Hybrid
ZAR 1,000,000 - 1,600,000
Salary and bonus structure
Health, dental, and vision insurance
Paid time off
+3
AI Infrastructure Engineer
AI Infrastructure Engineer

Placements24 • Stellenbosch

On-site
ZAR 900,000 - 1,300,000
Competitive salary
Hybrid work arrangement
Professional development opportunities
+3
Senior AI Platform Engineer
Senior AI Platform Engineer

Placements24 • Sandton

On-site
ZAR 1,000,000 - 1,600,000
Fully remote work arrangement
Flexible working hours
Career growth in AI tech
+1
AI/ML Infrastructure Engineer
AI/ML Infrastructure Engineer

Placements24 • Benoni

Hybrid
ZAR 900,000 - 1,300,000
Competitive salary + bonuses
Hybrid work model
Health, dental, vision insurance
+2
AI/ML Platform Engineer
AI/ML Platform Engineer

Placements24 • Randburg

Hybrid
ZAR 900,000 - 1,500,000
Competitive compensation
Health and wellness benefits
Professional development in AI/ML
+2
Senior AI Engineer - Cloud Platforms
Senior AI Engineer - Cloud Platforms

Placements24 • Randburg

Hybrid
ZAR 1,200,000 - 1,600,000
Health insurance
Professional development opportunities
Retirement plan
+1
Machine Learning Operations (MLOps) Engineer
Machine Learning Operations (MLOps) Engineer

Placements24 • Vereeniging

Hybrid
ZAR 720,000 - 900,000
AI Software Engineer
AI Software Engineer

Placements24 • Pretoria

On-site
ZAR 600,000 - 900,000
Competitive salary
Hybrid work model
Professional development