AI-Driven DevOps & Infra Model Evaluator

Mercor

Brussel Hoofdstad

Sur place

EUR 179 000 - 239 000

Plein temps

14 jours+
Générateur de candidature

Obtenez une réponse de cet employeur — un CV et une lettre de motivation adaptés exactement à ce qu’il recherche.

Passez les filtres ATS

Résumé du poste

Mercor is partnering with a leading AI research lab to support a Frontier Code Agents project. Contributors help evaluate and improve frontier AI coding models through structured technical assessments.

You will review model-generated implementations across cloud platforms, Kubernetes, CI/CD, observability, and infrastructure automation, applying professional engineering judgment to realistic scenarios.

Qualifications

  • 2+ years of professional DevOps, SRE or Cloud Engineering experience.
  • Experience with AWS, Azure, GCP, Kubernetes, Terraform, CI/CD pipelines or observability tooling.
  • Regular use of AI coding agents such as Cursor, Claude Code, Codex, Windsurf, Gemini CLI or similar tools.
  • Ability to evaluate model-generated infrastructure and reliability engineering solutions.
  • Experience supporting production-scale systems is preferred.

Responsabilités

  • Use frontier AI coding agents to complete and evaluate complex infrastructure engineering tasks.
  • Review model-generated implementations involving cloud platforms, Kubernetes, CI/CD systems, observability, and infrastructure automation.
  • Identify bugs, edge cases, reliability issues, and failure modes.
  • Compare outputs from multiple frontier models and assess their strengths and weaknesses.
  • Apply professional engineering judgment to realistic infrastructure engineering scenarios.

Connaissances

DevOps / SRE experience
Cloud engineering
2+ years

Outils

AWS
Azure
GCP
Kubernetes
Terraform
CI/CD pipelines
Observability tooling

Description du poste

Mercor is partnering with a leading AI research lab to support a Frontier Code Agents project. Contributors help evaluate and improve frontier AI coding models through structured technical assessments.

You will review model-generated implementations across cloud platforms, Kubernetes, CI/CD, observability, and infrastructure automation, applying professional engineering judgment to realistic scenarios.

Obtenez votre examen gratuit et confidentiel de votre CV.
ou faites glisser et déposez votre fichier ici.
Similar jobs

Postes similaires à comparer

DevOps Engineer - AI Model Evaluator
DevOps Engineer - AI Model Evaluator

Mercor • Brussel Hoofdstad

Sur place
EUR 179 000 - 239 000
Frontier AI Safety Evaluator: Expert Reviewer
Frontier AI Safety Evaluator: Expert Reviewer

Mercor • Brussel Hoofdstad

Sur place
EUR 70 000 - 110 000
Frontier AI Safety Evaluator
Frontier AI Safety Evaluator

Mercor • Brussel Hoofdstad

À distance
EUR 70 000 - 110 000
Frontier AI Safety Evaluator & Alignment Specialist
Frontier AI Safety Evaluator & Alignment Specialist

Mercor • Brussel Hoofdstad

Sur place
EUR 70 000 - 110 000
Frontier AI Safety Red Team Lead
Frontier AI Safety Red Team Lead

Mercor • Brussel Hoofdstad

Sur place
EUR 90 000 - 130 000
AI Content Evaluator & Quality Reviewer (UK/EU)
AI Content Evaluator & Quality Reviewer (UK/EU)

Mercor • Brussel Hoofdstad

Sur place
EUR 42 000 - 64 000
AI Safety Specialist - Fully Remote
AI Safety Specialist - Fully Remote

Mercor • Belgique

À distance
USD 90 000 - 150 000
Frontier AI Safety Evaluator
Frontier AI Safety Evaluator

Obsidian • Brussel Hoofdstad

À distance
EUR 70 000 - 110 000
Remote AI Safety Red Team Specialist
Remote AI Safety Red Team Specialist

Mercor • Brussel Hoofdstad

Sur place
EUR 78 000 - 104 000
Frontier AI Safety Evaluator & Alignment Specialist
Frontier AI Safety Evaluator & Alignment Specialist

Obsidian • Brussel Hoofdstad

Sur place
EUR 70 000 - 110 000