Remote AI Prompt Architect & Evaluation Engineer

Yonder Media Mobile Inc.

Poland

Remote

USD 120,000 - 180,000

Full time

14 days+
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Benefits offered by this job

Flexible working hours
Corporate equipment for work
Competitive salary
Real opportunity for personal and.prof

Job summary

House of YO is seeking a prompt and evaluation specialist for YOlanda, the AI concierge. You will own the system prompts, ensure smooth handoffs to the compute cooperative, and build evaluation sets with measurable scores.

You will also develop data and training pipelines, test robustness against adversarial prompts, and maintain versioned prompts in git. The role is remote, global, full-time, requiring strong English writing and bilingual capabilities, plus Python and LangChain/ LangGraph

Qualifications

  • 2+ years building with large language models in production.
  • Proficient Python for shipping scripts, API clients, data handling, tests.
  • Experience with LangChain and LangGraph; ability to read and modify a graph you didn’t write.
  • Experience with retrieval in production: chunking, embeddings, reranking, grounding.
  • Exceptional written English and bilingual Spanish and English.
  • Methodology: iteratively test one variable, record results, and move forward.

Responsibilities

  • Own YOlanda's system prompt — persona, tone, safety rules, refusals — and prompts behind every user action.
  • Decide what stays in YOlanda's frame of reference and what is handed off to YOnC, ensuring a seamless handoff.
  • Build a few-shot library and manage context so the model doesn’t start from a blank prompt or repeat answered questions.
  • Create an evaluation set with tasks, pass criteria and release-traceable scores.
  • Rank and critique outputs (RLHF), rewrite weak answers into fine-tuning data, and fact-check plan prices, balances and YOYO$ maths.
  • Run tests against adversarial prompts, jailbreak attempts, and abuse of top-up/rewards flows.
  • Write Python to run and score evaluations at volume, work with the API (streaming, function calling, structured output), and version prompts in git.

Skills

LLMs in production
Python
LangChain
LangGraph
Retrieval in production
English writing
English/Spanish bilingual
Git versioning

Tools

Git

Job description

House of YO is seeking a prompt and evaluation specialist for YOlanda, the AI concierge. You will own the system prompts, ensure smooth handoffs to the compute cooperative, and build evaluation sets with measurable scores.

You will also develop data and training pipelines, test robustness against adversarial prompts, and maintain versioned prompts in git. The role is remote, global, full-time, requiring strong English writing and bilingual capabilities, plus Python and LangChain/ LangGraph

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

AI Training and Prompt Engineering Specialist
AI Training and Prompt Engineering Specialist

Yonder Media Mobile Inc. • Poland

On-site
USD 120,000 - 180,000
Flexible working hours
Corporate equipment for work
Competitive salary
+1
Remote Presentation Design & AI Training Evaluator
Remote Presentation Design & AI Training Evaluator

YO AI Labs • Wrocław

Remote
PLN 192,000 - 307,000
AI Prompt Engineer
AI Prompt Engineer

SoftSnow AI • Poland

On-site
PLN 180,000 - 240,000
Comprehensive Training
Growth Opportunities
Collaborative Environment
+3
Remote AI Data Annotator & Content Evaluator
Remote AI Data Annotator & Content Evaluator

YO AI Labs • Wrocław

Remote
PLN 34,000 - 62,000
AI Prompt Architect - Remote, Growth-Oriented
AI Prompt Architect - Remote, Growth-Oriented

SoftSnow AI • Poland

Remote
PLN 180,000 - 240,000
Comprehensive Training
Growth Opportunities
Collaborative Environment
+3
Financial AI Prompt Engineer: Reporting & Analytics
Financial AI Prompt Engineer: Reporting & Analytics

Algoteque • Warszawa

On-site
PLN 180,000 - 240,000
Remote AI Data Annotator & Quality Analyst
Remote AI Data Annotator & Quality Analyst

YO AI Labs • Warszawa

Remote
PLN 78,000 - 130,000
AI Agent / Prompt Engineer
AI Agent / Prompt Engineer

Algoteque • Warszawa

On-site
PLN 180,000 - 240,000
AI Agent / Prompt Engineer - PART-TIME
AI Agent / Prompt Engineer - PART-TIME

Algoteque • Warszawa

On-site
PLN 83,000 - 165,000
Remote AI Data Annotation Specialist
Remote AI Data Annotation Specialist

YO AI Labs • Wrocław

Remote
PLN 34,000 - 48,000