AI Harness Engineer - Agentic & Autonomous Operations (m/f/x)

xHeron Solutions

Berlin

Sur place

Vertraulich

Plein temps

14 jours+
Générateur de candidature

Démarquez-vous pour ce poste — générez un CV et une lettre de motivation personnalisés en environ une minute.

Passez les filtres ATS

Avantages offerts par ce poste

Real ownership
Shape direction
Career growth & leadership

Résumé du poste

xHeron Solutions is seeking an AI Engineer to build the agentic capabilities at the heart of our service. You will shape how our agents reason, act, and learn, delivering dependable operations rather than just generating text.

You’ll own architecting the harness around models, managing state, tools, and execution loops, and connecting to real-world APIs and systems. Production ownership and observability will be key focus areas.

Qualifications

  • Strong software engineer able to design, build, test, and operate production systems in Python.

Responsabilités

  • Build the agent harness architecture for models and tools, including state management and retries.
  • Design decision and execution behavior for bounded operational tasks, with emphasis on reliability.
  • Define context, memory, and evidence strategies to maintain relevance and cost.
  • Connect agents to APIs and real-world systems with validation and idempotency.
  • Develop evaluation suites and automated tests for performance and safety in production.

Connaissances

Python
LLM systems
Workflow orchestration
Observability
Distributed systems
API integration

Outils

APIs

Description du poste

xHeron is building the Service-as-Software company for short-term rental operators. We take responsibility for guest communication, team coordination and issue resolution, so customers can hand over the work without handing over their business.

Today, we deliver that service through people, software and AI. We’re turning the expertise behind it into autonomous systems that understand situations, take action and follow through. Not another tool to supervise. A service that gets the job done.

We’re looking for an AI Engineer to build the agentic capabilities at the heart of xHeron. You’ll help shape how our agents reason, act and learn, and build a company where AI doesn’t just support the work, but increasingly delivers it.

Tasks

We are looking for an AI Harness Engineer to build the software layer that turns language models into dependable operational systems.

You will work on orchestration, context assembly, tool execution, state, permissions, evaluations, observability, failure recovery, and human handoffs. The goal is not merely to generate a good response.

What You’ll Own

  • Agent harness architecture: Build the control layer around models: task decomposition, state management, model and tool selection, execution loops, checkpoints, retries, stopping conditions, and human handoffs.
  • Decision and execution behavior: Design how agents interpret situations, choose actions, manage uncertainty, and complete bounded operational tasks—not only how they generate text.
  • Context and memory: Determine what an agent needs to know at each step. Build context assembly, retrieval, memory, evidence selection, and knowledge-maintenance strategies while managing relevance, latency, and cost.
  • Tools and real-world effects: Connect agents to APIs and operational systems through explicit tool contracts, permissions, validation, idempotency, and confirmation that the intended effect actually occurred.
  • Evaluation and regression testing: Build representative evaluation suites and automated tests for decision quality, tool use, task completion, safety, escalation behavior, and recurring failure patterns.
  • Observability and failure analysis: Make agent behavior inspectable through structured traces, state, outcomes, and error classification. Diagnose failures across models, prompts, context, tools, and surrounding software.
  • Continuous improvement: Use evaluations, production traces, operator feedback, and controlled experiments to improve prompts, context, tools, model choices, and agent architecture.
  • Production ownership: Ship and operate production Python software alongside our CTO and engineering team. Work closely with systems engineering on integrations, authorization, durable state, recovery, and safe execution.
Requirements

What Makes You a Great Fit

  • You are a strong software engineer who can design, build, test, and operate reliable production systems in Python.
  • You have built a substantial LLM-powered, automation, workflow, or decision system: not only notebooks, chat interfaces, prompt experiments, or a basic retrieval pipeline.
  • You understand the engineering layer around a model: orchestration, state, context, tools, structured outputs, permissions, retries, fallbacks, observability, and evaluations.
  • You can connect intelligence to real work: selecting relevant evidence, making a bounded decision, invoking tools or APIs, and verifying that the intended result occurred.
  • You design for partial failure and uncertainty. You think deliberately about when an agent should act, retry, wait, stop, ask for help, or elevate.
  • You evaluate behavior systematically. You use representative cases, traces, outcome checks, comparisons, and regression tests rather than treating a plausible response as proof of success.
  • You have shipped software that people or operational processes depend on and have worked through debugging, timeouts, inconsistent data, changing APIs, and production incidents.
  • You can investigate ambiguous problems independently, define a useful system boundary, explain trade-offs clearly, and turn decisions into working software.
  • You care about simplicity and control. You know when a deterministic workflow is better than an agent and when additional autonomy is justified by evidence.
  • You bring useful expertise and independent perspective. Strong adjacent experience in workflow engines, developer tooling, distributed systems, automation platforms, or integration-heavy backend software can be highly relevant when paired with credible LLM-system understanding.

Around 2+ years of professional software‑engineering experience is a useful guide, not a hard requirement. What matters is the technical depth of what you built, the decisions and production outcomes you owned, and your ability to learn unfamiliar systems.

You May Not Be a Great Fit If

  • Your LLM experience is mainly prompt engineering, chatbots, basic RAG, or connecting a model API to a user interface.
  • Your background is primarily offline data analysis, model training, or experimentation, and you do not want to own production software and operational behavior.
  • You treat retrieval as the complete agent architecture rather than one possible source of context within a larger execution system.
  • You consider a coherent model response successful without verifying the decision, tool call, state change, or real-world outcome.
  • You rely on an agent framework to provide the system design and are not comfortable reasoning about the control flow, state, permissions, and failure behavior underneath it.
  • You are not interested in owning evaluations, trace analysis, regression testing, and production learning alongside feature development.
  • You prefer fully scoped implementation tasks with stable requirements and limited responsibility for product or operational outcomes.
Benefits

Real ownership: you're not maintaining someone else's codebase, you're building a core, novel part of the product from the ground up.

  • Build at the core. Own a defining part of xHeron’s technology, from the first architectural decisions to production.
  • Shape the direction. Work directly with our CTO, challenge assumptions and help decide what we build, not just how.
  • Grow beyond the job description. We invest in your development, with room to expand your ownership and take on technical leadership as xHeron grows.
  • Share in the upside. Salary plus VSOP participation, so you can benefit from the value you help create.

Our Process

  • Intro call (15-30 min): a conversation with Friedrich (CTO) about your background, what you've built, and why this role.
  • Technical take-home case study: a short, real engineering problem, close to what you'd actually be building.
  • Case study interview (1 hour): you walk us through your thinking, not just the final code.
  • Culture fit (30 min): meet both Founders and make sure it's a fit both ways.
Obtenez votre examen gratuit et confidentiel de votre CV.

ou faites glisser et déposez votre fichier ici.

Similar jobs

Postes similaires à comparer

AI Engineer - Agentic & Autonomous Operations (m/f/x)
AI Engineer - Agentic & Autonomous Operations (m/f/x)

xHeron • Berlin

Sur place
EUR 90 000 - 130 000
Build at the core.
Shape the direction.
Grow beyond the job description.
+1
Backend / Systems Engineer, Agent Infrastructure (m/f/x)
Backend / Systems Engineer, Agent Infrastructure (m/f/x)

xHeron Solutions • Berlin

Sur place
Confidential
Real ownership
Direct founder access
Ownership growth
+1
Founding Marketer (m/f/x)
Founding Marketer (m/f/x)

xHeron Solutions • Berlin

Sur place
EUR 60 000 - 90 000
Ownership
Founder access
VSOP equity
Founding GTM (m/f/x)
Founding GTM (m/f/x)

Join • Berlin

Sur place
EUR 50 000 - 80 000
Direct founder access
VSOP equity
Ownership of GTM engine
Founding GTM (m/f/x)
Founding GTM (m/f/x)

Xheron • Berlin

Sur place
EUR 60 000 - 100 000
Founding Engineer
Founding Engineer

opus • Berlin

Sur place
EUR 90 000 - 150 000
28 vacation days
Phone Screen
Discovery Call
+3
Founding AI Engineer
Founding AI Engineer

Alago • München

Hybride
EUR 90 000 - 120 000
Equity: 0.5%–1.5% with four-year vest
Hybrid: three days a week in Munich
Early equity stake and direct impact
+1
Founding Marketer (m/f/x)
Founding Marketer (m/f/x)

xHeron • Berlin

Sur place
EUR 60 000 - 90 000
Senior Forward Deployed AI Architect (GenAI, AWS)
Senior Forward Deployed AI Architect (GenAI, AWS)

Embedded Shishya • Allemagne

Sur place
EUR 120 000 - 150 000
Remote-friendly culture
Medical insurance
Educational budget
AI Engineer
AI Engineer

Atira GmbH • München

Hybride
EUR 90 000 - 140 000
Competitive salary & equity package
Wellpass
JobRad
+1