AI Engineer

Rad Hires

United States

Remote

USD 83,000 - 165,000

Full time

3 hours ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

Rad Hires is seeking mid-to-senior LATAM engineers to place into US-based tech startups building AI-forward products. You will design and ship harnesses that enable AI agents to operate safely and effectively in production.

You will work directly with founding teams and product/ops, shaping MCP servers, tools, and orchestration flows. Remote in LATAM with a competitive USD hourly rate for this full-time placement.

Qualifications

  • Engineering judgment to decompose ambiguous business problems into a system (data, interfaces, control flow, failure modes).
  • Experience with multi-agent or orchestrated systems in production and clear boundaries for roles.

Responsibilities

  • Design codebases so AI agents can work in them with high success rates and robust guardrails.
  • Build MCP servers, tools, and sub-agents that expose business logic to AI clients and non-technical teammates.
  • Set up retrieval, schemas, and context systems to keep models from guessing.
  • Create internal admin panels, CMS layers, and configuration UIs for safe non-ticket changes.
  • Define evals and autonomy boundaries; own observability for production systems.
  • Take features from blank page to first customer iteration across the stack.

Job description

Remote, LATAM. Full-time placement with a US-based tech startup. Mid-to-senior.

About the role

Rad Hires places senior LATAM engineers into US tech startups that need someone who can take an ambiguous problem from blank page to shipped MVP, with AI integrated at every stage.

We are hiring architects and harness builders. The model is the engine. The harness is everything around it: repo structure, MCP servers, tools, sub-agents, evals, guardrails, internal UIs, documentation, traces, and automation pipelines. A good harness makes a mediocre model useful in production. A bad harness wastes a great one. We hire the people who build the harness.

This is explicitly a mid-to-senior role. AI tooling amplifies the engineering instincts you already have. It does not install them where they are missing. We are looking for engineers whose judgment about systems, maintainability, and tradeoffs was earned before the model could help.

Who you will work with

Our clients are US-based tech startups, typically seed through Series B, building AI-forward products. Small teams, fast cycles, real customers from day one. The kind of place where one engineer can ship a feature from blank page to first user without three meetings to get permission. You will work directly with the founding team, product, and operations, not behind layers of management.

What you will do
  • Design codebases so AI agents can work in them with high success rates. Repo conventions, AGENTS.md or CLAUDE.md hierarchies, machine-readable structure, and the guardrails (tests, linters, hooks) that keep agents honest while they edit.
  • Build MCP servers, custom tools, and focused sub-agents that wrap business logic and expose it to both AI clients and non-technical teammates.
  • Set up retrieval, schemas, and context systems so models do not have to guess.
  • Build internal tools (admin panels, CMS layers, visual editors, configuration UIs) that let marketing, ops, and support make routine changes safely, without filing tickets.
  • Design evals, define autonomy boundaries (always-do, ask-first, never-do), and own the observability that keeps agent-driven systems honest in production.
  • Take features from blank page to first-customer iteration, across the stack, orchestrating AI tooling deliberately to compress the build cycle without losing the plot.
What we look for
  • Engineering judgment, first. You can decompose an ambiguous business problem into a system: data, interfaces, control flow, failure modes, who owns what. You reason cleanly about tradeoffs between deterministic code, agent loops, retrieval, and human-in-the-loop. You surface non-functional requirements (latency, cost, blast radius, recoverability) without being asked.
  • Active harness experience. You have built and shipped at least one MCP server, custom tool surface, or production agent workflow. You can defend the tool boundaries you chose and what changed after real usage. You have opinions on hooks, sub-agents, and orchestration frameworks (LangGraph, CrewAI, AutoGen, or custom), and you are not married to any one stack.
  • Static harness instincts. You have a point of view on repo layout, naming, and module boundaries that make context windows tractable. You maintain honest AGENTS.md or CLAUDE.md files and treat agent failure as usually an environment-and-context problem, not a reasoning problem.
  • Safety harness discipline. You treat evals as first-class engineering work: offline sets, regression gates in CI, LLM-as-judge with human calibration. You have set up traces, tool-call logs, cost and latency budgets, and dashboards a non-engineer can read. You plan for model outages, regressions, and cost spikes.
  • End-to-end product velocity. You have taken a product or feature from blank page to a customer using it: data model, backend, frontend, deploy, monitoring, basic ops. You do not need permission to touch any layer. You know when an MVP is real versus when it is a demo that will not survive contact with reality.
  • AI enablement for non-engineers. You have shipped internal tools that non-engineers depend on, with preview, staging, approvals, and audit log designed in from the start, not bolted on. You believe the goal is to make non-engineers more capable, not more dependent on engineering.
  • Verification over delegation. You treat AI-generated code as a draft, never as authoritative. You can walk us through a specific time you rejected or rewrote AI-generated code: what you spotted (hallucinated APIs, silent error swallowing, scope creep from the agent), and how you caught it.
  • Engineering fundamentals (hard floor). You have shipped real production systems and can talk about what you learned operating them over time. Comfortable in TypeScript or Python plus at least one backend stack. You can model data in relational and document stores, understand HTTP, auth, and security basics, and have shipped with containers, CI, and either cloud or self-hosted infrastructure.
  • Communication. You can describe an architectural decision in plain language and in technical language, and pick the right register for the audience. You write clearly in chat, docs, and commits.
Signals we love
  • A portfolio of small, weird, useful internal tools, not just big projects.
  • A multi-agent or orchestrated system in production with explicit role boundaries, segmentation, and a story about what broke first.
  • Examples of helping non-engineers on past teams do more of their own work.
  • Strong opinions on hooks, slash commands, custom skills, and sub-agents, and when each is the right primitive.
  • Tried at least one new AI tool in the last sixty days. Can name two or three things that meaningfully changed in the field this quarter.
This role is not for you if
  • Your value pitch is "I write code fast."
  • You ship AI-generated code without reading it, testing it, or being able to defend it.
  • You have never built anything a non-developer actually uses.
  • You build "god mode" agents with broad toolsets and no clear role boundaries.
  • You treat AI as autopilot, or as an existential threat. Neither posture works here.
  • You have not touched a new AI tool in months.
Logistics
  • Location: Remote, anywhere in LATAM.
  • Hours: Minimum 4 hours of daily overlap with US Pacific or US Eastern business hours.
  • Language: Professional English fluency required, written and spoken. You will be in daily contact with US-based product, design, and operations teammates.
  • Engagement: Full-time placement with a US-based tech startup client. Rad Hires sources, vets, and matches. You work directly with the client team.
  • Compensation: Competitive hourly rate in USD based on experience and technical evaluation.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AI Engineer, Agent Builder (Remote).
AI Engineer, Agent Builder (Remote).

Catalyst Wayfare • United States

Remote
USD 120,000 - 170,000
AI Engineer - Forward Deployed
AI Engineer - Forward Deployed

Moring • Atlanta (GA)

On-site
USD 120,000 - 160,000
Applied AI Engineer
Applied AI Engineer

Human Agency • United States

On-site
USD 100,000 - 150,000
Applied AI Engineer
Applied AI Engineer

Human Agency • Boston (MA)

On-site
USD 100,000 - 130,000
Senior AI Engineer - Forward Deployed (FDE)
Senior AI Engineer - Forward Deployed (FDE)

Moring • Atlanta (GA)

On-site
USD 150,000 - 220,000
AI Agent Engineer
AI Agent Engineer

ChaiOne • Houston (TX)

Hybrid
USD 120,000 - 150,000
Competitive compensation
Growth opportunities
Leadership growth through mentoring
Senior Applied AI Engineer
Senior Applied AI Engineer

Level • Austin (CO)

On-site
USD 180,000 - 240,000
Relocation assistance
Senior Software Engineer (GO/Rust/Python) AI Team
Senior Software Engineer (GO/Rust/Python) AI Team

Staffing Science • San Francisco (CA)

Remote
USD 180,000 - 280,000
AI Engineer - Core
AI Engineer - Core

Hilbert's AI • San Francisco (CA)

On-site
USD 120,000 - 160,000
Competitive salary
Equity package
Performance-based bonuses
AI Workflow Engineer
AI Workflow Engineer

Ultimate Knowledge • Tennessee

On-site
USD 150,000 - 200,000
Remote work
Competitive salary
Career growth