Staff AI Engineer - Agent Architecture & Behavior

Artisan

San Francisco (CA)

On-site

USD 250,000 - 325,000

Full time

2 days ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Benefits offered by this job

Visa sponsorship

Job summary

Artisan is a hands-on technical-lead role in San Francisco focusing on building an ambitious agentic AI architecture. You will own the AI system, design agent behavior, multi-agent coordination, and robust production execution. Expect high-impact engineering challenges and leadership in a private, frontier project.

The role emphasizes architectural ownership, tough problem solving, and collaboration with product and infrastructure teams to ship reliable software.

Qualifications

  • Experience building and shipping substantial agentic systems.
  • Deep practical experience with LLM tool use, planning, context engineering, and evaluations.
  • Hands-on experience with browser or computer automation in an agentic system.

Responsibilities

  • Design and implement agent execution loops, planning, tool interfaces, and verification.
  • Build multi-agent delegation, coordination, context sharing, and result synthesis.

Skills

Python
TypeScript
LLM tool use
Planning
Context engineering
Evaluations
Agent architecture

Tools

LLM tooling
Browser automation

Job description

Artisan · Full-time · In person in San Francisco

Base salary: $250,000–$325,000 USD annually. Equity: 0.15%–0.30%. US visa sponsorship available.

Build something new at the frontier of applied AI

At Artisan, we're working on a new, ambitious project that will push the boundaries of what agentic AI can do. We're keeping the product details private ahead of launch, but we can tell you this: the technical problems are substantial, the scope for invention is real, and this hire will shape the core technology.

We're looking for a hands‑on technical lead to design and build the underlying AI architecture. You'll work across agent behavior, complex multi‑agent systems, tool use, context, and evaluation, taking promising ideas through to dependable production software.

This is an individual-contributor role with broad technical ownership. You'll make consequential architecture decisions, write the hardest parts of the system, and work closely with our existing engineers and leadership. You should enjoy both exploring an uncertain problem and doing the detailed engineering required to make a solution work.

What you'll own
  • Agent architecture and behavior. Design and implement agent execution loops, planning strategies, tool interfaces, and verification. Turn ambiguous technical requirements into clear system boundaries and working code.

  • Multi‑agent systems. Build delegation, coordination, context sharing, and result synthesis. Handle concurrent work, conflicting updates, cancellation, and stale results. Establish when a multi‑agent approach improves on a simpler baseline.

  • Reliable execution. Make complex, stateful workflows resilient to interruptions and partial failures. Build checkpoints, recovery strategies, and appropriate human intervention into the architecture.

  • Context, memory, and reusable methods. Improve retrieval, context construction, persistent state, and skill representation. Investigate how systems can use feedback and prior experience to perform better without introducing regressions.

  • Evaluation and experimentation. Build realistic evaluations, analyze task trajectories, and turn observed failures into measurable improvements. Compare approaches using quality, reliability, latency, and cost.

  • Model and tooling decisions. Evaluate models and emerging techniques, prototype promising approaches, and make informed build‑versus‑buy decisions. Choose tools because they solve the problem, and be willing to replace them when the evidence changes.

  • Technical leadership. Set engineering standards, review important design decisions, and help the team implement a coherent AI system. Stay close to the product and accountable for what ships.

You'll partner with product and infrastructure engineers on production services, integrations, secure execution, and observability. You'll own the AI architecture and its effectiveness, with implementation shared across the team.

What we're looking for
  • You have personally built and shipped a substantial agentic system. Production use or rigorous, reproducible open‑source work matters more than the name of a framework or employer.

  • You have deep practical experience with LLM tool use, planning, context engineering, and evaluations. You have implemented multi‑agent coordination or substantial parallel agent/tool execution and can explain its failure modes.

  • You have hands‑on experience with browser or computer automation in an agentic system, including observing state, verifying effects, and recovering when an interface or execution path fails.

  • You are an excellent software engineer in Python, TypeScript, or a comparable language. You are comfortable with asynchronous services, state machines, persistence, concurrency, retries, and cancellation.

  • You know which decisions belong to a model and which guarantees must be enforced in code. You can reason carefully about permissions, untrusted inputs, uncertain external outcomes, and human approvals.

  • You can design meaningful experiments, debug real system behavior, and explain what improved, why it improved, and where the evidence is still weak.

  • You can take technical ownership of an unclear problem, work effectively with other engineers, and ship with urgency and care.

Useful additional experience

Depth in agent memory and retrieval, skill acquisition, reinforcement learning or post‑training, trajectory datasets, sandboxed execution, distributed systems, inference optimization, or multimodal and voice models would be valuable. We expect strong foundations and particular depth in a few areas, rather than prior specialization in every one.

There is no required degree, publication record, previous employer, or agent framework. We're hiring for demonstrated engineering ability, judgment, and ownership.

How we work

This role is based in our San Francisco office. Expect a small team, short feedback cycles, direct communication, and high standards. We value people who move quickly, take responsibility for the result, surface problems early, and change their minds when the evidence calls for it.

You'll have substantial freedom to explore ambitious technical ideas and the responsibility to turn the best ones into software that works. We'll discuss the project in more detail during the interview process.

Interview process

Our conversations will center on systems you've built, a practical agent architecture and debugging exercise, and how you work with a team.

The base salary range is $250,000–$325,000 USD annually, plus equity. The final offer will reflect relevant experience, demonstrated skills, and the scope of responsibility.

About Artisan

Artisan builds AI employees that take on real work. Our first three are Ava, our outbound AI BDR; Aaron, our inbound AI SDR; and Aria, our AI account executive. Our broader mission is to build AI employees that can take responsibility for work across many roles and industries.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Staff Software Engineer
Staff Software Engineer

Artisan • San Francisco (CA)

On-site
USD 120,000 - 160,000
Competitive salary
Equity in the company
High impact on product development
Builder
Builder

Artisan • San Francisco (CA)

On-site
USD 130,000 - 230,000
Equity 0.08–0.20%
US visa sponsorship
Staff Devops Engineer
Staff Devops Engineer

Artisan • San Francisco (CA)

On-site
USD 130,000 - 180,000
Competitive salary
Meaningful equity
Flexible working conditions
Applied AI Engineer
Applied AI Engineer

Human Agency • United States

On-site
USD 100,000 - 150,000
Applied AI Engineer
Applied AI Engineer

Human Agency • Boston (MA)

On-site
USD 100,000 - 130,000
Lead AI Architect: Agent Systems & Behavior
Lead AI Architect: Agent Systems & Behavior

Artisan • San Francisco (CA)

On-site
USD 250,000 - 325,000
Visa sponsorship
Software Engineer, Agent
Software Engineer, Agent

OpenArt AI • San Carlos (CA)

On-site
USD 320,000 - 420,000
Visa sponsorship available
Hybrid work option
Equity ownership
Software Engineer, Agent
Software Engineer, Agent

re-zoo-me • San Francisco (CA), Northern (KY)

Hybrid
USD 300,000 - 380,000
Applied AI Engineer
Applied AI Engineer

daydream Labs • San Francisco (CA)

On-site
USD 180,000 - 220,000
Medical, dental, and vision insurance
Lunch on in-office days
Wellness and learning stipends
+2
Artificial Intelligence Engineer
Artificial Intelligence Engineer

Oxbow Talent • San Francisco (CA)

On-site
USD 175,000 - 220,000
Meals provided in office
Competitive equity
In-office SF workspace