Member Of Technical Staff — AI Systems & Benchmarking

Aionia Group

San Francisco (CA)

On-site

USD 130,000 - 220,000

Full time

2 days ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Benefits offered by this job

Equity
On-site

Job summary

AI Systems & Benchmarking in San Francisco is seeking a Member of Technical Staff to explore how AI systems perform in the real world and define how that performance is measured. You’ll shape evaluation frameworks, datasets, and metrics used by the industry, at the intersection of systems, analysis, AI, and strategy.

This is an on-site role with equity upside for early-stage impact. You’ll work with frontier models and real-world deployment contexts, translating complex behavior into actionable

Qualifications

  • Must-have — strong Python proficiency with recent, hands-on production work.
  • Experience with data analysis and building analytical frameworks.
  • Ability to operate as a technical generalist across systems and domains.
  • Clear, structured communication — translate complexity into insight.
  • High ownership and comfort working in ambiguous environments.

Responsibilities

  • Design and build AI evaluation and benchmarking systems.
  • Analyze how models and agents perform across real-world use cases.
  • Develop frameworks, datasets, and metrics to measure AI capabilities.
  • Translate complex system behavior into clear, actionable insights.
  • Work closely with engineers, product teams, and external partners.
  • Contribute directly to product direction and overall strategy.

Skills

Python
Data analysis
Technical generalist
Clear communication
Ownership

Job description

AI Systems · Benchmarking · San Francisco

Member of Technical Staff

Most AI roles focus on building AI systems. This one focuses on something more fundamental — understanding how AI systems actually perform in the real world, and defining how that performance is measured. You’ll operate at the intersection of systems, analysis, AI, and strategy.

$130K – $220K + Equity On-Site · San Francisco AI Systems & Benchmarking Early-Stage Equity

This is not a typical AI role. You won’t just build AI systems — you’ll work on the problems that define how modern AI is evaluated, understood, and deployed. You’ll shape the frameworks, datasets, and metrics that the industry uses to make sense of model and agent behavior at scale.

“You’ll operate at the intersection of systems, analysis, AI, and strategy — on problems that most engineers never get to touch.”

Who This Is For
This Role Tends to Resonate With

This role attracts a specific type of person. You might be a fit if you:

Think deeply about how AI systems behave, not just how to build them

Have built real systems — APIs, pipelines, integrations, or products

Enjoy breaking down complex problems into structured, reusable frameworks

Are comfortable moving between technical detail and big-picture thinking

Have used AI tools in practice — LLMs, agents, and automated workflows

Still enjoy coding and working hands-on at the implementation level

Responsibilities
What You’ll Do
  • Design and build AI evaluation and benchmarking systems
  • Analyze how models and agents perform across real-world use cases
  • Develop frameworks, datasets, and metrics to measure AI capabilities
  • Translate complex system behavior into clear, actionable insights
  • Work closely with engineers, product teams, and external partners
  • Contribute directly to product direction and overall strategy
Requirements
What Matters Most

Must-Have — Non-Negotiable

  • Strong Python proficiency with recent, hands-on production work
  • Experience with data analysis and building analytical frameworks
  • Ability to operate as a technical generalist across systems and domains
  • Clear, structured communication — you translate complexity into insight
  • High ownership and comfort operating in ambiguous environments
Backgrounds That Tend to Work Well
  • Product-minded software engineers with meaningful AI exposure
  • Engineers who’ve built systems involving LLMs, pipelines, or automation
  • Technical professionals from top-tier consulting with real coding ability
  • Founding or early engineers with broad, cross-functional ownership

The following profiles are unlikely to thrive in this role.

  • Focused exclusively on training models or academic research
  • Removed from hands-on coding and implementation work
  • Purely management-focused without technical execution
Why This Role
Why This Role Is Different
Define how AI is measured, not just built

The work here shapes the frameworks and metrics the industry relies on to understand model and agent behavior — a rare, foundational problem space.

Direct exposure to cutting-edge AI systems

You’ll work with frontier models and real-world deployment contexts that most engineers don’t have visibility into.

High ownership in a small, fast-growing team

The team is expected to scale rapidly. Joining now means meaningful equity upside and the ability to shape how the function is built.

A rare blend of technical depth, strategy, and impact

This isn’t a pure engineering role or a pure strategy role. It’s both — for someone who can operate at that intersection.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Member of Technical Staff (Applied AI Research)
Member of Technical Staff (Applied AI Research)

Artificial Analysis, Inc. • San Francisco (CA)

On-site
USD 170,000 - 210,000
Member of Technical Staff
Member of Technical Staff

Artificial Analysis • San Francisco (CA)

On-site
USD 180,000 - 260,000
Equity
Competitive compensation
Member of Technical Staff (Commercial) - Strategy Consulting track
Member of Technical Staff (Commercial) - Strategy Consulting track

Artificial Analysis • San Francisco (CA)

On-site
USD 150,000 - 190,000
Member of Technical Staff
Member of Technical Staff

Artificial Analysis, Inc. • San Francisco (CA)

On-site
USD 100,000 - 150,000
Competitive compensation including equity
Opportunity to shape AI development
Senior AI / ML Engineer
Senior AI / ML Engineer

Artificial Analysis, Inc. • San Francisco (CA)

On-site
USD 150,000 - 220,000
Equity
Applied AI Engineer — Well-Funded AI Platform
Applied AI Engineer — Well-Funded AI Platform

Aionia Group • New York (NY), Northern (KY)

Hybrid
USD 185,000 - 325,000
Equity
Senior AI Benchmarking & Systems Architect
Senior AI Benchmarking & Systems Architect

Aionia Group • San Francisco (CA)

On-site
USD 130,000 - 220,000
Equity
On-site
Principal Staff Software Engineer
Principal Staff Software Engineer

Harnham • Oakland (CA)

On-site
USD 280,000 - 350,000
Strong compensation
Meaningful equity
Opportunity to build from scratch
Member of Technical Staff (Language Model Evaluations)
Member of Technical Staff (Language Model Evaluations)

Artificial Analysis • San Francisco (CA)

On-site
USD 180,000 - 260,000
Equity
Senior / Staff Software Engineer, AI Platform
Senior / Staff Software Engineer, AI Platform

satoriq • New York (NY)

On-site
USD 120,000 - 160,000
Competitive Equity Package
401(k)
Health, Dental, and Vision Coverage
+2