Senior Backend Engineer (AI Platform) AZX · Remote · US · ML Platform & Ops $140,000–$230,000 4w ago

Aimlroles

Northern (KY)

Hybrid

USD 120,000 - 210,000

Full time

14 days+
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Benefits offered by this job

Health insurance
Equity
Bonus eligibility
Flexible time off
Fully remote culture

Job summary

AZX is building a scalable LLM infrastructure platform for client-facing engineers. As a Backend Engineer, you will own the LLM gateway, retrieval service, and async framework, shaping how models are accessed, billed, and governed.

You will work in a fully remote role with a fast-growing, mission-driven team, collaborating across energy, real estate, utilities, and climate initiatives, with opportunities to travel to Seattle for company summits.

Qualifications

  • 5+ years of production async Python experience.
  • Experience with PostgreSQL, Redis, and distributed systems.
  • Experience shipping API surfaces with versioning and docs.
  • Knowledge of LLM gateway or retrieval engineering.
  • Comfort with async frameworks and modern Python stack.

Responsibilities

  • Own the LLM gateway: routing, metering, budgets, guardrails and adapters.
  • Manage retrieval service: connectors, chunking, hybrid search and reranking.
  • Build an eval harness with confidence intervals and breakdowns.
  • Own the async framework API, rate-limit algorithms, and admin tooling.
  • Design authorization across the surface: per-tenant isolation and governance.
  • Define done criteria: evals, traces, cost accounting and dashboards.

Skills

Async Python (5+ years)
API design / versioning
Distributed systems experience
OpenTelemetry familiarity
Energy/real estate/utilities industrya

Tools

PostgreSQL
Redis/Dragonfly
FastAPI/Starlette
SQLAlchemy
Lua in Redis

Job description

About AZX

Our mission is to accelerate positive impact in critical industries through AI transformation. We specialize in physics-informed ML and enterprise AI solutions that directly address climate and sustainability challenges.

We’re growing quickly and already work with category-leaders in real estate (CBRE), energy (LevelTen Energy), logistics (Flexe) and utilities.

We’re a public benefit corporation, founded in 2024, and have been profitable from inception.

We work on challenges in clean energy, decarbonization, climate risk, energy systems, and global economics. We’re building our company for long-term success and aim to create the ultimate place to work for those passionate about AI and making a positive impact.

About the Role

As a Backend Engineer you will build the services our products and client solutions run on. The center of gravity is LLM infrastructure: the layer that puts one governed front door over many models, hosted and self-hosted, and answers the questions enterprises actually ask — who used which model, for what, at what cost, under whose rules. The platform's first customers are our own client‑facing engineers, the people delivering client outcomes with what you build, so requirements arrive concrete, feedback arrives same‑day, and the people you support are in the same meeting.

Responsibilities:
  • Own the LLM gateway: routing, metering, budgets, guardrail composition, and provider/backend adapters across hosted and self-hosted models.
  • Manage the retrieval and knowledge service: connectors, chunking, hybrid search and reranking, grounded answers, and the MCP surface.
  • Build the eval harness that turns retrieval tuning from judgment into evidence, with confidence intervals and per‑segment breakdowns.
  • Own the async framework — its public API, rate-limit algorithms, and admin tooling — and lead its evolution into durable execution for agent runs.
  • Design authorization across the whole surface: token hierarchies, capability grammars, and per‑tenant isolation tested as a regression, not asserted in a doc.
  • Own your own definition of done on everything you ship: evals, traces, cost accounting, and the dashboard that would wake you up.
Core Qualifications:
  • 5+ years of deep, production async Python — cancellation scopes, streaming lifecycle, and connection pooling
  • Postgres as infrastructure — comfortable reasoning about MVCC, advisory locks, and vacuum discipline, with Redis/Dragonfly used (and not used) where it belongs; real distributed‑systems experience, not just familiarity.
  • Experience shipping API surfaces other engineers build on — versioning, idempotency, error vocabularies, migration discipline, and documentation to match.
  • Depth in at least one of LLM gateway concern (routing, metering, guardrails, provider failover) or retrieval engineering (hybrid search, reranking, eval methodology, audit‑ready RAG), with credibility on the other.
  • A cost accounting and evals first mindset — you’ve shipped a gate or harness that caught a real regression, ideally one of your own.
  • Production depth in some modern stack, and the aptitude to ramp quickly on unfamiliar tools
  • Practical familiarity with our core stack — Python 3.12+ (async, type‑strict, FastAPI/Starlette, Pydantic v2), asyncpg/SQLAlchemy, Postgres (incl. pgvector), and Redis/Dragonfly with Lua — with willingness to research into the rest.
  • Exposure to the surrounding ecosystem: OpenTelemetry (incl. GenAI conventions), SSE/streaming lifecycles, OpenAI/Anthropic provider APIs, job frameworks (Celery/Dramatiq/Prefect‑class), rate‑limit algorithms (token bucket, GCRA), and integer‑cents money handling.
  • Past work in energy, real estate, utilities, climate or related fields is a plus
  • Experience in both startup and enterprise environments is a plus
Why AZX!
  • Be part of a fast‑growing, profitable, mission‑driven company with industry‑leading clients tackling the massive opportunity of AI transformation in critical industries.
  • Competitive early‑stage startup compensation (based on capabilities, experience, and location)
  • Bonus eligibility
  • Health insurance with meaningful coverage for dependents
  • Flexible paid time off
  • Equity
  • Fully remote culture with a cluster of teammates in Seattle
Additional Information:
  • Must be willing to travel to Seattle area for final interview and travel 2x/year for company summits
  • Applicants must be currently authorized to work in the United States on a full‑time basis.
  • We are unable to sponsor or take over sponsorship of employment visas at this time.
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Forward Deployed AI Engineer (Senior)
Forward Deployed AI Engineer (Senior)

careers.azx.io • United States

Remote
USD 150,000 - 210,000
Bonus eligibility
Health insurance
Flexible paid time off
+2
Senior ML Engineer (Client Solutions) at AZX
Senior ML Engineer (Client Solutions) at AZX

Matcha • Northern (KY)

On-site
USD 140,000 - 200,000
Bonus eligibility
Health insurance
Flexible paid time off
+2
Senior ML Engineer (Energy & Utilities) AZX · Remote · US · Machine Learning Engineering $140,000–$230,000 4w ago
Senior ML Engineer (Energy & Utilities) AZX · Remote · US · Machine Learning Engineering $140,000–$230,000 4w ago

Aimlroles • Seattle (WA), Northern (KY)

Hybrid
USD 140,000 - 190,000
Health insurance
Equity
Bonus eligibility
+2
Senior Software Engineer (AI Inference & Runtime Platform) at AZX
Senior Software Engineer (AI Inference & Runtime Platform) at AZX

Matcha • Northern (KY)

On-site
USD 150,000 - 190,000
Health insurance with dependents
Bonus eligibility
Equity
+2
Senior ML Engineer (Energy & Utilities) at AZX
Senior ML Engineer (Energy & Utilities) at AZX

Matcha • Northern (KY)

On-site
USD 150,000 - 210,000
Bonus eligibility
Health insurance
Equity
+1
Senior Product Engineer (AI, Full-Stack) at AZX
Senior Product Engineer (AI, Full-Stack) at AZX

Matcha • Northern (KY)

On-site
USD 120,000 - 180,000
Health insurance
Flexible PTO
Equity
+2
Senior Product Engineer (AI, Full-Stack)
Senior Product Engineer (AI, Full-Stack)

AZX • Seattle (WA)

On-site
USD 140,000 - 210,000
Health insurance
Equity
Fully remote culture
+2
Founding Backend Engineer (AI Focus) - Atrix
Founding Backend Engineer (AI Focus) - Atrix

Praxis, Inc. • New York (NY)

On-site
USD 140,000 - 190,000
Equity package
Health insurance
Unlimited PTO
+1
Founding Backend Engineer (AI Focus) - Atrix
Founding Backend Engineer (AI Focus) - Atrix

Pear VC • New York (NY)

On-site
USD 140,000 - 190,000
Competitive salary + equity package
Health and wellness support
Unlimited PTO
+1
Sr Software Engineer
Sr Software Engineer

Forward • Austin (TX), Northern (KY)

Hybrid
USD 180,000 - 240,000
Competitive salary and equity package
Flexible work arrangements
Generous PTO