Senior Platform Engineer

Rifa AI

United States

Remote

USD 140,000 - 190,000

Full time

14 days+
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

Rifa AI is building an AI agents platform for contact centers in regulated industries. You will own the platform underneath client deliveries—runtime, voice infrastructure, and enterprise capabilities—ensuring fast, reliable, and auditable operations.

Join a small, capable team driving governance, security, and scalable AI agent solutions that transform millions of conversations into actionable insights. Work with Kubernetes, ArgoCD, and observability tooling to keep systems robust.

Qualifications

  • 5+ years of hands-on experience building and operating production systems.
  • Strong backend and distributed systems fundamentals.
  • Proven experience with cloud platforms, infrastructure as code, and container orchestration.
  • Real-time systems depth: latency budgets, streaming, and live audio differences.
  • Experience with observability tooling and incident response.

Responsibilities

  • Build the agent platform primitives: runtime, orchestration, and Agent Studio tooling.
  • Own the voice infrastructure for real-time audio and low-latency processing.
  • Form the foundation of enterprise trust: RBAC, audit logs, data isolation.
  • Build the evaluation layer with DeepEval-driven suites and AB testing.
  • Build observability: traceable agent decisions and proactive monitoring.
  • Keep the platform fast and reliable: Kubernetes, GitOps, incident response.

Skills

Backend systems
Distributed systems
Cloud platforms
Observability tooling
Incident response
High autonomy

Education

BS in Computer Science or related field

Tools

Kubernetes
ArgoCD
GitOps
OpenTelemetry
SigNoz
OpenFGA

Job description

About us

Rifa AI is building the AI agents platform for contact centers in regulated industries.

Enterprises in these industries want AI agents handling their customer operations and mostly can't deploy them. It's not a model problem. Horizontal platforms lack governance, release processes, and change management, and in a domain where every call can be reviewed by a regulator, that's disqualifying. Building an AI agent has never been easier. Deploying one an enterprise can trust has never been harder. That harder problem is the one we work on.

Our platform turns a company's written procedures into AI agents: voicebots that hold real-time conversations, take actions in the client's CRM, and stay within the limits the client has set. Every release is gated by an automated testing suite, and every conversation feeds a post-analysis platform. We're live in production today, handling debt collection calls for US financial services clients.

The engineering convictions behind it: evaluation methodology, not model capability, is the bottleneck. Every rule and decision trace becomes part of a company's context graph. Observability goes beyond logging. And our Agent Studio lets engineers, non-engineers, and auditors collaborate to build and improve AI agents with the right guardrails to achieve the expected business outcomes.

Rifa was founded by Sameer Fulzele (IIT Bombay). We're a small team of exceptionally capable, passionate engineers with paying enterprise clients and growing revenue, backed by Seaborne Capital, a founder-first firm of exceptional industry operators that works closely with its founders, and by angel investors who are veterans of enterprise software and operators in the accounts receivable industry.

About the role

While Agent Engineers own individual client deliveries, you own the platform underneath all of them: the runtime that agents execute on, the voice infrastructure that keeps live calls fast and stable, the enterprise capabilities that get us through security reviews, and the intelligence layer that turns millions of conversations into insight. This is the highest-leverage engineering seat in the company: every improvement you ship reaches every client at once.

You'll work directly with the founder on architecture and own your decisions end to end, from design through production and the incident channel.

What you'll do
  • Build the agent platform primitives. Design and evolve the runtime, orchestration engine, and the systems behind Agent Studio: how agents are defined from written procedures, how they reason and take actions, and how guardrails make out-of-bounds behavior impossible rather than unlikely. Building agents is its own engineering discipline, and you'll be building the tooling that discipline needs.

  • Own the voice infrastructure. Real-time audio streaming over WebSockets, STT and TTS integration at low latency, and the hard parts of live conversation: interruptions, disconnects, transfers, and telephony integrations that pick up and place calls reliably at growing volume.

  • Form the foundation of enterprise trust. Clients hand us regulated conversations with their own customers. You'll build and uphold the capabilities that make that possible: RBAC, access management, audit logs, data isolation, and integrations that embed into complex client environments while surviving their security reviews.

  • Build the evaluation layer. We believe evaluation methodology, not model capability, is the bottleneck, and you'll own the system that proves it: the DeepEval-driven suite that replays real conversations against every change, simulation of the scenarios we haven't seen yet, and experimentation frameworks that let us A/B test agent behavior and make changes with evidence instead of intuition. Every release gates on what you build here.

  • Build the observability layer. Observability goes beyond logging. You'll extend Reflect so that every agent decision is traceable: what the system knew at each moment, which instructions the model was given, and why it responded the way it did. Add proactive monitoring that surfaces regressions, drift, and new conversation patterns before a client notices, and turn millions of calls into insight the whole company acts on.

  • Close the feedback loop. Connect what evaluation and observability find back into how agents improve: recurring failure modes become new eval cases, production patterns reshape procedures, and agents get measurably better over time.

  • Keep the platform fast, reliable, and boring to operate. Own Kubernetes, ArgoCD-driven GitOps deployments, CI/CD, and observability with SigNoz. Build the self-serve infrastructure that lets the rest of engineering ship without waiting on you, and lead incident response and postmortems when things break.

What we work with

Python with FastAPI, PostgreSQL, Temporal for background workflows, WebSockets holding real-time conversations open, and STT and TTS providers on the speech side, with LLMs via OpenAI and similar underneath.

Everything runs on Kubernetes with ArgoCD-driven GitOps deployments, SigNoz for observability, DeepEval driving the eval suite our CI runs before anything reaches production, Ory and OpenFGA for identity and access management, and a lot of beautifully built internal agent architecture underneath.

What you'll bring
  • 5+ years of hands-on experience building and operating production systems, with strong backend and distributed systems fundamentals.

  • Proven experience with cloud platforms, infrastructure as code, and container orchestration: you've run Kubernetes in production, not just in a tutorial.

  • Real-time systems depth: you understand latency budgets, streaming, backpressure, and what makes a live audio conversation different from a request-response API.

  • Experience with observability tooling (SigNoz, OpenTelemetry, or similar) and with incident response: you've been paged, found root cause, and made the pager quieter afterward.

  • You care that a system stays reliable, secure and maintainable as it grows, and the judgment to know when boring technology is the right answer.

  • High agency: you drive outcomes in a high-autonomy environment, find creative ways around obstacles, and don't wait to be told what matters.

  • Degree in computer science or a related field, or equivalent professional experience.

Even better...
  • Experience building enterprise features: SSO, RBAC, IAM, audit logs, data isolation, or compliance-adjacent systems

  • Production experience with LLMs, agent frameworks, retrieval, or evaluation systems

  • Telephony or streaming audio experience: Twilio, WebRTC, SIP, or contact center integrations

  • Experience with large-scale data systems, analytics platforms, or ML-powered product features

  • Experience building developer platforms, SDKs, or internal tooling other engineers love

  • Leadership experience on technical projects or teams

How we work

Small teams, real ownership. You'll have peers to pair with and review your work, and you'll review theirs. We review all code, we write things down, and we'd rather hear an honest "I don't know yet" than a confident wrong answer.

Our values
  • Trust: We do the right thing, especially when nobody is checking. Clients hand us regulated conversations with their own customers, and we earn that every day.

  • Transparency: We write things down, share the real numbers, and say "I don't know yet" out loud, with each other and with clients.

  • Technically best solution: We choose what's right, not what's easiest or trendiest. When we notice we got it wrong, we fix it.

  • Decisiveness: We decide quickly with the information we have, commit, and correct course fast when reality disagrees.

  • Simplicity: We keep systems, processes, and words simple. Complexity is a cost we pay only when it clearly buys something.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Agent Engineer
Agent Engineer

Rifa AI • United States

Remote
USD 130,000 - 180,000
GTM Lead
GTM Lead

Rifa AI • United States

Remote
USD 120,000 - 190,000
AI Solutions Engineer (Delivery Lead)
AI Solutions Engineer (Delivery Lead)

UNKNOWN • Raleigh (NC), Northern (KY)

On-site
USD 140,000 - 200,000
Implementation Manager
Implementation Manager

AUI™ (Augmented Intelligence) • New York (NY)

On-site
USD 120,000 - 180,000
Founding AI Engineer
Founding AI Engineer

Bonfirevc • New York (NY)

On-site
USD 120,000 - 160,000
100% employer-paid health insurance premiums
Flexible PTO
Seed-stage equity grant
+2
Software Engineer, Agent (New Grad 2027)
Software Engineer, Agent (New Grad 2027)

United States Digital Space LLC • San Francisco (CA)

On-site
USD 180,000 - 240,000
Flexible PTO
Medical, dental, vision
Parental leave
+1
Senior / Staff Backend Engineer
Senior / Staff Backend Engineer

Hamming • Austin (TX)

On-site
USD 120,000 - 160,000
Flexible work hours
Career development opportunities
Software Engineer, Agent (Spanish speaking)
Software Engineer, Agent (Spanish speaking)

United States Digital Space LLC • San Francisco (CA)

On-site
USD 170,000 - 250,000
Flexible PTO
Medical, dental, and vision benefits
Life insurance and disability benefits
+6
Founding Engineer
Founding Engineer

Empathic, Inc. • New York (NY)

Hybrid
USD 170,000 - 250,000
5 weeks PTO
Parental leave
Health (BCBS Platinum)
AI Infrastructure Engineer
AI Infrastructure Engineer

Percepta • New York (NY)

On-site
USD 140,000 - 190,000