Senior Platform Engineer

Rifa AI

San Francisco (CA)

On-site

USD 150,000 - 210,000

Full time

10 days ago

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Competitive salary
Healthcare
Stock options

Job summary

Rifa AI in San Francisco is building an AI agents platform for contact centers in regulated industries. You will own the platform underneath client deliveries: the runtime that agents execute on, and the voice infrastructure that keeps live calls fast and stable.

You'll work with the founder on architecture, guiding decisions from design to production, and oversee the evolution of observability, security, and evaluation layers to support scalable, auditable deployments across enterprise clients.

Qualifications

  • 5+ years building production systems.
  • Experience with cloud platforms, infrastructure as code, and container orchestration.
  • Strong backend and distributed systems fundamentals.

Responsibilities

  • Build and operate production-grade agent platform underneath client deliveries.
  • Own runtime and voice infrastructure for real-time conversations.
  • Collaborate with founders on architecture and end-to-end decisions.
  • Drive incident response, observability, and postmortems to improve systems.
  • Lead end-to-end delivery and ensure enterprise security and governance.

Skills

Distributed systems
Backend systems
Cloud platforms
Real-time audio
Observability tools
High autonomy

Education

Bachelor's degree in CS or related field

Tools

Kubernetes
ArgoCD
GitOps
SigNoz
OpenTelemetry
Twilio

Job description

About us

Rifa AI is building the AI agents platform for contact centers in regulated industries.

Enterprises in these industries want AI agents handling their customer operations and mostly can't deploy them. It's not a model problem. Horizontal platforms lack governance, release processes, and change management, and in a domain where every call can be reviewed by a regulator, that's disqualifying. Building an AI agent has never been easier. Deploying one an enterprise can trust has never been harder. That harder problem is the one we work on.

Our platform turns a company's written procedures into AI agents: voicebots that hold real-time conversations, take actions in the client's CRM, and stay within the limits the client has set. Every release is gated by an automated testing suite, and every conversation feeds a post-analysis platform. We're live in production today, handling debt collection calls for US financial services clients.

The engineering convictions behind it: evaluation methodology, not model capability, is the bottleneck. Every rule and decision trace becomes part of a company's context graph. Observability goes beyond logging. And our Agent Studio lets engineers, non-engineers, and auditors collaborate to build and improve AI agents with the right guardrails to achieve the expected business outcomes.

Rifa was founded by Sameer Fulzele (IIT Bombay). We're a small team of exceptionally capable, passionate engineers with paying enterprise clients and growing revenue, backed by Seaborne Capital, a founder-first firm of exceptional industry operators that works closely with its founders, and by angel investors who are veterans of enterprise software and operators in the accounts receivable industry.

About the role

While Agent Engineers own individual client deliveries, you own the platform underneath all of them: the runtime that agents execute on, the voice infrastructure that keeps live calls fast and stable, the enterprise capabilities that get us through security reviews, and the intelligence layer that turns millions of conversations into insight. This is the highest-leverage engineering seat in the company: every improvement you ship reaches every client at once.

You'll work directly with the founder on architecture and own your decisions end to end, from design through production and the incident channel.

What you'll do
  • Build the agent platform primitives. Design and evolve the runtime, orchestration engine, and the systems behind Agent Studio: how agents are defined from written procedures, how they reason and take actions, and how guardrails make out-of-bounds behavior impossible rather than unlikely. Agent development is an evolution of software development, and you'll be building the tools it requires.

  • Own the voice infrastructure. Real-time audio streaming over WebSockets, STT and TTS integration at low latency, and the hard parts of live conversation: interruptions, disconnects, transfers, and telephony integrations that pick up and place calls reliably at growing volume.

  • Form the foundation of enterprise trust. Clients hand us regulated conversations with their own customers. You'll build and uphold the capabilities that make that possible: RBAC, access management, audit logs, data isolation, and integrations that embed into complex client environments while surviving their security reviews.

  • Build the evaluation layer. We believe evaluation methodology, not model capability, is the bottleneck, and you'll own the system that proves it: the DeepEval-driven suite that replays real conversations against every change, simulation of the scenarios we haven't seen yet, and experimentation frameworks that let us A/B test agent behavior and make changes with evidence instead of intuition. Every release gates on what you build here.

  • Build the observability layer. Observability goes beyond logging. You'll extend Reflect so that every agent decision is traceable: what the system knew at each moment, which instructions the model was given, and why it responded the way it did. Add proactive monitoring that surfaces regressions, drift, and new conversation patterns before a client notices, and turn millions of calls into insight the whole company acts on.

  • Close the feedback loop. Connect what evaluation and observability find back into how agents improve: recurring failure modes become new eval cases, production patterns reshape procedures and agents get measurably better over time.

  • Keep the platform fast, reliable, and boring to operate. Own Kubernetes, ArgoCD-driven GitOps deployments, CI/CD, and observability with SigNoz. Build the self-serve infrastructure that lets the rest of engineering ship without waiting on you, and lead incident response and postmortems when things break.

What we work with

Python with FastAPI, PostgreSQL, Temporal for background workflows, WebSockets holding real-time conversations open, and STT and TTS providers on the speech side, with LLMs via OpenAI and similar underneath.

Everything runs on Kubernetes with ArgoCD-driven GitOps deployments, SigNoz for observability, DeepEval driving the eval suite our CI runs before anything reaches production, Ory and OpenFGA for identity and access management, and a lot of beautifully built internal agent architecture underneath.

What you'll bring
  • 5+ years of hands-on experience building and operating production systems, with strong backend and distributed systems fundamentals.

  • Proven experience with cloud platforms, infrastructure as code, and container orchestration: you've run Kubernetes in production, not just in a tutorial.

  • Real-time systems depth: you understand latency budgets, streaming, backpressure, and what makes a live audio conversation different from a request-response API.

  • Experience with observability tooling (SigNoz, OpenTelemetry, or similar) and with incident response: you've been paged, found root cause, and made the pager quieter afterward.

  • Pride in building reliable, maintainable, and secure systems that scale, and the judgment to know when boring technology is the right answer.

  • High agency: you drive outcomes in a high-autonomy environment, find creative ways around obstacles, and don't wait to be told what matters.

  • Degree in computer science or a related field, or equivalent professional experience.

Even better...
  • Experience building enterprise features: SSO, RBAC, IAM, audit logs, data isolation, or compliance-adjacent systems

  • Production experience with LLMs, agent frameworks, retrieval, or evaluation systems

  • Telephony or streaming audio experience: Twilio, WebRTC, SIP, or contact center integrations

  • Experience with large-scale data systems, analytics platforms, or ML-powered product features

  • Experience building developer platforms, SDKs, or internal tooling other engineers love

  • Leadership experience on technical projects or teams

How we work

Small teams, real ownership. You'll have peers to pair with and review your work, and you'll review theirs. We review all code, we write things down, and we'd rather hear an honest "I don't know yet" than a confident wrong answer.

Our values
  • Trust: We do the right thing, especially when nobody is checking. Clients hand us regulated conversations with their own customers, and we earn that every day.

  • Transparency: We write things down, share the real numbers, and say "I don't know yet" out loud, with each other and with clients.

  • Technically best solution: We choose what's right, not what's easiest or trendiest. When we notice we got it wrong, we fix it.

  • Decisiveness: We decide quickly with the information we have, commit, and correct course fast when reality disagrees.

  • Simplicity: we keep systems, processes, and words simple. Complexity is a cost we pay only when it clearly buys something.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Agent Engineer
Agent Engineer

Rifa AI • San Francisco (CA)

On-site
USD 120,000 - 180,000
GTM Lead
GTM Lead

Rifa AI • Miami (FL)

On-site
USD 150,000 - 210,000
Senior Backend Engineer – Agents (USA Only - 100% Remote)
Senior Backend Engineer – Agents (USA Only - 100% Remote)

close • United States

Remote
USD 140,000 - 190,000
Senior AI Engineer — Voice
Senior AI Engineer — Voice

Mira Mace • San Francisco (CA)

On-site
USD 140,000 - 200,000
Software Engineering Superbuilder, AI-DNA, $200k/year USD
Software Engineering Superbuilder, AI-DNA, $200k/year USD

IgniteTech • United States

On-site
USD 120,000 - 160,000
Staff Engineer
Staff Engineer

Mira Mace • San Francisco (CA)

On-site
USD 180,000 - 240,000
Head of AI Engineering (AIOS)
Head of AI Engineering (AIOS)

Fella Health • United States

On-site
USD 180,000 - 240,000
Work coaching
Equipment provided (Macbook)
Health coaching
+4
Founding Product Engineer
Founding Product Engineer

Worky • San Francisco (CA)

On-site
USD 200,000 - 250,000
Health insurance
Pension/retirement plan
Generous paid time off
+1
Senior Site Reliability Engineer, AI-DNA, $100k/year USD
Senior Site Reliability Engineer, AI-DNA, $100k/year USD

IgniteTech • United States

On-site
USD 120,000 - 160,000
Founding Design Engineer
Founding Design Engineer

Worky • San Francisco (CA)

On-site
USD 120,000 - 180,000