An application made for this job — a tailored resume and cover letter that speak straight to the posting.
VideoSDK is a developer-first real-time communication platform building real-time audio and voice AI infrastructure. The Voice AI Infrastructure Engineer will implement services that process audio across ASR, LLM, and TTS stages, and ensure sub-second latency in a production environment.
You will work with Go, Rust, or Python on the real-time audio path, deploy on Kubernetes, and use Redis for state and signaling, while instrumenting performance with Prometheus, Grafana, and OpenTelemetry.
Company: VideoSDK
Website: Visit Website
LinkedIn: Visit LinkedIn
Business Type: Startup
Company Type: Product
Business Model: B2B
Funding Stage: Seed
Industry: Information Technology
Position: Voice AI Infrastructure Engineer
Location: Surat, Gujarat (Onsite)
Education: BE/BTech in CS/IT.
VideoSDK is a developer-first real-time communication platform building infrastructure for real-time video, audio, live streaming, and AI voice applications. We provide low-latency, scalable APIs and SDKs that power voice AI agents, conversational experiences, and real-time communication at scale, solving complex problems across real-time audio, WebRTC, media infrastructure, distributed systems, and cloud scalability.
We build voice agents that talk to people in real time. A phone call has no loading spinner — every millisecond of latency is audible. You'll help build and run the infrastructure that moves audio through speech recognition, an LLM, and speech synthesis, and back again, inside a sub-second budget. This is an infrastructure role, not a modeling role.