Software Engineer II - Serverless Inference

Jobtailor

Bengaluru

On-site

INR 2,400,000 - 4,200,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

DigitalOcean in Bengaluru is seeking a Software Engineer to design and optimize serverless inference infrastructure and APIs, focusing on throughput, GPU utilization, and fault tolerance for AI workloads. You will build scalable, multi-tenant services and operate high-scale distributed systems in a production environment.

You will collaborate with platform, GPU, and product teams to deliver production-grade systems and highly available APIs, while advancing observability, reliability, and

Qualifications

  • 2+ years of experience building and operating multi-tenant platforms or distributed backend systems.
  • Strong experience operating high-scale distributed services in production environments.
  • Deep understanding of SRE principles, including observability, incident management, reliability engineering, capacity planning, and operational automation.
  • 1+ years of hands-on experience with Go / Golang in production systems.
  • 1+ years of experience with Kubernetes.
  • Strong understanding of cloud-native architectures, microservices, and distributed systems fundamentals.
  • Experience debugging performance, scalability, and reliability issues in production systems.
  • Observability Proficiency: Experience tracking infrastructure and inference metrics like TTFT, TPOT, and GPU utilization.

Responsibilities

  • Design and build scalable, multi-tenant services that power AI inference and intelligent routing workloads.
  • Develop and operate high-scale distributed systems with strong reliability, availability, and performance goals.
  • Strengthen platform resiliency through improved observability, capacity management, automation, and operational tooling.
  • Partner closely with platform, GPU infrastructure, and product engineering teams to deliver production-grade systems and highly available APIs.
  • Raise the engineering bar through strong software design, operational discipline, incident management, and continuous improvement practices.
  • Contribute to architecture decisions around traffic management, service orchestration, reliability, and platform scalability.
  • Participate in on-call rotations and lead efforts to reduce operator pain, improve service health, and prevent recurring incidents.

Skills

Go
Kubernetes
Distributed systems
Observability
SRE
Multi-tenant platforms
Incident management

Job description

Dive in and do the best work of your career at DigitalOcean. Journey alongside a strong community of top talent who are relentless in their drive to build the simplest scalable cloud. If you have a growth mindset, naturally like to think big and bold, and are energized by the fast-paced environment of a true industry disruptor, you ll find your place here. We value winning together while learning, having fun, and making a profound difference for the dreamers and builders in the world.

We are seeking a Software Engineer to implement and contribute to the design and optimization of our Serverless Inference infrastructure and APIs. In this role, you will tackle the challenges of large-scale AI workloads, focusing on throughput, GPU utilization, and fault tolerance to support next-generation inference needs of AI native enterprises.

What Youll Do:
  • Design and build scalable, multi-tenant services that power AI inference and intelligent routing workloads.
  • Develop and operate high-scale distributed systems with strong reliability, availability, and performance goals.
  • Strengthen platform resiliency through improved observability, capacity management, automation, and operational tooling.
  • Partner closely with platform, GPU infrastructure, and product engineering teams to deliver production-grade systems and highly available APIs.
  • Raise the engineering bar through strong software design, operational discipline, incident management, and continuous improvement practices.
  • Contribute to architecture decisions around traffic management, service orchestration, reliability, and platform scalability.
  • Participate in on-call rotations and lead efforts to reduce operator pain, improve service health, and prevent recurring incidents.
What You ll Add to DigitalOcean:
  • 2+ years of experience building and operating multi-tenant platforms or distributed backend systems
  • Strong experience operating high-scale distributed services in production environments
  • Deep understanding of SRE principles, including observability, incident management, reliability engineering, capacity planning, and operational automation
  • 1+ years of hands-on experience with Go / Golang in production systems
  • 1+ years of experience with Kubernetes
  • Strong understanding of cloud-native architectures, microservices, and distributed systems fundamentals
  • Experience debugging performance, scalability, and reliability issues in production systems
  • Observability Proficiency: Experience tracking infrastructure and inference metrics like Time To First Token (TTFT), Time Per Output Token (TPOT), and GPU utilization.
Nice to Have
  • AI/ML Framework Knowledge: Understanding of modern LLM serving architectures and familiarity with engines like vLLM or Triton.
  • Experience with API gateways, traffic routing, or service mesh technologies
  • Familiarity with LLM serving stacks such as vLLM, TensorRT-LLM, or similar technologies
  • Experience building systems for inference optimization, rate limiting, routing, or workload orchestration.

This is a hybrid role based out of Bengaluru, India.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Forward Deployed Engineer I (AI Inference)
Senior Forward Deployed Engineer I (AI Inference)

DigitalOcean • Bengaluru

On-site
INR 4,500,000 - 7,000,000
Senior Software Engineer I (Full Stack)
Senior Software Engineer I (Full Stack)

DigitalOcean • Hyderabad

On-site
INR 3,000,000 - 5,000,000
Senior Forward Deployed Engineer I (AI Infra)
Senior Forward Deployed Engineer I (AI Infra)

DigitalOcean • Bengaluru

On-site
INR 3,000,000 - 5,400,000
Senior Software Engineer I, AI/ML
Senior Software Engineer I, AI/ML

DigitalOcean • Bengaluru

On-site
INR 4,000,000 - 7,000,000
Senior Director Engineering - Serverless Compute Engineering
Senior Director Engineering - Serverless Compute Engineering

DigitalOcean • Bengaluru

Hybrid
INR 4,200,000 - 7,000,000
Equity compensation
Bonus eligible
Staff Software Engineer-AI
Staff Software Engineer-AI

DigitalOcean • Hyderabad

Hybrid
INR 4,000,000 - 7,000,000
Equity compensation
Employee Stock Purchase Program
Senior Director, Engineering - Serverless Compute Engineering
Senior Director, Engineering - Serverless Compute Engineering

DigitalOcean • Bengaluru

On-site
INR 4,000,000 - 6,000,000
Inference Engineer
Inference Engineer

Binaire Private Limited • New Delhi

On-site
INR 800,000 - 1,200,000
Senior Inference Engineer
Senior Inference Engineer

Binaire Private Limited • New Delhi

On-site
INR 2,600,000 - 4,800,000
Distributed Training & Inference Optimization Engineer
Distributed Training & Inference Optimization Engineer

Winzons • India

On-site
INR 3,000,000 - 5,000,000