Senior Backend Engineer

FlexAI

San Jose (CA)

On-site

USD 180,000 - 240,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Competitive salary
Cutting-edge AI infra
High ownership
Collaborative team
Developer tools access

Job summary

FlexAI in San Jose, CA is seeking a Senior Backend Engineer (Infrastructure & AI Platform) with deep Golang expertise to architect and build core backend systems powering our AI compute and PaaS platform. This in-person role focuses on distributed systems, cloud infrastructure, and AI platform engineering for large-scale model training, inference, and orchestration across heterogeneous compute.

You will drive backend architecture, scale platform services, and create high-performance

Qualifications

  • 5+ years of Backend or Infrastructure Engineering experience.
  • Expert-level proficiency in Golang (must-have).
  • Strong experience building production-grade distributed systems.
  • Experience with infrastructure platforms, PaaS, or deep-tech systems.
  • Familiarity with cloud-native architectures and containerized environments.

Responsibilities

  • Architect and develop high-performance Golang services for FlexAI’s AI PaaS and infrastructure platform.
  • Build internal APIs for model deployment, job scheduling, and compute lifecycle management.
  • Develop components interfacing with GPU/compute infrastructure and AI runtimes.
  • Design microservices and event-driven systems for high-throughput AI workloads.
  • Drive reliability, observability, and resilience across services.

Skills

Golang
Python
Distributed systems
Cloud platforms
Performance optimization

Tools

Kubernetes
Docker
PostgreSQL
Redis
Kafka
gRPC
OpenTelemetry
Prometheus
Grafana

Job description

Role Overview

FlexAI is looking for a Senior Backend Engineer (Infrastructure & AI Platform) with deep Golang expertise to architect and build the core backend systems powering our next-generation AI compute and PaaS platform. This role sits at the intersection of distributed systems, cloud infrastructure, and AI platform engineering — enabling large-scale model training, inference, and orchestration across heterogeneous compute. This is not a traditional backend role; you will be building platform-grade systems that support AI runtimes, scheduling, resource orchestration, and multi-tenant cloud infrastructure.


As a Senior Backend Engineer , you’ll drive backend architecture, scale platform services, and build high-performance infrastructure components that power AI workloads in production environments — influencing how the platform evolves from Beta to enterprise-grade deployment. Expect high ownership and technical autonomy in a research-driven, deep-tech environment — not SaaS CRUD apps.


This position is In-Person and located at our San Jose, CA Office.


What You’ll Do

Core Platform & Infrastructure Backend:



  • Architect and develop high-performance Golang services for FlexAI’s AI PaaS and infrastructure platform

  • Build internal APIs powering model deployment, job scheduling, and compute lifecycle management

  • Develop components interfacing with GPU/compute infrastructure and AI runtimes


Distributed Systems & Scalability:


  • Design and scale microservices and event-driven systems for high-throughput AI workloads

  • Optimize for low latency, high concurrency, and fault tolerance

  • Implement service-to-service communication (gRPC/REST, message queues, async pipelines)

  • Drive reliability, observability, and resilience across services


AI Platform Integration:


  • Collaborate with AI/ML and Runtime teams to integrate systems with training pipelines, inference infrastructure, experimentation workflows, and dataset/artifact management

  • Enable orchestration across cloud and on-prem environments

  • Build abstractions that simplify AI infrastructure consumption


Cloud-Native & Platform Engineering:


  • Design cloud-native, Kubernetes-native services

  • Work with DevOps/SRE on CI/CD, deployment automation, and scalability

  • Contribute to architecture decisions for multi-region, multi-cloud infrastructure

  • Improve monitoring, logging, and diagnostics


Technical Leadership:


  • Lead architecture reviews and set engineering standards

  • Mentor engineers and guide complex problem-solving

  • Drive long‑term roadmap for backend infrastructure and AI platform capabilities

  • Partner with Product, Runtime, and Infra leadership to translate requirements into scalable systems


Tech Stack (Indicative):


  • Languages: Golang (Primary), Python (Secondary)

  • Infrastructure: Kubernetes, Docker, Cloud (AWS/GCP/Azure)

  • Architecture: Microservices, gRPC, Event-driven systems

  • Data: SQL + NoSQL databases, caching, streaming systems

  • Observability: Prometheus, Grafana, OpenTelemetry (or similar)


What You’ll Need to Be Successful

Core Engineering:


  • 5+ years of Backend or Infrastructure Engineering experience

  • Expert‑level proficiency in Golang (must‑have, heavy hands‑on)

  • Strong experience building production‑grade distributed systems

  • Proven track record on infrastructure platforms, PaaS, or deep‑tech systems


Infrastructure & Systems:


  • Deep understanding of cloud‑native architectures and containerized environments

  • Strong experience with Kubernetes, Docker, and cluster orchestration

  • Familiarity with compute scheduling, resource management, or platform runtimes is a strong plus


Databases & Data Systems:


  • Experience with distributed databases (PostgreSQL, Cassandra, DynamoDB, etc.)

  • Strong understanding of caching, queues, and streaming systems (Redis, Kafka, etc.)


AI / Platform Exposure (Highly Preferred):


  • Experience on AI/ML platforms, model infrastructure, or data platforms

  • Familiarity with ML pipelines, inference systems, or GPU‑backed workloads

  • Exposure to PyTorch, TensorFlow infrastructure, or model serving systems is a plus


Ideal Candidate Profile (Who Will Thrive Here)



  • Infra‑first backend engineers (not just API developers)

  • Background in AI infra, cloud platforms, developer platforms, or deep‑tech systems

  • Strong systems thinkers who enjoy low‑level performance, scalability, and architecture challenges

  • Startup‑minded builders comfortable in ambiguous, high‑ownership environments


What We Offer


  • Competitive salary and benefits package

  • Work on cutting‑edge AI infrastructure

  • Build products used by developers and enterprises

  • High ownership, fast execution, real impact

  • Collaborative, high-caliber team

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Backend Engineer
Senior Backend Engineer

FlexAI • Santa Clara (CA)

On-site
USD 120,000 - 150,000
Competitive salary and benefits
Work on cutting-edge AI infrastructure
High ownership and fast execution
Senior Backend Infra Architect - AI Platform (Golang)
Senior Backend Infra Architect - AI Platform (Golang)

FlexAI • San Jose (CA)

On-site
USD 180,000 - 240,000
Competitive salary
Cutting-edge AI infra
High ownership
+2
Senior Member of Technical Staff
Senior Member of Technical Staff

DeepRec.ai • Palo Alto (CA)

Hybrid
USD 180,000 - 240,000
Competitive salary + meaningful equity
Flexible hybrid working model
High-performing, collaborative team environment
Staff AI Runtime Engineer
Staff AI Runtime Engineer

FlexAI • Santa Clara (CA)

On-site
USD 180,000 - 240,000
Competitive salary and benefits package
Work on cutting-edge AI infrastructure
High ownership and fast execution
Senior Full Stack Engineer
Senior Full Stack Engineer

FlexAI • Santa Clara (CA)

On-site
USD 120,000 - 160,000
Competitive salary and benefits package
Opportunity to work on cutting-edge AI infrastructure
High ownership and real impact
Senior / Lead / Principal Platform Engineer (DevOps / Cloud Infrastructure)
Senior / Lead / Principal Platform Engineer (DevOps / Cloud Infrastructure)

CB Smart Recruit • Los Angeles (CA)

On-site
USD 200,000 - 300,000
Competitive sign-on bonus
Comprehensive benefits package
Long-term career growth opportunities
Senior / Lead / Principal Platform Engineer
Senior / Lead / Principal Platform Engineer

CB Smart Recruit • Los Angeles (CA)

On-site
USD 200,000 - 300,000
Competitive sign-on bonus
Comprehensive benefits package
Opportunities for career growth in a high-growth AI company
Software Engineer, Platform
Software Engineer, Platform

Recruiting from Scratch • San Francisco (CA)

On-site
USD 140,000 - 210,000
Bi-annual bonuses
Relocation assistance
Housing stipend
+5
Senior Software Engineer
Senior Software Engineer

Xcede • San Francisco (CA)

On-site
USD 180,000 - 260,000
Senior Backend Engineer (Go)
Senior Backend Engineer (Go)

CloudFactory • Romania (PA)

Hybrid
USD 90,000 - 120,000