Senior Backend Engineer

flexai

Santa Clara (CA)

On-site

USD 140,000 - 190,000

Full time

4 days ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

FlexAI is seeking a Senior Backend Engineer (Infrastructure & AI Platform) with deep Golang expertise to architect and build core backend systems powering our AI compute and PaaS platform. You will lead architecture and scale services for AI workloads across cloud and on-prem environments in a high-ownership, deep-tech setting.

The role sits in-person at our San Jose, CA office, with collaboration across AI/ML, Runtime, and Infra teams.

Qualifications

  • 5+ years of Backend or Infrastructure Engineering experience.
  • Expert-level proficiency in Golang (must-have, heavy hands-on).
  • Strong experience building production-grade distributed systems.
  • Proven track record on infrastructure platforms, PaaS, or deep-tech systems.

Responsibilities

  • Architect and develop high-performance Golang services for FlexAI's AI PaaS and infrastructure platform.
  • Build internal APIs powering model deployment, job scheduling, and compute lifecycle management.
  • Develop components interfacing with GPU/compute infrastructure and AI runtimes.
  • Drive reliability, observability, and resilience across services.

Skills

Golang
Distributed Systems
Backend Engineering
Cloud Platforms
Performance & Scalability

Tools

Kubernetes
Docker
gRPC
OpenTelemetry
Prometheus
Grafana

Job description

About FlexAI

Build and Deploy AI the right way, anywhere.

The FlexAI Compute Infrastructure Platform provides an \"end-to-end AI compute layer\" for running and managing workloads across any cloud, any GPU, and any deployment model (public, hybrid, or on-prem). It brings together \"1-click simplicity\" for users with \"enterprise-grade orchestration, security, and automation\" under the hood.

Founded by Brijesh Tripathi, who bring experience from Nvidia, Apple, Tesla, Intel and Zoox, FlexAI is not just building a product – we’re shaping the future of AI. Our teams are strategically distributed across Silicon Valley and Bengaluru, united by a shared mission: to deliver more compute with less complexity.

If you're passionate about shaping the future of artificial intelligence, driving innovation, and contributing to a sustainable and inclusive AI ecosystem, FlexAI is the place for you !

Role Overview

FlexAI is looking for a Senior Backend Engineer (Infrastructure & AI Platform) with deep Golang expertise to architect and build the core backend systems powering our next-generation AI compute and PaaS platform. This role sits at the intersection of distributed systems, cloud infrastructure, and AI platform engineering — enabling large-scale model training, inference, and orchestration across heterogeneous compute. This is not a traditional backend role; you will be building platform-grade systems that support AI runtimes, scheduling, resource orchestration, and multi-tenant cloud infrastructure.

As a Senior Backend Engineer, you'll drive backend architecture, scale platform services, and build high-performance infrastructure components that power AI workloads in production environments — influencing how the platform evolves from Beta to enterprise-grade deployment. Expect high ownership and technical autonomy in a research-driven, deep-tech environment — not SaaS CRUD apps.

This position is In-Person and located at our San Jose, CA Office.

What You'll Do
Core Platform & Infrastructure Backend
  • Architect and develop high-performance Golang services for FlexAI's AI PaaS and infrastructure platform
  • Build internal APIs powering model deployment, job scheduling, and compute lifecycle management
  • Develop components interfacing with GPU/compute infrastructure and AI runtimes
Distributed Systems & Scalability
  • Design and scale microservices and event-driven systems for high-throughput AI workloads
  • Optimize for low latency, high concurrency, and fault tolerance
  • Implement service-to-service communication (gRPC/REST, message queues, async pipelines)
  • Drive reliability, observability, and resilience across services
AI Platform Integration
  • Collaborate with AI/ML and Runtime teams to integrate systems with training pipelines, inference infrastructure, experimentation workflows, and dataset/artifact management
  • Enable orchestration across cloud and on-prem environments
  • Build abstractions that simplify AI infrastructure consumption
Cloud-Native & Platform Engineering
  • Design cloud-native, Kubernetes-native services
  • Work with DevOps/SRE on CI/CD, deployment automation, and scalability
  • Contribute to architecture decisions for multi-region, multi-cloud infrastructure
  • Improve monitoring, logging, and diagnostics
Technical Leadership
  • Lead architecture reviews and set engineering standards
  • Mentor engineers and guide complex problem-solving
  • Drive long-term roadmap for backend infrastructure and AI platform capabilities
  • Partner with Product, Runtime, and Infra leadership to translate requirements into scalable systems
Tech Stack (Indicative)
  • Languages: Golang (Primary), Python (Secondary)
  • Infrastructure: Kubernetes, Docker, Cloud (AWS/GCP/Azure)
  • Architecture: Microservices, gRPC, Event-driven systems
  • Data: SQL + NoSQL databases, caching, streaming systems
  • Observability: Prometheus, Grafana, OpenTelemetry (or similar)
What You'll Need to Be Successful
Core Engineering
  • 5+ years of Backend or Infrastructure Engineering experience
  • Expert-level proficiency in Golang (must-have, heavy hands-on)
  • Strong experience building production-grade distributed systems
  • Proven track record on infrastructure platforms, PaaS, or deep-tech systems
Infrastructure & Systems
  • Deep understanding of cloud-native architectures and containerized environments
  • Strong experience with Kubernetes, Docker, and cluster orchestration
  • Familiarity with compute scheduling, resource management, or platform runtimes is a strong plus
Databases & Data Systems
  • Experience with distributed databases (PostgreSQL, Cassandra, DynamoDB, etc.)
  • Strong understanding of caching, queues, and streaming systems (Redis, Kafka, etc.)
AI / Platform Exposure (Highly Preferred)
  • Experience on AI/ML platforms, model infrastructure, or data platforms
  • Familiarity with ML pipelines, inference systems, or GPU-backed workloads
  • Exposure to PyTorch, TensorFlow infrastructure, or model serving systems is a plus
Ideal Candidate Profile (Who Will Thrive Here)
  • Infra-first backend engineers (not just API developers)
  • Background in AI infra, cloud platforms, developer platforms, or deep-tech systems
  • Strong systems thinkers who enjoy low-level performance, scalability, and architecture challenges
  • Startup-minded builders comfortable in ambiguous, high-ownership environments
What We Offer
  • Competitive salary and benefits package
  • Work on cutting-edge AI infrastructure
  • Build products used by developers and enterprises
  • High ownership, fast execution, real impact
  • Collaborative, high-caliber team
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Full Stack Engineer
Senior Full Stack Engineer

flexai • Santa Clara (CA)

On-site
USD 140,000 - 210,000
Competitive salary
Cutting-edge AI infrastructure
High ownership & impact
Senior Backend Engineer for AI Platform & Infra (Golang)
Senior Backend Engineer for AI Platform & Infra (Golang)

flexai • Santa Clara (CA)

On-site
USD 140,000 - 190,000
Software Engineer - AI & Cloud Engineering (Early to Mid Career)
Software Engineer - AI & Cloud Engineering (Early to Mid Career)

Flexcompute • Watertown (MA)

On-site
USD 110,000 - 160,000
Competitive salary
Meaningful equity of startup
401K contribution
+1
Software Engineer - AI & Cloud Engineering (Early to Mid Career, WI Based)
Software Engineer - AI & Cloud Engineering (Early to Mid Career, WI Based)

Flexcompute, inc. • Madison (WI)

On-site
USD 80,000 - 110,000
Competitive salary
Meaningful equity of early-stage startup
401K contribution
+1
Presales Engineer (AI Workload/ Cloud Infrastructure)
Presales Engineer (AI Workload/ Cloud Infrastructure)

FlexAI • Santa Clara (CA)

On-site
USD 120,000 - 160,000
Competitive Compensation
Performance-Based Incentives
Comprehensive Benefits
+2
Software Engineer - AI & Cloud Engineering (Early to Mid Career
Software Engineer - AI & Cloud Engineering (Early to Mid Career

Flexcompute • Watertown (MA)

On-site
USD 100,000 - 150,000
Competitive salary
Meaningful equity of early-stage start
401K contribution
+1
Staff Product Manager AI
Staff Product Manager AI

FlexAI • New York (NY)

On-site
USD 100,000 - 180,000
Competitive salary
Continuous growth opportunities
Support for personal and professional development
Senior Backend Engineer | AI Infrastructure & Distributed Systems
Senior Backend Engineer | AI Infrastructure & Distributed Systems

The Cypress Group • New York (NY)

On-site
USD 160,000 - 240,000
Member of Technical Staff - Compute Platform
Member of Technical Staff - Compute Platform

Primeintellect • San Francisco (CA)

On-site
USD 150,000 - 230,000
Software Engineer (Backend-Focused) $140,000 - $225,000 Posted 12 hours ago
Software Engineer (Backend-Focused) $140,000 - $225,000 Posted 12 hours ago

Fuel Talent LLC • Northern (KY)

On-site
USD 140,000 - 225,000