Software Engineer, Kubernetes/Cloud

SB Telecom America Corp.

Sunnyvale (CA)

On-site

USD 120,000 - 180,000

Full time

14 days+
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

SoftBank invites experienced practitioners to join the Infrinia.ai team in Silicon Valley, building the next-generation cloud-native platform for AI. You will extend Kubernetes with Operators and Controllers to manage AI workloads on state-of-the-art GPU infrastructure.

Requirements include a Bachelor's in CS/EE, 3+ years in distributed systems, Go or C++, Kubernetes experience, and willingness to contribute to cloud orchestration and bare-metal provisioning topics.

Qualifications

  • Bachelor's degree in CS, EE, or related field.
  • 3+ years of software development with a focus on distributed systems/cloud infra.
  • Strong proficiency in Go (Golang) or C/C++.
  • Experience building and extending Kubernetes (custom controllers, Operators, CRDs, or API aggregation).

Responsibilities

  • Design and implement high-performance, production-quality code in Go to build custom Kubernetes Operators and Controllers.
  • Develop software APIs and microservices that abstract complex infrastructure primitives for AI users.
  • Architect and build features for multi-cluster management, scheduling optimization, and resource isolation.
  • Collaborate with kernel/system engineers to expose hardware capabilities (GPU, NICs) up to the orchestration layer.
  • Write comprehensive unit and integration tests to ensure software reliability and stability.
  • Contribute to Product Definition (PRD) and program execution (sprint) planning.
  • Role model and foster a culture of humility and innovation for product delivery.

Skills

Go (Golang)
C/C++
Distributed systems
Algorithms & data structures
Kubernetes knowledge

Education

Bachelor's degree in CS/EE
Master's or PhD in a relevant field

Tools

Kubernetes (Operators, Controllers, CRDs)
Cluster API
Tinkerbell
container runtimes (containerd, CRI-O)

Job description

About Infrinia.ai, powered by SoftBank: SoftBank is making significant investments in infrastructure for AI. Through its wholly owned US subsidiary, SoftBank Corp. has established Infrinia team in Silicon Valley, focused on infrastructure software for AI and AI foundations for mobile networks. Our goals are to challenge the norms and create products making use of our SOTA infrastructure (like Nvidia GB200, MGX and DGX Grace & Hopper platforms) and cloud-native software. These products are geared towards centralized AI data centers as well as distributed AI Radio Access Network (AI RAN) data centers. We are looking for experienced practitioners who are inspired to bring innovation and build transformative products.

Minimum Qualifications:

  • Bachelor's degree in Computer Science, Electrical Engineering, or related field.
  • 3+ years of software development experience with a focus on distributed systems or cloud infrastructure.
  • Strong proficiency in Go (Golang) or C/C++.
  • Deep understanding of data structures, algorithms, and software design patterns.
  • Experience building and extending Kubernetes (e.g., custom controllers, Operators, CRDs, or API aggregation).

Preferred Qualifications:

  • Master's or PhD in a relevant field.
  • Experience contributing to the Kubernetes open-source ecosystem (k8s upstream, CNCF projects).
  • Deep knowledge of Kubernetes internals (scheduler, kubelet, etcd, networking).
  • Experience building software for bare-metal provisioning or cloud orchestration (Cluster API, Tinkerbell, etc.).
  • Familiarity with OCI standards, container runtimes (containerd, CRI-O), and Linux kernel features (eBPF, cgroups etc is a plus).

Role: Be a key member of the software engineering team responsible for building the next-generation cloud-native platform for AI. You will write code to extend and customize Kubernetes, enabling it to orchestrate massive AI workloads on our SOTA GPU infrastructure. You will move beyond simply using Kubernetes APIs to designing and implementing the software logic (Operators, Controllers) that automates the lifecycle of our compute, network, and storage resources.

Responsibilities:

  • Design and implement high-performance, production-quality code in Go to build custom Kubernetes Operators and Controllers.
  • Develop software APIs and microservices that abstract complex infrastructure primitives for AI users.
  • Architect and build features for multi-cluster management, scheduling optimization, and resource isolation.
  • Collaborate with kernel/system engineers to expose hardware capabilities (GPU, NICs) up to the orchestration layer.
  • Write comprehensive unit and integration tests to ensure software reliability and stability.
  • Contribute to Product Definition (PRD) and program execution (sprint) planning.
  • Role model and foster a culture of humility and innovation for product delivery.

Salary: The base salary for this position ranges from ($120,000-$180,000), with additional attractive biannual bonus and benefits.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AI Cloud-Native Kubernetes Engineer
AI Cloud-Native Kubernetes Engineer

SB Telecom America Corp. • Sunnyvale (CA)

On-site
USD 120,000 - 180,000
Senior Software Engineer, Infrastructure Software for AI
Senior Software Engineer, Infrastructure Software for AI

SB Telecom America Corp. • Sunnyvale (CA)

On-site
USD 150,000 - 250,000
Senior AI Infra Engineer - GPU & Kubernetes
Senior AI Infra Engineer - GPU & Kubernetes

SB Telecom America Corp. • Sunnyvale (CA)

On-site
USD 150,000 - 250,000
AI Infra Engineer – SRE (Kubernetes)
AI Infra Engineer – SRE (Kubernetes)

Berrybytes • United States

On-site
USD 110,000 - 150,000
Member of Technical Staff (AI Infrastructure Engineer)
Member of Technical Staff (AI Infrastructure Engineer)

Pantera Capital • Palo Alto (CA)

Hybrid
USD 190,000 - 250,000
Comprehensive health insurance
Dental and vision insurance
401(k) plan
+1
Member of Technical Staff, Platform - AI Infrastructure
Member of Technical Staff, Platform - AI Infrastructure

Hamilton Barnes Associates Limited • United States

On-site
USD 213,000 - 288,000
Equity
Health care
Senior Software Engineer, Infrastructure Software for AI (Centralized AI Data Centers & Distrib[...]
Senior Software Engineer, Infrastructure Software for AI (Centralized AI Data Centers & Distrib[...]

Intelliswift - An LTTS Company • Sunnyvale (CA)

On-site
USD 120,000 - 150,000
Competitive salary
Health insurance
Flexible work hours
Software Engineer – Cloud Infrastructure
Software Engineer – Cloud Infrastructure

FriendliAI • San Francisco (CA)

On-site
USD 150,000 - 190,000
Flexible working hours
Lunch and dinner provided
Health check-up support with top-tier硬
+1
Software Engineer, AI Infra
Software Engineer, AI Infra

Makers Fund • Palo Alto (CA)

On-site
USD 180,000 - 240,000
Equity
Health benefits
Monthly stipends
+1
Staff Software Engineer - Managed Kubernetes
Staff Software Engineer - Managed Kubernetes

Cloudjobs • San Jose (CA)

On-site
USD 230,000 - 360,000
Cash compensation
Equity compensation
Health coverage
+6