Principal Software Engineer, AI Inference Cloud

Arm Limited

Seattle (WA)

Hybrid

USD 263,000 - 355,000

Full time

3 days ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

Arm Limited is seeking a Principal Engineer for the AI Inference Cloud team to steer architecture and deliver scalable AI inference services. You will lead Kubernetes-driven workloads, ensure reliability, and influence across AI compute, Inference Runtime, and product teams.

You will mentor engineers, drive cross-team alignment, and contribute to the design and rollout of production-grade cloud capabilities for AI workloads in a hybrid environment.

Qualifications

  • 8+ years of experience in distributed systems, cloud platforms, or production infrastructure.
  • Deep production experience with Kubernetes, including controllers, operators, scheduling, networking, and lifecycle management.
  • Strong software and production engineering skills in Go, C++, Rust, or Python.
  • Track record leading complex technical initiatives while remaining hands-on.
  • Ability to troubleshoot complex systems and communicate across teams.

Responsibilities

  • Define and build the architecture for cloud-based AI inference services.
  • Develop Kubernetes controllers and platform capabilities for workload deployment, scheduling, recovery, scaling, upgrades, and lifecycle management.
  • Establish production practices for health validation, progressive rollout, rollback, observability, and service objectives.
  • Lead production readiness reviews and resolve issues across services, Kubernetes, networking, and compute infrastructure.
  • Lead design and build reviews, mentor engineers, and drive technical alignment across teams.

Skills

Kubernetes
Go
C++
Rust
Python
Distributed systems
Leadership
Observability

Tools

Kubernetes

Job description

As a Principal Engineer on Arm’s AI Inference Cloud team, you will shape the technical direction and develop highly available, scalable services for running AI inference workloads. You will guide architecture and actively contribute to development across Kubernetes orchestration, workload management, service delivery, and observability. Partnering with AI compute, Inference Runtime, and product teams to enhance the performance and usability of Arm’s AI platform.

Responsibilities:
  • Define and build the architecture for cloud-based AI inference services.
  • Develop Kubernetes controllers and platform capabilities supporting workload deployment, scheduling, recovery, scaling, upgrades, and lifecycle management.
  • Establish production practices for health validation, progressive rollout, rollback, observability, and service objectives. Improve platform reliability, scalability, performance, and resource efficiency.
  • Lead production readiness reviews and resolve complex issues across services, Kubernetes, networking, and compute infrastructure. Turn incidents and operational bottlenecks into durable platform improvements.
  • Lead design and build reviews, mentor engineers, and drive technical alignment across teams.
Necessary Skills and Experience:
  • 8+ years of experience, or equivalent proven impact, building distributed systems, cloud platforms, or production infrastructure.
  • Deep production experience with Kubernetes, including controllers, operators, scheduling, resource management, networking, and workload lifecycle management.
  • Strong software and production engineering skills, including hands‑on programming in Go, C++, Rust, Python, or a similar language, and experience with reliable services, APIs, concurrency, observability, deployment safety, capacity planning, and incident response.
  • A track record of leading complex technical initiatives while remaining hands‑on, including architecture, implementation, debugging, mentoring, and influencing technical direction across teams.
  • Ability to troubleshoot complex systems and communicate clearly with engineers from different technical backgrounds.
Preferred Skills and Experience:
  • Experience with AI infrastructure, model serving, or accelerator-backed workloads.
  • Familiarity with frameworks such as PyTorch, Ray, vLLM, SGLang, or TensorRT-LLM, or experience qualifying accelerators and tuning distributed workloads.
  • Knowledge of inference performance and resource‑efficiency considerations.
In Return:

You will be part of our AI Platforms team - A driven and diverse group passionate about developing foundational production capabilities for AI inference at Arm. We provide a collaborative setting where your ideas can come to life quickly. Your work will directly impact the success of our AI projects, shape Arm’s AI inference capabilities, defining and operating production inference workloads. This is an outstanding opportunity to work with world‑class teams and contribute to groundbreaking advances in AI technology. Join us in building the next generation of AI inference infrastructure!

Additional Information:

Please note that a relocation package (including visa sponsorship support) is available for this role, for candidates who require it.

Salary Range:

$262,700-$355,400 per year

We value people as individuals and our dedication is to reward people competitively and equitably for the work they do and the skills and experience they bring to Arm. Salary is only one component of Arm's offering. The total reward package will be shared with candidates during the recruitment and selection process.

Accommodations at Arm

At Arm, we want to build extraordinary teams. If you need an adjustment or an accommodation during the recruitment process, please email accommodations@arm.com. To note, by sending us the requested information, you consent to its use by Arm to arrange for appropriate accommodations. All accommodation or adjustment requests will be treated with confidentiality, and information concerning these requests will only be disclosed as necessary to provide the accommodation. Although this is not an exhaustive list, examples of support include breaks between interviews, having documents read aloud, or office accessibility. Please email us about anything we can do to accommodate you during the recruitment process.

Hybrid Working at Arm

Arm’s approach to hybrid working is designed to create a working environment that supports both high performance and personal wellbeing. We believe in bringing people together face to face to enable us to work at pace, whilst recognizing the value of flexibility. Within that framework, we empower groups/teams to determine their own hybrid working patterns, depending on the work and the team’s needs. Details of what this means for each role will be shared upon application. In some cases, the flexibility we can offer is limited by local legal, regulatory, tax, or other considerations, and where this is the case, we will collaborate with you to find the best solution. Please talk to us to find out more about what this could look like for you.

Equal Opportunities at Arm

Arm is an equal opportunity employer, committed to providing an environment of mutual respect where equal opportunities are available to all applicants and colleagues. We are a diverse organization of dedicated and innovative individuals, and don’t discriminate on the basis of race, color, religion, sex, sexual orientation, gender identity, national origin, disability, or status as a protected veteran.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Staff Software Engineer, AI Inference Cloud
Staff Software Engineer, AI Inference Cloud

Arm Limited • Seattle (WA)

Hybrid
USD 209,000 - 283,000
Relocation package
Visa sponsorship
Principal Software Engineer, AI Inference Runtime
Principal Software Engineer, AI Inference Runtime

Arm Limited • Seattle (WA)

Hybrid
USD 263,000 - 355,000
Hybrid working
Accommodations during recruitment
Principal Software Engineer, AI Compute Infrastructure
Principal Software Engineer, AI Compute Infrastructure

Arm Limited • Seattle (WA)

Hybrid
USD 263,000 - 355,000
Relocation package
Visa sponsorship
Staff Software Engineer, AI Inference Runtime
Staff Software Engineer, AI Inference Runtime

Arm Limited • Seattle (WA)

Hybrid
USD 209,000 - 283,000
Technical Program Director, AI
Technical Program Director, AI

Arm • San Jose (CA)

On-site
USD 264,000 - 359,000
Staff Robotics Engineer
Staff Robotics Engineer

Arm Limited • Seattle (WA)

On-site
USD 209,000 - 283,000
Relocation package
Principal Project Manger, AI & Developer Platforms
Principal Project Manger, AI & Developer Platforms

Arm Limited • Seattle (WA)

Hybrid
USD 214,000 - 290,000
Director, Segment Marketing - Cloud AI
Director, Segment Marketing - Cloud AI

Arm Limited • San Jose (CA)

Hybrid
USD 260,000 - 352,000
Senior Strategic Regional Sales Manager– Cloud AI & Datacenter
Senior Strategic Regional Sales Manager– Cloud AI & Datacenter

Arm • San Jose (CA)

On-site
USD 185,000 - 250,000
Principal Forward Deployed Engineer, Performance
Principal Forward Deployed Engineer, Performance

Arm Limited • Austin (TX)

Hybrid
USD 208,000 - 281,000