Principal Software Engineer, AI Inference Cloud

Arm

Seattle (WA)

Hybrid

USD 263,000 - 355,000

Full time

30 hours ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Benefits offered by this job

Relocation package with visa Spons.&
Hybrid working options

Job summary

Arm is seeking a Principal Engineer for the AI Inference Cloud team to shape technical direction and build highly available, scalable AI inference services. You will guide architecture, develop Kubernetes Controllers, and drive platform reliability across multiple compute and networking domains.

The role requires 8+ years in distributed systems, strong Kubernetes experience, and hands-on programming in Go/C++/Rust/Python. Relocation/visa support is available.

Qualifications

  • 8+ years of experience building distributed systems, cloud platforms, or production infrastructure.
  • Deep production experience with Kubernetes controllers, operators, scheduling, and workload lifecycle management.
  • Strong software and production engineering skills in Go, C++, Rust, Python, APIs, observability, and incident response.
  • Track record of leading complex technical initiatives while staying hands-on and influencing direction.

Responsibilities

  • Define and build architecture for cloud-based AI inference services.
  • Develop Kubernetes controllers and platform capabilities for workload deployment, scheduling, and lifecycle management.
  • Establish production practices for health checks, progressive rollout, rollback, observability, and reliability improvements.
  • Lead production readiness reviews and resolve issues across services, Kubernetes, networking, and compute infrastructure.
  • Lead design and code reviews, mentor engineers, and drive technical alignment across teams.

Skills

Kubernetes
Go
C++
Rust
Python
Distributed systems
Leadership

Job description

As a Principal Engineer on Arm’s AI Inference Cloud team, you will shape the technical direction and develop highly available, scalable services for running AI inference workloads. You will guide architecture and actively contribute to development across Kubernetes orchestration, workload management, service delivery, and observability. Partnering with AI compute, Inference Runtime, and product teams to enhance the performance and usability of Arm’s AI platform.

Responsibilities
  • Define and build the architecture for cloud-based AI inference services.
  • Develop Kubernetes controllers and platform capabilities supporting workload deployment, scheduling, recovery, scaling, upgrades, and lifecycle management.
  • Establish production practices for health validation, progressive rollout, rollback, observability, and service objectives. Improve platform reliability, scalability, performance, and resource efficiency.
  • Lead production readiness reviews and resolve complex issues across services, Kubernetes, networking, and compute infrastructure. Turn incidents and operational bottlenecks into durable platform improvements.
  • Lead design and build reviews, mentor engineers, and drive technical alignment across teams.
Necessary Skills And Experience
  • 8+ years of experience, or equivalent proven impact, building distributed systems, cloud platforms, or production infrastructure.
  • Deep production experience with Kubernetes, including controllers, operators, scheduling, resource management, networking, and workload lifecycle management.
  • Strong software and production engineering skills, including hands-on programming in Go, C++, Rust, Python, or a similar language, and experience with reliable services, APIs, concurrency, observability, deployment safety, capacity planning, and incident response.
  • A track record of leading complex technical initiatives while remaining hands-on, including architecture, implementation, debugging, mentoring, and influencing technical direction across teams.
  • Ability to troubleshoot complex systems and communicate clearly with engineers from different technical backgrounds.
Preferred Skills And Experience
  • Experience with AI infrastructure, model serving, or accelerator-backed workloads.
  • Familiarity with frameworks such as PyTorch, Ray, vLLM, SGLang, or TensorRT-LLM, or experience qualifying accelerators and tuning distributed workloads.
  • Knowledge of inference performance and resource-efficiency considerations.
In Return

You will be part of our AI Platforms team - A driven and diverse group passionate about developing foundational production capabilities for AI inference at Arm. We provide a collaborative setting where your ideas can come to life quickly. Your work will directly impact the success of our AI projects, shape Arm’s AI inference capabilities, defining and operating production inference workloads. This is an outstanding opportunity to work with world-class teams and contribute to groundbreaking advances in AI technology. Join us in building the next generation of AI inference infrastructure!

Additional Information

Please note that a relocation package (including visa sponsorship support) is available for this role, for candidates who require it.

Salary Range

$262,700-$355,400 per year

We value people as individuals and our dedication is to reward people competitively and equitably for the work they do and the skills and experience they bring to Arm. Salary is only one component of Arm's offering. The total reward package will be shared with candidates during the recruitment and selection process.

Accommodations at Arm

At Arm, we want to build extraordinary teams. If you need an adjustment or an accommodation during the recruitment process, please email accommodations@arm.com . To note, by sending us the requested information, you consent to its use by Arm to arrange for appropriate accommodations. All accommodation or adjustment requests will be treated with confidentiality, and information concerning these requests will only be disclosed as necessary to provide the accommodation. Although this is not an exhaustive list, examples of support include breaks between interviews, having documents read aloud, or office accessibility. Please email us about anything we can do to accommodate you during the recruitment process.

Hybrid Working at Arm

Arm’s approach to hybrid working is designed to create a working environment that supports both high performance and personal wellbeing. We believe in bringing people together face to face to enable us to work at pace, whilst recognizing the value of flexibility. Within that framework, we empower groups/teams to determine their own hybrid working patterns, depending on the work and the team’s needs. Details of what this means for each role will be shared upon application. In some cases, the flexibility we can offer is limited by local legal, regulatory, tax, or other considerations, and where this is the case, we will collaborate with you to find the best solution. Please talk to us to find out more about what this could look like for you.

Equal Opportunities at Arm

Arm is an equal opportunity employer, committed to providing an environment of mutual respect where equal opportunities are available to all applicants and colleagues. We are a diverse organization of dedicated and innovative individuals, and don’t discriminate on the basis of race, color, religion, sex, sexual orientation, gender identity, national origin, disability, or status as a protected veteran.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Principal Software Engineer, AI Inference Cloud
Principal Software Engineer, AI Inference Cloud

Arm Limited • Seattle (WA)

Hybrid
USD 263,000 - 355,000
Staff Software Engineer, AI Inference Cloud
Staff Software Engineer, AI Inference Cloud

Arm Limited • Seattle (WA)

Hybrid
USD 209,000 - 283,000
Relocation package
Visa sponsorship
Principal Software Engineer, AI Compute Platform
Principal Software Engineer, AI Compute Platform

Arm Limited • Seattle (WA)

Hybrid
USD 263,000 - 355,000
Principal Software Engineer, AI Inference Runtime
Principal Software Engineer, AI Inference Runtime

Arm Limited • Seattle (WA)

Hybrid
USD 263,000 - 355,000
Hybrid working
Accommodations during recruitment
Staff Software Engineer, AI Inference Runtime
Staff Software Engineer, AI Inference Runtime

Arm Limited • Seattle (WA)

Hybrid
USD 209,000 - 283,000
Staff Software Engineer, AI Inference Runtime
Staff Software Engineer, AI Inference Runtime

Arm • Seattle (WA)

Hybrid
USD 209,000 - 283,000
Hybrid working
Recruitment accommodations
Principal Software Engineer, AI Compute Platform
Principal Software Engineer, AI Compute Platform

Arm • Seattle (WA)

Hybrid
USD 263,000 - 355,000
Principal Software Engineer, AI Inference Runtime
Principal Software Engineer, AI Inference Runtime

Arm • Seattle (WA)

Hybrid
USD 263,000 - 355,000
Hybrid working
Accommodations during recruitment
Staff Software Engineer, AI Compute Platform
Staff Software Engineer, AI Compute Platform

Arm • Seattle (WA)

Hybrid
USD 209,000 - 283,000
Staff Software Engineer, AI Compute Platform
Staff Software Engineer, AI Compute Platform

Arm Limited • Seattle (WA)

Hybrid
USD 209,000 - 283,000