Stand out for this role — generate a tailored resume and cover letter in about a minute.
Arm is seeking a Principal Software Engineer to lead the AI Inference Runtime team in Seattle, shaping technical direction for distributed inference workloads and memory systems. You will collaborate with AI Infra, compute, and product groups to optimize performance across platforms and models.
You will architect scheduling, batching, KV‑cache management, and kernel development, driving end‑to‑end model support and production validation with a focus on latency, throughput, and resource
Arm is seeking a Principal Software Engineer to lead the AI Inference Runtime team in Seattle, shaping technical direction for distributed inference workloads and memory systems. You will collaborate with AI Infra, compute, and product groups to optimize performance across platforms and models.
You will architect scheduling, batching, KV‑cache management, and kernel development, driving end‑to‑end model support and production validation with a focus on latency, throughput, and resource