Stand out for this role — generate a tailored resume and cover letter in about a minute.
Semiconductor Engineering in Cambridge is seeking a C++ software engineer with experience in ML inference engines, runtime systems, or backend integration frameworks. Experience with LiteRT, TensorFlow Lite, ONNX Runtime or similar technologies is highly valued.
You will understand how AI models execute in practice, including graph processing, operator execution, and memory management, and you will analyse, investigate and resolve complex issues across frameworks, runtimes, compilers, drivers
We’re looking for strong C++ software development experience, ideally gained while building machine learning inference engines, runtime systems, or backend integration frameworks. Experience working with a machine learning inference framework such as LiteRT, TensorFlow Lite, ONNX Runtime, or a similar technology is also important.
A good understanding of how AI models execute in practice is essential, including graph processing, operator execution, memory management, and backend integration. This role involves analysing, investigating, diagnosing, and resolving complex functional or performance issues across frameworks, runtimes, compilers, drivers, and hardware abstraction layers.