An application made for this job — a tailored resume and cover letter that speak straight to the posting.
G-Research is seeking an exceptional NLP Performance Engineer to own large-scale LLM inference performance within our NLP Engineering team at the London HQ. You will profile, optimise and implement techniques to maximise inference throughput and cost-efficiency, collaborating with researchers and infrastructure engineers to shape scalable tooling and compute stacks.
The role blends high-impact hands-on work with systems design, offering a competitive package and a dynamic research-driven
G-Research is seeking an exceptional NLP Performance Engineer to own large-scale LLM inference performance within our NLP Engineering team at the London HQ. You will profile, optimise and implement techniques to maximise inference throughput and cost-efficiency, collaborating with researchers and infrastructure engineers to shape scalable tooling and compute stacks.
The role blends high-impact hands-on work with systems design, offering a competitive package and a dynamic research-driven