Stand out for this role — generate a tailored resume and cover letter in about a minute.
Evollabs is building the next generation of AI infrastructure and seeks engineers focused on LLM systems, architecture, and performance. You will work at the intersection of model internals and hardware to ensure state-of-the-art models run efficiently on our platform.
Responsibilities include profiling, benchmarking, and optimizing LLM inference across distributed systems, designing attention and quantization optimizations, and collaborating with hardware teams to co-design efficient pipelines.
Evollabs is building the next generation of AI infrastructure and seeks engineers focused on LLM systems, architecture, and performance. You will work at the intersection of model internals and hardware to ensure state-of-the-art models run efficiently on our platform.
Responsibilities include profiling, benchmarking, and optimizing LLM inference across distributed systems, designing attention and quantization optimizations, and collaborating with hardware teams to co-design efficient pipelines.