Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.
Applied Intuition in Sunnyvale, CA is seeking a software engineer specializing in optimizing ML models for production-grade embedded runtime environments. You will influence the entire ML framework stack across PyTorch, JAX, ONNX, TensorRT, CUDA, XLA and Triton.
You will collaborate with ML engineers and software teams to implement model pruning, quantization, and deployment strategies on memory-constrained embedded compute platforms, delivering efficient, low latency inference.
Applied Intuition in Sunnyvale, CA is seeking a software engineer specializing in optimizing ML models for production-grade embedded runtime environments. You will influence the entire ML framework stack across PyTorch, JAX, ONNX, TensorRT, CUDA, XLA and Triton.
You will collaborate with ML engineers and software teams to implement model pruning, quantization, and deployment strategies on memory-constrained embedded compute platforms, delivering efficient, low latency inference.