Get more replies from employers
Send a job-specific resume in minutes.
SambaNova Systems in San Jose, CA, seeks an Architect for the Inference Systems Performance team to own end-to-end performance for large-scale LLM inference, from tokenization to decode, across hardware and software. You will define strategy and drive reproducible workloads to optimize efficiency.
You will model, simulate, and optimize configurations to meet customer SLOs, mentor engineers, and represent our performance story to customers and partners.
SambaNova Systems in San Jose, CA, seeks an Architect for the Inference Systems Performance team to own end-to-end performance for large-scale LLM inference, from tokenization to decode, across hardware and software. You will define strategy and drive reproducible workloads to optimize efficiency.
You will model, simulate, and optimize configurations to meet customer SLOs, mentor engineers, and represent our performance story to customers and partners.