Get more replies from employers
Send a job-specific resume in minutes.
ByteDance is seeking a skilled engineer to architect and implement a scalable model inference system for large-parameter AI models. You will address challenges across deployment scenarios, optimize inference framework core modules, and drive improvements in performance, latency, and resource utilization.
The role requires handling high-concurrency distributed services, staying current with inference tech, and collaborating across teams to deliver system-wide improvements.
ByteDance is seeking a skilled engineer to architect and implement a scalable model inference system for large-parameter AI models. You will address challenges across deployment scenarios, optimize inference framework core modules, and drive improvements in performance, latency, and resource utilization.
The role requires handling high-concurrency distributed services, staying current with inference tech, and collaborating across teams to deliver system-wide improvements.