A pioneering AI lab is seeking an engineer to develop and support low-latency model inference platforms. You will engineer and scale data processing infrastructure while optimizing performance and cost. Ideal candidates should have strong programming skills, experience with container orchestration, and a passion for building efficient systems in a collaborative environment. This role offers the opportunity to shape the future of AI and its applications.
Qualifications
Strong programming skills in Python, Go, or similar.
Deep experience with containerization and orchestration.
Proven experience with GPU computational workloads.
Responsibilities
Develop and operate low-latency model inference platforms.
Engineer and scale core data processing infrastructure.
Design and maintain GPU-based training clusters.
Automate infrastructure provisioning and monitoring.
Drive performance tuning and cost optimization.
Skills
Programming skills (Python, Go)
Containerization (Docker)
Container orchestration (Kubernetes)
Infrastructure as Code (Terraform)
Experience with distributed systems
Performance tuning
Collaboration and communication skills
Job description
A pioneering AI lab is seeking an engineer to develop and support low-latency model inference platforms. You will engineer and scale data processing infrastructure while optimizing performance and cost. Ideal candidates should have strong programming skills, experience with container orchestration, and a passion for building efficient systems in a collaborative environment. This role offers the opportunity to shape the future of AI and its applications.