Get more replies from employers
Send a job-specific resume in minutes.
Sciforium, an AI infrastructure company, is hiring an ML Engineer to architect and optimize end-to-end multimodal GenAI systems. You will operate at the intersection of production software and Core AI/ML, building production-grade solutions across Serving, Post-Training, and Agentic frameworks.
The role emphasizes deep technical optimization, MLOps improvements, open-source contributions, and knowledge sharing through blogs and architecture breakdowns.
Sciforium is an AI infrastructure company developing next-generation multimodal AI models and a proprietary, high-efficiency serving platform. Backed by multi-million-dollar funding and direct sponsorship from AMD with hands-on support from AMD engineers the team is scaling rapidly to build the full stack powering frontier AI models and real-time applications.
As an ML Engineer at Sciforium, you will operate at the intersection of production software engineering and Core AI/ML to architect, scale, and optimize end-to‑end multimodal GenAI systems. In this role, you will build production‑grade solutions across Serving, Post‑Training and Agentic frameworks. You will also be responsible for driving deep technical optimizations and MLOps process improvements.
What You’ll Be Doing
Build and scale Agentic AI Systems:
Design and implement intelligent systems that can reason, plan, and execute complex multi‑step workflows. Develop architectures that combine LLMs, retrieval systems, memory, tools, and feedback loops.Build orchestration frameworks for multi‑agent and tool‑based systems. Develop evaluation frameworks that measure accuracy, reliability, latency, and task completion.
New Model Enablements, Automated Benchmarking, Profiling & Roofline Analysis: Rapidly benchmark, adapt, and integrate state‑of‑the‑art open‑weights models into production runtimes. Build automated MLOps tooling to profile deep learning workloads against theoretical hardware limits to drive optimization.
Open‑Source Leadership & Knowledge Sharing: Drive technical evangelism and elevate Sciforium’s presence in the global AI ecosystem through high‑impact community engagement. Actively contribute code, features, and optimizations to high‑visibility open‑source repositories. Author and publish deep‑dive technical blogs, whitepapers, and architecture breakdowns showcasing the novel innovations and complex problem‑solving happening at Sciforium.
Experience: 5+ years of professional ML/AI software engineering experience with a proven track record of architecting and shipping performance‑critical systems. Proven experience maintaining and developing model libraries or reusable ML components.
Education: BS, MS, or PhD in Computer Science, Computer Engineering, or a related technical field (or equivalent practical experience).
ML Systems: Strong knowledge of generative AI systems including Large Language Models, Transformers, Reinforcement Learning, RAG, and agentic patterns such as Chain‑of‑Thought, Tool Use, and Multi‑Agent orchestration
Machine Learning Expertise: Experience with one or more distributed ML training frameworks such as PyTorch, TensorFlow, or JAX, or Ray and inference engines like TensorRT, vLLM or SGLang. Good understanding of deep learning architectures across multiple domains (e.g., NLP, vision, speech, generative models).
Communication: Ability to articulate complex technical trade‑offs, write clear documentation, and collaborate smoothly across multidisciplinary engineering teams.
Sciforium is an equal opportunity employer. All applicants will be considered for employment without attention to race, color, religion, sex, sexual orientation, gender identity, national origin, veteran or disability status.