Overview
Architect, design, and implement state-of-the-art Generative AI models (auto-regressive, diffusion, Discrete Diffusion, Hybrid SSM, etc.) and advance the state-of-the-art in very small LLMs. Leverage expertise to train, fine-tune, and optimize models that impact LLMs (multimodal and multilingual), shaping them for optimal accuracy, naturalness, and performance. Collaborate with cross-functional teams to develop customized solutions that leverage the prowess of LLMs/MLLMs. Innovate and enhance applications by integrating LLMs/MLLMs to ensure optimal performance and speed on resource-constrained platforms. Stay abreast of latest advancements in Generative AI, Conversational AI, and language modeling, incorporating new techniques and best practices to continually improve system performance.
Responsibilities
- Architect, design, and implement state-of-the-art Generative AI models (auto-regressive, diffusion, Discrete Diffusion, Hybrid SSM, etc.) and advance the SOTA in very small LLMs.
- Leverage expertise to train, fine-tune, and optimize models that impact LLMs (Multimodal & Multilingual), shaping them for optimal accuracy, naturalness, and performance.
- Collaborate with cross-functional teams comprising agentic AI scientists, software engineers, and domain experts to develop customized solutions that leverage the prowess of LLMs/MLLMs.
- Innovate and enhance existing applications by integrating the power of LLMs/MLLMs, ensuring optimal performance and speed on resource-constrained platforms.
- Stay abreast of the latest advancements, trends, and research in Generative AI, Conversational AI, and language modeling, incorporating new techniques, technologies, and best practices to enhance system performance.
Qualifications
- Extensive experience and demonstrated mastery in developing LLM/MLLMs.
- Well-versed with commercial deployments, post-deployment support, model profiling and optimization needed for production environments. Experience deploying models in different languages preferred.
- Advanced degree (Ph.D. preferred) in computer science, electrical engineering, or a related field.
- Exceptional programming skills in Python with the ability to implement complex machine learning algorithms.
- Proven track record in developing conversational AI systems, including coordination with acoustic models, language models, and pronunciation models.
- Strong analytical thinking, problem-solving capabilities, and a collaborative mindset in fast-paced, innovative environments.
What we offer
- Salary range: $123,500 - $197,600 USD; actual salary determined by experience and other job-related factors.
- Annual bonus opportunity
- Flexible Time Off
- Choice-based medical coverage plus dental and vision plans
- Health Savings Accounts (HSA), Flexible Spending Accounts (FSA) and 401(k) with employer match
- Employee Stock Purchase Plan
- Paid parental leave and adoption assistance
- Disability and life insurance
- Employee Assistance Program, wellbeing platform, volunteer hours, discounts, tuition reimbursement, and rewards program
Equal Opportunity
Cerence is firmly committed to Equal Employment Opportunity (EEO) and to compliance with all laws prohibiting employment discrimination. All employees are expected to adhere to security, privacy, and safety policies and regulations.