Staff Research Engineer: LLM Efficiency Architect

Deepstreamtech

New York (NY)

On-site

USD 120,000 - 150,000

Full time

14 days+
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

Deepstreamtech is seeking a Staff Research Engineer to focus on improving large language model (LLM) efficiency. This role will involve developing and deploying innovative techniques aimed at boosting model performance during production. Candidates should have a PhD in Machine Learning and strong software engineering skills for a collaborative startup environment.

The ideal candidate will be passionate about mentorship and have a proven track record of publications in top-tier AI conferences.

Qualifications

  • Must have a PhD in Machine Learning or a related field.
  • Strong understanding of LLM architecture and optimization.
  • Experience with techniques to enhance model efficiency.
  • Proven software engineering skills.

Responsibilities

  • Develop, prototype, and deploy techniques to enhance LLM inference efficiency.
  • Work on model architecture and optimization for performance improvements.
  • Contribute to decoding and inference-time algorithm improvements.

Skills

Machine Learning expertise
Software engineering
Model efficiency techniques
Mentoring

Education

PhD in Machine Learning or related field

Job description

Deepstreamtech is seeking a Staff Research Engineer to focus on improving large language model (LLM) efficiency. This role will involve developing and deploying innovative techniques aimed at boosting model performance during production. Candidates should have a PhD in Machine Learning and strong software engineering skills for a collaborative startup environment.

The ideal candidate will be passionate about mentorship and have a proven track record of publications in top-tier AI conferences.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Staff Research Engineer (Model Efficiency)
Staff Research Engineer (Model Efficiency)

Cohere • New York (NY)

On-site
USD 120,000 - 150,000
LLM Researcher: Transformer R&D & Scalable Training
LLM Researcher: Transformer R&D & Scalable Training

Deepgram • San Francisco (CA)

On-site
USD 180,000 - 240,000
LLM Research Scientist – Transformer & RL
LLM Research Scientist – Transformer & RL

Deepgram • United States

Remote
USD 120,000 - 150,000
Member of ML Technical Staff
Member of ML Technical Staff

Pragmatike • San Francisco (CA)

On-site
USD 200,000 - 350,000
Senior Staff Research Scientist — LLMs & Post-Training AI
Senior Staff Research Scientist — LLMs & Post-Training AI

DeepMind Technologies Limited • Kirkland (WA)

On-site
USD 207,000 - 300,000
Staff Research Scientist - AI Modeling & LLMs
Staff Research Scientist - AI Modeling & LLMs

DeepMind Technologies Limited • Mountain View (CA)

On-site
USD 207,000 - 300,000
Health insurance
Dental insurance
Vision insurance
+4
Senior LLM Research Engineer — NLP & Production AI
Senior LLM Research Engineer — NLP & Production AI

Bloomberg • New York (NY)

On-site
USD 200,000 - 260,000
Senior Staff AI Scientist — Post-Training & LLM Research
Senior Staff AI Scientist — Post-Training & LLM Research

Google LLC • Mountain View (CA), Kirkland (WA)

On-site
USD 262,000 - 364,000
Health benefits
401(k) match
Paid time off 20 days/year
+4
Senior AI Engineer: LLM Evaluation & Production
Senior AI Engineer: LLM Evaluation & Production

LawPro.ai • Georgia

On-site
USD 140,000 - 210,000
ML Inference Engineer - LLM Deployment & Serving
ML Inference Engineer - LLM Deployment & Serving

Google DeepMind • Mountain View (CA)

Hybrid
USD 230,000 - 290,000