Machine Learning Engineer, LLM Fine-Tuning

First Soft Solutions LLC

San Jose (CA)

On-site

USD 150,000 - 180,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

A technology services company is seeking a Machine Learning Engineer specializing in LLM fine-tuning for Verilog/RTL applications. This mid-senior level role requires significant experience in ML/AI and leading projects. Responsibilities include owning the technical roadmap and designing ML pipelines on AWS. The ideal candidate will have a proven track record in deploying LLM-powered features. Full-time position located in San Jose, CA.

Qualifications

  • 10+ years total engineering experience, with 5+ in ML/AI or large-scale distributed systems.
  • 3+ years working directly with transformers/LLMs.
  • AWS expertise for secure enterprise deployments.
  • Strong software engineering fundamentals: testing, CI/CD, observability.

Responsibilities

  • Own the technical roadmap for Verilog/RTL-focused LLM capabilities.
  • Lead a team of applied scientists/engineers.
  • Fine-tune and customize models using state-of-the-art techniques.
  • Design privacy-first ML pipelines on AWS.

Skills

LLM fine-tuning
Verilog/RTL
AWS
Bedrock
SageMaker
PyTorch
Hugging Face Transformers
DeepSpeed
Python

Job description

Machine Learning Engineer, LLM Fine‑Tuning

We are actively hiring for a Machine Learning Engineer focused on LLM fine‑tuning for Verilog/RTL applications.

Location: San Jose, CA (Onsite)

Skills: LLM fine‑tuning, Verilog/RTL, AWS, Bedrock, SageMaker

Responsibilities
  • Own the technical roadmap for Verilog/RTL‑focused LLM capabilities—from model selection and adaptation to evaluation, deployment, and continuous improvement.
  • Lead a hands‑on team of applied scientists/engineers: set direction, unblock technically, review designs/code, and raise the bar on experimentation velocity and reliability.
  • Fine‑tune and customize models using state‑of‑the‑art techniques (LoRA/QLoRA, PEFT, instruction tuning, preference optimization/RLAIF) with robust HDL‑specific evals:
    • Compile‑/lint‑/simulate‑based pass rates, pass@k for code generation, constrained decoding to enforce syntax, and “does‑it‑synthesize” checks.
  • Design privacy‑first ML pipelines on AWS:
    • Training/customization and hosting using Amazon Bedrock and SageMaker (or EKS + KServe/Triton/DJL) for bespoke training needs.
    • Artifacts in S3 with KMS CMKs; isolated VPC subnets & PrivateLink (including Bedrock VPC endpoints), IAM least‑privilege, CloudTrail auditing, and Secrets Manager for credentials.
    • Enforce encryption in transit/at rest, data minimization, no public egress for customer/RTL corpora.
  • Stand up dependable model serving: Bedrock model invocation where it fits, and/or low‑latency self‑hosted inference (vLLM/TensorRT‑LLM), autoscaling, and canary/blue‑green rollouts.
  • Build an evaluation culture: automatic regression suites that run HDL compilers/simulators, measure behavioral fidelity, and detect hallucinations/constraint violations; model cards and experiment tracking (MLflow/Weights & Biases).
  • Partner deeply with hardware design, CAD/EDA, Security, and Legal to source/prepare datasets (anonymization, redaction, licensing), define acceptance gates, and meet compliance requirements.
  • Drive productization: integrate LLMs with internal developer tools (IDEs/plug‑ins, code review bots, CI), retrieval (RAG) over internal HDL repos/specs, and safe tool‑use/function‑calling.
  • Mentor & uplevel: coach ICs on LLM best practices, reproducible training, critical paper reading, and building secure‑by‑default systems.
Qualifications
  • 10+ years total engineering experience with 5+ years in ML/AI or large‑scale distributed systems; 3+ years working directly with transformers/LLMs.
  • Proven track record shipping LLM‑powered features in production and leading ambiguous, cross‑functional initiatives at Staff level.
  • Deep hands‑on skill with PyTorch, Hugging Face Transformers/PEFT/TRL, distributed training (DeepSpeed/FSDP), quantization‑aware fine‑tuning (LoRA/QLoRA), and constrained/grammar‑guided decoding.
  • AWS expertise to design and defend secure enterprise deployments: Bedrock, SageMaker, S3, EC2/EKS/ECR, VPC/Subnets/Security Groups, IAM, KMS, PrivateLink, CloudWatch/CloudTrail, Step Functions, Batch, Secrets Manager.
  • Strong software engineering fundamentals: testing, CI/CD, observability, performance tuning; Python a must (bonus for Go/Java/C++).
  • Demonstrated ability to set technical vision and influence across teams; excellent written and verbal communication for execs and engineers.
Seniority Level

Mid‑Senior level

Employment Type

Full‑time

Job Function

Engineering and Information Technology

Industries

IT Services and IT Consulting

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Staff Machine Learning Engineer, LLM Fine‑Tuning (Verilog/RTL Applications)
Staff Machine Learning Engineer, LLM Fine‑Tuning (Verilog/RTL Applications)

Highbrow Technology Inc • San Jose (CA)

On-site
USD 150,000 - 200,000
Staff ML Engineer: LLM Fine-Tuning for RTL/Verilog
Staff ML Engineer: LLM Fine-Tuning for RTL/Verilog

Highbrow Technology Inc • San Jose (CA)

On-site
USD 150,000 - 200,000
LLM Fine-Tuning Engineer for Verilog/RTL — Onsite San Jose
LLM Fine-Tuning Engineer for Verilog/RTL — Onsite San Jose

First Soft Solutions LLC • San Jose (CA)

On-site
USD 150,000 - 180,000
Senior Machine Learning Engineer (LLMs)
Senior Machine Learning Engineer (LLMs)

Albiware Inc. • Chicago (IL)

On-site
USD 140,000 - 210,000
Competitive salary
Generous PTO
Medical, dental, and vision coverage
+2
Senior Machine Learning Engineer (LLMs)
Senior Machine Learning Engineer (LLMs)

Albi • Chicago (IL)

On-site
USD 150,000 - 210,000
Competitive salary
Generous PTO
Medical, dental, and vision coverage
+6
AI Developer
AI Developer

Salvo Software • United States

On-site
USD 180,000 - 240,000
Remote LLM Engineer: Fine-Tuning & Production ML
Remote LLM Engineer: Fine-Tuning & Production ML

Bright Vision Technologies • Andover (MA)

On-site
USD 100,000 - 150,000
AI Engineer, LLMs
AI Engineer, LLMs

Logical Intelligence • San Francisco (CA)

On-site
USD 120,000 - 160,000
Machine Learning Engineer, LLM Post-Training
Machine Learning Engineer, LLM Post-Training

GoTo Meeting • Mountain View (CA)

On-site
USD 150,000 - 230,000
Health, dental, and vision care for you and your family
Top-tier 401(K) plan with company matching
Paid time off and paid holidays
+2
Machine Learning Engineer (LLM)
Machine Learning Engineer (LLM)

DeepRec.ai • Boston (MA)

Hybrid
USD 170,000 - 200,000