Member of Technical Staff, Architecture & Scaling

Hark

San Jose (CA)

On-site

USD 180,000 - 450,000

Full time

5 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Hark is seeking a researcher to lead the architecture and scaling of its largest multimodal models, shaping training design, optimization, and frontier compute strategy with direct production impact.

Working across teams, you will implement training pipelines on multi-GPU clusters, balancing performance and cost while advancing efficient, reliable AI at scale in a fast-growing startup environment.

Qualifications

  • Experience training large models at scale with significant compute and stability considerations.
  • Ability to design experiments that discriminate between hypotheses.
  • Familiarity with scaling laws and extrapolating small-scale results to frontier-scale runs.
  • Experience with large GPU clusters and modern training stacks.

Responsibilities

  • Research model architectures, optimization, and scaling to improve capability and efficiency of largest models.
  • Establish baselines and run controlled experiments to identify scalable ideas.
  • Set training recipes at scale: learning-rate schedules, context length, token budgets, and compute allocation.
  • Explore architectures including mixture-of-experts, hybrid attention, and long-context extension.
  • Diagnose instability in large runs: loss spikes, divergence, numerical issues, and related infrastructure failures.
  • Work across modalities since models are multimodal from pretraining forward.

Skills

Model training at scale
Empirical research mindset
Multimodal systems experience
Distributed training
Software engineering

Tools

PyTorch distributed
Profiling tools
Mixed-precision training

Job description

About Hark

Hark is an artificial intelligence company building advanced, personalized intelligence. One that is proactive, multimodal, and capable of interacting with the world through speech, text, vision, and persistent memory.

We’re pairing that intelligence with next-generation hardware to create a universal interface between humans and machines. While today’s AI largely operates through chat boxes and decade-old devices, Hark is focused on what comes next: agentic systems that interact naturally with people and the real world.

To get there, we’re developing multimodal models and next-generation AI hardware together - designed from the ground up as a single, unified interface for a new era of intelligent systems.

About the Role

You’ll work on the architecture and scaling of our largest models: what we train, how we train it, and how to spend the next order of magnitude of compute well. This is empirical research with a direct line to production. The recipes you set are the recipes our frontier runs use.

Responsibilities
  • Conduct research on model architecture, optimization, and scaling to improve the capability and efficiency of our largest models.
  • Establish strong baselines and run controlled experiments to determine which ideas actually scale to frontier training runs.
  • Set training recipes at scale: learning-rate schedules, context length, token budgets, and compute allocation.
  • Explore new architectures, including mixture-of-experts, hybrid attention, and long-context extension.
  • Diagnose instability in large runs: loss spikes, divergence, numerical issues, and the infrastructure failures that look like research problems.
  • Work across modalities, since our models are multimodal from pretraining forward.
Requirements
  • Hands-on experience training large models, at a scale where compute allocation and stability decisions carry real cost.
  • Strong empirical instincts: you design the experiment that distinguishes between two hypotheses instead of the one that confirms the first.
  • Fluency with scaling laws and how to use small-scale results to make a frontier-scale bet.
  • Deep familiarity with a modern training stack and distributed training across large GPU clusters.
  • Strong engineering. Research here means writing the code and reading the profiler, not handing off a spec.
  • A record of work that shipped into real models, whether that shows up as papers, systems, or production runs.
Bonus Qualifications
  • Experience with mixture-of-experts routing, sparse architectures, or long-context methods.
  • Work on data mixtures, curriculum, or tokenizer design.
  • Kernel-level optimization or mixed-precision training experience.
  • Experience with efficiency work aimed at constrained inference targets, including on-device.
Compensation

The US base salary range for this full-time position is between $180,000 - $450,000 annually.

The pay offered for this position may vary based on several individual factors, including job-related knowledge, skills, and experience. The total compensation package may also include additional components/benefits depending on the specific role. This information will be shared if an employment offer is extended.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Member of Technical Staff, Architecture & Scaling San Jose
Member of Technical Staff, Architecture & Scaling San Jose

Hark • San Jose (CA)

Hybrid
USD 180,000 - 450,000
Member of Technical Staff, Mid-training
Member of Technical Staff, Mid-training

Hark • San Jose (CA)

On-site
USD 180,000 - 450,000
Data Engineering Lead
Data Engineering Lead

Hark • San Jose (CA)

On-site
USD 170,000 - 450,000
Backend Engineer
Backend Engineer

Hark • San Jose (CA)

On-site
USD 170,000 - 400,000
Member of Technical Staff, Mid-training San Jose
Member of Technical Staff, Mid-training San Jose

Hark, Inc. • San Jose (CA)

On-site
USD 180,000 - 450,000
Platform Engineer
Platform Engineer

Hark • San Jose (CA)

On-site
USD 170,000 - 400,000
Infrastructure, Large-scale Training
Infrastructure, Large-scale Training

Hark • San Jose (CA)

On-site
USD 180,000 - 450,000
Data Engineering Lead San Jose
Data Engineering Lead San Jose

Hark, Inc. • San Jose (CA)

On-site
USD 170,000 - 450,000
Full-Stack Engineer
Full-Stack Engineer

Hark • San Jose (CA)

On-site
USD 170,000 - 400,000
Backend Engineer San Jose
Backend Engineer San Jose

Hark, Inc. • San Jose (CA)

On-site
USD 170,000 - 400,000