Founding ML Researcher

Base Compute

Berlin

Vor Ort

EUR 110.000 - 170.000

Vollzeit

14 Tage+
Bewerbungsgenerator

Verschicke keinen 08/15-Lebenslauf — erstelle einen Lebenslauf und ein Anschreiben, die genau auf diese Rolle zugeschnitten sind.

Schaffe es an den ATS-Filtern vorbei

Benefits dieser Stelle

Founding team equity
Strong base salary

Zusammenfassung

Base Compute is seeking a Founding ML Researcher to push the frontier of on-device AI. You will identify problems, design experiments and translate results into real-world performance, with significant ownership over our research agenda.

You’ll influence technical bets and work across inference efficiency, model routing and autonomous research pipelines. Strong PhD or equivalent research track is required; English proficiency is essential for collaboration.

Qualifikationen

  • PhD in ML or equivalent industry research experience.
  • Deep understanding of LLM architectures and AI inference.
  • Experience with speculative decoding, quantization, distillation or RL.
  • Track record of published papers or open-source projects.

Aufgaben

  • Lead inference research on on-device efficiency.
  • Develop and validate new approaches and hypotheses.
  • Design autoresearch pipelines for autonomous exploration.
  • Build rigorous evaluations and benchmarks in real-world settings.

Kenntnisse

ML research
Experiment design
LLM architectures
Communication

Ausbildung

PhD in ML

Tools

CUDA
Triton
ROCm

Jobbeschreibung

About Us

Base Compute is an AI inference lab. Our mission is to bring AGI on device. We believe in a world where everyone has access to intelligence: fast, private and always available on your device.

We’re building the infrastructure for the next generation of on-device AI, from silicon-level optimizations to distributed inference systems.

We’re working on hard problems at the intersection of inference efficiency, model intelligence and autonomous research.

The Role

We’re looking for a Founding ML Researcher to work at the frontier of on-device AI. This role is for someone who identifies problems and potentials, designs and executes experiments and derives insights that translate into real-world performance.

You’ll have significant ownership over our research agenda and direct influence on the technical bets the company makes.

What You’ll Work On
  • Inference research: Identifying and validating new approaches to on-device efficiency, including speculative decoding variants, novel quantization schemes and entirely new techniques yet to be discovered

  • Model routing research: Building the intelligence that decides how requests are served between on-device vs. frontier API models

  • Autoresearch pipelines: Designing systems that can autonomously explore, hypothesize and evaluate research ideas that accelerate our R&D loop

  • Evaluations and benchmarks: Developing rigorous evals that measure performance in the real world, outside of clean academic settings

What We’re Looking For
  • PhD in ML or equivalent industry research experience

  • Deep understanding of LLM architectures and the principles of AI inference

  • Expertise in a relevant topic, such as speculative decoding, quantization theory, model distillation, reinforcement learning

  • A track record of producing results that people build on: research papers, open-source projects or blog posts that prove out novel ideas

  • Good communication: the ability to explain complex ideas simply, give honest feedback and document findings in a reproducible way

  • Nice-to-haves:

    • Familiarity with GPU and accelerator architectures and kernel optimization (CUDA, ROCm, Metal, Triton, etc.)

    • Experience deploying models under on-device constraints (memory bandwidth, latency budgets, and thermal and power ceilings)

What We Offer
  • Founding team equity and strong base salary

  • Direct influence on technical direction: your ideas will shape the roadmap

  • Work on genuinely hard problems that haven't been solved yet

  • Small team, fast iteration, low bureaucracy

Location

The team is based in Melbourne and Berlin and works in-person from the office most days. We require strong written and spoken English, since the team collaborates across time zones.

Hol dir deinen kostenlosen, vertraulichen Lebenslauf-Check.

oder ziehe deine Datei hierhin.

Similar jobs

Ähnliche Jobs, die dir auch gefallen könnten

Research Engineer - Inference
Research Engineer - Inference

Artificial Intelligence Jobs • Deutschland

Remote
EUR 110.000 - 170.000
Senior Research Scientist
Senior Research Scientist

adaption • Berlin

Hybrid
EUR 90.000 - 130.000
Flexible work
Travel stipend
Lunch stipend
+2
[Nota AI GmbH] ML Researcher
[Nota AI GmbH] ML Researcher

Nota AI • Berlin

Hybrid
EUR 90.000 - 120.000
Founding Machine Learning Engineer
Founding Machine Learning Engineer

Clera • München

Vor Ort
EUR 88.000 - 176.000
Senior ML Engineer - Real-Time Inference & Systems
Senior ML Engineer - Real-Time Inference & Systems

Inworld AI • Deutschland

Vor Ort
USD 120.000 - 180.000
Staff+ Software Engineer, Inference Velocity
Staff+ Software Engineer, Inference Velocity

Anthropic • Deutschland

Vor Ort
EUR 349.000 - 418.000
Competitive compensation
Equity donation matching (optional)
Generous vacation and parental leave
+2
Germany Senior / Lead Research Scientist - Germany
Germany Senior / Lead Research Scientist - Germany

Inworld AI • Deutschland

Vor Ort
USD 136.550 - 204.825
Senior ML Engineer (Token Factory)
Senior ML Engineer (Token Factory)

Meyandy LLC • Berlin

Hybrid
EUR 110.000 - 150.000
Competitive compensation
Career growth and learning oppor tunun
Flexible working arrangements
Senior ML Research Developer
Senior ML Research Developer

European Tech Recruit • Berlin

Hybrid
EUR 85.000 - 125.000
Radar Signal Processing & AI Engineer - Ground-Up Systems
Radar Signal Processing & AI Engineer - Ground-Up Systems

THRYVE • Berlin

Vor Ort
EUR 120.000 - 190.000
Competitive compensation
Stock options
Flexible work environment
+1