Skip to content

Founding ML Researcher

basecompute

Berlin

Applying for this one?

We write the CV against this exact posting — its wording, its requirements — not a template with your name in it.

Get my CV for this job

$25, one-time. No subscription.

ABOUT US

Base Compute is an AI inference lab. Our mission is to bring AGI on device. We believe in a world where everyone has access to intelligence: fast, private and always available on your device.

We’re building the infrastructure for the next generation of on-device AI, from silicon-level optimizations to distributed inference systems.

We’re working on hard problems at the intersection of inference efficiency, model intelligence and autonomous research.

THE ROLE

We’re looking for a Founding ML Researcher to work at the frontier of on-device AI. This role is for someone who identifies problems and potentials, designs and executes experiments and derives insights that translate into real-world performance.

You’ll have significant ownership over our research agenda and direct influence on the technical bets the company makes.

WHAT YOU’LL WORK ON

- Inference research: Identifying and validating new approaches to on-device efficiency, including speculative decoding variants, novel quantization schemes and entirely new techniques yet to be discovered

- Model routing research: Building the intelligence that decides how requests are served between on-device vs. frontier API models

- Autoresearch pipelines: Designing systems that can autonomously explore, hypothesize and evaluate research ideas that accelerate our R&D loop

- Evaluations and benchmarks: Developing rigorous evals that measure performance in the real world, outside of clean academic settings

WHAT WE’RE LOOKING FOR

- PhD in ML or equivalent industry research experience

- Deep understanding of LLM architectures and the principles of AI inference

- Expertise in a relevant topic, such as speculative decoding, quantization theory, model distillation, reinforcement learning

- A track record of producing results that people build on: research papers, open-source projects or blog posts that prove out novel ideas

- Good communication: the ability to explain complex ideas simply, give honest feedback and document findings in a reproducible way

  • Nice-to-haves:
  • Familiarity with GPU and accelerator architectures and kernel optimization (CUDA, ROCm, Metal, Triton, etc.)
  • Experience deploying models under on-device constraints (memory bandwidth, latency budgets, and thermal and power ceilings)

WHAT WE OFFER

- Founding team equity and strong base salary

- Direct influence on technical direction: your ideas will shape the roadmap

- Work on genuinely hard problems that haven't been solved yet

- Small team, fast iteration, low bureaucracy

LOCATION

The team is based in Melbourne and Berlin and works in-person from the office most days. We require strong written and spoken English, since the team collaborates across time zones.

Seen 4 hours ago.

Original posting on basecompute's site ↗

Posting text belongs to the employer. Removal requests: contact us.

Nearby

Live postings like this one

Same employer first, then the same role elsewhere.

One job at a time

One posting. One CV. $25.

Pick the job you actually want and we write for it.

Get my CV for this job