Lemura AI Labs

Research

What we are working on.

Six threads, from expert pruning through to population genomics. Each one has its own page with the detail, the numbers, and the published work it sits alongside.

How we work

What we hold ourselves to

Publish the weights, then the method.Open weights on the Hub from day one, with the write-up following once the evaluation is broad enough to support the claim. Shipping first is not a substitute for measuring — it is an admission of what we have measured so far.

Report the number that could embarrass us.Refusal rate without KL divergence, or WER without parameter count, tells you the half of the story that flatters the model. Both halves or neither.

Efficiency is a research direction.A model that cannot be served on the hardware someone owns is a result, not a tool. Compression and quantization are the thread every other thread depends on.

Document what failed.Negative results and failure modes go in the write-up. A defence you cannot verify is a defence you cannot trust.