NewsLayer.com

PrismML hopes its tiny LLM will change how we all use AI

If AI lab PrismML isn't on your radar yet, it should be.

Julie Bort

Publisher TechCrunch AI

Sep 17, 2026 at 10:34 PM UTC · 3 Min. Lesezeit

PrismML hopes its tiny LLM will change how we all use AI
NewsLayer editorial artwork

Key Signal

5.9 GB Compressed model size

Last Updated

vor einem Tag

Übersetzung…

If AI lab PrismML isn’t on your radar yet, it should be — not because it’s raised gobs of money (it hasn’t yet, just a $22.25 million seed round), but because of the technical minds involved and the potentially industry-changing tech it’s developing.

PrismML is betting that capable, high-performing, reasoning large language models don’t, in fact, have to be large.

It is making reasoning models so small they can fit on PCs and smartphones. (It’s even rumored to be in talks with Apple, though CEO Babak Hassibi declined to comment on that to TechCrunch.)

On Thursday, PrismML released Bonsai 2 27B, its latest in a family of models, which compresses Qwen3.8 27B, a widely used open-source model from Alibaba, down to 5.9 GB. That’s small enough to fit on a PC and, possibly, a high-end smartphone. It’s a 9x to 10x reduction in memory versus the original.

PrismML was founded by a group of Caltech researchers and is led by Hassibi, a Caltech professor and an expert in compression technologies. The startup also counts Ion Stoica as an advisor. Stoica is a co-founder of Databricks (and other companies) and the director of Berkeley’s famed Sky Computing Lab, which has birthed many technologies and startups, from Letta to SGLang.

Article Intelligence

Topics

Sponsored

Ad
House — Advertise on NewsLayer
NewsLayerLearn more

NewsLayer Premium

Unlock deeper intelligence.

Ad-free reading, exclusive research, and real-time onchain insights.

Go Premium