← All stories
● Covered by 1 source · 1 reportMedium impact1 neutral

PrismML Releases Bonsai 2 27B, a Compressed LLM for PCs and Smartphones

🔄 Updated 6d ago
New to BrevFeed? We gather this story from every outlet covering it into one summary — ranked by real-world impact, not just the latest headline — so you never miss what matters. What is BrevFeed? →

Key points

  • PrismML released Bonsai 2 27B, a compressed LLM.
  • Bonsai 2 27B is 5.9 GB, fitting on PCs and smartphones.
  • It retains 98% of Qwen3.8 27B's benchmark performance.
  • The model offers a 9x to 10x memory reduction.

PrismML Introduces Bonsai 2 27B

PrismML, an AI lab, released Bonsai 2 27B, its latest compressed large language model. This model is a compact version of Qwen3.8 27B, an open-source model from Alibaba.

Technical Specifications and Performance

Bonsai 2 27B is compressed to 5.9 GB, which allows it to run on personal computers and potentially high-end smartphones. This represents a 9x to 10x reduction in memory footprint compared to the original Qwen3.8 27B model. PrismML states that Bonsai 2 matches 98% of Qwen's aggregate benchmark scores, an improvement from the first Bonsai model which achieved 95%.

Company Background and Funding

PrismML was founded by Caltech researchers, with CEO Babak Hassibi, a Caltech professor and compression expert, leading the company. Ion Stoica, co-founder of Databricks, serves as an adviser. The company has raised a $22.25 million seed round and is backed by investors including Khosla Ventures, Cerberus Capital, and Caltech.

Implications for On-Device AI

The development of highly compressed LLMs that maintain performance could expand the possibilities for running AI models directly on user devices. This approach reduces reliance on cloud infrastructure and could enable more private and responsive AI applications on PCs and mobile devices.

✨ This summary was generated by AI from the outlets' reporting listed below. It is not independently verified and may contain errors — check the original sources. How BrevFeed works →

The daily brief

One email each morning: the day's tech stories, clustered across outlets and summarized. No account needed.

One email a day. Unsubscribe in one click, any time.

Today's brief

Spend a few minutes, get the whole day. Every topic's top stories in one hands-free rundown — listen, watch, or read the transcript.

~26 min · 21 stories · Sep 23

▶ Play today's brief Listen on Spotify

New every morning, and the back catalogue is archived by date.

Reporting from

PrismML released Bonsai 2 27B, a large language model compressed to 5.9 GB, making it suitable for PCs and high-end smartphones. This model achieves 98% of the benchmark performance of the original Qwen3.8 27B model, representing a 9x to 10x memory reduction. The development of highly compressed yet performant LLMs could enable broader on-device AI applications.