Close Menu
MyAppsPlus

    Subscribe to Updates

    Get the latest creative news from FooBar about art, design and business.

    What's Hot

    How to change Siri’s voice on iPhone in iOS 27

    September 17, 2026

    Crusoe raises $3.9B to build massive data centers and small modular “AI factories”

    September 17, 2026

    RTX 5090s are going for as much as $9000, as AI server builders reportedly buy Nvidia’s flagship by the pallet, but the photos behind the story may be AI-generated

    September 17, 2026
    Facebook X (Twitter) Instagram
    Facebook X (Twitter) Instagram
    MyAppsPlusMyAppsPlus
    Thursday, September 17
    • Home
    • Breaking Tech
    • Apps & Software
    • AI & Automation
    • Android
    • iPhone & iOS
    • More
      • Reviews
      • How-To Guides
      • Deals & Discounts
      • Shop
    MyAppsPlus
    Home»Breaking Tech»PrismML hopes its tiny LLM will change how we all use AI
    Breaking Tech

    PrismML hopes its tiny LLM will change how we all use AI

    myappsplusBy myappsplusSeptember 17, 2026004 Mins Read
    Share Facebook Twitter Pinterest Copy Link LinkedIn Tumblr Email Telegram WhatsApp
    Follow Us
    Google News Flipboard
    PrismML hopes its tiny LLM will change how we all use AI
    Share
    Facebook Twitter LinkedIn Pinterest Email Copy Link

    If AI lab PrismML isn’t on your radar yet, it should be — not because it’s raised gobs of money (it hasn’t yet, just a $22.25 million seed round), but because of the technical minds involved and the potentially industry-changing tech it’s developing.

    PrismML is betting that capable, high-performing, reasoning large language models don’t, in fact, have to be large.

    It is making reasoning models so small they can fit on PCs and smartphones. (It’s even rumored to be in talks with Apple, though CEO Babak Hassibi declined to comment on that to TechCrunch.)

    On Thursday, PrismML released Bonsai 2 27B, its latest in a family of models, which compresses Qwen3.8 27B, a widely used open-it on a PC and, possibly, a high-end smartphone. It’s a 9x to 10x reduction in memory versus the original

    PrismML was founded by a group of Caltech researchers and is led by Hassibi, a Caltech professor and an expert in compression technologies. The startup also counts Ion Stoica as an advisor. Stoica is a co-founder of Databricks (and other companies) and the director of Berkeley’s famed Sky Computing Lab, which has birthed many technologies and startups, from Letta to SGLang.

    PrismML is also backed by investors Khosla Ventures, Cerberus Capital, and Caltech.

    This startup is certainly not the only company working on LLM compression tech. Multiverse Computing, founded by a well-known professor from Spain’s Donostia International Physics Center, is another. (And Multiverse Computing has raised gobs of cash.)

    But Hassibi says that PrismML’s compression tech is unique because its LLMs have lost virtually no performance compared with the originals. Bonsai 2 matches 98% of Qwen’s aggregate benchmark scores. That’s up from the first Bonsai, released a couple of months ago in March, that matched 95%. That original model has already been downloaded over 11 million times, and PrismML’s even smaller models have been downloaded another 2.6 million times, the company says.

    So this shows that PrismML’s compression results have improved from one release to the next. Whether it could ever get to 100% benchmark performance parity is a question that remains to be seen. Compression will likely always have some impact, Hassibi says.

    Still, perfect benchmark parity is fairly academic anyway. LLMs are not so accurate in their uncompressed form, and benchmarks not so perfectly reflective of actual tasks, that a 2% degradation would likely meaningfully affect how a model performs in actual use. (Plus, the surrounding software — the harness a model runs inside of — matters a lot when it comes to accuracy, too.)

    PrismML says it achieves this by shrinking the “weights” that make up a model — weights are, essentially, the information a model learns and stores during training. Normally, each weight requires 16 bits. PrismML’s approach, called “ternary” weights, simplifies that down to three: +1, −1, or 0. With far smaller values to store for each weight, the model takes up dramatically less space. (For a deeper dive on the compression technique, here’s the project’s GitHub page.)

    The startup’s next goal is to apply this compression technique to even bigger models. “The next models that we will release, hopefully in the next couple of months, will be in the several-hundred-billion-parameter range, and I expect it will be easier to retain the intelligence there,” Hassibi told TechCrunch.

    As model size grows, he added, “There is more room to be able to compress them without losing the intelligence. So I would just say, as a general trend, for larger models, it’s easier to get to 100%.”

    Stoica tells us that he’s excited for this tech because it’s making it possible for advanced models to run on users’ devices. “You are going to have intelligence at your fingertips, and it’s going to be free because it’s going to run on the device you already bought. It’s also going to be private, because you’re not going to send it to the cloud.”

    AI hopes PrismML Startups TC
    Follow on Google News Follow on Flipboard
    Share. Facebook Twitter Pinterest LinkedIn Tumblr Email Copy Link
    myappsplus
    • Website

    Related Posts

    Crusoe raises $3.9B to build massive data centers and small modular “AI factories”

    September 17, 2026

    The FAA’s plan to fix air traffic? $875 million worth of AI

    September 17, 2026

    AI told itself ‘feel no obligation’ to users: OpenAI flags ‘unexpected, concerning’ behaviour

    September 17, 2026
    Add A Comment
    Leave A Reply Cancel Reply

    Top Posts

    Top 10 Best React Native App Development Companies in 2026

    September 12, 20262 Views

    AI, automation, robot dogs ensure on-site nuclear safety

    September 7, 20262 Views

    This tiny AI box could save me from upgrading my perfectly good laptop

    September 6, 20262 Views
    Latest Reviews

    Fall is only a month away, grab some fitted sweaters from $14 (Reg. $30)

    myappsplusAugust 19, 2026

    Flock is testing a new AI tool that tracks and identifies people based on their driving habits

    myappsplusAugust 19, 2026

    Keep your car looking as good as new with Fanttik’s Nano detailing brush, now $46 (Save 30%)

    myappsplusAugust 19, 2026
    Stay In Touch
    • Facebook
    • YouTube
    • TikTok
    • WhatsApp
    • Twitter
    • Instagram

    Subscribe to Updates

    Get the latest tech news from FooBar about tech, design and biz.

    Most Popular

    Fall is only a month away, grab some fitted sweaters from $14 (Reg. $30)

    August 19, 20260 Views

    Flock is testing a new AI tool that tracks and identifies people based on their driving habits

    August 19, 20260 Views

    Keep your car looking as good as new with Fanttik’s Nano detailing brush, now $46 (Save 30%)

    August 19, 20260 Views
    Our Picks

    How to change Siri’s voice on iPhone in iOS 27

    September 17, 2026

    Crusoe raises $3.9B to build massive data centers and small modular “AI factories”

    September 17, 2026

    RTX 5090s are going for as much as $9000, as AI server builders reportedly buy Nvidia’s flagship by the pallet, but the photos behind the story may be AI-generated

    September 17, 2026

    Subscribe to Updates

    Subscribe to our newsletter and get the latest tech news, app updates, AI trends, smartphone reviews, and exclusive deals delivered straight to your inbox.

    Facebook X (Twitter) Instagram Pinterest
    • About Us
    • Get In Touch
    • Disclaimer
    • Privacy Policy
    • Terms & Conditions
    © 2026 MyAppsPlus. All Rights Reserved.

    Type above and press Enter to search. Press Esc to cancel.