Today, we are announcing a brand-new series of SupraLabs models: Supra2 This series will feature various models, including such as: - ๐ Supra2-Nano (0.4M) โ The smallest Supra2 model. - ๐ค Supra2-Small (1.4M) โ The tiny model that runs everywhere. - ๐ช Supra2-Medium (25M) โ Our medium class model in the Supra2 family. The powerful midsizer. - ๐ฅ Supra2-Pro (100M): base, instruct, reasoning, code, math and more! โ The most capable model yet! A real allrounder for all your everyday tasks. - ๐จ Supra2-IMG โ our generative text-to-image model ...and many more...
Current progress: - Nano (0.4M) and Small (1.4M): in training; almost done. Baseline set. - Medium (25M): coming soon... - Pro (100M): in training; finishes in 66 hours - Monday, 3rd August 2026, 12:00AM - IMG: coming soon...
You can support us with a like and follow if you want! Don't miss our next release! Stay tuned...
I went into this expecting to find a ~Q2 garbage dumpster but prism-ml/Ternary-Bonsai-27B-gguf is a slick feat of QAT engineering.
It is weaker then FP16 on a handful of tasks where quants usually degrade, as per their own paper the loss is "concentrated on sustained chains of reasoning / agentic" and in ReasonScape this bites on Sort, Shuffle and Dates, but counter-acting this are some noticeable improvements to thinking length without accuracy loss on several other tasks (Shapes, Cars).
I haven't had a chance to run the Binary yet, but PQ2 + Bonsai QAT are confirmed to be pretty darn impressive.