I mean most SLM labs are doing the same thing though, go look at BananaMind, Axiomic Labs, and FromZiro. So i wasnt doing anything out of the normal, and its the same training data as Pebble-25M and Pebble-10M so it isnt 100% the dataset.
๐ฝ Can AI go rogue?
Ash
Hoglet-33
AI & ML interests
Open source AI, datasets, parameter efficiency, SLMs, AI for the betterment of humanity. Contact at ash@basicallyai.co
Recent Activity
liked a Space 2 days ago
AxiomicLabs/Open_SLM_Leaderboard new activity 2 days ago
AxiomicLabs/Open_SLM_Leaderboard:Adding Pebble. repliedto their post 2 days ago
Today, we planned to release Pebble-50M and Pebble-50M-Chat to the world. Unfortunately, due to a few issues, that didn't go quite as planned.
What happened:
- Some data and benchmark results were lost or corrupted
- The models performed worse on benchmarks than our other Pebble models
Despite that, you can still find both models here:
Pebble-50M-beta: https://huggingface.co/basically-experimental/Pebble-50M-beta
Pebble-50M-Chat-beta: https://huggingface.co/basically-experimental/Pebble-50M-Chat-beta
There are still some interesting improvements in these models:
- Compatible with non-CUDA devices
- Vocabulary increased to 16K tokens
- Context length increased to 16K tokens
For now, there won't be any more Pebble releases for a while. We're going to take some time to experiment with other approaches and hopefully make the next generation a monumental leap over this one.
Follow for updates:
@Hoglet-33
https://huggingface.co/basically-ai
https://huggingface.co/basically-experimental