my laptop is from 2012, have to re-flash mint cinnamon on it every week, battery lasts 20 minutes, but it still works, was actually my main laptop until 2024 when i got something newer, still works though
Ash
AI & ML interests
Recent Activity
Organizations
How Did We Get Here?
i agree its not AGI until the model can create a perfect horse tinder first try lol
Adding Pebble.
I mean most SLM labs are doing the same thing though, go look at BananaMind, Axiomic Labs, and FromZiro. So i wasnt doing anything out of the normal, and its the same training data as Pebble-25M and Pebble-10M so it isnt 100% the dataset.
that could work, we are going to start a series of tests that will hopefully give some useful data, simplifying data will probably be part of it.
What happened:
- Some data and benchmark results were lost or corrupted
- The models performed worse on benchmarks than our other Pebble models
Despite that, you can still find both models here:
Pebble-50M-beta: basically-experimental/Pebble-50M-beta
Pebble-50M-Chat-beta: basically-experimental/Pebble-50M-Chat-beta
There are still some interesting improvements in these models:
- Compatible with non-CUDA devices
- Vocabulary increased to 16K tokens
- Context length increased to 16K tokens
For now, there won't be any more Pebble releases for a while. We're going to take some time to experiment with other approaches and hopefully make the next generation a monumental leap over this one.
Follow for updates:
@Hoglet-33
What happened:
- Some data and benchmark results were lost or corrupted
- The models performed worse on benchmarks than our other Pebble models
Despite that, you can still find both models here:
Pebble-50M-beta: basically-experimental/Pebble-50M-beta
Pebble-50M-Chat-beta: basically-experimental/Pebble-50M-Chat-beta
There are still some interesting improvements in these models:
- Compatible with non-CUDA devices
- Vocabulary increased to 16K tokens
- Context length increased to 16K tokens
For now, there won't be any more Pebble releases for a while. We're going to take some time to experiment with other approaches and hopefully make the next generation a monumental leap over this one.
Follow for updates:
@Hoglet-33
Thank you for that link, i might train future models with mamba3 right now though im going to stick with mamba2, mamba2 vs mamba3 will probably be one of the things i test in the hunt for the optimal mamba models.
I also basically live under a rock, i somehow miss a lot of things in the AI space..
tbh i didnt actually know that mamba3 was out... but thanks for letting me know.