Join the conversation

Join the community of Machine Learners and AI enthusiasts.

Sign Up
appvoidย 
posted an update about 21 hours ago
Post
829
I don't know if it was us or one of you guys or maybe all of us at once but lately we have seen a finetuning/pretraining explosion of models below 200m params and we can't be more happy about it keep coming tinkerers all of this is possible because of you!

I've seen this wave as well and it's great to see!

it might be me
my models are <50m
fun fact: making small models uses less storage on hf

  • BananaMind 2
  • Supra2
  • SupraBrain
  • GPT-X2.5
  • Dillion-v2
  • Glint 2
  • Qana-mini
  • void.0
  • BarunLM and Action 35M
  • Min-Spark
  • Rose-Medium

these are all of them i've been keeping track of

ยท

great! don't forget rose-mini!