The final shootout - shoutout @nanbeige for birthing the current best small LLM.
Nanbeige4.2-3B is our winner ๐
Mad respect to our runners up @SparkLLM and @TheInclusionAI . Both amazing models, and honestly more truly "small"
https://t.co/ezDWLXR2jQ
@notkingofchids I was surprised how big the gap was too. Plenty of it's failures were verbosity. I had theorized it to be the serving config, but we'd see the same gap with 4B if that were the case. Likely worth a re-test now that Spark X2.5 support is in llama.cpp main (gguf)