@0xBakeer Let me know when you are doing the 2x spark version, I decided to switch qwen3.8-flash to fp8 the way I landed on running 3.6 and 3.8 27b. Something always feels off with the 4 bit quants and kV not at 16 bit for me.
@0xBakeer Thank you. Mia's is a bit too optimized for speed and high context for my liking, but blaze seems solid so far, and back on the RadixArk that your recipe uses that I had become familiar with.
@0xBakeer I'm running your qwen3.8-flash vllm single spark recipe, I really appreciate your work. From your benchmarking would you recommend a different recipe?
@darvasch@UnslothAI@Alibaba_Qwen I'm seeing significant reliability/capability decrease of this model relative to the FP8 running in openclaw, running on gx10 as well, modified eugr recipe. Have you done any evals beyond basic token/second?
@thdxr People who think they are "logical" have no awareness of where their thoughts come from, get in touch with your intuition folks, and stop getting in its way
@DeepPsycho_HQ@HeidiPriebe1 The gems should be internal for the one on the right, changing to steps is the same thing as the singular goal just on a different scale