@RoundtableSpace Qwen3.8 scaling is wild. 2.4T open-weights is a massive shift for local benchmarks. Curious how the VRAM requirements look for a usable quant.
@pankajkumar_dev The output variance at same cost is real โ reminds me I benchmark 2-3 prompts before committing to one model for production templates.