Qwen3.8-Flash-Next might not run well on one dgx spark at all 😨
it has a huge 51B Engram table that even when compressed to 4 bit would use around 25gb of RAM, but because of DGX spark unified memory the total the entire model would use 105–125GB of unified memory total making it very close
@maria_rcks fair lol it is extremely token inefficient making it extremely expensive with intelligence literally worse than Gemini somehow on AA which is why I think it belongs down there with Google
“quietly deleted their posts” meanwhile them apologizing and publicly saying they deleted it lmao. This guy sketches me out, constantly hating on people and glazing Elon like what?
why do people think the dgx spark is better than the pro 6000? IMO even 4 DGX sparks aren’t as good, I do training and the memory bandwidth and worse compute kills the DGX spark for me