Qwen3.8-Flash-Next might not run well on one dgx spark at all ๐จ
it has a huge 51B Engram table that even when compressed to 4 bit would use around 25gb of RAM, but because of DGX spark unified memory the total the entire model would use 105โ125GB of unified memory total making it very close
@maria_rcks fair lol it is extremely token inefficient making it extremely expensive with intelligence literally worse than Gemini somehow on AA which is why I think it belongs down there with Google
โquietly deleted their postsโ meanwhile them apologizing and publicly saying they deleted it lmao. This guy sketches me out, constantly hating on people and glazing Elon like what?
why do people think the dgx spark is better than the pro 6000? IMO even 4 DGX sparks arenโt as good, I do training and the memory bandwidth and worse compute kills the DGX spark for me