you have to shovel an absurd amount of total garbage out of your head just to clear the pipes. you must create a lot of bad art in order to create good art. if you sit there waiting for pure gold on the first swing, you’re going to die holding an empty shovel. you gotta make the ugly stuff - let the bad ink bleed out of your system until the stream runs clear.
2024: A gaming PC with a RTX 4090 ran Llama 3 70B at ~2 tok/s in 4-bit.
2026: A gaming PC with a RTX 5090 runs DeepSeek-V4-Flash 284B at ~24 tok/s with native mxfp4.
Prediction: by the end of 2027, a gaming PC will run Fable 5-class intelligence locally at decent speed.
God is not going to drag your bitter half formed soul into a promised land, it would spoil the milk and spoil the honey, and you would wreck it within a year. the thirst is teaching you. the rations are building someone who can receive a promise without choking on it. blessed are those who walk it without cursing the sand, for every grain grinds a slave into a son
We’re releasing new Qwen3.8-27B GGUFs with 10% higher accuracy.
Unsloth Dynamic V3 outperforms others by >10% on Div-300, KLD & more benchmarks.
We also release 1-bit quants that retain 77% accuracy. Run on 8GB RAM.
Blog: https://t.co/tHsBexyh2K
GGUF: https://t.co/xIdNwm7CLQ
Introducing GLM-5.3: Built to Code. Ready for Cyber Defense.
- Top-tier coding and agentic capabilities, achieved through post-training on the 743B base model
- A major leap in cybersecurity, setting a new standard among open models
Tech Blog: https://t.co/ekQkO83jCv
true neuroplasticity requires a massive ego. you have to be delusional enough to think you can master a complex system overnight, so you push past the exact wall of mental exhaustion where normal people quit.
🧩 DeepSeek Harness v0.1 is now available in Developer Preview!
🔹 We’re opening it up to developers building agent harnesses worldwide and open-sourcing the codebase in MIT license.
🔹 Powered by the Cordis meta-framework, DeepSeek Harness is an agent harness built around one core idea: Everything is a plugin. Models, tools, skills, sessions, sandboxes, filesystems, loops, orchestration, and UI are ALL implemented as plugins, and can be mixed, matched, replaced, and extended.
Try it now!
https://t.co/2YWSvJHhKA
my mom has been telling me my whole life that you can stratify every person into the categories of giver or taker i used to argue with her about nuances and such but no she was completely right.
If you just follow your heart you'll end up in hell, because hearts are liars and feelings are fools and the road to destruction is paved with authentic self-expression, which is why wise men follow their minds and holy men follow God and only idiots follow their hearts
The entire internet is going to be a ravaged soon. Not by an individual misaligned powerful AI model, but millions upon millions of relatively inexpensive models, each incredibly aligned, aligned to the human they serve
Iraq is offering effectively a $25-$29 a barrel payment to anyone willing to risk loading at its port inside the Persian Gulf and take the barrels out.
On a single VLCC, that’s a ~$50-$58 million potential profit, minus costs (and the risk of getting sunk by Iran, that’s it).
What are the best models you can run on your @NVIDIAAI DGX Spark? ✨
Aug 2026 Edition
1× DGX Spark
• DeepSeek v4 Flash - 1M ctx, 26 tok/s - recommended!
• Qwen 3.6 35b NVFP4 - 256k ctx, 81 tok/s
• Qwen 3.6 27b NVFP4 - 256k ctx, 33 tok/s
Qwen 3.8 27b - should be released soon and this could change my recommendation!
2× DGX Sparks ← sweet spot!
• DeepSeek v4 Flash 0731 - 1M ctx, 82 tok/s
• Inkling-Small - 1M context, Full Omni, 33 tok/s
• MiMo-V2.5 - 1M ctx, Full Omni, 31 tok/s
• Step-3.7-Flash — 256K ctx, 30 tok/s
3× DGX Sparks
• GLM-5.2 with Vision - 348k context, 25 tok/s - still the best intelligence you can run locally if you have 3 sparks.
• DeepSeek v4 Flash on 2 units + smaller models on the 3rd spark for images, ComfyUI and other things. This setup gives you DeepSeek v4 Flash speeds for coding, plus strong agentic workflows and image support from smaller models.
4× DGX Sparks
• GLM 5.2 NVFP4 across all 4 units - still think this is the way to go if you have 4 units!
• The alternative is running DeepSeek v4 Flash on 2 units and any other 2x setup on the other units. The choice is your.
Links and repos below 👇