🚨 DeepSeek Code Is Coming
> DeepSeek is building a dedicated AI coding agent powered by its new Harness framework
> The project is designed for autonomous software engineering, including planning, tool use, and code execution
> Harness will provide long-running agent workflows with memory and repository awareness
> Recent V4-Flash benchmarks were already evaluated using DeepSeek Harness, signaling it’s a core part of DeepSeek’s roadmap
> The company is positioning it as a direct competitor to Claude Code and OpenAI Codex
> Closed beta testing is expected to begin soon
Can DeepSeek Code become the biggest open-source coding agent? 👀
You can now run a 2.8 TRILLION parameter model on a 4GB GPU.
Someone open-sourced a tool that uses "Layer-wise Inference." It only loads one layer onto your GPU at a time. so the VRAM you need depends on the layer size, not the model size.
no quantization. no distillation. no pruning.
→ DeepSeek-V3 (671B) on 12GB
→ Llama 3.1 405B on 8GB
→ Kimi K3 (2.8 TRILLION params) on under 4GB
→ works with almost every open model
the biggest model on it needs the LEAST VRAM. K3 is sparse MoE, so it streams only the experts a token actually routes to instead of a whole dense layer.
2.8 trillion parameters running in less VRAM than a 70B.
Codex users: do this right *now*.
Max reasoning effort is off by default.
Luna at Max reasoning is ~ Sol Medium / Opus 5 Medium level at 1/6th the cost.
Change your life today.
DeepSeek V4 Flash 0731 is now open weights!
@deepseek_ai has just released the weights for its new flash tier model, DeepSeek V4 Flash 0731. With a score of 50 on the Artificial Analysis Intelligence Index, it lands among the top 3 open weights models on the leaderboard. The weights are released under the MIT license, allowing unrestricted commercial use and modification.
DeepSeek V4 Flash 0731 shares identical architecture and pricing with the earlier DeepSeek V4 Flash. At a size of 284B total parameters (13B active), released in mixed FP4/FP8 precision at ~167GB total file size, it lands on our Pareto frontier for Intelligence Index vs. Total Parameters. Among open weights models, DeepSeek V4 Flash 0731 delivers a significant leap in intelligence for its size class. DeepSeek V4 Flash 0731 is also available now through DeepSeek's first-party API.
Check out Artificial Analysis to compare DeepSeek V4 Flash 0731 with other leading open weights and proprietary models: https://t.co/zeUIrzHIOC