@rez0__ Well, some people decided to ruin it for everyone and share on how to find critical using gpt-5.6 , which also I believe openai got bombarded with new Trusted Cyber Access so they were forced to block cyber related works.
All the ads from Chinese labs are pro-human life.
No AI gonna wipe out everyone’s job.
No AI escaping fear marketing.
Just better life to enjoy and chill because of AI.
I can live with these messages.
We're starting to leave the territory where you'd test an LLM by e.g. "create an svg of pelican on a bicycle". As one idea to generalize it, I was interested what Opus 5 would do if I gave it the first paragraph of the Lord of the Rings, a 1M token budget (~$10) and asked for three js render of it. Opus went off for ~2 hours and wrote 5500 lines of code that (procedurally) rendered the story. It's kind of janky but fun. But it's a bit mindboggling that the LLM has to place and orchestrate various polygon assets in (x,y,z) coordinates and write code that animates it all, and that it even does anything at all.
I also like this kind of examples because no one in their right mind would ever spend the time to write something this custom but LLMs have all the stamina and patience in the world, so it's an example where we go from "no one would ever do this" to "sure, why not, it's ~free". There might be a lot more. But I'm excited about creating hyper custom worlds that you can imagine dropping players into, e.g. here to participate in the LoTR story as a spectator NPC, or one of the characters, or etc. Something like an ephemeral GTA of X on demand.
Last thought is that the domain of worlds/games exposes a weakness in LLMs: they can't easily audit their work because they aren't able to efficiently and natively perceive videos or play games within them. Here, Opus 5 had to very slowly and painstakingly take screenshots at different points, and it messed up a few times and created a bunch of jank. An example of raw capability (multimodal, gameplay) that I think is still quite lacking.
🚀 DeepSeek-V4-Flash Official API is now LIVE in public beta!
🔷 We’ve massively upgraded its Agent capabilities—benchmark scores are now far surpassing the V4-Pro-Preview. Check out the massive performance leap below! 👇
🔷 The official V4-Flash now natively supports the Responses API format and is fully adapted for Codex!
Check out the configuration details in our official API docs: https://t.co/smCwQZMeiq
Metasploit 6.5 is out just in time for Hack Summer Camp. This release comes with Malleable C2 support for Meterpreter, more relaying improvements and an integrated MCP server. Check out all the details here: https://t.co/GcjZVboiUT
We compared 1-bit Kimi K3 to Claude Opus 5 and GPT 5.6.
We gave 4 models the same prompt:
Create a glass aquarium whose side panel develops a visible crack and then bursts.
1-bit Kimi K3 GGUF ran locally on 4x B200s at 36 tok/s.
GitHub repo: https://t.co/aZWYAtakBP
We've open-sourced AgentENV in collaboration with kvcache-ai.
AgentENV is a distributed system for running agent environments at scale. Its components power agentic RL training for Kimi K3, with fast snapshot, resume, and fork support for large-scale parallel agent workflows.
Explore on GitHub:
https://t.co/Dsxyw5rLqm
We've open-sourced MoonEP, our high-performance communication library for distributed MoE workloads.
Built to make expert-parallel communication more efficient at scale, MoonEP helps reduce communication overhead in large MoE training and inference systems.
Explore on GitHub:
https://t.co/h2Rcwg88dQ
@Mononofu@JensenHuang This is a false equivalence. Everyone has the right to keep their code private. The problem is when someone tries to stop OTHERS from open sourcing.