Releasing the model weights and technical report of Kimi K3.
Kimi K3 is our most capable model: a 2.8T MoE model with native visual understanding and a 1M-token context window.
New model architecture: 2.5x the intelligence per unit of compute, not just more params.
Alongside Kimi K3, we're opening up more of the stack behind it — high-performance attention kernels, MoE communication library, and infrastructure for running agent environments at scale.
Model weights: https://t.co/7m7eEg6Y0B
Tech report: https://t.co/yeu6cjpMCT
Tech blog: https://t.co/YTfiMSNM1f
I was going to write a whole thing about information monopolies and why open source has to win, but I got lazy. Assume I made several excellent points.
It’s time ComfyUI had real VFX capabilities.
Real-time preview while you adjust, multiple effects chained together, changes propagating down the chain, proxy loading for large files…
No more switching between workflows, what you see is what you get, and everything happens on a single canvas.
Color grading, curves, LUTs, stylization, green-screen keying, frame interpolation, blur, sharpen…
if you can think of it, it’s here.
All of it, in ComfyTV.
Claude is better and I’m prepared to fight for this completely independent opinion. Also, entirely unrelated: thank you Anthropic for the free 6 months of 20x Max.
This is me talking to my computer without making a sound.
After just a month of collecting data, our model is already approaching dictation in accuracy. We were surprised to see that it generalizes to unseen participants as well!
(1/n)
Stupid question: what if real art starts when we stop writing prettier prompts and start messing with the activations until the model hallucinates on purpose? I don’t mean the hallucination after it becomes an output. I mean the weird internal moment before the model remembers how to behave. I’d write more, but my English ran out.
New Anthropic research: A global workspace in language models.
Of everything happening in your brain right now, only a tiny fraction is consciously accessible—thoughts you can describe, hold in mind, and reason with.
We found a strikingly similar divide inside Claude.
Dropped GreyZone. For Krea 2.
Just that strange early 3D atmosphere.
Not clean, not modern, not pretty in the normal way. Just old 3D mood and weird places.
https://t.co/7gsD7A2d5h
🎨 I trained 999 Style LoRAs with the @fal Krea 2 Trainer.
Every single image in this video is a different LoRA.
That's 999 unique Style LoRAs, each trained for just 100 steps at 0.00035 LR, taking an average of ~30 seconds to train.
As I said before, @krea_ai 2 continues to prove that it's an incredible model for open-source Style LoRAs. Next week I'll be sharing even more LoRAs, along with a few surprises.