Dear NVIDIA RTX 16 GB GPU users 🫵
You'll soon be able to easily run Qwen3.8-27B with large context + MTP. And no, it's not Q2 GGUF 👀
That includes all RTX 5060 ti/5070 ti/5080 & older NVIDIA 16 GB GPUs.
Linux & Windows.
Coming soon
Introducing Gemini 3.8, our best reasoning & coding model yet.
By leveraging long-running agentic loops, we’re building on the momentum of 3.7 Flash from just three weeks ago to release two new 3.8 variants:
Meet Gemini 3.8 Flash and Gemini 3.8 Flash Cyber 🧵