🚀 Meet Qwen-Image — a 20B MMDiT model for next-gen text-to-image generation. Especially strong at creating stunning graphic posters with native text. Now open-source.
🔍 Key Highlights:
🔹 SOTA text rendering — rivals GPT-4o in English, best-in-class for Chinese
🔹 In-pixel text generation — no overlays, fully integrated
🔹 Bilingual support, diverse fonts, complex layouts
🎨 Also excels at general image generation — from photorealistic to anime, impressionist to minimalist. A true creative powerhouse.
Blog:https://t.co/FZ415nwcAp
Hugging Face:https://t.co/XSLwjJ6lYa
ModelScope:https://t.co/mcZ0dHeD64
Github:https://t.co/A9yvJZ6TJc
Technical report:https://t.co/tld83OrOUP
Demo: https://t.co/MgEAr4uqoR
🎨 Meet Qwen-Image-3.0 — the third generation of our foundational image generation model.
If 1.0 was about "Precision," and 2.0 added "Variety, Completeness, Beauty & Authenticity," then 3.0 comes down to a single word: Real (实).
Three dimensions of "Real":
📰 Rich Content ��� prompts up to 4.5k tokens. One-pass generation of complex layouts: newspapers, storyboards, exam papers — even a 3×3 infographic grid or picture-in-picture-in-picture UIs.
🔬 Authentic Details — text legible down to 10px, full LaTeX paper pages, pores, hair strands & near-photographic skin texture.
🌏 Deep Knowledge — native rendering in 12 languages, 100+ art styles, realistic UIs (web / games / livestreams), plus world knowledge & live web retrieval.
Not just "good-looking" — genuinely useful. Image generation as a real productivity tool for design, content, education & e-commerce.
Go create 🏃🎨
💬Qwen Chat: https://t.co/941HmITJ2W
📝Blog: https://t.co/5mnS4uI9Ar
🚀 Introducing the Qwen 3.5 Small Model Series
Qwen3.5-0.8B · Qwen3.5-2B · Qwen3.5-4B · Qwen3.5-9B
✨ More intelligence, less compute.
These small models are built on the same Qwen3.5 foundation — native multimodal, improved architecture, scaled RL:
• 0.8B / 2B → tiny, fast, great for edge device
• 4B → a surprisingly strong multimodal base for lightweight agents
• 9B → compact, but already closing the gap with much larger models
And yes — we’re also releasing the Base models as well.
We hope this better supports research, experimentation, and real-world industrial innovation.
Hugging Face: https://t.co/wFMdX5pDjU
ModelScope: https://t.co/9NGXcIdCWI
🚀 Introducing the Qwen 3.5 Medium Model Series
Qwen3.5-Flash · Qwen3.5-35B-A3B · Qwen3.5-122B-A10B · Qwen3.5-27B
✨ More intelligence, less compute.
• Qwen3.5-35B-A3B now surpasses Qwen3-235B-A22B-2507 and Qwen3-VL-235B-A22B — a reminder that better architecture, data quality, and RL can move intelligence forward, not just bigger parameter counts.
• Qwen3.5-122B-A10B and 27B continue narrowing the gap between medium-sized and frontier models — especially in more complex agent scenarios.
• Qwen3.5-Flash is the hosted production version aligned with 35B-A3B, featuring:
– 1M context length by default
– Official built-in tools
🔗 Hugging Face: https://t.co/wFMdX5pDjU
🔗 ModelScope: https://t.co/9NGXcIdCWI
🔗 Qwen3.5-Flash API: https://t.co/82ESSpaqAF
Try in Qwen Chat 👇
Flash: https://t.co/UkTL3JZxIK
27B: https://t.co/haKxG4lETy
35B-A3B: https://t.co/Oc1lYSTbwh
122B-A10B: https://t.co/hBMODXmh1o
Would love to hear what you build with it.
🚀 Qwen3.5-397B-A17B is here: The first open-weight model in the Qwen3.5 series.
🖼️Native multimodal. Trained for real-world agents.
✨Powered by hybrid linear attention + sparse MoE and large-scale RL environment scaling.
⚡8.6x–19.0x decoding throughput vs Qwen3-Max
🌍201 languages & dialects
📜Apache2.0 licensed
🔗Dive in:
GitHub: https://t.co/NzNdS9joAT
Chat: https://t.co/bg4tAU0Rhw
API:https://t.co/YiiyKTnHoU
Qwen Code: https://t.co/qqwj5nAger
Hugging Face: https://t.co/wFMdX5p5um
ModelScope: https://t.co/9NGXcId57a
blog: https://t.co/AW8UQStXaL
🖼️ AI Slides Just Got a Brain Upgrade!
🎯Meet Qwen AI Slides — your personal presentation designer that actually thinks. Powered by Qwen3 Agent + Qwen-Image 2.0, it turns chaos into clarity, beautifully.
✅ Type a simple idea, throw in a wall of text, or upload a document — it just works
✅ Search Agent researches, organizes & builds the story structure for you
✅ Qwen-Image 2.0 directly generates each slide as a polished visual — text, layout, color scheme & graphics, all in one shotStop designing slides. Start designing impact. 💡✨
Try it now: https://t.co/R1P4UxCAD3
🚀 Introducing Qwen-Image-2.0 — our next-gen image generation model!
🎨 Your imagination, unleashed.
✨ Type a paragraph → get a pro slides
✨ Describe a scene → get photoreal 2K magic
✨ Add text → it just works (no more glitchy letters!)
✨ Key upgrades:
✅ Professional typography (1K-token prompts for slides, posters & comics)
✅ 2K native resolution with stunning detail
✅ Flawless text rendering + unified generation/editing
✅ Lighter architecture = faster inference
Try it now → https://t.co/941HmITJ2W
Full details → https://t.co/OtKi6vuU4S
Qwen3-ASR and Qwen3-ForcedAligner are now open source — production-ready speech models designed for messy, real-world audio, with competitive performance and strong robustness.
● 52 languages & dialects with auto language ID (30 languages + 22 dialects/accents)
● Robust in noisy and complex settings (yes, singing and songs too)
● Long audio support: up to 20 minutes per pass
● Word/phrase-level timestamps: high-precision alignment for 11 languages via Qwen3-ForcedAligner, stronger than MFA/CTC/CIF-style aligners
Also included: a full open-source inference & finetuning stack with vLLM batch, streaming, and async serving.
GitHub: https://t.co/yxxUTTIjLr
Hugging Face: https://t.co/38fny0K2Qz
ModelScope: https://t.co/0fAIFksAwv
Hugging Face Demo: https://t.co/GbLHOn7hdD
ModelScope Demo: https://t.co/Jvb2kfnHr5
Blog: https://t.co/yP6VbGc7bg
Paper: https://t.co/zzlFgwcP1E
🚀 Introducing Qwen3-VL-Embedding and Qwen3-VL-Reranker – advancing the state of the art in multimodal retrieval and cross-modal understanding!
✨ Highlights:
✅ Built upon the robust Qwen3-VL foundation model
✅ Processes text, images, screenshots, videos, and mixed modality inputs
✅ Supports 30+ languages
✅ Achieves state-of-the-art performance on multimodal retrieval benchmarks
✅ Open source and available on Hugging Face, GitHub, and ModelScope
✅ API deployment on Alibaba Cloud coming soon!
🎯 Two-stage retrieval architecture:
📊 Embedding Model – generates semantically rich vector representations in a unified embedding space
🎯 Reranker Model – computes fine-grained relevance scores for enhanced retrieval accuracy
🔍 Key application scenarios:
Image-text retrieval, video search, multimodal RAG, visual question answering, multimodal content clustering, multilingual visual search, and more!
🌟 Developer-friendly capabilities:
• Configurable embedding dimensions
• Task-specific instruction customization
• Embedding quantization support for efficient and cost-effective downstream deployment
Hugging Face:
https://t.co/QBTP0XEVmk
https://t.co/c0DB96xxP5
ModelScope:
https://t.co/MhPATTPjYJ
https://t.co/JyNIYQqLAm
Github: https://t.co/qG5Khxk7o9
Blog: https://t.co/K52IC34oNV
Tech Report:https://t.co/dtqAOkMprp
🎁 A New Year gift from Qwen — Qwen-Image-2512 is here.
🚀 Our December upgrade to Qwen-Image, just in time for the New Year.
✨ What’s new:
• More realistic humans — dramatically reduced “AI look,” richer facial details
• Finer natural textures — sharper landscapes, water, fur, and materials
• Stronger text rendering — better layout, higher accuracy in text–image composition
🏆 Tested in 10,000+ blind rounds on AI Arena, Qwen-Image-2512 ranks as the strongest open-source image model, while staying competitive with closed-source systems.
👉 Try it now in Qwen Chat: https://t.co/941HmITJ2W
🤗 Hugging Face: https://t.co/mP4AFvdvH1
📦 ModelScope: https://t.co/Jq34O0RGQw
💻 GitHub: https://t.co/A9yvJZ6TJc
📝 Blog: https://t.co/mr4UVRvQlT
🤗 Hugging Face Demo: https://t.co/MrnQEn44zx
📦 ModelScope Demo: https://t.co/NCvu7M4Z6k
✨API: https://t.co/9N3jB1f8Ll
🎆 Start the New Year with better images.
Merry Qwristmas! 🎄🎁
Huge thanks for all the love and support this year.
Get ready for Qwen’s New Year surprises! 🎆✨
See you next year — with even more Qwen magic. ✨
🏆 We are incredibly honored to announce that our paper, "Gated Attention for Large Language Models: Non-linearity, Sparsity, and Attention-Sink-Free" has received the NeurIPS 2025 Best Paper Award!
A huge congratulations to our dedicated research team for pushing the boundaries of AI.
Read more: https://t.co/qu3ERa3pH5
🚀 Update Next Scene V2 only 10 days after last version, now live on Hugging Face
👉 https://t.co/NbjAWj5wO3
🎬 A LoRA made for Qwen Image Edit 2509 that lets you create seamless cinematic “next shots” — keeping the same characters, lighting, and mood.
I trained this new version on thousands of paired cinematic shots to make scene transitions smoother, more emotional, and real.
🧠 What’s new:
• Much stronger consistency across shots
• Better lighting and character preservation
• Smoother transitions and framing logic
• No more black bar artifacts
Built for storytellers using @ComfyUI or any diffusers pipeline.
Just use “Next Scene:” and describe what happens next , the model keeps everything coherent.
🧩 Try it directly in ComfyUI, or check the thread to launch it on @fal .
Open-source, no restrictions, made for filmmakers, animators, and dreamers.
@ComfyUI #AIcinema #LoRA #Flux #Qwen #ComfyUI #AIart #GenerativeVideo
you can test on comfyui or to try on https://t.co/SpUceQW1mb, you can go here :
https://t.co/K3amauYhT2
and use my lora link :
https://t.co/Vp8bBEr5Ol
start your prompt with "Next Scene:" and lets go !!
Qwen Image Edit 2509 is the new leading open weights image editing model, ranking #3 overall in the Artificial Analysis Image Editing Arena and introducing multi-image editing capabilities!
The latest release from Alibaba Qwen trails only Gemini 2.5 Flash (Nano-Banana) and Seedream 4.0 in our Image Editing Arena, while leading all open weights alternatives.
Qwen Image Edit 2509 supports multiple image inputs combined with text prompts, enabling finer control of the output image and a notable upgrade from the single image input of the original Qwen Image Edit model.
Released under Apache 2.0 license with weights available on @huggingface, it maintains competitive pricing at $30/1k images on @fal and @replicate (listed as "Qwen Image Edit Plus"), matching its predecessor while delivering enhanced capabilities.
See below for comparisons between Qwen Image Edit 2509 and other leading models in our Image Editing Arena 🧵