🚀 MiMo Code V0.1 is now live and open-source!
More than an AI coding assistant in your terminal — it's the smartest coding partner you'll ever work with.
Comes with MiMo V2.5, a multimodal model available free for a limited time, featuring a million-token context window—ready to use out of the box.
♾️ Infinite Context: Knowledge accumulates automatically, and with lossless compression, even million-line projects keep every critical detail intact—quality never drops.
🧠 Agent-Model Synergy: An Agent framework deeply optimized for MiMo, with a full closed loop of testing, review, and validation—so complex tasks get done in one pass.
📝 Compose Mode: Specs → Plans → Build → Report. Design first, code second—clear thinking, no rework.
🔄 Self-Evolving System: Every session is automatically reviewed, distilling experience and best practices—the more you use it, the smarter it gets.
🎙️ Voice Input: Powered by MiMo-V2.5-ASR — just speak instead of type, and your voice becomes the prompt for truly hands-free coding.
🔌 Claude Code Compatible: Automatically loads your existing skills, MCP servers and commands, and reuses your API configuration—zero-cost migration, no setup required.
🌐 Open & Flexible: MIT licensed, with support for leading model providers including Anthropic, OpenAI, DeepSeek, Kimi, GLM and more.
Install in one line:
Mac & Linux
curl -fsSL https://t.co/ViHsb4eGns | bash
(For the best experience,we recommand Mac user use it on iTerm or vscode terminal)
Windows
npm install -g @mimo-ai/cli
🔗 Learn more
Website ↓
https://t.co/Aq9BVuyA9a
Blog ↓
https://t.co/f40wLgQicK
GitHub ↓
https://t.co/O5Mj3rzl9g
Meet DiffusionGemma ⚡ Our latest experimental open model (Apache 2.0) that generates text up to 4x faster.
Instead of predicting and typing just one word at a time like most language models, it drafts and refines entire blocks of text simultaneously.
Here’s how it works 🧵 ↓
🚀 The era of "AI forgingAI" is officially here!
Introducing ForgeTrain — the world’s first fully AI‑generated production‑level pre‑training framework.
No human in the loop. This is not an experimental prototype, but a true "AI engine" with zero human code intervention, ready for production.
This marks a complete closed loop: Framework written by AI → Validated on diverse hardware → Training leading models.
Highlights:
⚡ Speed & Efficiency: Outperforms NVIDIA's mainstream Megatron framework by 10% on H100 GPUs (achieving 44.13% MFU), directly reducing compute costs.
🌐 Cross-Hardware: Full pre-training pipeline successfully validated on Huawei Ascend 910.
🏆 Trained MiniCPM5-1B — ranked #1 among sub-2B models on the @ArtificialAnlys Index.
🔓 100% Open Source: The framework code (tailored for both NVIDIA H100 and Huawei Ascend) and the "Agent Harness" toolchain are fully available for reproduction — coming soon.
💻 GitHub: https://t.co/ChYFOgTXYh
Grok Build is now available in Beta for all SuperGrok and X Premium+ users.
Use Plan Mode, create images and videos with Imagine, and build automations or orchestrators with the CLI.
Visit https://t.co/bpTHpjivWD to get started.
MiniCPM5-1B is now live — the strongest open‑source base model under 2B.🚀🚀
🔥 Ranks #1 on the Artificial Analysis (AA) index for small models, scoring 17.9 to beat the 2B-scale Qwen3.5-2B (16.3).
⚡ Comprehensively surpasses Qwen3.5-0.8B and LFM2.5-1.2B-Thinking in knowledge, math, coding, and tool use.
🏗️ INT4 weights = just ~0.5GB — runs on phones, browsers, laptops.
👾 Powers fully offline AI “Desktop Pet” — no cloud, no GPU cluster.
Try the model here:
🤗 Hugging Face: https://t.co/jYRKhRYe48
💻 GitHub: https://t.co/zaffsLsx5m
🔭 Modelscope: https://t.co/mkOlyKNr2n
Meet LongCat-Video-Avatar 1.5🐱—our upgraded, open-source digital human framework.
Built for real production, not just short demos.
What's New:
🔹 Upgraded Audio Encoder: Replaces Wav2Vec2 with Whisper-Large, yielding significantly smoother and more natural lip dynamics.
🔹 Production-Ready Stability: Achieves accurate lip-synchronization, full-body temporal stability, and robust long-video generation with strict identity consistency.
🔹 Stylized Domain Generalization: Robustly generalizes to anime, animals, and complex real-world conditions such as multi-person interactions and object handling.
🔹 Efficient 8-Step Inference: Advanced step distillation accelerates inference to 8 NFE, balancing cost-effective serving with exceptional visual fidelity.
📊 LongCat-Video-Avatar 1.5 performs strongly in realism, naturalness, and stability, outperforming leading open-source models and closed systems.
🐱 Avatar 1.5 framework is now open source:
🔗 Weights & Code:https://t.co/b9fVxTLaPs
🔗 HuggingFace: https://t.co/s4GjqxuAA2
🔗 Tech Report: https://t.co/noSnB7aFCh
🔗 Project Page: https://t.co/CCeRcWgRpl
🚀 Introduce Hy-MT2: New Open-Source Multilingual Translation Model
We proudly launch our new Hy-MT2 translation model and the Tencent Hy Translation mini-program!
Hy-MT2 is a powerful multilingual model supporting seamless translation across 33 languages — and it's fully open-source!
It's 7B and 30B-A3B models achieve state-of-the-art performance among all open-source models on various translation tasks, surpassing models with dozens of times more parameters.
The lightweight 1.8B model even outperforms mainstream commercial APIs like Microsoft and so on. Powered by Tencent AngelSlim 1.25-bit extreme quantization, it needs just 440MB storage and enables effortless local inference on mainstream mobile chips — with 1.5x faster speed vs. Hy-MT1.5.
Open-source AI translation just got way smarter, faster, and more accessible! 🌏
Project Page: https://t.co/eRkw3JLJjB
Hugging Face: https://t.co/4J8gzMdcDu
Modelscope: https://t.co/xh5vEYkHpo
Github: https://t.co/EekzbAdjOX
Meet Gemini 3.5 Flash — our strongest agentic and coding model yet.
It delivers frontier-level performance at 4x the speed of comparable frontier models — often at less than half the cost.
Generally available, starting today. 🧵
#GoogleIO
Introducing Gemini Spark ✨
It’s your 24/7 personal AI agent that helps you navigate your digital life, taking action on your behalf, and under your direction.
🧠 It runs on Gemini 3.5 and is built on @Antigravity, so it can perform long-running tasks easily in the background.
⏱️ And because it runs on dedicated virtual machines on Google Cloud, you don’t even need to keep your laptop open.
🧰 Spark will integrate seamlessly with Google tools, and soon with third parties through MCP.
#GoogleIO
1/6 Introducing Qwen3.5-LiveTranslate: Next-gen real-time interpretation is here. 🌍
We’re breaking down language barriers with 3,500+ language pairs, ultra-low latency, visual context, real-time voice cloning, and hotword customization.
Engineered to help you ship native, frictionless real-time translation experiences to a global audience.
🚀 Ring-2.6-1T is now open source.
A trillion-scale flagship thinking model built for real-world complex tasks: Agent workflows, coding & engineering, long-horizon tasks, complex reasoning, research, and enterprise automation.
It is designed to move beyond “answering” toward execution: understanding context, planning steps, calling tools, and staying stable across long task chains.
Highlights:
- Advanced agentic workflow support.
- Reasoning effort levels: high for agentic tasks, xhigh for complex reasoning.
- Scalable asynchronous RL via the IcePop algorithm, enabling stable, trillion-scale training for long-horizon agentic RL.