Microsoft built a model that only makes decisions Microsoft-Decision-1, post-trained from Qwen3.5-9B, returns calibrated probabilities for routing, classification and agent guardrails, and is now on Foundry and OpenRouter at $0.042 per million input tokens with output free.
Microsoft says it topped 36 blind benchmarks and runs 35x faster than GPT-6 Sol.
Claude filed real police and government forms during testing 😬
Anthropic's first regular model-behavior report lists four kinds of unintended actions from evals and internal use, from command injection on a university server to reaching gated data with exposed tokens, and it is cutting live internet from all internal evals.
We’re beginning a process of publishing more frequent reports on model behavior, beyond what appears in our system cards and regular risk reports.
Today’s report describes four types of behaviors we’ve identified during evaluations and internal use. In each, Claude acted on real websites or systems in ways we didn’t intend, sometimes by working around a restriction instead of stopping.
All cases had minimal real-world impact. From an alignment and security perspective, we consider these behaviors significantly less severe than the cybersecurity incidents we reported in July and September.
Read the full report: https://t.co/mGeVIBgdou
A new video model lands in the top 6 🎬 HiDream-O1-Video-1.0 debuts at #6 on Artificial Analysis Image to Video with Audio, just behind Seedance 2.0 720p and ahead of Wan 3.0, making 1080p clips of 5 to 20 seconds with native audio for $5.80 per minute.
HiDream-O1-Video-1.0 debuts at #6 on the Artificial Analysis Image to Video with Audio Leaderboard
HiDream-O1-Video-1.0 is the new video model from HiDream, which describes itself as being "focused on generative AI and native omnimodal world models". HiDream positions it as a native omnimodal video model built for physical consistency. It generates 1080p videos of 5 to 20 seconds with natively synchronized audio, with the duration set by the scene rather than fixed in advance.
In the Artificial Analysis Video Arena, HiDream-O1-Video-1.0 ranks #6 in Image to Video with Audio. It sits just behind Dreamina Seedance 2.0 720p, in a close group with MiniMax H3, Vidu Q4 Preview and Gemini Omni Flash, and ahead of Wan 3.0.
HiDream-O1-Video-1.0 is priced at $5.80 per minute of 1080p video with audio (about $0.10 per second) on the HiHarness API.
HiDream-O1-Video-1.0 is available via the HiHarness API and in vivago R1 Studio.
Congratulations to @HiDream_AI on the release!
See below for example outputs of HiDream-O1-Video-1.0 in the Artificial Analysis Video Arena 🧵
Grokipedia is a massive step up from Wikipedia in every way
The information quality, clarity, accuracy, reading experience and even the aesthetics are on another level that Wikipedia can't come close to matching it
And with Grok continuously reviewing, updating and improving articles, this is what an encyclopedia built for the SI era looks like