Congrats to the @ornith_ team on their Ornith 1.5 release today!
Use the models via API at 50% off: https://t.co/FYbxJcvzjm
ornith-1.5-397b
ornith-1.5-35b-a3b
ornith-1.5-9b
Aloha! 🌺Introducing Ornith-1.5, a family of open-source LLMs spanning 9B Dense, 35B MoE, and 397B MoE, trained with self-improving strategies.
It achieves state-of-the-art performance among open-source models of comparable size and delivers performance comparable to Claude Opus 4.8 across reasoning, agentic, and coding tasks:
✅Terminal-Bench 2.1 (86.1)
✅SWE-Bench (86 on verified, 65.1 on pro, 79.6 on Multilingual)
✅DeepSWE (56)
✅HLE (44.6)
✅ClawEval (81.4)
✅Tool Decathlon (71.2)
Ornith-1.5 takes a major step toward training foundation models through end-to-end self-improvement, extending the self-scaffolding strategies introduced in Ornith-1.0 into a more complete self-improvement loop: the model proposes new tasks, generates task-specific scaffolds, and produces solution rollouts for reinforcement learning, continuously creating new learning experiences from which it can improve.
All models, along with their quantized versions (FP8, GGUF, MLX, and NVFP4), have been released under the MIT License, enabling unrestricted commercial and research use.
📘Tech Blog: https://t.co/OZ63scRWLB
🤗Huggingface: https://t.co/mGJLwhrQOM
Aloha! 🌺Introducing Ornith-1.5, a family of open-source LLMs spanning 9B Dense, 35B MoE, and 397B MoE, trained with self-improving strategies.
It achieves state-of-the-art performance among open-source models of comparable size and delivers performance comparable to Claude Opus 4.8 across reasoning, agentic, and coding tasks:
✅Terminal-Bench 2.1 (86.1)
✅SWE-Bench (86 on verified, 65.1 on pro, 79.6 on Multilingual)
✅DeepSWE (56)
✅HLE (44.6)
✅ClawEval (81.4)
✅Tool Decathlon (71.2)
Ornith-1.5 takes a major step toward training foundation models through end-to-end self-improvement, extending the self-scaffolding strategies introduced in Ornith-1.0 into a more complete self-improvement loop: the model proposes new tasks, generates task-specific scaffolds, and produces solution rollouts for reinforcement learning, continuously creating new learning experiences from which it can improve.
All models, along with their quantized versions (FP8, GGUF, MLX, and NVFP4), have been released under the MIT License, enabling unrestricted commercial and research use.
📘Tech Blog: https://t.co/OZ63scRWLB
🤗Huggingface: https://t.co/mGJLwhrQOM
Deepseek V4 Flash
$0.02/m input | $0.04/m output and $0.006/m cache
https://t.co/Pn8E7MySXc
Things are moving quickly and we're happy to still be able to provide Deepseek v4 Flash at 85% off launch rates with high tps and low latency!
Deepseek V4 Flash 0731 is being offered free by plenty providers today. Fine if you want constant timeouts!
We're offering it for ALMOST FREE (85% off) with low latency and high tps on our dedicated servers.
Get away from the freeloaders https://t.co/SChn4CbfmT