Sonnet 5.5 improves on Sonnet 5 across benchmarks, in some cases dramatically.
It’s a faster, lower-cost complement to Claude Opus 5.5, strongest at well-scoped everyday tasks, fixing bugs, and creating polished documents, slides, and spreadsheets.
Grok 4.7 by @SpaceXAI just landed in Agent Arena at #16, with a net improvement score of +3.96%
Grok 4.7 (xHigh) shows improvement over previous versions, but comes with a comparable lift in price vs. Grok 4.6 (High):
- Grok 4.7 (xHigh): $1.14 median per task / +3.96 net improvement
- Grok 4.6 (High): $0.74 median per task / +1.22 net improvement
- Grok 4.5: $0.43 median per task / +1.50 net improvement
By signal, Grok 4.7 (xHigh) ranks:
- #6 Confirmed Success (+10.41%)
- #12 Bash Recovery (+5.98%)
- #16 Praise vs Complaint (+4.93%)
- #28 Steerability (-1.89%)
- No issues with Tool Hallucination (+0.36%)
Congrats to the @SpaceXAI team on this solid release!
Penamaan badan usaha beda-beda di tiap negara :
1. Indonesia = PT
2. Malaysia = Sdn. Bhd.
3. Singapura = Pte. Ltd.
4. Jepang = K.K.
5. Korea Selatan = Co., Ltd.
6. Jerman = GmbH / AG
7. Finlandia = Oy
8. Prancis = SARL / SA
9. Selandia Baru = Limited
10. Turki = A.Ş. / Ltd. Şti.
Kota yang baik itu dimulai dari keputusan yang berbasis data dan berpihak pada warganya. Prinsip ini yang kami pegang di Karsa City Lab bersama Google.
Awal minggu ini, kami telah resmi teken MoU-nya di kantor Google, dan ini jadi awal kerja bareng untuk sembilan kota dan wilayah, yaitu Jakarta, Bandung, Solo, Medan, Malang, Yogyakarta, Semarang, Surabaya, dan Bali.
Tujuannya sederhana dan konkret. Jalan yang lebih mudah dilalui, lingkungan yang lebih terjaga, dan ekonomi kota yang bisa menghidupi warganya. Google bawa wawasan teknis dan data geospasial, Karsa City Lab bawa suara pemerintah kota dan komunitas. Semuanya dipertemukan di satu meja.
Kami yakin kota-kota Indonesia layak sejajar dengan kota-kota global lainnya. Melalui Karsa City Lab, kami akan membuktikannya.
Fusion in Devin Desktop & CLI
Fusion achieves frontier intelligence while being up to 39% cheaper.
Use your favorite frontier model for planning and a cost-effective model for execution.
Fable/Astra + SWE-2 is very cost-effective during SWE-2’s free promo.
The result is a model that is way more efficient than SWE-1.7.
On FrontierCode 1.1 Main, SWE-2 medium scores higher than SWE-1.7 while taking 58% fewer turns and costing 81% less on average.
We observe that SWE-2 learns to explore in a more focused way before making changes.
Perplexity has joined the Rust Foundation.
We believe in supporting the people who build reliable open-source software. By joining the Rust Foundation, our goal is to improve how people and agents build with Rust.
GPT-5.6 Sol is now the most affordable frontier model on Devin Desktop and CLI.
Today, OpenAI reduced Sol’s prices by 20% for the next 3 months. Last week's 70% discount still applies, so through October 3 you now get a 76% discount off the list price.