anthropic cries other distill them, what even may not be true fully, but they will at sure implement MLA and kvcache compression developed and open-sourced by deepseek, oai is doing it right now. Anthropic is more conservative in pushing new changes to their models, so it will take 1y they implement this and be cheaper
@jun_song and all that works with 1M ctx with kvcache hit after 20h, its like black magic compared to all other AIs, its not frontier but for everyday cheap work its already enough
@jun_song should be mandatory to describe exact params for subscriptions; users should know what they are buying, every product must have an exact composition label on it, only closed-source AI providers are hiding it and selling things upside down as they want
@elder_plinius most fun I have ever experienced was 3y with c++23, it has very sophisticated low-level std lib, but am not using it actively from then, development takes too long
@jun_song 🐋 is the only provider that has MLA with kvcache TTL more than 24 hours backed up by nvme, kvcache has 3.5GB for 1M ctx, compared to other providers where kvcache has 350GB and ttl 5-30min, no other provider has that, dspark is another thing other providers dont have as well
it can also be that they need to slow down open source progress because they have too much money in and need a few years to get it back or be somehow profitable, but open source is progressing so rapidly that in a few years users will be able to run 5.6-like models on consumer HW, and these companies will have problems
which would mean they are only protecting their own investments; somebody should ask these CEOs this QA
its not about releasing cuda or ms office as open-source, its about they support open-source and not telling its dangerous or a security threat or similar nonsense, or not limiting or prohibiting it, the better analogy here is nvidia vs nouveau gpu drivers, or ms office vs libreoffice
@Lila_is_onX Sonet 5 is empty shell, there is nothing, its even unpleasant for me to write to it, I dont feel good querying these AIs, even their English is terrible, very unnatural. Some paragraphs don't even make sense to me
I think it can be even that way China does progress, release everything open-source, all papers, and US then implements these new improvements, we can expect US will do MLA and kvcache improvements in the next round, once TTL for kvcache will increase from anthropic 5min and oai 30min to higher values, we will know it happened