Continuing to push @deepseek_ai's V4 hard on the M3U Studio.
There is more speed on the table, and I will continue to update the repo.
https://t.co/GR2MPbEhqS
We've improved our serving efficiency and increased rate limits by +10x. The update is live on @OpenRouter@vercel AI Gateway and https://t.co/Pin6Re29Gr.
Thank you to everyone who has used the model, shared feedback and stuck with us while we improved the experience!
Usage is already climbing. On OpenRouter alone, Laguna S 2.1 is on pace for ~250B tokens today, 4x our daily average this week.
We are also taking 10% off our paid endpoint on OpenRouter. It's a dedicated deployment with the full 1M context window for the best performance on harder tasks.
Run Laguna S 2.1 in pool or Poolside Desktop Assistant, or plug it into @opencode, @NousResearch Hermes Agent, @kilocode, @cline or @pidotdev and let it run over the weekend.
We'll be watching the graphs.
I've set myself a challenge to see if I can go a week without using frontier models as my go to.
Caveat: I am only using them to refine the plugins I have for @pidotdev, but overall, I am pleasantly surprised.
Here's a cheeky dashboard I put together for my Spark.
@lais_bsc@pidotdev With local LLM’s they tend hallucinate or doom poop or do wild things.
I’ve created a bunch of plugins to stop the doom loop, a neural memory, chain of thought scratchpad, a proper goal system.
My best one is Cortheon which is a validation engine. I plan to OOS it soon.