@angelus513@TeksEdge --spec-type ngram-mod
--spec-ngram-mod-n-match 32
--spec-ngram-mod-n-min 48
--spec-ngram-mod-n-max 64
these are very helpful for decode speed in repetitive output
@appariciojunior É patético. Eu cancelei meu 5x. Agora vou de IA local e algumas chamas API para funções específicas. Deepseek4flash e laguna s 2.1 mudaram o jogo
@gnukeith it's amazing because i've upgraded to 128gb of ram last night. i woke up testing laguna s 2.1 and now i can't sleep because DS4flash is too damn good <3
@MichaelHutu i've running a q3 on a 5080 + 128gb of ddr5. i'm getting ~18t/s. i'm 100% in love. cant stop working with it, probably wont sleep tonight lol. just canceled my anthropic max x5 subscription.
@eric_alcaide To me it’s clear you didn’t! The XS outperformed all similar models in my agentic coding workflow, despite scoring less in benchmarks. I’m buying ram to run the S version as my daily driver.