New @timbutterly special from @gasdigital is absolutely must watch 🤣🤣 might even have my wife watch it for the second view. But will it reveal too much truth? https://t.co/QicIjw0QQw
@quimedesu@faizrazz Thanks - how can someone increase Token per second when running shared ram/vram? Is there any thing I can do? Speed isn’t critical, but hours between prompts would make it impractical for some things, still usable tho for fun
@quimedesu@faizrazz Will try both suggestions thanks. Running with opencode but it could be that which is not effective with this model. Will try pi and Hermes 🙏🤞
@faizrazz@quimedesu Thanks I have downloaded both but only tested the iq2_xs and it was underwhelming (but hopeful and promising), I am spoiled by my subs to codex and Claude at the moment and trying to pivot to home run LLMs but my gpu power is my bottleneck. 128gb ram tho, but 1 tps is too slow
@faizrazz@quimedesu I’m pretty much finding it unusable with opencode and 12gb gpu. Could be a skill issue but prompts basically fail to recall the the previous (ie changed our port and next prompt rewrote with old port over and over). Nothing built has been good 😭
@sudoingX I have tried using it and it’s been frustratingly useless compared to Claude or Codex. Forgets things immediately for the next prompt, constant errors in the deliverable. How do you avoid it? I’m using opencode and lmstudio to run it
2 session on Max, five hour blocks using “Ultracode” completely burned through without a result to test yet. Feeling more like an ultrascam. What gives @AnthropicAI@claudeai - I will never use higher than low effort again.