@thsottiaux@TJeparskis Work harder then, this is 2 sol max sessions, each hit 500k context and used 3-5 luna max sub agents.
This is not the kind of usage i expected considering luna and sol prices.. anthropic 20x was giving me much more usage, 2 500k on opus 5 was barely 1% of the weekly..
@TJeparskis Why lie ? This is 2 sol max sessions, each hit 500k context and used 3-5 luna max sub agents, so explain what "half an hour" means for you please.
Well, reasoning loops are a known problem, i just turned off the reasoning effort level and i use my own harness to make them think as necessary, it achieves the "todo" style reasoning of Ox Alpha the exact same way, but with real output tokens and actions not "i read X file" without ever reading it
huihui-ai/Huihui-Ornith-1.5-9B-abliterated — huihui-ai's standard abliteration (weights orthogonalization, minimal KL).
zaakirio/Ornith-1.5-9B-Uncensored — decensored with the Heretic method (refusal direction removal), plus a GGUF quant (zaakirio/Ornith-1.5-9B-Uncensored-GGUF).
junafinity/Ornith-1.5-9B-uncensored — another uncensored finetune, also with GGUF-8bit and MLX quants.
You know the model only sees whats on his context and generally effort level isnt unless you put it in there + wtf is fowlcode ? that low, 10 out of 100 values are hallucinated or you made smt to tell the model its effort level and that is not working properly, or you are just attention farming.
@Da7_Tech Are you using a local proxy to save all data like tokens you use etc so after sometime you can publish how much the 200 usd was worth it in api prices
@1camey_@matt_kaschel@dan0madpro@elder_plinius full self improvement loop ? bro it updates its skills and create useless memory files, .md spaguetti, its lazy because you guys dont care if its good or bad, you see words you dont understand and you think its magic
30 tokens per second is too slow, 10 hours to use 1m tokens, how do you expect us to use 100T per day ? if you have the capacity make it FASTER. it is so slow it triggers idle stop warnings and i needed to modify my local proxy just to run it.
Ox Alpha (stealth model) is free for the next week
- 1M Context
- Multi-modal
- Zero Data Retention
Generous rate limits, near unlimited usage
We have capacity for 100T tokens per day, lets see what you can do
Yeah like the 20x plan of anthropic only gives the 20 on the monthly limit, but 5h limits are like 2-3x more than the 5x plan, and weekly i am not sure, but yeah basically they play us with those empty terms, we need all limits, weekly etc to be fixed credits and we need transparency about what each model cost on those credits per 1m input/output etc if it differs from API
@matt_kaschel@dan0madpro@elder_plinius That just means that people who uses it are lazy and dont bother doing their own tooling/workflow, so that doesnt mean anything really
@opencode@OpenRouter No problem guys, i created 10 more accounts + keys in openrouter/opencode, now i have real unlimited usage on my local proxy with key rotation, i even added command code on top just to make sure
@opencode UNLIMITED HUH ? @OpenRouter
i WILL OPEN YOUR A-Z
Rate limited (429) — stream error (upstream_rate_limit): upstream status 429: {"error":
{"message":"Rate limit exceeded: free-models-per-day-stealth. ","code":429,"metadata":{"headers":
{"X-RateLimit-Lim. Try again later.