@cstanley the longer the character length the more it happen but yeah it's not totally stable and it take a long time to process this. Completely destroying your 1/RTF for the batch ( if that was implemented)
A few weeks ago , while doing advanced networking configuration, Chatgpt 5.x explained to me why the 2 first port of my routeur had an hardware problem. When I pushed back multiple time on it, it confidently put in the trace that I gave unreliable information and he would trust the cmd prompt that he could analyze and that I should buy another routeur. That's enough context root for you.
@VOX_of_SU you need to go into personalization and put "Efficient" and "less" on everything other than header and list.
+ Add a custom sentence asking to tone down emotion, be analytics etc.
Then it's usable
there is not hardcoded max on qwen3 TTS but longer text tend to increase the low chance of the model allucinating and giving you total bullshit. Something around 40-90s is often efficient enought and not too memory intensive when batching properly.
Not sure how this lib did the batching
@nrehiew_ video are truly insane, probably most of what you have seen is real. I tested only the fast model before loosing access and the results were already quite impressive.
@melvynx I think the rating of chinese AI are still too high on this test. In prod it's clearly way way worse than that when you compare 5.3 Codex or opus 4.6 to any chinese model
@jelanifuel opencode is probably a better comparison as it's way more customizable but yeah ... Basically it's all about the plugin which are just skill + api/ws loop + skills( md files). If you decide to not trust them with Google suite and account automation I see zero points.
@ItakGol I did the test and couldnt get opus 4.6 to pass it but sonnet 4.5 did. cgatgpt 5.2 Thinking in the app didnt pass in standard more but worked in extended mode. chatgpt 5.2 default in agentic passed it though . Rest of my result are the same as yours