@EmperoAI I tested it, and at present it deserves further testing.
Pay special attention, please do not turn on the `preserve_thinking` function(default value is false).
When this function is turned on, the quality of the model will be seriously reduced.
@sojufx@NaraharikripaM If this model is used in the work related to the browser_use(MCP). There is basically no improvement, or there is a little retrogression. But, it may perform better for the related tasks of the terminal.
@sojufx@NaraharikripaM bro, you can try Grm3.2-Sky. I think he works well on my Hermes Agent. Although the upper limit is not as good as Ornith1.5, it has the stability of Qwen3.6.
Sometimes I feel that I am using Qwen3.7 35B-A3B. (Because there is still some gap with Qwen3.8)
@johnny_everson@bnjmn_marie You can limit the number of thinking token to 8192. This can cut off the thinking cycle to some extent.
8192 can meet most job requirements. It is enough to simply execute 2560/4096.(But I measured 35B-A3B)