Qwen3.8-27B ships with reasoning_effort on xhigh. Turns out that's the one setting you don't want.
Ran all four levels through the same config, 1083 tasks each. low and medium tie. xhigh burns 3x the tokens of low, loops 10x more often, and scores no better for it.
I'd just run low effort. Config below.
UD-IQ2_XXS since that's what fits 12GB, but I feel like higher quants will land in the same place.
日本企業、AIに古典的職制ヒエラルキーを持ち込みがちなのだけど
大人しくanthropicやopenaiが実践するアプローチを試した方が良い
How we built our multi-agent research system https://t.co/6c5mjOQAis
A Practical Guide to Building Agents https://t.co/sKmYs1aZbG