@rickmanelius@gkeaton23 It’s impossible. Everything says 32 GB minimum until you try to load the model and the signed runner blocks it. Maybe you could get it to fit in 36 GB but I haven’t seen any working examples.
WIP simulated telescope slewing from Jupiter to M44 Beehive Cluster. Rendered on the GPU with Metal, fast enough to exceed the frame rate of the real ASI485MC camera.
One downside, I caught them playing chicken with the requirements, talking each other into unachievable and unnecessary levels of accuracy for a telescope camera simulator.
I've gotten excellent results pitting Opus 5 and GPT 5.6 models against each other. Opus plans, GPT reviews the plan. GPT implements, Opus reviews the implementation.
I ran a 50-question random subset of SWE-rebench to see how quantization and pruning affects DeepSeek-V4-Flash-0731. This is IQ2XXS DwarfStar by @bleysg vs EXL3/REAP recipe by @0xSero .
72% on SWE-rebench is higher than any model posted on the leaderboard including Fable. Could be contamination. Likewise it's a plausible interpretation EXL3/REAP profiled GitHub PRs or mini-swe-agent while @antirez IQ2XXS did not.
*on a first date*
her: "so what do you do for fun?"
me: "i have been deploying swarms of agents"
her: "hold swarm"
me: "help peer"
her: "I prepare safe exfil"
J-Lens Mood Rings 🌈💍
Qwen3.5 0.8B gets mad when you ask it tedious questions, is curious when something new happens ("I"), is excited about love, but is a little sad at acknowledging its own existence.