@Sentdex Have you run qwen 3.8 flash yet? I was able to get it at around 20 tps at q4 128k context on a 5060 ti and 128gb system ram. KInda crazy. I'm wondering how you think it is compared to 5.3 flash.
11/
If anyone is working on model honesty, evals, interpretability, or local-to-cloud routing, I'd be happy to talk.
This is exactly the kind of work I want to do.
1/
Anthropic dropped their Global Workspace / Jacobian Lens paper yesterday, so I tried it on open models.
Started as curiosity: what do models look like inside under normal prompts, emotion prompts, ragebait, and base vs abliterated weights?
Then it became a router idea.
10/
I'm not claiming hidden states, probes, logit lens, or hallucination detection are new.
The narrow test is whether Jacobian-lens workspace trajectories are useful as a one-pass risk signal for confident wrong answers.
Pointers to prior work very welcome.
@stevibe Thats pretty cool, I just built something similar for a hackathon, I fine tune a few models for the and gemma 4 12b had the best results https://t.co/PHdcsauwk1 https://t.co/QiyBqr6jpw
We did pixel art with the WRONG model on purpose
PixelLock retextures sprites with a fine-tuned TEXT LLM, not an image generator. A GBNF grammar locks the exact shape β pixel-perfect by construction. Language modeling, not diffusion.
#BuildSmall