I don't think Anthropic realizes how disruptive these changes are to users. I appreciate the extension, but please stop playing games. Either keep it under the subscriptions or put it under the API already.
A "global workspace" is a real cognitive-science concept, not marketing. If it holds up, the interesting part isn't that Claude thinks silently, it's that we can now point to where. Interpretability is quietly becoming the an important AI subfield.
Perf lives in the software stack, not the silicon. Baseten just raised $1.5B building a company around exactly this... squeezing more out of the same NVIDIA chips everyone else has. Benchmarks measure ceilings; the stack decides what you actually hit.
AI researcher breaks down why your $5,000 Mac Studio runs local AI slower than benchmarks promise:
"Performance isn't about the hardware - it's about three software layers stacked on top of it."
in 15 minutes he tests MLX, Ollama, llama.cpp, and the new vllm-mlx head to head
no GPU needed, just the Mac you probably already own
- why Ollama, the most popular tool, is actually one of the slower options on Mac
- the real bottleneck nobody benchmarks: 'prefill' time before the first token even appears
- the exact quantization setting (4-bit) that's the best speed-to-quality tradeoff
- why unified memory makes Macs uniquely good for big models compared to Nvidia GPUs
complete guide to choosing the right hardware for local AI is below 👇
The comparison cuts both ways. Dot-com had revenue projections that never showed up; semis today have real earnings and real demand. The risk isn’t fake growth, it’s that even real growth gets priced past perfection. ‘Extremes tend to rhyme’ is the right framing.
Semiconductor stocks are now up 246% over the last 14 months, surpassing the 234% surge during the peak of the dot-com bubble.
The AI mania has taken semis to a place only seen once before.
History doesn’t repeat, but extremes tend to rhyme.
Video: https://t.co/WStwCcfM0J