this guy tried running LLMs locally to save on API costs and waited 13 minutes for a single response
lets be honest we've all thought about it
"why am i paying for Claude when i can just run an open source model locally for free?"
so he tried it and ran Gemma 4 to avoid API costs
13 minutes to get this response: "I am a large language model, trained by Google."
tools like Claude Code and OpenClaw have system prompts over 20,000 tokens. so even your first message isn't starting from a clean slate.
your local model is choking on context before you even ask it anything
the API bill hurts but time is money
Liftoff.
The Artemis II mission launched from @NASAKennedy at 6:35pm ET (2235 UTC), propelling four astronauts on a journey around the Moon.
Artemis II will pave the way for future Moon landings, as well as the next giant leap — astronauts on Mars.