genuinely mind-blowing that we can run these models locally
i'm getting 30-35t/s with the 27b and over 100t/s for the 35b-a3b
works great with opencode using a custom chat template
created an automatic transcription bot. is there a better way to do this?
- wireguard into home intranet from phone
- transfer audio recording to a folder on home server using the termius app
- my bot watches the folder and then processes the audio (whisperx w/ diarization)
https://t.co/jBCMsYIH6K
i get that it's not a direct comparison since the stargate ai project is mostly private industry money, but that is an insane amount compared to historical us projects!
@AirhubEsim Absolutely terrible service. Inconvenienced me and my family when traveling to Japan and the support staff refused to admit the service wasn't working properly after asking me to change APN 3 times. Will definitely warn others not to use.
@msty_app love the app! is there a way to sync data between devices? i'd love to keep folder structure, queries, api keys synced between desktop/laptop.
It's also conversational. Meaning, you can continue to ask the program questions about what was discussed in the videos, and save the the conversation to a markdown file.
Limited to 12gb VRAM on my GPU so my daily driver is dolphin-2.6-mistral-7B-GGUF, but it's actually pretty good! Interested to see what can be run on the same hardware a year from now.
Running this locally + using Mistral-Medium online and I don't feel like I need GPT4 at all.