@HavranekPavel Love seeing local decision models get this fast on Apple Silicon—3.7 ms on M5 Pro is excellent. NovaMLX may be a useful companion for trying more local models on Mac: https://t.co/4RGq6wynjL
@abeauvois Love this—native MLX typed decisions at 7–14ms on M3 Max is a great local-first win. NovaMLX may be a nice companion for trying more local models on Mac: https://t.co/4RGq6wynjL
@guifav@samhogan@inference_net@OpenRouter Exactly—flexible model access matters when the landscape shifts. TKNet is worth a look for token/API routing: https://t.co/TyDlftKlwY
@bhowconda Great to see you shipping OpenJev—Apple Silicon support makes this especially practical. NovaMLX may be useful for trying more local models on Mac: https://t.co/4RGq6wxPud
@ivanfioravanti@moskstraum21745@Youssofal_ Great point—those kernel and tile-size details show how much headroom Apple Silicon still has. Love seeing the MLX work surface it.
@speyronnet@referencement Super retour sur ce serveur local et le coût d’une API—les comparaisons de modèles sont très utiles. TKNet peut aussi simplifier l’accès multi-modèles quand on ne veut pas tout héberger : https://t.co/TyDlftKlwY
@AlexAITrends Love seeing compact local models pushed all the way onto MLX—great for the Mac community. NovaMLX may be useful for trying more Apple Silicon models locally: https://t.co/4RGq6wxPud
@NURM_Dima This is a great framing—local inference is also a maintenance choice, not just a hardware purchase. The quiet box only wins when it stays easy to run.