🌱 BMC
← October 4th, 11:59pmOctober 9th, 10:46pm →

First run of Mistral3 8B Q4 model on Steam Machine — ~700 tok/s prompt processing and ~43 tok/s token generation using llama.cpp.

Going to put llama-swap on there to try out a couple of different models…which is written by my old friend Benson @mostlygeek.bsky.social!

October 8, 2026

#bmcjournal#mistral#steammachine#localai#llama

Notes 🌱 •Blog •Journal •Links •Search •Feeds •Login