More context that the 10k model Fast, affordable (lol) q8 quantization 3x slower than the 10k model, though :( https://turbowarp.org/1370285710
Part of the PocketLM Family: https://github.com/Greninja9257/PocketLM