favorite-llms
Favorite LLMs
My favorite LLM's which are below 15GB:
(the QAT quantizations seem to work the best/fastest on my machine)
- Gemma 4 E2B / E4B
- Gemma 4 26B A4B
- GLM 4.6V Flash (about half the speed of Gemma 4 26B A4B)
- Bonsai 8B
- Ornith 1.5 35B A3B (about half the speed of Gemma 4 26B A4B)
- Qwen 3.8 27B
Some of my favorite LLM's above 15 GB:
- Seed OSS 36B Instruct (Q3_K_M, GGUF, 16.41GB)
Good for creating a roleplaying game with multiple players. Not the fastest though, needs to think for some time, but one of the few "small" LLM/AI's that can actually keep track of multiple persons/players. - Qwen 3.8 27B