Articles by author
SLM vs LLM Compared on Quality, Speed and Memory
SLM vs LLM compared across nine local model configurations: quality on 74 tasks, generation speed, memory use, and energy for a classification task.
9/14/26
13 min read
EXL3 Quantization Compared with GGUF on Quality, Speed and VRAM
EXL3 vs GGUF tested on an RTX 5090: compare quantization quality, prompt speed, token generation, and VRAM, with a working ExLlamaV3 setup.
9/14/26
13 min read
No items found.
