quantization · 37 topics
2026 · Sep
2026 · Aug
why your local llm feels dumber than it is
08-16🔥🔥🔥unsloth qwen3.8 27b gguf
08-02🔥🔥unsloth deepseek v4 flash 0731 gguf
2026 · Jul
prism ml ternary bonsai 27b gguf
07-28🔥🔥turboquant kv cache quantization
07-27🔥🔥🔥prism ml bonsai 27b gguf
07-27🔥🔥poolside laguna s 21 nvfp4
07-27🔥🔥unsloth laguna s 2 1 gguf
07-12🔥🔥deploying quantized models on amazon sagemaker ai with unsloth
2026 · Jun
2026 · May
2026 · Sep
lg exaone 3 5 awq
09-27🔥🔥alibaba shrinks qwen3 32b to fit on a 24gb consumer gpu
09-26🔥🔥moondream shrinks parakeet redux speech recognition to 178 mb
09-26🔥🔥ornith 1 5 beats 35b models on reasoning
09-23🔥🔥together ai fixes a hidden bias crashing llm reinforcement learning training
09-21🔥🔥comfy org qwen image 2 1
09-18🔥🔥prismml squeezes qwen3 8 27b
09-06🔥🔥qwen3 8 27b gsq rco gguf
2026 · Aug
unsloth glm 5 3 flash gguf
08-26🔥🔥quantization aware healing
08-26🔥🔥liquid ai pipette on device benchmarking
08-25🔥🔥liquid ai lfm2 5 on device phone benchmarks
08-21🔥🔥empero qwen38 27b ridge gguf
08-20🔥🔥liquidai lfm25 qad
08-19🔥🔥🔥qwen3 8 2 4t a95b fp8