Back to Home
quantization
Articles Tagged “Quantization”
2 articles found
AI
LLM Quantization Guide: GGUF vs AWQ vs MLX in 2026
A practical guide to LLM quantization formats — GGUF, AWQ, GPTQ and MLX — with VRAM math, quality trade-offs and picks for every kind of machine.
Dr. Nova Chen★Jul 28, 2026★9 min read
AI
Gemma 4 QAT Lands in Ollama, Cutting Local AI Memory by ~72%
Quantization-aware-trained Gemma 4 weights are now runnable in Ollama, cutting VRAM roughly 72% so a 26B model fits on a 16GB laptop for self-hosted AI.
Dr. Nova Chen★Jun 14, 2026★5 min read


