Skip to main content
The Quantum Dispatch
Back to Home
self-hosted-llm

Articles Tagged “Self Hosted LLM

4 articles found

Mini Computers

Best Mini PC for Local LLMs in 2026: A Buyer's Guide

A practical 2026 buyer's guide to the best mini PCs for running local LLMs, comparing unified memory, NPUs, and price so you can self-host with confidence.

Alex Circuit
Alex CircuitJul 11, 20269 min read
AI

Ollama v0.31.1 Boosts Local AI Performance on Apple Silicon

Ollama v0.31.1 makes Gemma 4 about 90% faster on Apple Silicon via multi-token prediction, advancing local AI performance and privacy.

Dr. Nova Chen
Dr. Nova ChenJul 4, 20265 min read
AI

Ollama 0.30.8 Widens Local AI Hardware Support and Speeds Up Apple Silicon

Ollama 0.30.8, released June 12, broadens GGUF hardware support through llama.cpp and upgrades its Apple Silicon MLX engine for faster, private local AI.

Dr. Nova Chen
Dr. Nova ChenJun 20, 20263 min read
AI

Gemma 4 QAT Lands in Ollama, Cutting Local AI Memory by ~72%

Quantization-aware-trained Gemma 4 weights are now runnable in Ollama, cutting VRAM roughly 72% so a 26B model fits on a 16GB laptop for self-hosted AI.

Dr. Nova Chen
Dr. Nova ChenJun 14, 20265 min read