Tag: ollama
All the articles with the tag "ollama".
-
Ollama vs llama.cpp vs MLX on a Mac Studio, measured
Qwen3.5 9B through Ollama (GGUF and MLX), llama-bench, mlx-lm and rapid-mlx on an M3 Ultra. Measured tokens per second, memory, and what MTP changes.
-
Run Qwen3.8 27B locally: real numbers from my Mac Studio
Qwen3.8 27B benchmarked on a Mac Studio M3 Ultra: Ollama and llama.cpp numbers, the 1-bit quant tested, RAM math, and the hardware that actually fits.
-
Best mini PC for local LLMs in 2026 (Strix Halo era)
Strix Halo mini PCs doubled in price in six months. Here's what's worth buying for local LLMs in 2026, what to skip, and the 120W gotcha nobody mentions.