Tag: qwen
All the articles with the tag "qwen".
-
Ollama vs llama.cpp vs MLX on a Mac Studio, measured
Qwen3.5 9B through Ollama (GGUF and MLX), llama-bench, mlx-lm and rapid-mlx on an M3 Ultra. Measured tokens per second, memory, and what MTP changes.
-
Run Qwen3.8 27B locally: real numbers from my Mac Studio
Qwen3.8 27B benchmarked on a Mac Studio M3 Ultra: Ollama and llama.cpp numbers, the 1-bit quant tested, RAM math, and the hardware that actually fits.