Tag: mac-studio
All the articles with the tag "mac-studio".
-
Ollama vs llama.cpp vs MLX on a Mac Studio, measured
Qwen3.5 9B through Ollama (GGUF and MLX), llama-bench, mlx-lm and rapid-mlx on an M3 Ultra. Measured tokens per second, memory, and what MTP changes.
-
Mac mini alternatives for local LLMs: M6, M5 and Strix Halo
Apple's M6 Mac mini and M5 Ultra Mac Studio versus 128GB Strix Halo mini PCs for local LLMs, from a 256GB M3 Ultra owner. Bandwidth, price, measured t/s.
-
Run Qwen3.8 27B locally: real numbers from my Mac Studio
Qwen3.8 27B benchmarked on a Mac Studio M3 Ultra: Ollama and llama.cpp numbers, the 1-bit quant tested, RAM math, and the hardware that actually fits.
-
How to run DeepSeek V4 Flash locally (Mac, Linux, Windows)
How to run DeepSeek V4 Flash locally on Mac, Linux, or Windows: hardware that fits 167GB, llama.cpp setup, and realistic speed expectations.