Apple Silicon and macOS VMs: 11–16× Faster LLM Inference with Llama.cpp

(github.com)

99 points | by frabonacci  2 hours ago

12 comments