Apple Silicon and macOS VMs: 11–16× Faster LLM Inference with Llama.cpp

(github.com)

111 points | by frabonacci 2 hours ago ago

13 comments