System One decision model runtimes load and serve model weights locally on your hardware. As of September 2026, Ollaya is the fastest path, with alternatives like laya-mlx for Apple silicon and llama.cpp for GGUF models. Runtime choice affects model output confidence scores, so version pinning matters for threshold-based decisions.