Laya is an open-weight model optimized for Apple Silicon that runs locally on macOS with Core ML, achieving 49-50 decisions per second in a Snake game demo with no generated tokens. The multilingual variant processes decisions in ~5ms on M3 Max with 2.78× better energy efficiency than MLX, and requires no PyTorch or external dependencies for inference.