Ollama now supports decision models via the Jev API, enabling fast local inference for tasks like ticket triage and content moderation. Three new models (nimble, tev1, tev1:0.8b) are available through the /v1/systemone endpoint in Ollama 0.35, delivering near-instant decisions with no additional costs or network latency.