JevEval is a custom LLM evaluation metric that separates evaluation logic, decision-making, and scoring. Instead of asking an LLM to generate a single score, it uses Jev to answer bounded questions with calibrated probabilities, then applies fixed math to produce deterministic, reproducible evaluation scores.