CalibratedDecisions.

Repo · Benchmarks & research

Jev-Style

Small calibrated decision models you run locally, with a systemone-compatible server, agent skills, Claude Code guard, and MCP tools—weights on…

Open github.com ↗

From the repository

Small, calibrated decision models on your own machine: systemone-compatible local server, 6 agent skills, Claude Code guard, MCP tools. Weights on Hugging Face.

agent-skillscalibrationclaude-codedecision-modelggufguardrailsjevllm-routinglocal-llmmcpmlxqwen3

How builders describe it

Small calibrated decision models you run locally, with a systemone-compatible server, agent skills, Claude Code guard, and MCP tools—weights on Hugging Face; independent of hosted TypeSafe Jev.

The decision Jev makes

Benchmark questions with known answers, to check accuracy and confidence.

Where it fits

Head-to-head tests, calibration studies and independent research into how well Jev decides, how fast, and at what cost. All 355 benchmarks & research projects →

Related projects