CalibratedDecisions.

Repo · Benchmarks & research

typesafe-ai-benchmark

This is a LLM Gateway that mimics typesafe ai structured output.

Open github.com ↗

How builders describe it

This is a LLM Gateway that mimics typesafe ai structured output. Like an imposter Jev.
I built a TypeSafe alternative but with Qwen 3.8 27b on Cerebras. Similar quality, Similar speed, but TypeSafe is way way cheaper.

The decision Jev makes

Benchmark questions with known answers, to check accuracy and confidence.

Where it fits

Head-to-head tests, calibration studies and independent research into how well Jev decides, how fast, and at what cost. All 332 benchmarks & research projects →

Related projects