Repo · Benchmarks & research
Security Bench
Evaluate injection and code-security detection.
Open github.com ↗From the repository
Blind security benchmarks for Jev, TypeSafe's System One model: prompt injection and vulnerable code detection, built on jev-go
How builders describe it
Blind security benchmarks for Jev, TypeSafe's System One model: prompt injection and vulnerable code detection, built on jev-go.
Blind security benchmarks for TypeSafe Jev (prompt-injection and vulnerable-code detection) built on jev-go, with a terminal dashboard over saved results.
the code is opensource if you wish to have a look inside!
The decision Jev makes
Benchmark questions with known answers, to check accuracy and confidence.
Where it fits
Head-to-head tests, calibration studies and independent research into how well Jev decides, how fast, and at what cost. All 332 benchmarks & research projects →