Benchmarks & research

jev-eval

@4esv1Pythonupdated 2026-09-19

Benchmark TypeSafe Jev against any OpenRouter model on your own labelled classification data: accuracy, calibration, latency, cost

4esv/jev-eval

Where it calls Jev

JEV_URL = "https://api.typesafe.ai/v1/systemone"

evaljev/runners.py:15

The link points at the commit we read, so the line number still holds.

This one cannot run here

Its question set is assembled at runtime, or never written out literally in the code, so there is nothing to lift. The source link above will show you.

Other projects in this category