Benchmarks & research

jev-benchmarks

@AbdelStark11PythonApache-2.0updated 2026-09-17

Probability-aware evaluation for typed decision models: calibration, selective risk, latency, and reproducible benchmarks.

AbdelStark/jev-benchmarks

Where it calls Jev

from typesafe_sdk import Choice, TypeSafeClient

src/jev_benchmarks/adapters/jev.py:5

The link points at the commit we read, so the line number still holds.

This one cannot run here

Its question set is assembled at runtime, or never written out literally in the code, so there is nothing to lift. The source link above will show you.

Other projects in this category