Skip to content
MrJev

jev-as-a-judge

Uses Jev through langchain-typesafe as the judge in an eval suite, asking typed quality questions instead of asking a larger model to grade. No licence file.

View on GitHub →

We haven't reviewed this project hands-on yet. The description above is our own summary of its README. Numbers it reports are the author's, not ours.

More in Evaluation & Benchmarks

jev-align

★ 297▲ 49

sutro-sh/jev-align

CLI from Sutro that finds the examples a Jev function is least sure about, asks you to label them, and uses GEPA to improve the question.

PythonReviewed

JevBench

★ 152▲ 120

fstandhartinger/jevbench

Benchmark for typed decision models across several suites, with confidence cascades and committees reported separately.

PythonReviewed

jevals

★ 97▲ 39

openlayer-ai/jevals

Agent evals and guardrails as typed questions instead of an LLM judge, packing every eval for a trace into one request. From Openlayer, with a mock backend so the whole library runs without a key.

PythonReviewed

Get new Jev projects every week

New Jev releases, pricing changes, and the best new projects, once a week. No spam; unsubscribe anytime.

Powered by Buttondown. See our privacy policy.