Community project / Evaluations and independent research

When a Judgment Layer’s Self-Reported Fields Lie

Independent measurement of Jev’s typed decisions as one of three judgment layers run on identical items (Laya locally, Jev and a frontier model remotely), with per-call ledger-based cost accounting and Murphy/Brier/ECE calibration against binary ground truth; the Jev-specific findings are a verdict vocabulary reaching three values where six are documented, and a sufficient field that does not separate thin evidence from contradictory evidence, and the claim about contradiction rests on seven live readings that have no JSON artifact behind them, which the write-up states in its own errata rather than leaving to be discovered.

Awesome Jev share card for When a Judgment Layer’s Self-Reported Fields Lie Download share card

Share this listing

Maintaining this project? Share its link, badge, or image. Listed means included, not endorsed.

This is an independent community listing. Check the source, license, data handling, and evaluation caveats before relying on a project. Inclusion is not an endorsement or security review.

Found an outdated or inaccurate detail? Report a correction →

Explore more in Evaluations and independent research

README-DERIVED DIRECTORY

Find a Jev project

Loading projects…

Every listing comes from the README. Inclusion is not an endorsement.