Community project / Evaluations and independent research
jevos
MIT-licensed 1B yes/no model released as a q4_k_m GGUF that runs on CPU with llama.cpp behind a Jev-compatible /v1/systemone endpoint; it answers only Noul questions, and Choice and Score requests return 422. The authors report 0.815 accuracy against 0.927 for hosted Jev on 2,000 questions from unseen policies, but the training code and evaluation set are not published, so that result cannot be reproduced from the repository.
Download share card
Share this listing
Maintaining this project? Share its link, badge, or image. Listed means included, not endorsed.
This is an independent community listing. Check the source, license, data handling, and evaluation caveats before relying on a project. Inclusion is not an endorsement or security review.
Found an outdated or inaccurate detail? Report a correction →