# jev-benchmarks Reproducible evaluation for calibration, selective risk, and latency. Source: https://github.com/AbdelStark/jev-benchmarks Category: Evaluation and calibration Snapshot: 2026-09-19. Inclusion is not an endorsement. - cobanov/awesome-jev: https://github.com/cobanov/awesome-jev/blob/ef507d7935aeac9240d4e68c8d0392ad11aa4122/README.md#L202