Jev 101
English ▾
Tools/Evaluation and calibration
Community

typesafe-ai-benchmark

iammrduncan/typesafe-ai-benchmark · Evaluation and calibration

LLM gateway that mimics the System One output shape for comparison work.

EvaluationGateway
Read Chinese translation

模拟 System One 输出结构的大语言模型网关,用于对比研究。

Results belong to each project's dataset, prompts, model version, and measurement setup. Inclusion means the evidence is inspectable, not that benchmarks were independently rerun.

Directory and research sources

Provenance is retained from the previous edition. This edition did not install or re-audit the project.

cobanov/awesome-jev →fatwang2/awesome-jev →