Jev 101
English ▾
Tools/Evaluation and calibration
Community

jev-eval-agent

vinilana/jev-eval-agent · Evaluation and calibration

Compares LLM tool selection with Jev routing in a personal-assistant harness containing 100 mocked tools.

Tool callsMock evaluation
Read Chinese translation

在包含 100 个模拟工具的个人助理测试框架中,对比大语言模型工具选择与 Jev 路由。

Results belong to each project's dataset, prompts, model version, and measurement setup. Inclusion means the evidence is inspectable, not that benchmarks were independently rerun.

Directory and research sources

Provenance is retained from the previous edition. This edition did not install or re-audit the project.

cobanov/awesome-jev →Pinned source evidence: README.md →Pinned source evidence: jev-router.ts →