github / Research & data
jev-benchmark
AI-assisted summaries and translations. Check original sources for context and performance claims.
このリポジトリは、TypeSafe AIのJevに対する再現可能なベンチマークを提供し、エージェントのツール呼び出しリスク分類における性能を評価します。明確なタスク、曖昧なタスク、敵対的なタスクの各シナリオにおいて、精度、レイテンシ、および信頼度スコアの妥当性を検証します。
翻訳された要約 · AI-assisted; check the original.
Source notes
This is a linked resource, not an independent verification of performance, cost or results. Check the original for current details.
- Collected
- 2026-09-21
- Discovered via
- awesomejev.com
トップ10
Sponsor
- Wallpets 94 クリック数$125
W
- Walltank 109 クリック数$85
W
- CChowder 61 クリック数$50
- Falconer 52 クリック数$45
F - Lemonpod 92 クリック数$40
L
- Menta 38 クリック数$35
M - Sway 36 クリック数$30
S
- Peon-Ping 32 クリック数$25
P - Zeron 94 クリック数$20
Z
- Guideless 27 クリック数$15
G