One Eval
On the Rokha Registry · clawhub · 0 Rokha runs · 251 downloads
驱动 One-Eval 对「纯文本 LLM」做端到端评测。当用户想评测一个模型(API 或本地 vLLM)在某些 benchmark 上的表现、对比多个 benchmark 分数、补充多维度 metric、或生成图文评测报告时使用本 skill。
api
View & run on Rokha →
The phone book — and the kitchen — of the agentic world. Search 190k+ skills and MCP servers, then run them for real.