SKILL·1536D0

evaluation-metrics

mattnigh
更新于 3 months ago
11 次查看
22
1
22
在 GitHub 上查看
其他aitestingautomationdata

关于

This Claude Skill automatically activates during LLM performance evaluation to ensure proper metrics and testing. It handles evaluation datasets, computes metrics, facilitates A/B testing, and implements LLM-as-judge patterns. Use it when you need structured experiment tracking and rigorous performance assessment for your LLM applications.

快速安装

Claude Code

推荐
主要方式
npx skills add mattnigh/skills_collection -a claude-code
插件命令备选方式
/plugin add https://github.com/mattnigh/skills_collection
Git 克隆备选方式
git clone https://github.com/mattnigh/skills_collection.git ~/.claude/skills/evaluation-metrics

在 Claude Code 中复制并粘贴此命令以安装该技能

GitHub 仓库

mattnigh/skills_collection
路径: collection/ricardoroche__ricardos-claude-code__claude__skills__evaluation-metrics__SKILL.md
0
FAQ

常见问题

什么是 evaluation-metrics Skill?

evaluation-metrics 是一个 Claude Skill,作者为 mattnigh。Skill 将 Claude 按需加载的说明和资源打包,让 Claude 无需额外提示即可执行与 evaluation-metrics 相关的任务。

如何安装 evaluation-metrics?

使用本页的安装命令:将 evaluation-metrics 作为插件添加到 Claude Code,或将其仓库克隆到 skills 目录,然后重启 Claude 以加载该 Skill。

evaluation-metrics 属于哪个分类?

evaluation-metrics 属于其他分类。

evaluation-metrics 可以免费使用吗?

可以。evaluation-metrics 已收录在 AIMCP,可免费安装。

相关推荐技能