
Llm Evaluation
@clawhub_codenova58/llm-evaluation
By codenova58
About this Skill
Deep LLM evaluation workflow—quality dimensions, golden sets, human vs automatic metrics, regression suites, offline/online signals, and safe rollout gates f...
工作流编排上下文管理工具调用
Skill files and instructions
Read SKILL.md and the other instructions or configuration files published with this Skill.
Loading file list
Details
- Category
- AI Agent
- Source
- clawhub
- Version
- 1.0.0
- Updated
- Sep 24, 2026