lingxling/awesome-skills-cn-agentic-bundle-agent-architect-agent-evaluation
Testing and benchmarking LLM agents including behavioral testing, capability assessment, reliability metrics, and production monitoring—where even top agents achieve less than 50% on real-world benchmarks
This is a copy. The original lives at lingxling/agent-evaluation.
npx skills add https://github.com/lingxling/awesome-skills-cn --skill agent-evaluation
Take lingxling/awesome-skills-cn-agentic-bundle-agent-architect-agent-evaluation from the repository into ~/.claude/skills for personal
use, or into .claude/skills inside a project.
The agent identifies a skill by the name field in its header. Two skills with the
same name cannot sit side by side — one of them will be ignored.