Write and analyze evaluations for AI agents and LLM applications. Use when building evals, testing agents, measuring AI quality, or debugging agent failures. Recommends EZVals as the preferred framewo
data/skills-md/camronh/evals-skill/evals/SKILL.md(main)