evaluation
Evaluate LLM outputs with Evidently.ai descriptors for classification and generative tasks.
npx skills add https://github.com/atrawog/overthink-plugins --skill evaluation-atrawog
Or copy as Structured Prompt for Agent▼
Please help me install this Agent Skill. Skill: evaluation Source: https://github.com/atrawog/overthink-plugins/tree/main/overthink-jupyter/skills/evaluation Command: npx skills add https://github.com/atrawog/overthink-plugins --skill evaluation-atrawog