LLM Evaluation skills
Free agent skills tagged llm evaluation, ready to install into any SKILL.md-compatible agent.
4 skills
Advanced Evaluation
muratcankoylan
Evaluate LLM outputs with precision and reliability.
AI & AgentsintermediatePython17.7k repo
LLM Eval Harness
daymade
Evaluate LLM endpoints for reliability and performance.
AI & AgentsintermediatePython · Shell1.3k repo
Promptfoo Evaluation
daymade
Efficiently evaluate LLM outputs with Promptfoo.
AI & AgentsintermediatePython · Node · Shell1.3k repo
S
Senior Prompt Engineer
alirezarezvani
Optimize prompts and evaluate LLM outputs efficiently.
AI & AgentsintermediatePython · Shell24.3k repo
