Personal
~/.agents/skills/Project
.continue/skills/One-line install
npx skills add owner/repo/skill --agent continueSkills for Continue
Free agent skills compatible with Continue by Continue. Search within them or filter by category, level and popularity.
17 skills
RAG Evaluation
nvidia
Evaluate retrieval-augmented generation benchmarks efficiently.
NV-Reason-CXR
nvidia
Run smoke tests for chest X-ray reasoning models.
NeMo Evaluator SDK
davila7
Enterprise-grade LLM benchmarking across 100+ tasks.
LLM Evaluation Harness
davila7
Benchmark LLMs across 60+ academic tasks.
Advanced Evaluation
muratcankoylan
Evaluate LLM outputs with precision and reliability.
DeepEval
confident-ai
Streamline evaluation workflows for AI applications.
Obliteratus
nousresearch
Remove refusal behaviors from LLMs without retraining.
Hugging Face Community Evals
huggingface
Evaluate Hugging Face models locally with ease.
Agent Platform Eval Flywheel
Evaluate and enhance AI models on Google Cloud.
Phoenix Observability
davila7
Open-source observability for LLM applications.
Agent Evaluation
muratcankoylan
Systematically assess agent performance and quality.
Ollama CLI Interface
hkuds
Manage models and generate text from the command line.
LLM Eval Harness
daymade
Evaluate LLM endpoints for reliability and performance.
Behavioral X-Ray
sickn33
Probe AI models for hidden behavioral patterns.
AI Ethics Validator
jeremylongshore
Ensure fairness and compliance in AI models and datasets.
Promptfoo Evaluation
daymade
Efficiently evaluate LLM outputs with Promptfoo.
OpenClaw Model Switch
daymade
Easily manage OpenClaw model configurations.
