Skip to main content
Agent Skills are modular knowledge packages that teach your AI coding agent how to evaluate effectively. They follow the open Agent Skills standard and work with Claude Code, Cursor, Windsurf, Codex, and 40+ other agents.

Install

This installs all Coval skills into your agent’s skills directory. Skills are loaded on demand — only the name and description are in memory until activated. To install only the recommended starter collection, pass the skill names to the installer’s --skill selector:
Choose the target agent and the project or global installation scope in the installer. Skill names are instructions to your coding agent, not terminal commands.

Skills vs Connector vs CLI

We recommend Skills + CLI for the most complete experience. Skills teach your agent what to create, and the CLI executes it with structured JSON output.

Available Skills

Start with a product question: what must your agent get right, and what evidence would show it? The starter collection guides a small first evaluation, failure discovery, metric calibration, and regression checks. Install it, then ask your coding agent:
Use coval-eval-start to help me evaluate my agent with Coval. Reuse what I already have, propose a small execution budget, and distinguish working setup from evidence that my agent performs well.
A first voice run starts small: one case, one persona, one iteration, concurrency one. The skills present the concrete plan before spending and count reruns and base + mutation variants against an agreed session budget. These are workflow controls, not server-side spending limits. Human calibration requires real human labels; without them, the output is a review-ready plan, not fabricated validation.

Full catalog

Every skill in the repository, grouped by category. You can install any skill individually; router handoffs (for example from coval-eval-start) require the destination skill to be installed too.

Evaluation

Onboarding

Runs

Simulations

Reports

Agents

Personas

Test Cases

Metrics

Dashboards

Traces

See Tracing Skills for copy-paste prompts and validation guidance.

Human Review

Migrations

Resources

Sofia

How Skills Work

Skills use progressive disclosure to stay lightweight:
  1. At startup (~100 tokens per skill): Only the name and description are loaded
  2. When activated (under 5000 tokens): The full skill instructions load when your agent detects a relevant task
  3. On demand: Reference files (templates, examples) load only when needed
This means having all Coval skills installed adds minimal overhead to your agent’s context.

Skill Structure

Each skill follows the Agent Skills spec:

Source Code

All skills are open source: github.com/coval-ai/coval-external-skills