Install
--skill selector:
Skills vs Connector vs CLI
We recommend Skills + CLI for the most complete experience. Skills teach your agent what to create, and the CLI executes it with structured JSON output.
Available Skills
Recommended starter collection
Start with a product question: what must your agent get right, and what evidence would show it? The starter collection guides a small first evaluation, failure discovery, metric calibration, and regression checks. Install it, then ask your coding agent:
Use coval-eval-start to help me evaluate my agent with Coval. Reuse what I already have, propose a small execution budget, and distinguish working setup from evidence that my agent performs well.
A first voice run starts small: one case, one persona, one iteration, concurrency one. The skills present the concrete plan before spending and count reruns and base + mutation variants against an agreed session budget. These are workflow controls, not server-side spending limits. Human calibration requires real human labels; without them, the output is a review-ready plan, not fabricated validation.
Full catalog
Every skill in the repository, grouped by category. You can install any skill individually; router handoffs (for example fromcoval-eval-start) require the destination skill to be installed too.
Evaluation
Onboarding
Runs
Simulations
Reports
Agents
Personas
Test Cases
Metrics
Dashboards
Traces
See Tracing Skills for copy-paste prompts and validation guidance.
Human Review
Migrations
Resources
Sofia
How Skills Work
Skills use progressive disclosure to stay lightweight:- At startup (~100 tokens per skill): Only the
nameanddescriptionare loaded - When activated (under 5000 tokens): The full skill instructions load when your agent detects a relevant task
- On demand: Reference files (templates, examples) load only when needed