Common Workflows
Build test sets
Ask: “Create a test set for a billing support agent with refund, cancellation, and payment-failure scenarios.” The assistant can create the test set and add individual test cases through separate write tools.Launch and monitor evaluations
Ask: “Run the billing test set against my support agent, then tell me the new run ID.” The assistant resolves the named resources before creating the run.Inspect results
Ask: “Which of my recent evaluations were unsuccessful?” The assistant reads your run data and summarizes the result without changing anything.Consult Sofia
Ask: “Have Sofia inspect my latest failed run and recommend what I should test next.” Sofia combines your organization’s evaluation evidence with Coval product and voice-agent evaluation knowledge. This consultation is read-only.Connect
- Install Coval from the ChatGPT or Claude directory, or add
https://mcp.coval.dev/mcpto Codex or another MCP client. - Sign in to Coval in the browser window that opens.
- Select the organization you want the assistant to access.
- Approve access and return to your client.
- Start with: “List my Coval agents and three most recent runs.”
Working Safely
- Name the exact resource when requesting a write or update.
- Review the assistant’s proposed inputs before approving costly work such as a run.
- Use disposable names while testing the connector.
consult_sofiacan analyze and recommend, but it cannot mutate your organization.- Disconnect and reconnect when you need to switch organizations.