- Test Cases: reusable caller scenarios and their expected outcomes
- Test Runs: current and previous executions of those cases
- Assertions: additional criteria evaluated during every applicable test
Create test cases
Create cases in any of four ways:- New Test Case: write a scenario manually.
- Generate with AI: let Avoca analyze the selected agent, optionally provide focus instructions, and generate up to 20 cases.
- Add from library: attach existing reusable cases to the selected agent.
- Import from call: select real calls and queue their recordings or transcripts for conversion into test cases.
Configure tool mocks
Use Tool Call Overrides on a case to return a controlled response instead of allowing that tool to use its normal response. Mock responses can reference scenario facts and relative dates, which helps keep a case reusable. Fill Default Mocks previews missing mocks across the selected agent’s cases and fills gaps without overwriting mocks you already configured. Review the preview before applying it.Build a regression suite from real calls
When regression testing is enabled for your team, you can turn any production call into a repeatable case and re-run the whole collection alongside your other suites. Add a call to the suite- Open the call from Channels → Calls and go to the call details view.
- In the Actions row (next to Debug Call, Open Call Debugger, and Copy Call Link), select Add to regression suite.
- Write the Success criteria — a concise description of the ideal agent behavior for this call. Future runs are judged against these criteria, so the field is required.
- Save. The button now shows Remove from regression suite, so you can take the call back out without overwriting its stored outcome.
Run selected cases
- Select an agent.
- Choose individual cases, filter by tags, or select cases from a test suite.
- Click Run Tests.
- Give the run a recognizable name.
- Choose Voice or Chat when both are available. Chat availability depends on the selected agent and its configured provider key.
- For voice testing, choose how many times to sample each case, from 1 to 10.
- Review the final case list and start the run.
Review Test Runs
The Test Runs table shows the run name, date, number of cases, outcome score, and status. You can cancel an active run or compare two completed runs. Canceling removes the active run and cannot be undone. Open a run to review each case. The details can include:- Expected-outcome result and its checklist
- Assertion results, including skipped checks
- Recording and transcript
- Scenario facts
- Latency, talk-ratio, interruption, and other available metrics
Additional assertions are shown in run details as guardrails. Their failures are diagnostic and do not change the case’s overall pass or fail result; the expected outcome is the sole overall result.
Maintain Assertions
Open Assertions and select the agent.- Create an assertion with a clear name and importance.
- Choose when it should be evaluated, or write a custom trigger.
- State one specific evaluation criterion.
- Define explicit pass and fail criteria.
- Keep Enabled on only while the assertion is relevant.
Schedule recurring runs
When scheduling is available for the selected agent, select the cases and choose Schedule. Configure:- A recognizable schedule name
- Weekly, every-two-weeks, or monthly frequency
- The day of the week, or days 1–28 for a monthly schedule
- Voice or chat mode
- Whether to run immediately as well as create the schedule
Diagnose a failed case
Trace the behavior to its source, make one focused change, and rerun the same case:- Incorrect business answer: review the Knowledge Base.
- Missing or invalid appointment: review Booking Windows, CRM mappings, Functionality, and Bookability.
- Wrong transfer: review the destination, schedule, reasons, and pre-call rules in Call transfers.
- Wrong outcome: review Call Classification.
- Wrong greeting or voice: review Settings → Agents.