The words for reading back what an agent actually did, and what it cost to do, one section per term.

Conversations

Transcript

The record of one conversation the agent actually had, with every turn and trace. The transcripts table shows each at a glance; filters narrow it by date, platform, user ID, environment and more.

Also asked as:

  • read back a conversation the bot had
  • find a chat by the customer’s ID
  • the history of calls and chats

Related: Transcript property · Evaluation · PII redaction · Call recording

Documented in: Transcripts › Understanding the transcripts table · Transcripts › Filtering transcripts · Transcripts › Reviewing individual transcripts · Observe (API)

Transcript property

A label you define once and then set on individual conversations, so your own outcome data can be filtered alongside Voiceflow’s. Unlike variables, properties aren’t used during the conversation.

Also asked as:

  • tag a conversation with our own outcome
  • attach custom metadata from a workflow
  • filter transcripts by a label we set

Related: Transcript · Variable

Documented in: Transcripts › Attaching custom properties

Transcript settings

Preferences for the transcripts view: whether test conversations are saved to transcripts, and which columns the table shows.

Also asked as:

  • hide test chats from the transcript list
  • add an evaluation column to the table
  • exclude test conversations from transcripts
  • which columns show in the transcripts table

Related: Transcript · Test

Documented in: Transcripts › Configuring transcript settings · Transcripts › Customizing the table view

Evaluations

Evaluation

A criterion a model scores past conversations against, turning quality into a number you can track. Each has a name, a model, a metric type and instructions; results feed the Analytics dashboard.

Also asked as:

  • score conversations automatically
  • rate how well the assistant handled each chat
  • a rubric the AI applies to transcripts

Related: Default evaluations · Resolution · Batch evaluation · Analytics

Documented in: Evaluations › Understanding evaluations · Evaluations › Creating an evaluation · Evaluations › Viewing evaluation results · Insights (API)

Default evaluations

The three evaluations included out of the box: Customer satisfaction (a 1–5 rating from tone and content), Deflection rate (was the issue resolved without a human) and Resolution rate (the outcome category). Edit them or create a custom evaluation for different criteria.

Also asked as:

  • the scores Voiceflow gives every conversation
  • built-in quality metrics
  • customer satisfaction, deflection and resolution scores
  • the metrics that come pre-configured

Related: Evaluation · Resolution

Documented in: Evaluations › Default evaluations

Resolution

The default evaluation that sorts every conversation into one of five outcomes; these feed the Resolution chart on the Analytics dashboard.

Also asked as:

  • was the customer’s problem actually solved
  • resolved versus escalated conversations
  • outcome categories on the dashboard

Related: Evaluation · Analytics

Documented in: Evaluations › Resolution categories · Analytics › Drilling into evaluation results

Batch evaluation

Running an evaluation over transcripts that already exist. Evaluations run automatically only on new transcripts, so scoring history means selecting transcripts and running a batch.

Also asked as:

  • apply a new scoring rule to older transcripts
  • apply an evaluation retroactively
  • backfill scores on old chats
  • grade older conversations with a new evaluation
  • run an evaluation on last month’s transcripts

Related: Evaluation · Transcript

Documented in: Evaluations › Running evaluations on past transcripts · Transcripts › Running batch evaluations

Evaluation costs

Each evaluation consumes a small amount of credits because a model analyses the transcript; the exact cost depends on the model chosen.

Also asked as:

  • do evaluations use up credits
  • how much does scoring cost
  • how many credits an evaluation uses
  • does automatic scoring cost anything

Related: Evaluation · Credits

Documented in: Evaluations › Evaluation costs

Tests

Test

A scripted conversation replayed against the agent, with checks asserting what it does. Tests never run automatically: you run one, or a batch, yourself.

Also asked as:

  • a simulated conversation to check the bot’s behaviour
  • regression tests before shipping a change
  • make sure the assistant still escalates when it should

Related: Test turn · Test check · Test run · Persona

Documented in: Tests › Key concepts · Tests › Creating a test · Tests › Running tests · QA (API)

Test turn

One side of a scripted conversation: either what the user says, or the agent reply the checks assert against.

Also asked as:

  • one message in a test script
  • a user message or agent reply in a test
  • the steps of a scripted test conversation

Related: Test · Test check

Documented in: Tests › Creating a test

Test check

One assertion about an agent turn. Compares the reply itself, or the behaviour behind it such as which tool was called.

Also asked as:

  • verify the bot called the right tool
  • assert what the reply must contain
  • assert that the agent escalated
  • verify the reply contains a phrase

Related: Test · Test turn

Documented in: Tests › Adding checks to an agent turn

Test run

One execution of a test, carrying how many of its checks passed. Runs are asynchronous, so poll until the status settles. Past runs are listed per environment.

Also asked as:

  • see which tests failed last time
  • run every test at once
  • results of the last test batch
  • which checks failed in my test

Related: Test · Credits

Documented in: Tests › Running a single test · Tests › Running multiple tests · Tests › Reviewing past runs

Analytics and cost

Analytics

Aggregate usage and cost: tokens, calls, conversations, unique users. The dashboard is filtered by environment and date range, compares each metric with the preceding period, and breaks conversations down by evaluation outcome.

Also asked as:

  • how many people talked to the bot this week
  • usage numbers per day
  • a dashboard of conversation volume
  • daily count of visitors who chatted with the assistant
  • unique users per day
  • conversation counts over time

Related: Evaluation · Resolution · Business impact · Credits

Documented in: Analytics › Filtering your view · Analytics › Customizing the dashboard · Analytics › Trends · Analytics (API)

Business impact

The Time saved and Dollars saved widgets, which estimate the value the agent provides from figures you enter.

Also asked as:

  • put a dollar figure on what the assistant saves us
  • report ROI from the dashboard
  • estimate hours and dollars saved by the agent
  • show ROI on the analytics dashboard

Related: Analytics

Documented in: Analytics › Estimating business impact

Credits

The unit Voiceflow bills usage in. Agents consume credits for AI responses, phone calls, evaluations and tests; paid plans include a monthly allotment, and a bigger bundle or auto top-ups add more.

Also asked as:

  • what does each message cost us
  • does a bigger model burn more credits
  • we ran out of credits mid-month

Related: Plan · Auto top-up · Faster processing

Documented in: Credits pricing table · Billing overview › Credits · Tests › Test costs

Use the up and down arrow keys to select a result, Enter to open it, and Escape to close the search.