Тест качества ответов

Раздел База знаний → Тест качества отвечает на вопрос «а что бот вообще говорит клиентам?». Вы задаёте список вопросов, агент отвечает на каждый — точно так же, как ответил бы посетителю сайта: та же модель, тот же промпт, та же база знаний.

После создания агента здесь фоном готовятся 10 типичных вопросов ваших покупателей по уже собранной информации сайта. Ассистент направляет в этот раздел: вопросы и ответы больше не нужно запускать в его чате. Подготовка вопросов не запускает тест. Если подготовка не удалась, можно повторить её или добавить свои вопросы. Удалённые вопросы не появляются заново.

Как пользоваться

  1. Вставьте вопросы в поле — по одному на строку. Удобно взять реальные вопросы клиентов из «Диалогов» или из «Пропусков».
  2. Нажмите «Добавить вопросы» — только после этого вопросы попадают в список. Пока список пуст, кнопка «Запустить тест» неактивна.
  3. Нажмите «Запустить тест». Рядом с кнопкой видно, сколько вопросов в списке и сколько примерно уйдёт кредитов.
  4. Посмотрите таблицу ответов. Плохой ответ — повод дополнить базу знаний или поправить промпт.
  5. Поменяли настройки — запустите тест ещё раз. Вопросы сохраняются, прогоны копятся в истории, их можно открывать и сравнивать.

Что важно знать

  • Ответы расходуют кредиты по обычному тарифу — это настоящие ответы агента, а не имитация. Сколько именно, зависит от выбранной модели (см. «Тарифы и кредиты»).
  • Тест не создаёт лидов, не отправляет данные в CRM, не дёргает оператора и не попадает в «Диалоги».
  • Каждый вопрос задаётся «с чистого листа» — агент не помнит предыдущие вопросы теста.
  • Если агент решил позвать оператора, у вопроса появится пометка — это тоже полезный сигнал.
  • Прогон можно остановить кнопкой «Остановить»: всё, что успело ответиться, сохранится. Если закрыть вкладку, тест остановится сам.
  • Одинаковые вопросы не дублируются: повторная вставка того же списка ничего не добавит.
  • Если вопрос не получилось задать, в таблице вместо ответа будет причина, а рядом — сколько времени ушло. «Ответ не уложился в отведённое время» значит, что модель думала слишком долго: выберите модель побыстрее или уменьшите глубину рассуждений в разделе AI.
  • У агента можно держать до 200 вопросов. Больше добавить не получится — удалите лишние.
  • Вопросы хранятся постоянно, ответы — 30 дней: они всё равно устаревают, как только вы поменяли модель, промпт или базу знаний. Старые прогоны убираются при следующем запуске теста, а не по расписанию: пока вы не запускали тест, история остаётся на месте.

Доступно на всех тарифах, включая бесплатный.

When a background onboarding check is enabled and available, its result appears beside the checked answer as a separate automatic assessment. It does not replace your expected answer. The assessment belongs to that saved run and answer; changing the answer does not carry the old assessment forward. Review the explanation and the cited business facts before deciding whether the agent needs adjustment.

Что дальше

Virtual buyer sales audits

In Knowledge base → Quality test → Virtual buyers, Standard and Pro owners can start a background audit of the saved agent. The audit infers the concrete goal from the agent prompt and settings (for example, a showroom visit, demo booking, order or resolved consultation), defines observable success criteria, describes two to four audience segments prioritizing sanitized customer questions and existing widget materials separately from private instructions, and runs twelve adaptive scenarios across them. Each scenario has up to eight buyer messages. Each journey has its own stable shopper situation, personal constraints and decision criteria. The buyer keeps qualitative working memory of visible claims, remaining concerns and readiness, without numerical psychology scores. A detailed report trace shows how these changed. Journeys may end earlier when the simulated buyer ignores the widget, leaves the conversation, or chooses a next step.

The report shows audience assumptions, the funnel stage at each simulated ending, exact quotes from problem interactions, concrete suggested changes, successful behaviors, and full conversations. A separate review removes unsupported criticisms before they appear in the report. If that review fails, the dialogue is retained, unreviewed findings are hidden, and the audit continues with partial coverage. A citation opens the exact referenced reply. These are hypotheses for investigation, not measured visitor psychology, conversion rates, or a promise of increased sales. Configured copy and conversation behavior are tested; actual hint timing, visual layout, and traffic exposure are not.

Before launch, the screen shows a credit ceiling and the required available balance. AI replies use the agent model's normal rate; scripted answers cost one credit. Buyer reactions and report preparation do not debit additional client credits. A reserved answer can consume credits even if execution later fails; the report preserves that charge and does not repeat it after an interruption.

You can close the page and return later. Stop requests take effect before the next operation; an already running operation may finish. Interrupted audits retain a partial report with the completed-scenario count; conclusions apply only to those completed scenarios. The hourly maintenance task releases an expired run if queue delivery was lost; paid operations are never replayed. Changing the agent settings, knowledge sources or catalog during a run stops it so one report does not mix configurations. Reports remain visible for thirty days; expired finished reports are removed by the daily cleanup. Synthetic conversations do not create leads, send CRM messages, contact operators, or appear in live conversation analytics. Access requires Knowledge base, Playground and Settings and Dialogs permissions because the report may quote configuration and customer-derived evidence.

Buyer profiles prioritize recent customer questions already in the widget. The bounded sample includes up to forty website/share-link conversations from the last ninety days and up to sixteen recent visitor messages per conversation. Contact messages, declared names, test sessions, recorded button taps and exact configured button labels are excluded or cleaned with the existing corpus privacy rules. Only sanitized visitor questions enter audience profiling; historical assistant answers and private agent instructions do not become the new buyer's knowledge. The report shows how much usable history supported the profiles. When there is no usable history, widget materials support explicit hypotheses; adding a website is not required. This sample cannot establish why actual people left or their overall conversion rate. Buyer/planner/analyst messages do not debit customer credits.