Skip to main content
The By Question metric evaluates how well agents respond to specific questions during customer interactions. Supervisors can apply this metric to all conversations (Static) or trigger it only when specific conditions occur (Dynamic). This metric helps teams measure response quality, enforce compliance, and provide targeted coaching. By Question is available in two versions, By Question V1 and By Question V2. You choose the version when you create the metric. V2 is an addition, not a replacement: existing V1 metrics continue to work unchanged, and you can still create new V1 metrics.

Key Benefits

  • Standardized quality assessment framework.
  • Customizable evaluation criteria based on business needs.
  • Multi-language support for global operations.
  • Automated evaluation suggestions through AI.

When to Use This Metric


Choose a Version: V1 or V2

When you add a By Question metric, select the version that matches the complexity of the rule you need to evaluate.
  • By Question V1: Separate trigger and response rules. Limited logic.
  • By Question V2: Combined rule with condition and response. Clear justifications, supports Not Applicable and complex logic.
By Question V2

How the Metric Works

The metric operates through a question-driven evaluation process with two main adherence approaches. Both approaches are available in V1 and V2.
  • Static Adherence: Evaluates every response across all conversations; used for universal rules (for example, greetings). No triggers needed.
  • Dynamic Adherence: Evaluates responses only when specific triggers or conditions occur; used for conditional rules that apply in certain situations.

Configure a By Question Metric

These steps are common to both versions. The version-specific configuration follows in the sections below.
  1. Navigate to Quality AI > Configure > Evaluation Forms > Evaluation Metrics.
  2. Select + New Evaluation Metric.
  3. From the Evaluation Metrics Measurement Type dropdown, select By Question. Measurement Type
  4. Select the version — By Question V1 or By Question V2.
  5. Enter a descriptive Name (for example, Agent's Warm Greeting).
  6. Select a Language used for evaluation.
  7. Enter an evaluation Question, which is a reference prompt for supervisors during audits and reviews. Question and Adherence Type
  8. Select the Adherence Type (Static or Dynamic).
For Static, configure at least one agent answer utterance. For Dynamic, configure at least one trigger and one agent answer utterance.

By Question V1 Configuration

V1 configures the trigger and the agent answer as two separate rules. Use V1 for straightforward checks that do not require Not Applicable justifications or multi-condition logic.

Adherence Type Configuration

V1 supports two evaluation modes. Static Adherence: Applies to all calls without any condition or trigger. Use for mandatory, universal compliance items like greeting scripts or regulatory disclaimers.
  • Define acceptable utterances for the queue.
  • Set a similarity threshold to evaluate whether the agent’s response matches the predefined utterance.
  • Configure at least one agent utterance template.
Dynamic Adherence: Evaluates adherence only after detecting a configured trigger. Use for context-sensitive checks.
  • Define at least one trigger (customer or agent utterance) and one acceptable agent response.
  • Set the similarity threshold based on criticality:
    • Lower threshold (~60%, yellow) — for casual interactions or greetings.
    • Higher threshold (~100%, green) — for critical topics such as legal disclaimers or privacy policies.

Trigger Configuration

Choose the utterance source that initiates the trigger: You can add multiple trigger utterances and define AND/OR conditions to control when the trigger activates.

Trigger Detection Method

Agent Answer Configuration

GenAI-Based Adherence

Uses LLMs to detect meaning, context, and intent. Evaluates whether the agent’s answer fulfills the intent, even if phrased differently.
  1. Select GenAI-Based Adherence as the answer detection method.
  2. Enter a Description explaining the metric’s intent.
Before using GenAI-based adherence, ensure supported models and the GenAI features (GenAI-based agent answer adherence and customer trigger detection) are enabled. No sample utterances or thresholds are required — LLMs evaluate adherence using zero-shot prompts.

Deterministic Adherence

Evaluates responses based on semantic similarity to predefined reference utterances.
  1. Select Deterministic Adherence.
  2. Define an Answer — acceptable utterances per queue. Use Generative AI to suggest similar variations.
  3. Set the Similarity threshold:
    • Lower threshold (~60%) for soft skills like greetings and etiquette.
    • Higher threshold (~100%) for compliance-critical statements (Policy, Privacy, Disclaimers).

Count Type Configuration

Choose how the metric evaluates adherence across the conversation: Entire Conversation: Evaluates adherence throughout the entire interaction at any point. Time Bound: Focuses on specific timeframes — the first or last X seconds or messages.
  1. Select Parameter — First Part or Last Part of the conversation.
  2. Set the evaluation window:
    • Voice — enter number of seconds.
    • Chat — enter number of messages.
  3. Select Create to save and activate the metric.

By Question V2 Configuration

V2 replaces the split trigger/answer architecture with a single unified rule definition. Instead of two disconnected rules, V2 describes the trigger condition and the expected agent behavior in one natural-language description. Use V2 when you need Not Applicable justifications or multi-condition logic.

Unified Rule Definition

Define the condition and the expected agent behavior together in a single natural-language description rather than as separate trigger and answer rules. The auditor reads this description to return one of three verdicts for the conversation.

Write the Metric Description (Best Practices)

Write the description as a set of labeled sections. The auditor reads these sections in a fixed order, so a clear, well-structured description produces consistent verdicts. Follow the structure, connectors, and style rules below.

Description Structure

A description uses up to four sections. Always write them in this order. Only EXPECTED BEHAVIOR is mandatory; CONDITION, PROHIBITED, and EXCEPTION are optional.

Simple and Extended Form

Use the simple form when a section holds a single statement (see Patterns A and B). Use the extended form when a section holds multiple components joined by connectors:

Logical Connectors

Join multiple items with one of the following connectors. Use a single connector type per list.
Don’t mix AND and OR in the same list. Mixed connectors make the logic ambiguous. If you need branching logic, create separate metrics instead.

How the Auditor Evaluates the Description

The auditor applies the sections in a fixed order and returns the first verdict that matches:
  1. If any EXCEPTION matches, the verdict is NA.
  2. If any PROHIBITED action occurs, the verdict is NO.
  3. If the CONDITION is not met, the verdict is NA.
  4. If the EXPECTED BEHAVIOR is fully satisfied, the verdict is YES.
  5. Otherwise, the verdict is NO.
Because the auditor reads condition and behavior together, it produces one coherent explanation — and a clear reason when the metric doesn’t apply.

Style Rules

  • Write in the present tense and active voice.
  • Describe observable actions, not subjective qualities.
  • Refer to each speaker clearly so responsibilities are unambiguous.
  • Write section headers and connectors in uppercase.
  • End every statement with a period.
  • Keep each bullet to 20 words or fewer, and the full description to about 12 lines.

Do and Don’t

Avoid subjective language. Describe what the agent does, not how they seem.
Don’t mix connectors in one list. Mixed AND and OR make precedence ambiguous.
Don’t use conditional branching. Create separate metrics instead of “if/otherwise” logic.
Also, avoid internal tool references, message identifiers, and turn identifiers in the description.

Standard Patterns

Start from one of these patterns and adapt it to your requirements. Pattern A — Trigger and response
Pattern B — Unconditional requirement
Pattern C — Sequential workflow
Pattern D — Alternative acceptable responses
Pattern E — Positive requirement with guardrails

Description Template

Use this template as your starting point:

Not Applicable Handling

Not Applicable is a first-class output in V2. When an exception matches or the condition isn’t met, the auditor returns Not Applicable directly, along with its own justification explaining why the metric did not apply. The application no longer infers Not Applicable from a trigger result.

Evaluation Window Constraints

V2 evaluates the rule against a turn-based window. Choose where in the conversation the metric applies:
  • Entire conversation — evaluate at any point.
  • First X turns — evaluate only the opening turns.
  • Last X turns — evaluate only the closing turns.
A turn is one customer block and one agent block.

Audit Screen (V2)

For V2 metrics, the Audit Screen shows two separate justification fields, giving evaluators a clear, auditable explanation for every outcome:
  • Condition determination — why the condition did or did not apply.
  • Adherence verdict — whether the agent complied when the condition applied.

Enable GenAI-Based Features (Prerequisites)

To activate GenAI-based features:
  1. Navigate to Manage > Generative AI > GenAI Features.
  2. Enable and publish the following features:
    • GenAI-based agent answer adherence
    • GenAI-based customer trigger detection
    GenAI-based Features
GenAI-based adherence uses Large Language Models (LLMs) to evaluate responses based on intent and context. It doesn’t require example utterances or similarity thresholds, as adherence is evaluated using zero-shot prompts. For guidance on writing effective prompts, refer to AutoQA Prompting Guide.

Edit or Delete a By Question Metric

  1. Select the metric from the By Question category. Edit Metric
  2. Choose an option: Edit to modify the required details, or Delete to remove.
  3. Select Update to save changes.

Language Dependency Warnings

You can’t remove a language while any evaluation form or metric uses it. Remove the language from all linked forms and metrics before deleting it. Language Warning

Delete Warnings

Before deleting a metric:
  1. Remove it from all associated evaluation forms.
  2. If any attributes link to the metric, reassign them to a different metric first.
  3. Resolve all dependencies before deleting the metric.
If the metric is still in use, the system displays a warning and prevents deletion.