Skip to main content
Voice simulation evaluators assess how your agent sounds, listens, and converses. After a simulated voice interaction completes, evaluators score the agent’s audio output (TTS), transcription quality (STT), and conversational delivery — so you can catch robotic voices, mispronunciations, transcription errors, and unbalanced dialogue before they reach real users.

Why Voice Simulation Evaluators Matter

Voice agents fail in ways text agents do not. Voice simulation evaluators target these failure modes:

Evaluators Dashboard

Navigate to Library → Evaluators from the left navigation panel. Switch to the Library tab to browse all preconfigured evaluators. Netra organizes voice simulation evaluators into three categories: TTS, STT, and Conversational.

Library Evaluators

Netra provides 11 preconfigured library evaluators across three voice-specific categories: TTS, STT, and Conversational.

TTS Evaluators

TTS evaluators score the audio your agent produces.

STT Evaluators

STT evaluators measure how accurately your pipeline converts user speech to text.

Conversational Evaluators

Conversational evaluators assess how the agent behaves across the interaction.

Evaluator Configuration

LALM (Large Audio Language Model) evaluators listen to the recorded audio and score it with a judge model. Rule evaluators apply deterministic rules against measured values such as duration or latency.
You can adjust the pass criteria threshold for any evaluator based on your requirements. A higher threshold enforces stricter quality standards.

Configurable Parameters

Several voice evaluators accept parameters that you configure during evaluation creation:

Best Practices

Choosing Evaluators by Scenario Type

Getting Started with Evaluators

  1. Start with Humanness and Transcription Correctness — these cover the most critical aspects of any voice simulation
  2. Add scenario-specific evaluators — Language Accuracy for multilingual agents, Time to First Transcript for real-time use cases
  3. Adjust pass criteria if the default thresholds are too lenient or strict for your needs
  4. Monitor results across the first few test runs to ensure evaluators align with your expectations

Combining Voice and Text Evaluators

Voice evaluators assess audio quality, but they don’t evaluate what the agent said. For complete coverage, pair voice evaluators with text evaluators on the same evaluation:
When configuring a voice simulation evaluation, select both voice evaluators and text evaluators. Voice evaluators listen to the audio while text evaluators analyze the transcript — giving you end-to-end evaluation of the interaction.
Last modified on August 28, 2026