Why Voice Simulation Evaluators Matter
Voice agents fail in ways text agents do not. Voice simulation evaluators target these failure modes:Evaluators Dashboard
Navigate to Library → Evaluators from the left navigation panel. Switch to the Library tab to browse all preconfigured evaluators. Netra organizes voice simulation evaluators into three categories: TTS, STT, and Conversational.Library Evaluators
Netra provides 11 preconfigured library evaluators across three voice-specific categories: TTS, STT, and Conversational.TTS Evaluators
TTS evaluators score the audio your agent produces.STT Evaluators
STT evaluators measure how accurately your pipeline converts user speech to text.Conversational Evaluators
Conversational evaluators assess how the agent behaves across the interaction.Evaluator Configuration
LALM (Large Audio Language Model) evaluators listen to the recorded audio and score it with a judge model. Rule evaluators apply deterministic rules against measured values such as duration or latency.
Configurable Parameters
Several voice evaluators accept parameters that you configure during evaluation creation:Best Practices
Choosing Evaluators by Scenario Type
Getting Started with Evaluators
- Start with Humanness and Transcription Correctness — these cover the most critical aspects of any voice simulation
- Add scenario-specific evaluators — Language Accuracy for multilingual agents, Time to First Transcript for real-time use cases
- Adjust pass criteria if the default thresholds are too lenient or strict for your needs
- Monitor results across the first few test runs to ensure evaluators align with your expectations
Combining Voice and Text Evaluators
Voice evaluators assess audio quality, but they don’t evaluate what the agent said. For complete coverage, pair voice evaluators with text evaluators on the same evaluation:When configuring a voice simulation evaluation, select both voice evaluators and text evaluators. Voice evaluators listen to the audio while text evaluators analyze the transcript — giving you end-to-end evaluation of the interaction.
Related
- Evaluators Overview - Understand the full evaluator framework
- Text Evaluators - LLM-as-Judge, code, and rule-based text evaluators
- Image Evaluators - Multimodal and rule-based image evaluators
- Simulation Overview - Understand the full simulation framework
- Voice Evaluations - Create voice simulation evaluations
- Test Runs - View transcripts and synced call audio
