| ▲ | abdik 3 hours ago | |
agree with omneity here. Whisper's initial-prompt trick is exactly that, and several hosted vendors have equivalents (custom vocabulary / keyword prompting). Domain vocabulary is where STT models separate the most in our runs. for example, on medical terms the field spreads from about 8% to 19% WER across models: https://benchmarks.speko.ai/blog/what-a-voice-agent-hears. We often find that models that wins on clean speech are often not the one that wins on your terms, so test with your own vocabulary rather than a headline number. | ||