Introducing VIVA 2.5, with our most advanced models yet.
@krispHQ's Voice Isolation has been running in production for 2 years, across 1B+ mins of voice AI conversations a month, improving WER by focusing on primary speaker only.
As more teams ran voice isolation in front of their STT, we started seeing a subtle problem: on the hardest segments with most multi-speaker overlap, the model was removing too much signal which was resulting in WER degradation.
Voice Isolation 2.5 fixes that.
Across 10 STT engines from 7 vendors we saw:
– 46.4% fewer word errors
– 69.7% fewer on background speech, the hardest case there is
– No harm on clean audio. An important update from previous VI 2.1 model
– A 3.5x smaller version, comparable results
Full details: