AssemblyAI launches Universal-3.5 Pro with native code-switching and improved speaker diarization
AssemblyAI released Universal-3.5 Pro, a new flagship async speech-to-text model priced at $0.21/hr. It natively transcribes code-switched conversations across 18 languages without separate configuration, introduces jointly modeled speaker diarization built directly into the transcript (rather than stitched from a separate system), and supports contextual prompting to prime the model with domain knowledge. Benchmarks show lower word error rate on code-switched audio and higher cpWER accuracy on diarization compared to several competitor models and AssemblyAI's own prior Universal-3 Pro.