Universal 3.6 Pro Advances Real-Time Transcription (Full Transcript)

A new speech-to-text model delivers low-latency support for 32 languages, stronger noise handling, and better multilingual call accuracy.
Download Transcript (DOCX)
Speakers
add Add new speaker

[00:00:00] Speaker 1: Today we are releasing our most powerful speech-to-text model for real-time use cases yet. Let me show you with a quick demo. I'm going to first talk in English. You see four languages switched real-time with the lowest latency and we not just support four languages but now 32 of them. Now here are the ways our model performs better than ever before. Number one, during heavy noises. Think about a scenario when you have TV playing in the background or there's heavy noise. Here's an example. And the accuracy is proven to be way higher when the TV or the noise is louder. Number two, during one-word answers. For example, let's say an agent asks you can I have a have this meeting at 3 pm and now when you respond, it's sure. During those utterances, our word error rate has gone down as low as 1.45 percent. And number three, staying in the caller's language. Let's say you do a call in English and slowly sometimes switch to Spanish but the entire steering of that call will remain in the primary language that you have chosen. And here you can see noticeable accuracy improvement in Chinese, Arabic and other languages. And now we support 32 languages. And now we cannot wait to see how you use Universal 3.6 Pro through our API. Thank you.

ai AI Insights
Arow Summary
The speaker announces Universal 3.6 Pro, a new real-time speech-to-text model supporting 32 languages with low latency. The model improves transcription accuracy in noisy environments, for brief one-word responses, and when callers switch languages during a conversation. It maintains the selected primary call language despite occasional code-switching and shows notable gains in languages including Chinese and Arabic. The model is available through an API.
Arow Title
Universal 3.6 Pro: Real-Time Speech-to-Text
Arow Keywords
Universal 3.6 Pro Remove
speech-to-text Remove
real-time transcription Remove
32 languages Remove
low latency Remove
noise robustness Remove
word error rate Remove
code-switching Remove
API Remove
multilingual Remove
Arow Key Takeaways
  • Universal 3.6 Pro is positioned as the company’s most powerful real-time speech-to-text model.
  • It supports 32 languages and offers low-latency transcription.
  • Accuracy has improved in loud background-noise conditions.
  • For short, one-word answers, word error rate can be as low as 1.45%.
  • The model preserves the caller’s chosen primary language even when occasional language switching occurs.
  • The model is accessible through an API.
Arow Sentiments
Positive: The tone is enthusiastic and promotional, emphasizing product capabilities, accuracy improvements, and availability for developers.
Arow Enter your query
{{ secondsToHumanTime(time) }}
Back
Forward
{{ Math.round(speed * 100) / 100 }}x
{{ secondsToHumanTime(duration) }}
close
New speaker
Add speaker
close
Edit speaker
Save changes
close
Share Transcript