Sarvam's Saaras V4 transcribes 22 Indian languages

TL;DR

  • Sarvam published Saaras V4 on 24 August: a speech-to-text model for all 22 Indian languages and English, including Indian-accented English.
  • One model gives five outputs (verbatim, transcribe, codemix, translit, translate), takes up to 50 key terms, and streams with under 150 ms to the first token, Sarvam says.
  • It's the default model on Sarvam's API, with REST, batch and WebSocket access and SDKs for Python and Node.

Read the full story

Sign in with your email to read AI News. It’s free.

Share this story

Explain like I’m 15