DigestAI news desk

Cut through the AI noise.

Developing story · 2 episodes · 18 Sept to 26 Sept

Rapid Advances in Multilingual Speech Recognition Technology

The saga tracks the competitive release of advanced speech-to-text models by major AI companies. It currently stands with Sarvam AI launching its Saaras V4 model for Indian languages and English, following xAI's Grok Voice Transcribe 2.0.

  1. Sarvam AI launches Saaras V4 speech-to-text model for 22 Indian languages and global English

    Sarvam AI released Saaras V4, a speech-to-text model supporting all 22 scheduled Indian languages plus global English accents. The encoder-decoder system uses a 3B-parameter hybrid state-space…

    1 source
  2. xAI releases Grok Voice Transcribe 2.0 speech-to-text model

    On September 18, 2026, xAI introduced Grok Voice Transcribe 2.0, its newest speech‑to‑text model. The service keeps batch pricing at $0.10 per hour of audio and $0.20 per hour for streaming, and xAI…

    2 sources
Who and what