Overview
SpaceXAI’s transcription model for batch and real-time speech recognition, improving multilingual accuracy, diarization, timestamps, and noisy real-world audio.
SpaceXAI’s transcription model for batch and real-time speech recognition, improving multilingual accuracy, diarization, timestamps, and noisy real-world audio.
Grok Voice Transcribe 2.0 is available in the Speech-to-Text API and is now the default model; 1.0 remains available but will be deprecated in the coming weeks.
Only published specifications are shown.
SpaceXAI’s transcription model for batch and real-time speech recognition, improving multilingual accuracy, diarization, timestamps, and noisy real-world audio.
It turns the audio foundation behind Grok Voice into a dedicated transcription API and replaces the first-generation model without requiring integration changes.
SpaceXAI made Grok Voice Transcribe 2.0 available in the Speech-to-Text API on September 17 and announced it would become the default transcription model; Grok Voice Transcribe 1.0 entered a deprecation transition. The launch article is dated September 18.
View sourceSelect a card to view the related model.
Copy this template, add the field, proposed value, and primary source, and share it with the archive maintainer.