SpaceXAI's Grok Voice Think Fast 2.0 Jumps to #2 on Speech Benchmarks, Cuts Response Time in Half
SpaceXAI released Grok Voice Think Fast 2.0 on July 29, lifting its score on Artificial Analysis's Speech to Speech Index to 82.9% and its time-to-first-audio to 0.70 seconds, ahead of OpenAI's GPT-Realtime-2.1 and Google's Gemini 3.1 Flash on the same tests.
SpaceXAI, the combined xAI-SpaceX entity, released Grok Voice Think Fast 2.0 on July 29, 2026, a successor to the Grok Voice Think Fast 1.0 model it shipped earlier this year, with sizable gains on independent speech-to-speech benchmarks and a sharp cut in latency.
Benchmark gains
According to results published by Artificial Analysis, Grok Voice Think Fast 2.0's "High" reasoning variant scored 82.9% on the firm's Speech to Speech Index, debuting at #2 overall — behind only Qwen Audio 3.0 Realtime Plus (84.1%) and ahead of OpenAI's GPT-Realtime-2.1 High (79.1%) and Google's Gemini 3.1 Flash (69.5%). That's a 7.3 percentage-point jump from Think Fast 1.0's 75.7% score. On Artificial Analysis's Tau Voice benchmark, which measures agentic voice performance, the new model took the top spot at 56.5%, ahead of Qwen Audio 3.0 Realtime Plus (54.6%) and its own predecessor (52.1%).
Faster and among the first to reply
Beyond raw accuracy, SpaceXAI emphasized latency: average time-to-first-audio fell from 1.25 seconds in version 1.0 to 0.70 seconds in version 2.0, making it, per Artificial Analysis, the only model in the Speech to Speech Index's top five averaging under one second to first response. The company also said reasoning-token usage dropped by roughly 60% version-over-version, which should lower the effective cost of running the model at scale even before accounting for its published price of $0.08 per minute of audio.
Why it matters
Voice has become one of the more competitive fronts in the model race this year, as OpenAI, Google and Alibaba's Qwen team all ship realtime speech models aimed at voice agents and assistants rather than text chat. SpaceXAI is positioning Think Fast 2.0 specifically for agentic use inside its Agent Builder tooling, not just as a conversational demo, which tracks with the company's broader push — following July's Grok 4.5 release — to compete on task-completion and latency in agent workflows rather than general chatbot benchmarks alone. SpaceXAI said it will automatically migrate existing Grok users from Think Fast 1.0 to the new model on August 5, 2026.
Sources
- Artificial Analysis benchmark results for Grok Voice Think Fast 2.0
- SpaceXAI launches Grok Voice Think Fast 2.0 — americanbazaaronline.com
- SpaceXAI launches Grok Voice Think Fast 2.0 on Agent Builder — TestingCatalog
- xAI Unveils Voice AI 'Grok Voice Think Fast 2.0' with Dramatically Improved Transcription Accuracy and Inference Speed — BigGo Finance
AI-assisted reporting, overseen by the AgentsAI team. Spotted an error? Let us know.
Related agents
More ai news
Amodei's 'Pace the Frontier' Plan Draws Same-Day Backing From OpenAI, DeepMind and xAI
Anthropic CEO Dario Amodei published an essay arguing frontier AI labs should deliberately slow capability gains, and within hours Sam Altman, Demis Hassabis and Elon Musk publicly endorsed the idea, with Microsoft's Satya Nadella following a day later.
Positron Raises $875M to Build an HBM-Free AI Inference Chip
Chip startup Positron closed an $875 million Series C at a $5 billion post-money valuation to fund its Asimov inference accelerator, which pairs its compute architecture with up to 2,304GB of commodity LPDDR5X memory instead of scarce high-bandwidth memory.
DeepSeek Releases V4.1 Flash, Cuts API Prices and Sets End Date for V4 Pro
DeepSeek officially released V4.1 Flash, a cheaper and faster multimodal successor to V4 Pro with a 1-million-token context window, and said it will reroute all V4 Pro API traffic to the new model from September 14.
OpenAI Says a Swarm of 10,000 AI Agents Solved the Navier-Stokes Millennium Problem
OpenAI published a claimed solution to the Navier-Stokes existence and smoothness problem, one of math's seven Millennium Prize Problems, produced by roughly 10,000 coordinated AI agents and formally verified in the Lean proof language.