
SpaceXAI releases Grok Voice Transcribe 2.0 speech-to-text model
SpaceXAI released Grok Voice Transcribe 2.0, a speech-to-text model it says is twice as accurate as version 1.0 at the same price. It ranks first for accuracy among 32 streaming models on the Artificial Analysis leaderboard. The model is trained on live, noisy, multilingual audio and supports dozens of languages with automatic detection and mid-recording switches.
Sources and evidence
Attributed quotes
“Today we're releasing Grok Voice Transcribe 2.0, our latest speech-to-text model.”
“Across our real-world evaluations, Grok Voice Transcribe 2.0 is one of the most accurate transcription models available today and twice as accurate as Grok Voice Transcribe 1.0, at the same price.”
“On the public Artificial Analysis leaderboard, Grok Voice Transcribe 2.0 ranks first for accuracy among 32 streaming models.”
Summary last validated Oct 6, 2026
Reader actions
Report an issue
Use this for an incorrect summary, wrong source, duplicate story, or wrong category. Submissions are private and do not change the story automatically.
Related coverage
- models
Cloudflare adds Clef-omni multimodal model, speeds Clef and cuts Clef-flash pricing
Cloudflare expanded its Clef decision model family with Clef-omni, which natively processes audio, video, images, and text in a single pipeline. The company also increased Clef inference speeds by up to 2.0x and lowered pricing for Clef-flash.
- models
AWS Adds Bedrock Models, Faster AgentCore Agents, Knowledge Base Sync
AWS's September 2026 recap covers Amazon Bedrock, Bedrock AgentCore, and Strands updates: broader model choice, faster serverless agents with built-in evaluation, and automated knowledge base syncing via native enterprise connectors. The post aggregates the month's releases for AI builders across AWS's model and agent tooling.
- models
Anthropic's Claude Haiku 5.5 launches on Amazon Bedrock and Claude Platform on AWS
Claude Haiku 5.5 is now available on Amazon Bedrock and Claude Platform on AWS. Anthropic says it is the fastest, most efficient model in the Claude 5.5 family, built for subagents and high-volume, cost-sensitive work, and costs roughly 75% less than Claude Haiku 4.5 for most tasks.
- models
Anthropic ships Claude Haiku 5.5 as default Haiku model in Claude Code v2.1.293
Claude Code v2.1.293 adds Claude Haiku 5.5 (claude-haiku-5-5), now the default Haiku model on the Anthropic API, with 1M context and pricing of $0.10/$0.50 per million tokens, rising to $0.50/$2.50 for prompts over 100K tokens. The release also adds agentType to subagentStatusLine and isDeferred to tool registration.
- models
Liquid AI releases d1-3B and d1-omni-600M open decision models for edge
Liquid AI released two open decision models: d1-3B and experimental d1-omni-600M. d1-3B scores 48.57 on Decision Index 0.2.1, topping sub-10B models and Decider 35B-A3B, and answers in 16 ms on NVIDIA Jetson AGX Thor. d1-omni-600M handles text with image or audio. Both are built on Liquid Foundation Models.