G

Grok Voice Think Fast 2.0

SpaceXAI · Veröffentlicht Juli 2026

Bedingt

SpaceXAIs Sprach-zu-Sprache-Modell der nächsten Generation mit einer parallelen Reasoning-Architektur, die es deutlich intelligenter macht als herkömmliche Sprachmodelle.

Ist es das Richtige für dich?

Gut für

  • Parallel reasoning during speech — the model thinks while it speaks, improving intelligence without latency penalty
  • Transcription in noisy environments — 10x advantage over dedicated STT models when background noise is present
  • Voice agent deployment at scale — 25+ languages, 21 voices, no-code Voice Agent Builder for production

Nicht geeignet für

  • Consumer voice assistant use — model is developer/API-only, no end-user conversational product like ChatGPT Voice
  • Cost-sensitive high-volume telephony — at $0.08/min with $0.01/min telephony surcharge, budget STT-only pipelines are cheaper

Leistung nach Aufgabe

Speech-to-speech conversation

Excellent

SOTA parallel reasoning with 82.9% benchmark score, leading the category

Transcription accuracy

Excellent

1.5-2x better than dedicated STT models, 10x in noisy settings

Multilingual voice agents

Very Good

25+ languages with 21 voices, strong but real-world quality varies by language

Cost-effectiveness for high-volume

Good

$0.08/min premium pricing — dedicated STT+TTS pipelines are cheaper at scale

Preise

Eingabe

$0.08 / min of audio (speech-to-speech)

Ausgabe

Included in speech-to-speech rate

Kontext

API-only

Alle Preise ansehen

Benchmarks

BenchmarkWertQuelle
Artificial Analysis speech-to-speech82.9% Quelle
Artificial Analysis speech-to-speech (v1.0)75.7% Quelle
Conversational Dynamics (Full Duplex Bench)95.1% Quelle

Noch keine Urteilsänderungen

Die Uhr läuft ab dem ersten Tag — Änderungen erscheinen hier, sobald sich unser Urteil weiterentwickelt.

Prüfprotokoll

  • Preise— Keine Änderungen

    Automatisierter Agent

Wie wir bewerten