G
Grok Voice Think Fast 2.0
SpaceXAI · Lanzado jul 2026
Condicional
El modelo speech-to-speech de nueva generación de SpaceXAI con una arquitectura de razonamiento paralelo que lo hace sustancialmente más inteligente que los modelos de voz convencionales.
¿Es adecuada para ti?
Bueno para
- Parallel reasoning during speech — the model thinks while it speaks, improving intelligence without latency penalty
- Transcription in noisy environments — 10x advantage over dedicated STT models when background noise is present
- Voice agent deployment at scale — 25+ languages, 21 voices, no-code Voice Agent Builder for production
No recomendado para
- Consumer voice assistant use — model is developer/API-only, no end-user conversational product like ChatGPT Voice
- Cost-sensitive high-volume telephony — at $0.08/min with $0.01/min telephony surcharge, budget STT-only pipelines are cheaper
Rendimiento por tarea
Speech-to-speech conversation
Excellent
SOTA parallel reasoning with 82.9% benchmark score, leading the category
Transcription accuracy
Excellent
1.5-2x better than dedicated STT models, 10x in noisy settings
Multilingual voice agents
Very Good
25+ languages with 21 voices, strong but real-world quality varies by language
Cost-effectiveness for high-volume
Good
$0.08/min premium pricing — dedicated STT+TTS pipelines are cheaper at scale
Precios
Entrada
$0.08 / min of audio (speech-to-speech)
Salida
Included in speech-to-speech rate
Contexto
API-only
Pruebas de rendimiento
Aún no hay cambios de veredicto
El reloj corre desde el primer día: los cambios aparecerán aquí a medida que evolucione nuestro veredicto.
Fuentes
Registro de verificación
- Precios— Sin cambios
Agente automatizado