We learn to speak before we learn to read.
Voice is the most natural interface we have.
We just raised a $100M to make building voice AI as easy as a web app.
Gemini 3.5 Transcribe Live is now available in LiveKit Agents, bringing LLM-based realtime transcription built for alphanumerics, domain-specific vocabulary, multilingual speech, and code-switching.
Give it a try today on LiveKit Inference:
We built a patient intake agent with @SpaceXAI that runs end to end on Grok voice models.
Grok STT → Grok 4.3 → Grok TTS, cascaded through LiveKit Inference. Three model strings in one AgentSession. No separate API key, no separate billing, ZDR on every hop.
Talk to it: