LiveKit hosts Gemma 4 31B inference tuned for voice agents
LiveKit's own latency-first LLM SKU for Agents — 192 ms TTFT vendor claim versus multi-hundred-millisecond hosted peers on the same voice-shaped prompts.
Also Think
Details1 source
LiveKit Inference began serving Gemma 4 31B with voice-shaped serving (long system prompts, large tool catalogs, low queueing) billed on the same LiveKit API key as Agents.
Gemma 4 31B on LiveKit Inference
xAIPlatformlaunch2026-07-01
xAI opens Voice Agent Builder beta on Grok Voice
Packages Grok Voice S2S with telephony and tools under one meter — a platform SKU ahead of Think Fast 2.0 (29 Jul), not another raw model card.
Also ThinkConnectUse
Details2 sources
xAI launched Voice Agent Builder in beta: a no-code console that turns a plain-language call flow, documents, tools, and guardrails into a production phone agent on Grok Voice speech-to-speech.
Voice Agent Builder
VercelPlatformAPI2026-06-29net-new
Vercel AI Gateway adds realtime voice, TTS, and transcription
Puts OpenAI-class realtime voice behind the same Gateway metering Next.js teams already use for text — beta on 29 Jun, before OpenAI's 6 Jul gpt-realtime-2.1 API bump.
Also HearSpeakThink
Details1 source
Vercel put realtime voice agents, text-to-speech, and speech-to-text on AI Gateway in beta through AI SDK 7, with the same observability, spend controls, and bring-your-own-key path as its other modalities.
AI Gateway realtime voice
Retell launches Conductor, a graph-native copilot for voice agents
Moves Retell from no-code build into continuous ops: failed-call → simulation → reviewable patch, with human approval on every live-agent edit.
Details3 sources
Retell announced Conductor, a review-first AI copilot that proposes edits inside a production voice-agent workflow, generates simulation tests from failed calls, and ships nothing until a human accepts each change.
Conductor