hotwirednews Install bot

Week 30

20–26 Jul 20262026-W30 · 14 filings · 8 companies · 2 net-new

The Speak layer led the week with 4 of 14 filings, from LiveKit, Microsoft, Daily / Pipecat and Alibaba / Qwen.

Speaktext-to-speech4

LiveKitPlatformlaunch2026-07-25

LiveKit Agents 1.6.7 adds Speechify streaming TTS with word timestamps

Speechify streaming TTS addition.

Details1 source

livekit-agents 1.6.7 adds Speechify streaming TTS with word-level timestamps and Fish Audio loudness/generation-tuning options.

livekit-agents 1.6.7

MicrosoftModel providermodel release2026-07-23

MAI-Voice-2-Flash: 2× faster TTS at $15 / 1M characters

In-house TTS at $15/1M characters for Dynamics Contact Center and Azure Voice Live. Vendor also claims up to 89% GPU-cost reduction in Contact Center.

Details1 source

Microsoft put MAI-Voice-2-Flash in public preview: 2× faster and 32% cheaper than MAI-Voice-2, at $15 per 1M characters, for Dynamics 365 Contact Center and Azure Voice Live.

MAI-Voice-2-Flash

Daily / PipecatPlatformlaunch2026-07-21

Pipecat 1.6.0 adds Deepgram Flux TTS

Builders can pair Flux STT and Flux TTS inside one Pipecat cascade.

Details3 sources

pipecat-ai 1.6.0 adds DeepgramFluxTTSService for Deepgram's Flux TTS early access over WebSocket.

pipecat-ai 1.6.0

Alibaba / QwenModel providermodel release2026-07-21net-new

Qwen-Audio-3.0-TTS: Flash and Plus, 16 languages, on Alibaba Cloud Model Studio

Q3 TTS comparisons that stop at Cartesia, ElevenLabs, and Inworld miss a multilingual, promptable option already on Alibaba Cloud.

Details1 source

Alibaba Cloud shipped Qwen-Audio-3.0-TTS in Flash (about 300ms first packet) and Plus variants on Model Studio, covering 16 languages with natural-language style control and inline non-verbal tags.

Qwen-Audio-3.0-TTS

Showavatars, video1

LiveKitPlatformlaunch2026-07-25

LiveKit Agents 1.6.7 adds Spatius avatar and inference.AvatarSession

Avatar gateway provisioning lands with Spatius.

Details1 source

livekit-agents 1.6.7 adds a Spatius avatar plugin and inference.AvatarSession for avatar provisioning via the gateway.

livekit-agents 1.6.7

ThinkLLMs, speech-to-speech1

Daily / PipecatPlatformAPI2026-07-21

Pipecat 1.6.0 adds OpenAI Responses reasoning and Flows NO_RESPONSE

Silent tool outcomes and Responses reasoning controls land for agentic voice flows.

Details3 sources

pipecat-ai 1.6.0 adds reasoning support on OpenAIResponsesLLMService and NO_RESPONSE returns from Pipecat Flows consolidated functions.

pipecat-ai 1.6.0

Orchestrateframeworks, evals3

LiveKitPlatformlaunch2026-07-25

LiveKit Agents 1.6.7 adds adaptive interruption for realtime models

Realtime interruption parity with pipeline agents. Class-2: folded quiet patch(es) 1.6.9 into this row.

Details2 sources

livekit-agents 1.6.7 supports adaptive interruption for realtime models.

livekit-agents 1.6.7

VapiPlatformlaunch2026-07-21

Vapi Model Intelligence adds presets and production model metrics

Turns Vapi's model-agnostic catalog into guided defaults, so builders pick a goal (latency, cost, intelligence) instead of hand-picking every STT, LLM, and TTS pair.

Details1 source

Vapi shipped Model Intelligence: curated Model Presets for Balanced, High Intelligence, Ultra Fast, and Cost Saver stacks, plus dashboard cost, latency, and quality metrics drawn from live Vapi traffic.

Vapi Model Intelligence

Daily / PipecatPlatformAPI2026-07-21

Pipecat 1.6.0 adds absent expectations to the eval harness

Negative assertions catch bots that speak when they should stay silent.

Details3 sources

pipecat-ai 1.6.0 lets eval scenario expectations set absent: true so a turn passes only when no matching event arrives within the window.

pipecat-ai 1.6.0

Connecttelephony, channels2

LiveKitPlatformlaunch2026-07-25

LiveKit Agents 1.6.7 adds Krisp voice isolation telephony mode

Telephony noise isolation mode.

Details1 source

livekit-agents 1.6.7 adds Krisp voice isolation telephony mode on the noise-cancellation plugin.

livekit-agents 1.6.7

Daily / PipecatPlatformlaunch2026-07-21

Pipecat 1.6.0 adds Media over QUIC (MoQ) transport

MoQ is a new transport class beside WebRTC and WebSockets for fan-out media.

Details3 sources

pipecat-ai 1.6.0 adds MOQTransport for bidirectional low-latency Media over QUIC between bots and clients.

pipecat-ai 1.6.0

Useagents in market3

SoundHound AIPlatformpartnership2026-07-23

SoundHound Smart Ordering integrates with Deliverect for end-to-end voice ordering

Turns restaurant voice AI from a standalone answering SKU into kitchen-connected fulfilment via an existing digital-ordering backbone.

Details2 sources

SoundHound and Deliverect connected Smart Ordering voice agents to Deliverect menus and POS routing so drive-thru, phone, kiosk, and in-car orders reach the kitchen without re-keying.

Smart Ordering × Deliverect

AnthropicModel providerlaunch2026-07-23net-new

Claude Voice Mode expands to Opus and Sonnet, still turn-based

Opus and Sonnet on voice without barge-in. Date from @claudeai coverage; tweet URL not independently verified.

Details3 sources

Anthropic expanded the Claude voice mode beta to Opus 5 and Sonnet 5 (plus Haiku 4.5) on mobile, desktop, and web. It remains turn-based, not full-duplex.

Claude Voice Mode

OpenAIPlatformlaunch2026-07-22

OpenAI launches Presence for governed enterprise voice and chat agents

Moves OpenAI from API voice models into a managed contact-centre-style SKU, competing with Sierra and other agent platforms on governed production deployments rather than only on gpt-realtime and GPT-Live.

Also ThinkOrchestrate

Details3 sources

OpenAI released Presence, a managed enterprise product for deploying governed AI agents on voice and chat that answer questions, take approved actions in company systems, and escalate to people under customer-set policies.

OpenAI Presence