hotwirednews Install bot

Week 22

25–31 May 20262026-W22 · 9 filings · 4 companies · 0 net-new

The Hear, Speak, Think and Orchestrate layers led the week, with 2 filings each.

Hearspeech-to-text2

LiveKitPlatformlaunch2026-05-29

LiveKit Agents 1.5.15 adds Cartesia ink-2 STT

Cartesia STT model addition. Class-2: folded quiet patch(es) 1.5.17 into this row.

Details2 sources

livekit-agents 1.5.15 adds Cartesia ink-2 STT.

livekit-agents 1.5.15

Daily / PipecatPlatformlaunch2026-05-28

Pipecat 1.3.0 adds CartesiaTurnsSTTService

Cartesia's turn-aware ASR becomes selectable beside Deepgram Flux and AssemblyAI U3 Pro.

Details3 sources

pipecat-ai 1.3.0 adds CartesiaTurnsSTTService for Cartesia Streaming ASR with turn signals, plus max_endpoint_delay_ms on SonioxSTTService.

pipecat-ai 1.3.0

Speaktext-to-speech2

Daily / PipecatPlatformAPI2026-05-28

Pipecat 1.3.0 adds Rime coda TTS model support

Keeps Rime users on current model names after Arcana defaults in earlier releases.

Details3 sources

pipecat-ai 1.3.0 adds the Rime coda model to RimeTTSService and RimeHttpTTSService.

pipecat-ai 1.3.0

ElevenLabsModel providermodel release2026-05-28

ElevenLabs ships Dubbing v2 with performance-conditioned transfer

Performance-conditioned multilingual speech — not flat TTS from a transcript — is the gap studios and creators still pay humans to close; API still sales-gated at launch.

Details1 source

ElevenLabs launched Dubbing v2, which conditions on the original speaker's performance to carry tone, pacing, and emotion across 90+ languages with sync-aware translation, available in ElevenCreative and ElevenProductions with API access still forthcoming.

Dubbing v2

ThinkLLMs, speech-to-speech2

Daily / PipecatPlatformlaunch2026-05-28

Pipecat 1.3.0 adds Inception Mercury 2 LLM service

Diffusion LLM option for builders comparing cascade latency versus autoregressive peers.

Details3 sources

pipecat-ai 1.3.0 adds InceptionLLMService for Inception's Mercury 2 diffusion reasoning model.

pipecat-ai 1.3.0

KotobaModel providerAPI2026-05-28

Kotoba opens API and SDK alpha for S2S translation, streaming STT, and TTS

First developer-facing Kotoba surface in the corpus; pairs with the June seed and August Agentforce Japan filing.

Also HearSpeak

Details1 source

Kotoba Technologies released an alpha API and Python SDK exposing speech-to-speech translation plus streaming STT and TTS, aimed at builders evaluating East Asian voice workloads.

Kotoba API & SDK Alpha

Orchestrateframeworks, evals2

Daily / PipecatPlatformlaunch2026-05-28

Pipecat 1.3.0 ships multi-agent workers and UIWorker

Subagents and UI workers become a core runtime pattern rather than parallel pipelines glued by hand.

Details3 sources

pipecat-ai 1.3.0 makes pipelines multi-agent compatible by default and adds the pipecat.workers framework including UIWorker for observing and driving client web UIs.

pipecat-ai 1.3.0

ElevenLabsPlatformAPI2026-05-25

ElevenLabs ships Speech Engine for bring-your-own LLM voice

Splits the stack so ElevenLabs owns the speech loop and builders keep their own LLM — the missing middle between raw TTS/STT APIs and a fully hosted ElevenAgents runtime.

Also HearSpeak

Details2 sources

ElevenLabs released Speech Engine, a WebSocket path that handles STT, turn-taking, TTS, and browser playback while the developer's server owns the LLM logic and streams response text — an alternative to fully hosted ElevenAgents.

Speech Engine

Connecttelephony, channels1

Daily / PipecatPlatformlaunch2026-05-28

Pipecat 1.3.0 adds VonageVideoConnectorTransport

Vonage WebRTC joins Daily, LiveKit, and SmallWebRTC as a first-party transport.

Details3 sources

pipecat-ai 1.3.0 adds VonageVideoConnectorTransport for realtime Vonage WebRTC sessions.

pipecat-ai 1.3.0