hotwirednews Install bot

Week 24

8–14 Jun 20262026-W24 · 6 filings · 6 companies · 2 net-new

The Speak and Orchestrate layers led the week, with 2 filings each.

Hearspeech-to-text1

SonioxModel providermodel release2026-06-11

Soniox ships v5 Async STT for structured speech data

Opens the v5 stack on batch audio five days before Real-Time — speaker-aware, language-aware transcripts aimed at automation, not cleaned-up walls of text.

Details2 sources

Soniox released stt-async-v5 for recorded audio, with higher multilingual accuracy across 60+ languages, reengineered speaker separation, native language ID, session context injection, and stronger alphanumeric formatting of numbers, emails, codes, and names.

Soniox v5 Async (stt-async-v5)

Speaktext-to-speech2

Resemble AIModel providermodel release2026-06-10

Resemble ships Chatterbox Multilingual v3 with default PerTh watermark

Open-weight multilingual TTS that watermarks every self-hosted sample by default — compliance feature, not an optional add-on, timed ahead of Article 50.

Details2 sources

Resemble released Chatterbox Multilingual v3, an open-weight 0.5B MIT-licensed TTS covering 25 languages with PerTh watermarking embedded by default, plus an NVIDIA NIM path and Single-Language Pack checkpoints for priority dialects.

Chatterbox Multilingual v3

GradiumModel providermodel release2026-06-10net-new

Gradium upgrades default TTS for hard-case pronunciation

Default-path pronunciation fix for the phone-quality cases that break agents — emails and codes — weeks before the Aug hard-case eval packaging.

Details1 source

Gradium made a new TTS model the API default, targeting email, phone, acronym, and code read-out accuracy across English, French, Spanish, Portuguese, and German, with prior voices carrying over and no migration step.

Gradium TTS (June upgrade)

ThinkLLMs, speech-to-speech1

GoogleModel providermodel release2026-06-09

Google ships Gemini 3.5 Live Translate for continuous S2S

Puts continuous S2S translate into Gemini Live API, Meet, and consumer Translate on one day — 70+ languages with SynthID, and builder paths that name both Pipecat and LiveKit.

Also SpeakHear

Details1 source

Google released Gemini 3.5 Live Translate, a continuous speech-to-speech translation model covering 70+ languages with preserved intonation, SynthID watermarking, developer preview on the Gemini Live API and AI Studio, Meet private preview, and Google Translate app rollout.

Gemini 3.5 Live Translate

Orchestrateframeworks, evals2

LiveKitPlatformlaunch2026-06-11

LiveKit Agents 1.6.0 (Python + JS) introduces asynchronous tools

Major tool-runtime change for voice agents that call slow backends, filed once for Python and Node.

Details2 sources

livekit-agents and @livekit/agents 1.6.0 introduce asynchronous tools so long-running tool calls can stream progress updates and keep the conversation alive instead of silent waits.

livekit-agents / @livekit/agents 1.6.0

CrestaPlatformlaunch2026-06-11net-new

Cresta launches Conductor agent-building agent for CX

Turns Cresta's conversation data and Testing Suite into an agent that builds the CX agent — blueprint-before-code, not another prompt playground.

Also Use

Details2 sources

Cresta launched Cresta Conductor, a natural-language agentic engine that discovers requirements, produces a reviewable blueprint, generates prompts, subagents, configs, and custom code, then tests with Synthetic Customers and optimises from production transcripts.

Cresta Conductor