Embeddable Interactive Avatars
Synthesia launched Interactive Avatars, a generally available plugin that attaches to a LiveKit Agent and renders a real-time talking avatar as a video participant in a web or mobile app, priced at $0…
Synthesia launched Interactive Avatars, a generally available plugin that attaches to a LiveKit Agent and renders a real-time talking avatar as a video participant in a web or mobile app, priced at $0…
Instruction following remains the biggest unsolved challenge in production voice AI in 2026, according to a guide published by voice AI platform Famulor, because most production voice agents still run…
Pipecat released PhoneLLM Alpha 1, an open-weights LLM for voice agent use cases, claiming it performs on par with GPT 5.6 Terra but is 94% cheaper and has a 1,300ms faster P95 time-to-first-token. Th…
Ojin launched its Real-Time GenAI Platform, enabling developers to deploy conversational AI agents with lifelike avatars in about two minutes via a single widget embed. The platform offers two face mo…
Pipecat, an open-source framework for voice AI agents, now supports monitoring and observability through OpenTelemetry, enabling developers to export logs, traces, and metrics to SigNoz for real-time …
Speko launched on July 29th as a benchmark-based router for voice AI models, giving developers a single gateway that selects speech and language models using continuously updated benchmarks. Founder B…
A developer running a voice AI sales training platform intentionally slowed down their voice agent's pipeline to fix an interruption problem. The pipeline's p95 latency of 880 ms was fast enough that …
AssemblyAI's Universal-3.5 Pro Realtime streaming speech-to-text model achieves a 6.99% pooled word error rate on Pipecat's open STT benchmark, outperforming competitors like Deepgram Flux (15.58%), E…
Ojin, a German-based AI company, launches a platform for building real-time human-like AI agents with voice and face capabilities, offering APIs for easy website integration and starting with $10 in f…
A developer building a real-time voice agent asks how production platforms like Bland.ai, Retell, and Vapi orchestrate prompts and conversation flow without relying on huge hand-written system prompts…
Jargo, a Go-based conversational-AI framework built on Pipecat's architecture, aims to redefine audio-first interactions but faces skepticism over its lack of originality. The framework targets Go dev…
Developer jasonmoo released Jargo, a Go port of the Pipecat conversational-AI framework, enabling real-time voice agents with WebRTC audio, streaming transcription-to-speech pipelines, and turn-taking…
The San Francisco Voice Company launched a real-time alert system for video and audio applications, designed for teams handling over 10,000 calls per day. The platform detects acoustic features beyond…