LiveKit
PulseAugur coverage of LiveKit — every cluster mentioning LiveKit across labs, papers, and developer communities, ranked by signal.
4 day(s) with sentiment data
-
Fish Audio raises $50M seed for AI voice models, serving 8M users
Fish Audio, a startup specializing in AI voice models, has secured $50 million in seed funding. The company offers both open-source and hosted solutions, catering to creators and enterprises with over 15,000 natural lan…
-
AssemblyAI's Universal-3.5 Pro Realtime hits human parity in speed and accuracy
AssemblyAI's Universal-3.5 Pro Realtime model has achieved a significant milestone by being the sole entry within Coval's Human Parity Zone on an independent speech-to-text leaderboard. This zone signifies models that m…
-
AI pipeline analyzes interviews in real-time using Grok and Deepgram
This article details the architecture of TrueVoice HQ, an AI platform designed for real-time interview analysis. The system processes audio and video streams using LiveKit and Deepgram for transcription, with Supabase E…
-
AssemblyAI and LiveKit simplify voice agent development
AssemblyAI and LiveKit have collaborated to simplify the creation of voice agents. The first approach integrates AssemblyAI's Voice Agent API with LiveKit's WebRTC capabilities, handling the entire AI pipeline—speech-to…
-
Voice AI startup Rime raises $24M Series A for enterprise call handling
Rime, a voice AI startup, has secured $24 million in Series A funding to enhance its capabilities in handling enterprise customer calls. The company differentiates itself by collecting its own conversational data to red…
-
Build In-House HIPAA Compliant Voice Agent with LiveKit
This article provides a guide on building a HIPAA-compliant voice agent that operates entirely in-house, ensuring that Protected Health Information (PHI) remains within the user's control. It details the process of sett…
-
AssemblyAI offers framework-free voice agent architecture
AssemblyAI has introduced a new framework-free architecture for building voice agents, challenging the necessity of tools like Pipecat and LiveKit. Their approach consolidates speech-to-text, LLM, and text-to-speech fun…
-
Google launches Gemini 3.5 Live Translate for real-time voice translation
Google has launched Gemini 3.5 Live Translate, an advanced audio model designed for real-time speech-to-speech translation. This new model supports over 70 languages and can translate continuously, preserving the speake…
-
AethexAI raises $3M for voice AI in Africa and Middle East
AethexAI, a startup focused on developing voice AI for overlooked markets in Africa and the Middle East, has secured $3 million in pre-seed funding. The company differentiates itself by building its own small language m…
-
Glad Labs enhances MCP platform with improved error handling and testing
Glad Labs has significantly improved its MCP platform by addressing silent failures and enhancing observability. Key updates include fixing the voice bridge to fail loudly rather than silently, re-enabling previously sk…
-
xAI releases Grok's Text to Speech API with natural, low-latency voices
xAI has launched its Grok Text to Speech API, enabling developers to integrate natural and expressive voice capabilities into their applications. The API supports low-latency streaming and is multilingual, offering over…
-
Google DeepMind launches Gemini 3.1 Flash TTS, Live, and Lite models
Google DeepMind has unveiled a suite of Gemini 3.1 Flash models, including Flash TTS for advanced text-to-speech, Flash Live for real-time dialogue, and Flash-Lite for cost-efficient, high-volume workloads. These models…