LiveKit
PulseAugur coverage of LiveKit — every cluster mentioning LiveKit across labs, papers, and developer communities, ranked by signal.
- partners with Universal-3.5 Pro Realtime 90%
- partners with AssemblyAI 80%
- partners with Pipecat 70%
- used by Deepgram 70%
- competes with Pipecat 70%
- uses Deepgram 70%
- used by Universal-3.5 Pro Realtime 70%
- uses AssemblyAI 60%
- uses Universal-3.5 Pro Realtime 60%
- competes with Vapi Ai Voice Startup 60%
- competes with Twilio 60%
- used by Pipecat 50%
4 day(s) with sentiment data
-
Google launches Gemini 3.8 Live and Extended Thinking audio models
Google has launched Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking, new advanced audio models designed to enhance voice agent capabilities and create more natural AI conversations. Gemini 3.8 Live focuses on scal…
-
AssemblyAI guides voice agent creation with Twilio and advanced streaming models
AssemblyAI has released a guide detailing how to build voice agents using their Universal-3 Pro Streaming and Universal-3.5 Pro Realtime models, in conjunction with Twilio's communication platform. The guide offers two …
-
AssemblyAI, LiveKit Discuss Real-World Voice Agent Challenges
AssemblyAI and LiveKit recently hosted a meetup in New York focusing on the challenges of building real-world voice agents. Panelists from Boardy, a founder networking platform, and Flagler Health, an AI platform for mu…
-
AssemblyAI launches Voice Agent API with flexible integration options
AssemblyAI has introduced its Voice Agent API, offering two integration lanes for voice AI stacks. The first lane, Voice Agent API, is a fully managed service handling speech-to-text, LLM routing, and text-to-speech. Th…
-
Techpotions cuts LLM API costs by 60% with AI agent architecture changes
Techpotions successfully reduced the LLM API costs for their AI Calling Agent by 60% through architectural optimizations rather than just model selection. Key strategies included implementing model routing to use cheape…
-
Meta AI launches single real-time model for voice tasks
Meta AI has introduced Muse Voice Transcribe, a novel real-time audio perception model designed to handle speech recognition, speaker diarization, and endpointing within a single system. This model, which ranks highly o…
-
AssemblyAI guides building real-time voice agents with STT-LLM-TTS architecture
AssemblyAI is detailing how to build real-time voice agents using a chained architecture that connects speech-to-text (STT), large language models (LLMs), and text-to-speech (TTS) components. The company emphasizes the …
-
AI voice tutor Bhasha Academy built with Gemini and Murf Falcon
A developer has created Bhasha Academy, an AI-powered voice tutor designed to help students in India practice English and basic math. The application supports Hinglish conversations, offers personalized learning with pe…
-
AI voice tech advances from research to real-world applications
Recent advancements in AI voice technology are highlighted across several domains, from academic research on voice cloning and detection to practical applications in contact centers and personalized learning companions.…
-
AssemblyAI enhances voice agent transcription with context carryover
AssemblyAI has introduced a new feature called Agent Context Carryover, designed to improve the accuracy of voice agent transcriptions. This feature allows the speech-to-text model to retain awareness of the ongoing con…
-
Fish Audio raises $50M seed for AI voice models, serving 8M users
Fish Audio, a startup specializing in AI voice models, has secured $50 million in seed funding. The company offers both open-source and hosted solutions, catering to creators and enterprises with over 15,000 natural lan…
-
AssemblyAI's Universal-3.5 Pro Realtime hits human parity in speed and accuracy
AssemblyAI's Universal-3.5 Pro Realtime model has achieved a significant milestone by being the sole entry within Coval's Human Parity Zone on an independent speech-to-text leaderboard. This zone signifies models that m…
-
AI pipeline analyzes interviews in real-time using Grok and Deepgram
This article details the architecture of TrueVoice HQ, an AI platform designed for real-time interview analysis. The system processes audio and video streams using LiveKit and Deepgram for transcription, with Supabase E…
-
AssemblyAI and LiveKit simplify voice agent development
AssemblyAI and LiveKit have collaborated to simplify the creation of voice agents. The first approach integrates AssemblyAI's Voice Agent API with LiveKit's WebRTC capabilities, handling the entire AI pipeline—speech-to…
-
Voice AI startup Rime raises $24M Series A for enterprise call handling
Rime, a voice AI startup, has secured $24 million in Series A funding to enhance its capabilities in handling enterprise customer calls. The company differentiates itself by collecting its own conversational data to red…
-
Build In-House HIPAA Compliant Voice Agent with LiveKit
This article provides a guide on building a HIPAA-compliant voice agent that operates entirely in-house, ensuring that Protected Health Information (PHI) remains within the user's control. It details the process of sett…
-
AssemblyAI offers framework-free voice agent architecture
AssemblyAI has introduced a new framework-free architecture for building voice agents, challenging the necessity of tools like Pipecat and LiveKit. Their approach consolidates speech-to-text, LLM, and text-to-speech fun…
-
Google launches Gemini 3.5 Live Translate for real-time voice translation
Google has launched Gemini 3.5 Live Translate, an advanced audio model designed for real-time speech-to-speech translation. This new model supports over 70 languages and can translate continuously, preserving the speake…
-
AethexAI raises $3M for voice AI in Africa and Middle East
AethexAI, a startup focused on developing voice AI for overlooked markets in Africa and the Middle East, has secured $3 million in pre-seed funding. The company differentiates itself by building its own small language m…
-
Glad Labs enhances MCP platform with improved error handling and testing
Glad Labs has significantly improved its MCP platform by addressing silent failures and enhancing observability. Key updates include fixing the voice bridge to fail loudly rather than silently, re-enabling previously sk…