Voice AI
Speech recognition, synthesis and the voice agents built on them
Latest stories See all →
- 7 Healthcare Call Center Challenges (and How Voice AI Solves Them)
- Customer Voice Puts Kissflow at #1 on Gartner Peer Insights ahead of Salesforce, Appian and Outsystems
- OpenAI’s voice model doesn’t think. That’s the point.
- Build an AI voice agent for customer support that can look up orders
- Best AI Voice Agent for Solar Installers (2026)
- Big Update from the AI World!
- Build real-time voice applications with Gemini 3.8 Live and 3.5 Transcribe
- Best Tools & Use Cases For Real-Time Translation in 2026
- Teaching Old Apps New Tricks with Advanced Conversational AI
- Google Translate vs. DeepL (2026): Accuracy, Pricing, and Which to Use
- Conversational AI in Vault CRM
- It’s the speech synthesiser you wanted for your Apple II
- IVR vs. AI Phone Agent
- DeepL Voice now preserves your voice in real time across languages
- AWS and Salesforce Put CRM Data, AI Agents, and Model Choice Into the Tools Teams Use Every Day
- Getting voice infrastructure right when deploying voice AI for CX
- What it takes to zoom into a two-hour recording without losing your place
- AI legal transcription: 5 benefits of using AI to transform modern legal practice
- Restaurant Phone Call Automation: How AI Voice Assistants Crush the Lunch Rush
- The Cognitive Tax of Legacy Radiology Dictation
- Building voice agents for the real world: 7 takeaways from our NYC meetup with LiveKit
- The ANZ Legal AI Moment: What's Real, What's Required, What's Next
- Best AI voice agents for healthcare (2026)
- How does an AI medical receptionist run your front desk phones?
Videos
- Create AI Ads in Claude with ElevenLabs (Full Tutorial)
- From Mars Rovers to Voice AI: How AI Learns Hidden Meaning
- Why New AI Models Won't Fix Your Voice Agent | Carter Huffman, Modulate
- This Open-Weights TTS Beat ElevenLabs... Then I Read the License
- How To Create A Staging Environment for Your Voice Agent
- The Faster Way to Build Voice AI for Twilio and WhatsApp
- Breeze TTS 2 - Built for Real-Time Voice Design: Run Locally
- Install Pipecat PhoneLLM Locally for Free Voice AI Agent
- Train Your Own CPU TTS Model Locally in Any Language and Any Voice
- AI Call Center from Future: Building a FIRE VOICE AGENT: Retell AI
- We Put HBO's Euphoria Through the World's Most Powerful Voice AI 👀
- S1-Mini: 0.6B Model Fixes Messy Speech-to-Text Locally
Podcasts
Recent episodes
- Universal Music Group is collaborating with ElevenLabs on a new AI-powered creation platform
- Speech Recognition Is Not a Solved Problem — Pavan Muddireddy
- Anthropic caught scientists using Claude to further bio weapon research, CA has new laws on the use of social media and AI chatbots by kids, and UMG is collaborating with ElevenLabs
- SE Radio 737: Owen McGirr on Software Accessibility
- SaaStr 877 CRO Confidential: 0 to $600M in Under 4 Years. The ElevenLabs GTM Playbook with Carles Reina
- TikTok introduces voice comments and simplified polls
- GoPro says it's moving into AI data centers, BT's old copper network could be worth over $2 billion, and Meta's new AI transcription model can distinguish between multiple speakers and languages in real-time
- iOS 817: Digital Backpack – Best Apps for Back to School - iOS Tools for Students & Lifelong Learners
- OpenAI detailed the failures that led to its Hugging Face breach, Meta apparently abandoned an AI-focused restructuring plan, and Google’s latest Gemini transcription model can turn your ramblings into structured text
- 115. Data Collective with E.M. Lewis-Jong
- Threads is testing a podcast transcription feature, ESPN is raising prices on most of its streaming subscriptions, and Zillow settled its antitrust lawsuit with the FTC
- We ask Gemini and Alexa to track cats and give advice