Happy 6th Anniversary to Two Voice Devs! In this milestone episode, Mark Tucker and Allen Firstenberg look back at six years of podcasting and discuss how the industry is coming full circle. We started in the era of hardware assistants like Alexa and Google Assistant, shifted into text-based LLMs, and are now witnessing the return of voice-first interfaces through smart glasses and conversational LLM voice modes.
As developers rush to build the next generation of AI agents, are we repeating the painful mistakes of the past? Mark and Allen share six critical lessons today's agent developers must learn, covering the abysmally poor developer-to-consumer discovery experience, the puzzle of monetization for independent creators, the true meaning of "voice first, not voice only," the absolute necessity of concise responses, the lost art of crafted entertainment over just-in-time generation, and why we need asynchronous interactions modeled after the Star Trek computer.
If you're thinking about the next wave of agents in all sorts of form factors and modalities, this is the episode for you to watch. And we'd love to hear your take on what we've learned and what we still need to learn.
[00:00:00] Celebrating six years of Two Voice Devs!
[00:01:00] The full circle return of voice-first LLM interfaces
[00:03:59] Lesson 1: The discovery and installation bottleneck
[00:08:59] Lesson 2: The monetization puzzle for indie developers
[00:10:48] Platform plays: Android's advantage vs. Amazon's closed beta
[00:16:50] Lesson 3: Designing for "voice first, not voice only"
[00:18:45] Lesson 4: LLMs are too verbose—managing output conciseness
[00:20:09] Lesson 5: Tailored entertainment and crafted storytelling
[00:23:13] Lesson 6: Latency, response times, and the Star Trek computer
[00:24:42] Outro and looking ahead to another year
#VoiceFirst #AIAgents #GenerativeAI #SmartGlasses #IntelligentEyewear #VoiceUX #AmazonAlexa #GoogleAssistant #Gemini #AndroidXR #OpenAI
Episode 278