#local-inference
4 episodes
#5456: Parakeet vs Whisper: Picking a Phone ASR Model
Why Whisper loses on Android, why Parakeet v2 beat v3, and how to benchmark speech-to-text without any tooling.
#4966: Local vs Cloud: Running Hugging Face Models
Hugging Face's compatibility tracker, cache management, and the real difference between Inference Endpoints and the direct API.
#1540: Why Gnome 50 is Breaking Your Voice-to-Text Tools
Explore the engineering battle to bring low-latency AI voice input to Linux while navigating the strict security of Wayland and GNOME 50.
#857: The Cognitive Cost of Capitalization
Can local AI fix your messy typing in real-time? Explore the tech behind "transparent buffers" that turn sloppy drafts into polished prose.