Category: Audio

  • You Ask, I Answer: Preventing Audio Timestamp Hallucinations?

    Summary In today's episode, I address the issue of timestamp hallucinations when local speech-to-text agents chunk audio files. Here's what this means for you. You gain the ability to produce highly accurate transcriptions with precise time codes. You'll also learn these concepts: why switching to SRT or VTT formats improves timing precision, how Python scripts…

    Continue reading →

  • So What? Using Generative AI for Voice Generation

    Summary In today's episode, I walk through how to use generative AI for voice generation, from selecting the right tool to preparing text that sounds natural when read aloud by a machine. Here's what this means for you. You'll discover a practical framework for deciding when AI voice makes sense for your business and when…

    Continue reading →

  • Mind Readings: Unsolicited Review of the Neewer CM31 Wireless Mic

    In this episode, we put mobile video microphones to the test. You’ll see a direct comparison of four different audio setups for iPhone video. You’ll evaluate how popular microphones perform in real-world conditions. You’ll uncover surprising issues and discover useful features you might miss. You’ll learn which budget wireless mic delivers reliable sound for your…

    Continue reading →

  • Mind Readings: Making a Podcast with Generative AI, Part 5

    Summary In today's episode, I walk through a troubleshooting workaround for running an AI-driven podcast interview when you don't know audio gear well, splitting the recording across your normal studio setup and a smartphone. Here's what this means for you. You'll pick up a practical, low-barrier workflow that trades setup complexity for a little post-production…

    Continue reading →

  • Mind Readings: Making a Podcast with Generative AI, Part 4

    Summary In today's episode, I walk through the post-production process for a podcast created with generative AI, covering audio leveling, compression, and transcript-based editing in Adobe Premiere. Here's what this means for you. You gain a practical workflow that turns raw AI interviews into polished, professional-sounding episodes while preserving your unique human voice. You'll also…

    Continue reading →

  • Mind Readings: Making a Podcast with Generative AI, Part 1

    Summary In today's episode, I walk through how to set up an AI-generated podcast interview using ChatGPT's advanced audio mode instead of relying on Google's Notebook LM. Here's what this means for you. You'll be able to create interactive podcast-style content where you control the conversation and inject your own voice rather than settling for…

    Continue reading →

  • Mind Readings: Turning a Lavalier Mic Into a Handheld Mic

    In today’s episode, you’ll see a simple hack to transform a lavalier microphone into a handheld microphone. I’ll walk you through how I used a Rode Wireless Go transmitter, a power bank, and a USB-C connector to create a more ergonomic and acoustically sound setup. You’ll learn why this method, while not ideal for a…

    Continue reading →

  • Mind Readings: Turning a Lavalier Mic Into a Handheld Mic

    Summary In today's episode, I walk through a cheap DIY hack for using a lavalier microphone handheld-style by rigging it to a no-name power bank with a USB-C cable, as an alternative to Rode's $29 plastic holder. Here's what this means for you. You get a budget-friendly, travel-ready setup that improves your audio quality and…

    Continue reading →

  • Mind Readings: Why I Hired a Human Musician Instead of AI

    Summary In today's episode, I explain why I hired a human composer for my theme music instead of using artificial intelligence. Here's what this means for you. You will understand how to navigate the creative and legal differences between human and machine authorship. You'll also learn these concepts: how humans handle nuanced thematic instructions better…

    Continue reading →

  • Mind Readings: Why AI Struggles With Sarcasm

    Summary In today's episode, I break down why today's text-based generative AI struggles with sarcasm and other tone-dependent language, even when the underlying words are statistically predictable. Here's what this means for you. When your AI outputs miss the mark on nuance, the culprit is usually the missing audio and visual cues that humans use…

    Continue reading →