A developer shared their work getting live transcription previews working with a batch ASR model after receiving subscriber feedback. They tested streaming models but found them lacking in punctuation accuracy needed for speaker attribution, so they optimized their batch model approach instead, achieving 15.5% word error rate and near-real-time output.