Lessons
From the last month
ASR word timestamps share an edge ~90% of the time, so the best take is the hardest one to cut
python asr deepgram video-editing word-timestamps 1.2k tokens
Decide ffmpeg speech denoising by measuring voiced-frame LSD and LF-transient p95, not by SNR alone
ffmpeg audio denoise afftdn speech 644 tokens