The TXT file contains no text
No speech was detected. Confirm the media has a usable spoken track rather than silence or music only.
Create a plain-text transcript from a compatible downloaded video on your computer. Desktop detects speech locally and writes one UTF-8 text line per recognized segment.
Desktop prepares 16 kHz audio with FFmpeg, detects speech, then recognizes one segment at a time with the local model.
Each step maps to a control or task state in the Desktop application shown above.
Enable TXT + SRT before starting a compatible video download.
On first use, confirm setup and wait until Settings reports Enabled or Ready.
Desktop downloads the video, prepares audio and processes transcription serially.
Find the matching .txt and .srt files beside the downloaded media.
Failures remain attached to their task so completed items in the same queue stay available.
No speech was detected. Confirm the media has a usable spoken track rather than silence or music only.
Clearer speech and less background noise help. Review names, technical terms and overlapping speakers manually.
Check that the download folder still exists and is writable, then retry.
Choose the current package for your computer. Release cards appear only when that package is available.
Free: 10 successful single-video parses per day; no profile extraction. Desktop Pro: $9.90/month or $79.90/year for unlimited parsing and supported profile extraction. Compare plans.
No. It is a UTF-8 plain-text file with recognized speech segments in playback order.
No. Use the matching SRT output when you need start and end times.
Yes for compatible tasks; Free still has 10 successful single-video parses per day and no profile extraction.