Convert audio or video into Markdown with local FFmpeg and OpenAI Whisper, then minimally correct the transcript until it is readable without polishing or rewriting the speaker's expression. Use for MP3, WAV, M4A, MP4, MOV, MKV, WebM, long recordings, Chinese, English, mixed-language speech, ASR typo repair, broken sentence repair, and full-transcript readability correction.