Core reason: In v4.15 the sentence-splitting (VAD re-segmentation) stage was rewritten, and there is a known v4.15 bug where the re-split of subtitles does not strictly respect your 最长语音持续 / max speech duration and 最短语音持续 / min speech duration values — so subtitles come out far longer than the numbers you set. A corrected v4.15 patch build is already published for this, and the patch must be re-downloaded and overwritten even if you already have the v4.15 full package.
Do the following, in order:
1. Check for a no_resegment.txt file in the software root folder (same folder as sp.exe)
If that file exists, it disables re-splitting by word-level timestamps for faster-whisper / openai-whisper, and the app then uses the model's own long sentences directly — which cannot follow your max/min duration settings. Delete it (or rename it).
2. Re-download the latest v4.15 patch package and overwrite it into the v4.15 directory again
Overwrite all files and folders of the patch (not only sp.exe), especially sp.exe, videotrans/ and _internal/. If after overwriting you still get behaviour from the old code, search for every __pycache__ folder inside _internal/, delete them, then overwrite the patch again.
3. Verify your parameter combination is legal
最长语音持续(秒) × 1000 must be greater than 最短语音持续(毫秒). E.g. max = 2 s (2000 ms) with min = 3000 ms creates a conflict and unpredictable splitting.- The engine allows roughly a 1.5× buffer, so max = 5 s can still produce ~7.5–8 s subtitles. To force short lines, set max to 2–3 s.
- Lower
静音分割持续(ms) (e.g. 300) so it cuts on shorter pauses.
4. Re-run with 清理已生成 / Clear Output checked
Otherwise the previously generated SRT in the output folder is reused and it looks like your changes had no effect.
5. Note on the recognition channel
Fine-grained splitting depends on word-level timestamps; the whisper-family channels (faster-whisper, openai-whisper) give the finest control. Some other channels do not support this level of splitting.
If after the patch overwrite the sentences are still extremely long, please send the bottom ~30 lines of the newest log file in the software's logs/ folder, plus your exact values for 最长语音持续/最短语音持续/静音分割持续 and the STT channel used.
请查阅相关文档: