-
#7820 LLM纠错每批字幕行数: LLM大模型纠错时,每次发送多少条字幕,该值越大断句效果越好,一次性发送全部字幕最佳,但受限于最大输出token和上下文(max_ token),过长输入可能导致超出AI限制而失败,默认20条字幕LLM纠错所用AI......
💬 1[Last reply:1 hour ago] The token / max_token is not filled in the "Advanced Options" page. The Advanced Options only lets you choose the batch size (LLM Correction errors How many subtitles are sent each time) and the AI pr...2001**7c90 Posted at: 1 hour ago -
#7819 :Traceback (most recent call last): File...
💬 1[Last reply:1 hour ago] The core cause is that your OpenAI (compatible) translation channel has no API key configured — the error Missing credentials comes from the OpenAI SDK finding an empty key.Solution:Get an API key fro...194.124**4 Posted at: 1 hour ago -
#7818 翻译字幕阶段出错 [OpenAI ChatGPT] The origin web server did not return a complete response within the 120-second Proxy Read Timeout window. The connection was established, but the origin t...
💬 1[Last reply:3 hours ago] 错误码 524 是代理网关超时,说明翻译请求已发出,但 AI 服务端在 120 秒内未返回完整响应。核心原因通常是:单次发送的字幕行数过多、或模型为思考类模型导致耗时过长、或所选第三方中转 API 服务缓慢/不稳定。解决方案(按顺序尝试):减小单批发送行数菜单 工具/选项 → 高级选项 → 字幕翻译,将 AI翻译渠道每批字幕行数 改为 10 或更小。取消「AI翻译附带所有字幕做参考」该选项会让每次...125.33**3 Posted at: 3 hours ago -
#7817 Network connection failed:Traceback (most recent call last): File...
💬 1[Last reply:3 hours ago] This is a network connection failure, not a software bug. The error [WinError 10053] An established connection was aborted by the software in your host machine means the connection to the Gemini API s...194.124**4 Posted at: 3 hours ago -
#7816 语音识别阶段出错[faster-whisper(内置)] [ONNXRuntimeError] : 6 : RUNTIME_EXCEPTION : Non-zero status code returned while running Conv node. Name:'/encoder/feature_extractor/Conv' Status Messa...
💬 1[Last reply:4 hours ago] bad allocation 表示内存分配失败。这是 VAD 阶段(vad.py)在 ONNX 运行 Conv 节点时,因内存不足导致的崩溃,与 CUDA 加速无直接关系,VAD 是固定使用 CPU/ONNX 运行的。另外从日志看,您的视频信息中 width: 0, height: 0、video_streams: 0,这是纯音频 mp4,缺少视频流。此异常元数据也可能引发预处理阶段的资源分配异常...2409**ed93 Posted at: 4 hours ago -
#7814 这个软件自从411之后就很难调了.字幕识别缺斤短两.一条字幕横跨多个时间点.怎么调都调不好有没有最佳调配方案...
💬 7[Last reply:5 hours ago] 菜单 工具 高级选项 语音识别参数,这三个参数,这样调整试试最长语音持续秒:设为5最短语音持续(ms):设为 2000静音分割持续(ms):设为140语音识别使用 faster-whisper(内置),模型使用 large-v3,如果显存不足,就使用 large-v3-turbo...120.217**1 Posted at: 7 hours ago -
#7815 声音克隆真的有用吗,为什么选择克隆,最后生成的视频是没有声音的...
💬 1[Last reply:6 hours ago] 核心原因:clone(声音克隆)角色依赖从原始视频对应时间段动态裁切出 3-10 秒参考音频,若切出的参考音频不合规(过短、有背景噪声、语句被硬切断),会导致克隆失败、生成静音或无声输出。解决方案:先确认克隆本身有效性:clone 仅在“翻译视频”功能中可用(它需要原始视频音轨作参考),在单独的“文字配音”功能中无法使用,误用会报错或无声。拉长单条字幕时长(克隆硬性要求参考音频 3-10 秒):菜...2404**464c Posted at: 6 hours ago -
#7812 in pyvideotrans can i give the translated script i mean can i manually upload the translated version of the script and t......
💬 1[Last reply:8 hours ago] Yes. You can import your own pre-translated SRT, and pyVideoTrans will skip the ASR and/or translation stage and use your file directly.How it works (default output-location rule matters):Suppose the ...152.59**7 Posted at: 8 hours ago -
#7811 语音识别阶段出错[faster-whisper(内置)] mkl_malloc: failed to allocate memory:Traceback (most recent call last): File...
💬 1[Last reply:8 hours ago] 核心原因:mkl_malloc: failed to allocate memory 是内存分配失败,即物理内存/虚拟内存不足导致的,与模型本身无关。从日志看:你未启用CUDA加速,使用的是 faster-whisper large-v3-turbo 模型,并启用了降噪。纯CPU运行 large-v3-turbo 模型需占用大量内存(通常8GB以上),叠加降噪处理,很容易耗尽内存。解决方案(按优先...2602**b8 Posted at: 8 hours ago -
#7810 parallel_for failed: cudaErrorInvalidDevice: invalid device ordinal:Traceback (most recent call last): File...
💬 1[Last reply:9 hours ago] cudaErrorInvalidDevice: invalid device ordinal 是 CUDA 设备索引无效导致的——软件把语音识别任务分配到了 GPU 编号 N,但该编号在你的系统中不存在或不可用(多显卡、集显独显混合、驱动/环境变量异常时常见)。解决方案(按顺序尝试)强制指定重型任务运行设备菜单(Menu)-工具/选项(Tools/Options)-高级选项(Advanced Op...240e**1df Posted at: 9 hours ago -
#7808 F5-TTS 软件自动下载的模型,是不是只支持 中英文视频克隆 配音...
💬 11[Last reply:13 hours ago] 对的,这样才正确。你的完整路径是:D:\pyvideotrans413\models\models--hvoss-techfak--F5-TTS-German\里面的文件直接放在这一层,名字完全符合软件要求。同时请确认以下几点:1. F5-TTS 中英文默认模型仍在独立文件夹内,即 D:\pyvideotrans413\models\models--SWivid--F5-TTS\F5TTS_v1_...2603**b0 Posted at: 13 hours ago -
#7809 是否能提取出多音轨视频中的多个音轨...
💬 7[Last reply:13 hours ago] <p>命令行对新手确实不友好。但有两点需要先说明:</p><p><b>1. pyVideoTrans 本身没有「提取多音轨视频中指定音轨」这个功能</b>,知识库中也未收录相关能力,所以无法在软件界面里完成此操作。</p><p><b>2. 如果你只是想用现成工具完成提取</b>,不需要自己敲...2001**3d66 Posted at: 13 hours ago -
#7807 No audio file was generated during the listening test.self=HiggsVoice(uuid=None, proxy_str=None, last_down_time=1791064467.1035028, tts_type=11, len=1, language='fr', play=True, vo...
💬 1[Last reply:19 hours ago] The "No audio file was generated during the listening test" error with the Higgs-audio-v3 (HiggsVoice, tts_type=11) channel typically means the model failed to load or produce output — most commonly b...67.68**9 Posted at: 19 hours ago -
#7806 翻译字幕阶段出错 [兼容AI/本地模型] prediction aborted, token repeat limit reachedTraceback (most recent call last):...
💬 1[Last reply:1 day ago] `这是因为本地小模型(兼容AI/本地模型)指令跟随能力弱,在翻译时陷入死循环重复输出,触发了模型端 token 重复限制而被中止。解决方案:精简提示词:打开 软件目录/videotrans/prompts/srt/localllm.txt(你已勾选发送完整字幕,所以是srt目录下的文件),检查是否有过于复杂的要求,删减或简化,注意 {} 包裹的变量不要动。降低单批字幕行数:进入 菜单 → 工具/选...120.85**3 Posted at: 1 day ago -
#7805 请先选择目标语言=system:Windows-10-10.0.19045-SP0version:v4.14frozen:Truelanguage:zh_CNroot_dir:C:/桌面/win-pyvideotrans-v4.14...
💬 1[Last reply:1 day ago] 核心原因:未选择目标语言(Target Language)。目标语言是翻译和配音的输出语言,不选择无法继续。解决方案:在主界面第3行"翻译字幕(TranslateSrt)"区域,找到目标语言(Target)下拉框,选择你希望翻译成的语言(如"简体中文"、"英语"等)。同时确认发音语言(Spoken)已正确选择,不要选"自动检测"(若为自动检测,需确保语音识别渠道为 faster-whisper /...223.116**1 Posted at: 1 day ago -
#7804 Traceback (most recent call last): File...
💬 1[Last reply:1 day ago] 核心原因是 OpenAI 兼容渠道的 API Key 缺失。你当前使用的翻译渠道 Cheaper Inference 需要 API 密钥,但软件没有读取到,因此抛出 Missing credentials. Please pass an api_key...。解决方法:打开软件顶部菜单 翻译设置,找到 Cheaper Inference 对应设置项。填入从该服务商处获取的 API Key,并保存。...183.192**4 Posted at: 1 day ago -
#7803 选择LLM纠正错误,使用智谱AI渠道,但未填写密钥SK.可在菜单-工具/选项-高级选项-LLM纠错渠道中切换...
💬 1[Last reply:1 day ago] 核心原因:翻译渠道(字幕翻译)与 LLM 纠错(LLM Correction)在软件中是两套独立的配置入口,各自读取不同的密钥字段。你在「翻译设置-智谱AI」中填写的密钥仅对翻译渠道生效,而 LLM 纠错默认使用 DeepSeek 渠道,需单独指定并填写对应密钥,因此运行时仍提示"未填写密钥"。解决方案:打开 菜单(Menu)-工具/选项(Tools/Options)--高级选项(Advanced...240b**f912 Posted at: 1 day ago -
#7802 必须选择视频文件=system:Windows-10-10.0.19045-SP0version:v4.14frozen:Truelanguage:zh_CNroot_dir:C:/桌面/win-pyvideotrans-v4.14...
💬 1[Last reply:1 day ago] 请检查以下两个原因:软件路径包含中文:你的安装目录为 C:/桌面/win-pyvideotrans-v4.14,其中「桌面」为中文路径。软件对所上传视频的处理依赖 ffmpeg 等底层工具,解压路径强烈建议仅包含英文、数字等简单路径(推荐形如 D:\pyVideoTrans),否则易引发加载及处理失败。请将软件移动到非系统盘的纯英文目录下(如 D:\pyVideoTrans),再重新启动 sp.e...223.116**1 Posted at: 1 day ago -
#7801 必须选择视频文件=system:Windows-10-10.0.19045-SP0version:v4.14frozen:Truelanguage:zh_CNroot_dir:C:/桌面/win-pyvideotrans-v4.14...
💬 1[Last reply:1 day ago] 未选择任何待处理的视频/音频文件,或所选文件路径未被软件正确读取。解决步骤:点击左上角「选择音频或视频」按钮,选中至少一个 mp4/mkv/avi/mov/wav/mp3 文件后再点击「开始执行」。检查路径问题:软件安装在 C:/桌面/win-pyvideotrans-v4.14,路径中包含中文「桌面」。请将软件整体移动到纯英文、无空格、无特殊符号的目录下(如 D:\pyvideotrans),重...223.116**1 Posted at: 1 day ago -
#7800 模型下载失败,请点击 【模型下载地址】按钮,手动下载所有文件到:C:/Users/Administrator/AppData/Local/Temp/360zip$Temp/360$0/models/models--mobiuslabsgmbh--faster-whisper-large-v3-turbo...
💬 1[Last reply:1 day ago] 核心问题有两个:一是在360压缩包临时目录(360zip$Temp)内直接运行了 sp.exe,二是模型自动下载失败。关键根源分析从报错路径 C:/Users/istrator/AppData/Local/Temp/360zip$Temp/360$0/ 可以确认:你正在压缩包内直接双击运行 sp.exe,没有先解压。这会导致:所有模型、缓存、临时文件写入到路径中含有 $ 特殊符号的临时目...91.103**5 Posted at: 1 day ago
Open source and free maintenance is not easy. If this project is helpful to you, please consider making a small donation to help the project continue to maintain and update.