-
#7831 该模型始终思考,不支持关闭思考;请使用 low、high 或 max。https://open.bigmodel.cn/api/paas/v4/...
💬 5[Last reply:43 minutes ago] 核心操作就是在【智谱AI渠道的设置窗口里,把模型名从 glm-5.3 改成其他模型名】。步骤如下:打开软件顶部菜单:菜单 → 翻译设置 → 智谱AI(若是语音识别/配音用到智谱,则去对应的识别设置或TTS设置里找智谱AI)。在弹出的设置窗口中,找到 模型名称(model / 模型) 输入框或下拉框。把其中的 glm-5.3 删除,改填一个非强制思考的模型名。智谱常见的可填写模型名例如:glm-4-...2401**469 Posted at: 48 minutes ago -
#7830 该模型始终思考,不支持关闭思考;请使用 low、high 或 max。https://open.bigmodel.cn/api/paas/v4/...
💬 1[Last reply:50 minutes ago] 该模型(glm-5.3)为强制思考模型,必须在思考模式(thinking)下使用,且需指定思考强度参数(low/high/max),不支持关闭思考。软件中默认将思考关闭,因此报 400 错误。请按以下方式处理:打开 菜单 → 翻译设置 → 智谱AI(或你当前使用的 OpenAI兼容/智谱渠道设置页面)。在设置面板中勾选 Thinking(启用思考)。若面板中有思考强度选项,选择 low、high ...2401**469 Posted at: 50 minutes ago -
#7829 最后一步字幕配音嵌入时出错submitting a packet to the muxer: Invalid argument submitting a packet to the muxer: Invalid argument muxing a packetTraceback (most recent call last):...
💬 1[Last reply:1 hour ago] 底层原因:原视频(novoice.mp4,即无声视频流)本身的视频编码与封装容器不兼容,导致无法直接 -c:v copy(流复制)打包进 mp4 容器。此现象常见于原视频为 VP9/AV1 等编码却强行封装 mp4,或是视频流参数异常(如时间戳/帧率错乱)。此外容器带空格路径、软件版本较旧也可能是诱因。解决方案(按顺序尝试):将待翻译视频移动到浅层纯英文数字目录(例如 D:\video\2ce.m...2407**101c Posted at: 1 hour ago -
#7828 :{'text': '你好啊,我亲爱的朋友,希望你的每一天都是美好愉快的!', 'rate': '+0%', 'role': 'coral', 'filename': 'C:\\Users\\Administrator\\AppData\\Local\\Temp/pyvideotrans/22-1791190822.5382059.wav', 'tts_ty...
💬 1[Last reply:10 hours ago] 报错核心原因:tts_type: 22 对应的是 OpenAI-TTS 配音渠道,但你未在软件中配置 OpenAI 的 API Key,导致初始化 OpenAI 客户端时抛出 Missing credentials 错误。解决方案:配置密钥:点击顶部菜单 翻译设置(Translation Settings) 或 TTS设置,找到 OpenAI / OpenAI-TTS 设置项,填入你的 api_k...38.148**9 Posted at: 10 hours ago -
#7827 :{'text': '你好啊,我亲爱的朋友,希望你的每一天都是美好愉快的!', 'role': 'alloy', 'filename': 'E:/yinwenzhuan/win-pyvideotrans-v4.15/tmp/26484/1791190756.7714622-openai.wav', 'tts_type': 22}...
💬 1[Last reply:10 hours ago] 核心原因:未配置 OpenAI 的 API Key,导致 OpenAI-TTS 渠道在初始化客户端时因缺少凭据报错(Missing credentials)。解决方案:点击软件顶部菜单 翻译设置(或 TTS设置)→ 找到 OpenAI-TTS / OpenAI 渠道设置项。在 API Key(密钥/SK)输入框中填入你从 OpenAI 官方(https://platform.openai.com/...38.148**9 Posted at: 10 hours ago -
#7826 [FishTTS] 该渠道需另行部署并启动后可用:...
💬 3[Last reply:10 hours ago] pyVideoTrans 的知识库中不含 FishTTS 项目本身的部署与启动细节(它属于第三方开源项目,需你自行从其官方仓库获取)。FishTTS 属于本地API渠道,pyVideoTrans 只负责向其 API 地址发送请求,无法替你启动该服务。因此如何手动启动需参照 FishTTS 官方仓库的说明。请按以下方向操作:前往 FishTTS 官方开源仓库(GitHub 上搜索 fish-spee...38.148**9 Posted at: 10 hours ago -
#7825 firered3tts 该渠道需另行部署并启动后可用self.api_url='http://127.0.0.1:7860'...
💬 1[Last reply:10 hours ago] 核心原因:FireRedTTS3 属于「本地 API」渠道,必须你先在本地自行部署并启动其 Gradio WebUI 服务(监听 7860 端口),软件才能连接调用。当前报错 [WinError 10061] 由于目标计算机积极拒绝,无法连接 表示 http://127.0.0.1:7860 这个地址上根本没有服务在运行。解决方案:确认你已按照 FireRedTTS3 官方文档,在本地成功部署并启...38.148**9 Posted at: 10 hours ago -
#7824 请先选择目标语言=system:Windows-10-10.0.19045-SP0version:v4.15frozen:Truelanguage:zh_CNroot_dir:E:/yinwenzhuan/win-pyvideotrans-......
💬 1[Last reply:10 hours ago] 该提示意味着软件检测到「目标语言(Target Language)」下拉框为空,未选中任何值,因此任务无法启动。此提示本身不是报错,而是操作阻断。解决方案:在主界面第 3 行 翻译字幕(TranslateSrt) 区域,找到 目标语言(Target Language) 下拉框。展开下拉列表,选择一个具体的语言(例如:中文 zh-cn、英语 en 等),不要留空。同时确认同一行的 发音语言(Spok...38.148**9 Posted at: 10 hours ago -
#7823 runffmpeg() got an unexpected keyword argument 'state_dict'...
💬 21[Last reply:12 hours ago] 结论明确:勾选「配音加速」只修复了时长对齐,但没有修复视频画面本身——画面依然定格。这说明问题不在音频、不在对齐、不在最终编码,而在更前端的「视频读取/切片」环节:视频轨根本没有拿到完整的 16 分钟画面。核心证据锁定:你的多份日志中,self.raw_total_time 初始值稳定为 33033ms(约 33 秒),而原视频是 16 分 22 秒。这个 33 秒就是病灶——软件从原视频里只读出...107.172**1 Posted at: 13 hours ago -
#7822 API URL is error: :{'text': '你好啊,我亲爱的朋友,希望你的每一天都是美好愉快的!', 'role': 'zh_male_bj.wav', 'filename': 'D:/win-pyvideotrans-v4.15/tmp/28356/1791171818.0280275-12.wav', 'tts_type': 12}...
💬 1[Last reply:15 hours ago] This is because the Index-TTS local API service is not running or the API address is not filled in / is incorrect, causing the TTS channel to fail when connecting to the API URL (the error message sho...222.252**9 Posted at: 15 hours ago -
#7821 FunASRNano is not registered:Traceback (most recent call last): File...
💬 1[Last reply:15 hours ago] 核心原因是:阿里FunASR(内置) 渠道中的 Fun-ASR-Nano-2512 模型未在本地 FunASR 库中正确注册,通常由模型文件下载不完整/缺失或模型名称与版本不匹配导致。同时指出两个需注意的配置问题:你使用的是 翻译视频 模式,并将 发音语言 设为 自动检测。FunASR(内置) 渠道需要明确指定发音语言,不支持自动检测(自动检测仅适用于 faster-whisper / opena...101.71**3 Posted at: 15 hours ago -
#7820 LLM纠错每批字幕行数: LLM大模型纠错时,每次发送多少条字幕,该值越大断句效果越好,一次性发送全部字幕最佳,但受限于最大输出token和上下文(max_ token),过长输入可能导致超出AI限制而失败,默认20条字幕LLM纠错所用AI......
💬 1[Last reply:1 day ago] The token / max_token is not filled in the "Advanced Options" page. The Advanced Options only lets you choose the batch size (LLM Correction errors How many subtitles are sent each time) and the AI pr...2001**7c90 Posted at: 1 day ago -
#7819 :Traceback (most recent call last): File...
💬 1[Last reply:1 day ago] The core cause is that your OpenAI (compatible) translation channel has no API key configured — the error Missing credentials comes from the OpenAI SDK finding an empty key.Solution:Get an API key fro...194.124**4 Posted at: 1 day ago -
#7818 翻译字幕阶段出错 [OpenAI ChatGPT] The origin web server did not return a complete response within the 120-second Proxy Read Timeout window. The connection was established, but the origin t...
💬 1[Last reply:1 day ago] 错误码 524 是代理网关超时,说明翻译请求已发出,但 AI 服务端在 120 秒内未返回完整响应。核心原因通常是:单次发送的字幕行数过多、或模型为思考类模型导致耗时过长、或所选第三方中转 API 服务缓慢/不稳定。解决方案(按顺序尝试):减小单批发送行数菜单 工具/选项 → 高级选项 → 字幕翻译,将 AI翻译渠道每批字幕行数 改为 10 或更小。取消「AI翻译附带所有字幕做参考」该选项会让每次...125.33**3 Posted at: 1 day ago -
#7817 Network connection failed:Traceback (most recent call last): File...
💬 1[Last reply:1 day ago] This is a network connection failure, not a software bug. The error [WinError 10053] An established connection was aborted by the software in your host machine means the connection to the Gemini API s...194.124**4 Posted at: 1 day ago -
#7816 语音识别阶段出错[faster-whisper(内置)] [ONNXRuntimeError] : 6 : RUNTIME_EXCEPTION : Non-zero status code returned while running Conv node. Name:'/encoder/feature_extractor/Conv' Status Messa...
💬 1[Last reply:1 day ago] bad allocation 表示内存分配失败。这是 VAD 阶段(vad.py)在 ONNX 运行 Conv 节点时,因内存不足导致的崩溃,与 CUDA 加速无直接关系,VAD 是固定使用 CPU/ONNX 运行的。另外从日志看,您的视频信息中 width: 0, height: 0、video_streams: 0,这是纯音频 mp4,缺少视频流。此异常元数据也可能引发预处理阶段的资源分配异常...2409**ed93 Posted at: 1 day ago -
#7814 这个软件自从411之后就很难调了.字幕识别缺斤短两.一条字幕横跨多个时间点.怎么调都调不好有没有最佳调配方案...
💬 7[Last reply:1 day ago] 菜单 工具 高级选项 语音识别参数,这三个参数,这样调整试试最长语音持续秒:设为5最短语音持续(ms):设为 2000静音分割持续(ms):设为140语音识别使用 faster-whisper(内置),模型使用 large-v3,如果显存不足,就使用 large-v3-turbo...120.217**1 Posted at: 1 day ago -
#7815 声音克隆真的有用吗,为什么选择克隆,最后生成的视频是没有声音的...
💬 1[Last reply:1 day ago] 核心原因:clone(声音克隆)角色依赖从原始视频对应时间段动态裁切出 3-10 秒参考音频,若切出的参考音频不合规(过短、有背景噪声、语句被硬切断),会导致克隆失败、生成静音或无声输出。解决方案:先确认克隆本身有效性:clone 仅在“翻译视频”功能中可用(它需要原始视频音轨作参考),在单独的“文字配音”功能中无法使用,误用会报错或无声。拉长单条字幕时长(克隆硬性要求参考音频 3-10 秒):菜...2404**464c Posted at: 1 day ago -
#7812 in pyvideotrans can i give the translated script i mean can i manually upload the translated version of the script and t......
💬 1[Last reply:1 day ago] Yes. You can import your own pre-translated SRT, and pyVideoTrans will skip the ASR and/or translation stage and use your file directly.How it works (default output-location rule matters):Suppose the ...152.59**7 Posted at: 1 day ago -
#7811 语音识别阶段出错[faster-whisper(内置)] mkl_malloc: failed to allocate memory:Traceback (most recent call last): File...
💬 1[Last reply:1 day ago] 核心原因:mkl_malloc: failed to allocate memory 是内存分配失败,即物理内存/虚拟内存不足导致的,与模型本身无关。从日志看:你未启用CUDA加速,使用的是 faster-whisper large-v3-turbo 模型,并启用了降噪。纯CPU运行 large-v3-turbo 模型需占用大量内存(通常8GB以上),叠加降噪处理,很容易耗尽内存。解决方案(按优先...2602**b8 Posted at: 1 day ago
Open source and free maintenance is not easy. If this project is helpful to you, please consider making a small donation to help the project continue to maintain and update.