-
#7016 Audio file could not be read as PCM WAV, AIFF/AIFF-C, or Native FLAC; check if file is corrupted or in another formatTraceback (most recent call last):...
💬 1[Last reply:1 month ago] The root cause is the speech_recognition library used by the Google STT API (Free) channel cannot decode the audio file passed to it. This most often happens because the source video filename contains...2402**5081 Posted at: 1 month ago -
#7015 Video file selection required...
💬 1[Last reply:1 month ago] This error occurs because no input video/audio file was actually selected before you clicked "Start Execution." The video translation workflow requires a valid media file to be chosen and registered b...37.114**2 Posted at: 1 month ago -
#7014 语音识别阶段出错[faster-whisper(内置)] No transcription results returned. Please check the original audio/video [info.duration_after_vad=0.0s].info=TranscriptionInfo(language='en', language_...
💬 1[Last reply:1 month ago] 核心原因:duration_after_vad=0.0s 表示语音活动检测(VAD)没有从你这 48 秒音频中分离出任何人类语音段,导致 Whisper 无法产生转录结果。这通常是因为原始视频 没有有效的人声音轨、人声极微弱,或 VAD 参数设置过严,将语音误判为静音。按以下顺序排查:确认原视频本身有声音:用播放器播放该视频(D:/TASTIEST Creamy Chipotle Chicken ...188.253**1 Posted at: 1 month ago -
#7013 No transcription results returned. Please check the original audio/video [info.duration_after_vad=0.0s].info=TranscriptionInfo(language='en', language_probability=1, duration=48.46...
💬 1[Last reply:1 month ago] 该报错的核心原因是:VAD(语音活动检测)将该音频中所有片段判定为“非语音”,导致 duration_after_vad=0.0s,因此没有截取到任何有效语音片段用于识别。这不一定是模型损坏或CUDA问题,而是音频本身未检测到人声或VAD阈值过严。请按以下顺序排查:确认原始视频是否真的包含清晰人声:用播放器试听该 .webm 文件,确认有说话声、且声音音量正常。如果只有背景音乐、纯音效或人声被音乐...188.253**1 Posted at: 1 month ago -
#7012 自动下载模型失败,你可以点击 【查看帮助教程】 了解如何手动下载...
💬 1[Last reply:1 month ago] 这是模型自动下载时网络连接中断导致的,日志显示只接收到约80MB(期望约1.6GB),部分文件下载不完整。解决方法(优先推荐手动下载):删除已下载的不完整模型文件夹:软件目录/models/models--mobiuslabsgmbh--faster-whisper-large-v3-turbo(若文件夹不存在,则删除里面残留的临时文件)手动下载该模型:下载地址:https://huggingfa...192.254**8 Posted at: 1 month ago -
#7011 系统错误 (10054):远程主机强迫关闭了一个现有的连接。:{'text': 'Hello, my dear friend. I hope your every day is beautiful and enjoyable!', 'rate': '+0%', 'role': 'WilliamMultilingual(Male/AU)', 'filename...
💬 1[Last reply:1 month ago] 错误 [WinError 10054] 是由 Edge-TTS 微软免费接口在连接过程中被服务器强制断开导致的,通常是因为短时间内请求过于频繁(触发限流)或网络代理配置不当。解决方案降低并发并增加暂停进入 菜单 → 工具 → 高级选项 → 字幕配音调整,将 EdgeTTS配音渠道配音并发数 设为 1,将 配音后暂停秒 设为 5~10。以低频方式调用,可有效避免被微软限流。检查/关闭代理如果填写了代...192.254**8 Posted at: 1 month ago -
#7010 [12.0099g99gg9994%] 404 Cient Eror: Not Found for urt https://edge.microsoft.com/translate/authfasterwhis HV- Power Dist......
💬 1[Last reply:1 month ago] 这是微软免费翻译接口已变更导致的。微软已关闭旧版免费翻译API,导致 https://edge.microsoft.com/translate/auth 返回 404 Not Found。解决方案(按优先级):更新软件版本:请确保软件已升级到 v4.09 或更高版本。若已是 v4.09,请重新下载最新补丁包覆盖主目录后重试。切换翻译渠道(推荐):微软翻译已不可用,建议改用更好的替代方案:DeepS...2409**e5e6 Posted at: 1 month ago -
#7009 我翻译好的字幕 现在出现了一句话 分成了很短甚至只有一个字 现在咋办...
💬 11[Last reply:1 month ago] 在 2060S 上,beam_size 和 best_of 设为 5 会显著增加解码时的搜索计算量,尤其是 beam search,它会同时探索多个候选路径,消耗大量 GPU 时间。虽然 large-v3-turbo 在模型结构上比 large-v3 更精简、推理更快,但 beam search 的额外开销会部分抵消这种优势,导致两者最终耗时趋近。之前给出的 beam_size=5 是针对“准确度...171.91**7 Posted at: 1 month ago -
#7008 Error during speech recognition stage[faster-whisper(Built-in)] Subprocess Error: [WinError 1455] The paging file is too small for this operation to complete.Traceback (most recent...
💬 1[Last reply:1 month ago] The error [WinError 1455] The paging file is too small for this operation to complete means your Windows virtual memory (pagefile) is insufficient to spawn the subprocess that runs Faster-Whisper. Thi...2402**9e5e Posted at: 1 month ago -
#7007 parallel_for failed: cudaErrorInvalidDevice: invalid device ordinal:Traceback (most recent call last): File...
💬 1[Last reply:1 month ago] 这是 CUDA 设备索引无效导致的错误,常见于显卡驱动/CUDA 环境异常、多卡识别错误或子进程无法正确获取 GPU 序号。解决方案(按顺序尝试)先验证 CUDA 环境 打开命令行(CMD),执行:nvidia-smi nvcc -V确认能正常输出显卡信息且 CUDA 版本 ≥ 12.8。如果 nvidia-smi 报错或看不到显卡,说明驱动未正确安装,请更新 NVIDIA 驱动。关闭多显卡模式...104.28**3 Posted at: 1 month ago -
#7006 语音识别阶段出错[faster-whisper(内置)] Unable to allocate 695. MiB for an array with shape (182129152,) and data type float32:Traceback (most recent call last): File...
💬 1[Last reply:1 month ago] 此报错为系统内存(CPU RAM)分配失败,而非GPU显存不足。Unable to allocate 695. MiB for an array with shape (182129152,) 表示 numpy 在申请约 695MB 内存时失败,根本原因是视频时长过长(Huberman 实验室播客通常1-2小时)+ 当前 large-v3 模型 + 整段音频缓冲导致可用内存不足。解决方法(按优先级...240e**f45 Posted at: 1 month ago -
#7005 ASR Error[XiaoMi] File not found: D:/App/pyVideo/tmp/17332/7bd82eece8/audio_knowii-1788280620.0781882.wavTraceback (most recent call last):...
💬 1[Last reply:1 month ago] The error indicates that the ASR pipeline failed because the temporary WAV file (audio_knowii-1788280620.0781882.wav) was not created in the temp directory before the XiaoMi ASR channel tried to slice...139.144**2 Posted at: 1 month ago -
#7004 ASR Error[XiaoMi] Request failedhttps://api.xiaomimimo.com/v1...
💬 1[Last reply:1 month ago] The core cause is that the Xiaomi MiMo ASR API (api.xiaomimimo.com) rejected your request with a 400 Bad Request error. The message Connection prematurely closed BEFORE response indicates the server c...139.144**2 Posted at: 1 month ago -
#7003 翻译字幕阶段出错 [M2M100(内置)] 系统错误 (None):NoneTraceback (most recent call last):...
💬 1[Last reply:1 month ago] 这是 M2M100 本地翻译模型不完整或未正确存放导致的。报错提示缺少 sentencepiece.model 文件,说明 D:/翻译/models/m2m100_12b/ 目录下的模型文件缺失或下载失败。解决方法:先确认 D:/翻译/models/ 下是否存在 m2m100_12b 文件夹。若无,说明模型从未成功下载。若已有该文件夹,请删除整个 m2m100_12b 文件夹,然后手动重新下载完整...119.4**3 Posted at: 1 month ago -
#7002 翻译字幕阶段出错 [阿里百炼API] 需要纯文本模型,但 qwen3.7-max-2026-06-08 可能是多模态模型Traceback (most recent call last):...
💬 1[Last reply:1 month ago] 这是由于在【阿里百炼API】翻译渠道中选择了一个多模态大模型(qwen3.7-max-2026-06-08),而阿里百炼平台限制多模态模型不能用于文本翻译任务,必须使用纯文本模型。解决方案:打开软件顶部菜单 → 翻译设置 → 阿里百炼 API;在“模型名称”下拉框或输入框中,将当前多模态模型改为纯文本大模型,例如:qwen-plusqwen-maxqwen-turboqwen-longqwen3-...2408**15fb Posted at: 1 month ago -
#6990 为什么没找到说明文档中推荐的英伟达开源日语模型...
💬 39[Last reply:1 month ago] 根本原因是取消“发送完整字幕”后,软件默认按“每批字幕行数”逐段发送给 DeepSeek。你现在的“每批字幕行数”极大概率被设为了 1(为了防字幕合并),导致 500 行字幕需要发起约 500 次 API 请求。每次请求都包含网络往返和大模型推理延迟,时间自然成倍增加。请按以下步骤调整,速度会有显著提升:1. 调整翻译批次大小进入 菜单 → 工具 → 高级选项 → 字幕翻译调整,将 “AI翻译渠道...240e**fff0 Posted at: 1 month ago -
#7001 Reference audio is empty after silence removal. Try setting preprocess_prompt=False.,Traceback (most recent call last): File...
💬 1[Last reply:1 month ago] 核心原因:OmniVoice 在克隆音色时,会先从原始视频中截取与字幕时间轴对应的音频片段作为参考音频,并进行静音去除。报错表明截取的参考音频在去除静音后为空,说明该音频片段几乎全是静音(可能落在说话间隙),或该片段过短。解决步骤调整语音识别参数,让字幕片段足够长且非静音。打开 菜单 -> 工具 -> 高级选项 -> 语音识别参数:最长语音持续(秒):设为 6-10最短语音持续(...212.107**4 Posted at: 1 month ago -
#7000 语音识别阶段出错[字节语音大模型极速版] 413 Client Error: Payload Too Large for url: Traceback (most recent call last):File "videotrans\tas......
💬 1[Last reply:1 month ago] 这是字节语音识别大模型极速版接口限制了单次请求的音频体积/时长,导致请求体超过限制(HTTP 413)。你当前使用版本 v4.02 较旧,该渠道的音频切片逻辑可能未适配完整。解决方案(按顺序尝试)升级软件到最新版本 v4.11 旧版对在线语音识别接口的切片处理不完善,新版已优化分片与重试机制。 到 pyVideoTrans 官方下载页 重新下载完整包或补丁包覆盖升级。升级后仍报错时,切换识别渠...155.117**6 Posted at: 1 month ago -
#6999 语音识别阶段出错[faster-whisper(本地内置)] No transcription results returned. Please check the original audio/video or model and try again.Traceback (most recent call last):...
💬 1[Last reply:1 month ago] 该报错表示 faster-whisper 在语音识别阶段没有返回任何转录结果,核心原因是模型未能从音频中识别出有效语音内容,通常由以下原因之一导致:视频无音轨/无人声、原音频音量过低或噪声过大、模型文件下载不完整、以及路径/文件名包含中文或特殊符号。请按以下顺序排查:检查视频是否有声音:用播放器播放该视频确认能听到清晰人声。若无音轨,该软件无法处理此类文件。修正文件路径:你的路径包含中文和空格(D...115.216**9 Posted at: 1 month ago -
#6998 必须选择一个配音角色...
💬 1[Last reply:1 month ago] 这是未在“字幕配音”区域的“配音角色”下拉框中选择任何角色导致的。pyVideoTrans 的视频翻译流程要求必须指定配音角色(或明确选择“不配音”)。解决方法:在软件主界面找到第4行“字幕配音”。在“配音角色”下拉列表中,点击选择一个角色,例如 zh-CN-YunxiNeural(中文男声)或 en-US-GuyNeural(英文男声)等。如果此视频你不需要配音,请在“配音渠道”下拉框中选择 N...113.137**0 Posted at: 1 month ago
Open source and free maintenance is not easy. If this project is helpful to you, please consider making a small donation to help the project continue to maintain and update.