百科.dev
全部条目趋势榜开源项目技术资讯提交条目
登录
< 返回工具列表
S

so-vits-svc

> 编程语言
开源

SoftVC VITS 唱歌声转换

28.2K stars0 点赞0 次浏览
访问官网GitHub

工具介绍

SoftVC VITS 唱歌声转换

QQ工作室包含可见的f0编辑器,扬声器混合时间线编辑器等功能(Onnx模型使用的地方): MoeVoiceStudio QQA分叉,用户界面得到极大改进:34j/so-vits-svc-fork QQA客户端支持实时转换:w-okada/voice-changer 这个计划与VITS有着根本的不同,因为它侧重于唱出语音转换(SVC)而不是文本对语音(TTS). 在这个项目中,TTS功能不支持,而VITS无法执行SVC任务. 需要注意的是,这两个项目所使用的模型不能互换或普遍适用. 通知 这个项目的目的是使开发者能够让所爱的动画人物去完成歌唱任务. 开发者的意图是只关注虚构的人物,并避免任何与真实个体,任何与真实的individua相关的.

核心特点

  • •Feature input is changed to the 12th Layer of Content Vec Transformer output, And compatible with 4.0 branches.
  • •Update the shallow diffusion, you can use the shallow diffusion model to improve the sound quality.
  • •Added Whisper-PPG encoder support
  • •Added static/dynamic sound fusion
  • •Added loudness embedding
  • •Added Functionality of feature retrieval from RVC
  • •To support the 4.0 model and incorporate the speech encoder, you can make modifications to the config.json file. Add the speech_encoder field to the "model" section as shown below:
  • •ContentVec: checkpoint_best_legacy_500.pt
  • •Place it under the pretrain directory
  • •ContentVec: hubert_base.pt

> 标签

Pythonaiaudio-analysisdeep-learningflow

暂无评论,来聊聊你的看法吧

> 工具信息

发布日期2026年8月1日
最后更新2026年9月9日
分类编程语言
定价开源

> 相关工具

T
TypeScript
JavaScript 的超集,为前端与全栈提供静态类型
P
Python
通用编程语言,广泛用于 Web、数据与 AI
G
Go
Google 推出的简洁高效系统语言