Skip to content
View weijietan09's full-sized avatar
🎯
Focusing
🎯
Focusing
  • Shenzhen University
  • Shenzhen, China
  • Joined Jun 28, 2026

Block or report weijietan09

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
weijietan09/README.md

Cheng Jiawei · 程嘉伟

深圳大学研究生,做 声音克隆与 TTS 语音合成大模型。 关心一件事:如何用尽量少的参考音频,合成出音色像、情感对、韵律自然的语音。

我的研究从 零样本音色克隆 起步,逐步走向 情感与韵律可控 的语音生成。写代码守两条线:参考实现以纯 NumPy 为核心,import 时不拉深度学习框架,无 GPU、无网络也能跑通与单测;需要真正训练时,再按需接上可选的 PyTorch 后端。

开源项目

项目 方向 关键词
voxflow 零样本声音克隆 TTS 工具包 说话人编码器 · 扩散 / 流匹配声学模型 · 声码器 · 中英双语
prosodia 可控语音合成大模型框架 情感 / 风格提示 · 韵律控制 · 流式推理

一些取舍

  • 可复现优先 —— 固定随机种子、离线数据、CI 全绿。
  • 离线优先 —— 默认不联网、不下权重也能跑核心逻辑与测试。
  • 由小见大 —— 先用纯 NumPy 参考实现把算法讲清楚,再接框架做真实训练。

Shenzhen University · 声音克隆 / TTS · reproducible & offline-first

Pinned Loading

  1. prosodia prosodia Public

    可控情感与韵律的语音合成框架:文本 / 风格 / 情感提示驱动的 TTS,支持流式推理

    Python

  2. voxflow voxflow Public

    零样本声音克隆 TTS 工具包:说话人编码器 + 流匹配/扩散声学模型 + 声码器,支持中英双语

    Python

  3. xiaohanc/AnoShip xiaohanc/AnoShip Public

    Anoamly-detection-driven deployment safety for production AI/ML systems

    Python 163 17

  4. tuya/tuya-smart-control-cli tuya/tuya-smart-control-cli Public

    The official command-line tool for Tuya Smart Control -- manage your smart home devices directly from the terminal.

    JavaScript 123 7