仅需几秒到一分钟语音即可克隆音色并合成的 TTS 工具。
Whisper 精选
OpenAI 开源的通用语音识别模型,支持多语言转写与翻译。
- 分类
- 🎙️ 语音与音频
- 仓库
- openai/whisper
- Stars
- ★ 110,307
- Forks
- 13,368
- 语言
- Python
- 许可证
- MIT
- 最近更新
- 2026-08-31
- 访问次数
- 2
语音识别
同类项目
通义实验室的多语言语音合成大模型,支持零样本音色克隆。
VoiceStudio is the open-source, fully-local ElevenLabs alternative — voice cloning, voice design, video dubbing, dictation, transcription & audiobook creation in 646 languages.
The open-source AI voice studio. Clone, dictate, create.
Open-Source Frontier Voice AI
Whisper 的 C/C++ 移植,可在 CPU、手机和浏览器里运行。