ランキングに戻る

LuckyHookin/edge-TTS-record

HTML

一个可以录制 Microsoft Edge 浏览器的语音合成(TTS)语音并输出为 .wav 音频的(windows平台)工具。

ttsedgerecordvuejsaardio
スター成長
スター
1.4k
フォーク
172
週間成長
Issue
9
5001k
2021年5月2023年1月2024年10月2026年7月
README

edge-TTS-record

⚠⚠请勿由于违法犯罪用途,否则后果自负⚠⚠

新写的另一个工具发布:https://t.bilibili.com/694237238144270361

edge-TTS-record-img

一个可以录制 Microsoft Edge 浏览器的语音合成(TTS)语音并输出为 .wav 音频的(windows平台)工具。

Microsoft Edge 浏览器中有两款非常逼真的在线(Online)中文(zh-CN)语音:Xiaoxiao、Yunyang。

用法:

  1. 下载 edge-TTS-record.exe,运行并允许联网
  2. 在文本编辑框中输入文本,调整参数,点击试听
  3. 没问题就可以点击录制了,音频文件会保存在指定的目录下

演示:https://www.bilibili.com/video/BV1eK411c73s

注意:

  • 需要 Microsoft Edge 浏览器 Chromium 内核版,一般是 Windows 10 自带的,如果系统中没有安装,程序将自动为你下载安装
  • 不管是试听还是录制,使用在线(Online)语音都需确保电脑是联网的
  • 录制是全局的,应避免其他软件声音的干扰
  • 在线(Online)语音似乎无法调整音调
  • 当无法录制时,你可能需要额外安装 .NET Framework: https://www.microsoft.com/zh-CN/download/details.aspx?id=17851

TODO:

  • 路径配置
  • 可自定义选择要录制的设备
  • 软件更新检测

相关仓库:

界面(vue.js):https://github.com/LuckyHookin/tts-record-html

音频录制模块(C#,NAudio):https://github.com/LuckyHookin/RecordAudio

関連リポジトリ
unslothai/unsloth

Unsloth is a local UI for training and running Gemma 4, Qwen3.6, DeepSeek, Kimi, GLM and other models.

PythonPyPIApache License 2.0fine-tuningllama
unsloth.ai/docs
68.7k6.2k
CorentinJ/Real-Time-Voice-Cloning

Clone a voice in 5 seconds to generate arbitrary speech in real-time

PythonPyPIOtherdeep-learningpytorch
60k9.4k
RVC-Boss/GPT-SoVITS

1 min voice data can also be used to train a good TTS model! (few shot voice cloning)

PythonPyPIMIT Licensetext-to-speechtts
60k6.5k
mudler/LocalAI

LocalAI is the open-source AI engine. Run any model - LLMs, vision, voice, image, video - on any hardware. No GPU required.

GoGo ModulesMIT Licensellamaai
localai.io
47.7k4.3k
coqui-ai/TTS

🐸💬 - a deep learning toolkit for Text-to-Speech, battle-tested in research and production

PythonPyPIMozilla Public License 2.0pythontext-to-speech
coqui.ai
45.8k6.2k
2noise/ChatTTS

A generative speech model for daily dialogue.

PythonPyPIGNU Affero General Public License v3.0agenttext-to-speech
2noise.com
39.7k4.2k
myshell-ai/OpenVoice

Instant voice cloning by MIT and MyShell. Audio foundation model.

PythonPyPIMIT Licensetext-to-speechtts
research.myshell.ai/open-voice
37k4.1k
babysor/MockingBird

🚀Clone a voice in 5 seconds to generate arbitrary speech in real-time

PythonPyPIOtheraispeech
36.9k5.2k
OpenBMB/VoxCPM

VoxCPM2: Tokenizer-Free TTS for Multilingual Speech Generation, Creative Voice Design, and True-to-Life Cloning

PythonPyPIApache License 2.0audiodeeplearning
voxcpm.com
34k3.9k
fishaudio/fish-speech

SOTA Open Source TTS

PythonPyPIOtherllamatransformer
speech.fish.audio
31.3k2.7k
mastra-ai/mastra

Mastra is the modern TypeScript framework for AI-powered applications and agents.

TypeScriptnpmOtheragentsai
mastra.ai
26.4k2.5k
ATH-MaaS/Pixelle-Video

🚀 AI 全自动短视频引擎 | AI Fully Automated Short Video Engine

PythonPyPIApache License 2.0aigccomfyui
aidc-ai.github.io/Pixelle-Video/zh
25.8k3.7k