text-to-speech

47 progetti condividono questo topic GitHub

text-to-speech — GPT-SoVITS ★59.9ktext-to-speechOpenVoice — ★37kindex-tts — ★22kdia — ★19.3kpyvideotrans — ★18.4kedge-tts — ★11.5kvits — ★7.9kRealtimeTTS — ★4kelevenlabs-python — ★3kvall-e — ★3kMARS5-TTS — ★2.8kmarytts — ★2.6kChatTTS_colab — ★2.6kvall-e — ★2.2kAuto-Synced-Translated-Dubs — ★1.7kGenie-TTS — ★1.7kOuteTTS — ★1.4kGPA — ★1.3ksoprano — ★1.2kXZVoice — ★1.2kdia2 — ★1.2kbotium-speech-processing — ★943bark.cpp — ★866CloneTTS — ★738glow-tts — ★712bark-voice-cloning-HuBERT-quantizer — ★711voicebox-pytorch — ★698LLaSA_training — ★660f5-tts-mlx — ★640examples — ★617tiktok-voice — ★607vits2_pytorch — ★548orpheus-tts-local — ★544e2-tts-pytorch — ★516vixtts-demo — ★515ComfyUI-OmniVoice-TTS — ★509google-speech-v2 — ★469StyleTTS — ★465dectalk — ★451ProDiff — ★432FastDiff — ★423OpenVoice★ 37kindex-tts★ 22kdia★ 19.3kpyvideotrans★ 18.4kedge-tts★ 11.5kvits★ 7.9kRealtimeTTS★ 4kelevenlabs-python★ 3kvall-e★ 3kMARS5-TTS★ 2.8kmarytts★ 2.6kChatTTS_colab★ 2.6kvall-e★ 2.2kAuto-Synced-Translated-D…★ 1.7kGenie-TTS★ 1.7kOuteTTS★ 1.4kGPA★ 1.3ksoprano★ 1.2kXZVoice★ 1.2kdia2★ 1.2kbotium-speech-processing★ 943bark.cpp★ 866CloneTTS★ 738glow-tts★ 712bark-voice-cloning-HuBER…★ 711voicebox-pytorch★ 698LLaSA_training★ 660f5-tts-mlx★ 640examples★ 617tiktok-voice★ 607vits2_pytorch★ 548orpheus-tts-local★ 544e2-tts-pytorch★ 516vixtts-demo★ 515ComfyUI-OmniVoice-TTS★ 509google-speech-v2★ 469StyleTTS★ 465dectalk★ 451ProDiff★ 432FastDiff★ 423

Le linee collegano membri che sono misurabilmente correlati tra loro. La dimensione dei punti riflette le stelle.

🧬 Membri
GPT-SoVITS
1 min voice data can also be used to train a good TTS model! (few shot voice cloning)
★ 59.9k
OpenVoice
Instant voice cloning by MIT and MyShell. Audio foundation model.
★ 37k
index-tts
An Industrial-Level Controllable and Efficient Zero-Shot Text-To-Speech System
★ 22k
dia
A TTS model capable of generating ultra-realistic dialogue in one pass.
★ 19.3k
pyvideotrans
Translate the video from one language to another and embed dubbing & subtitles.
★ 18.4k
edge-tts
Use Microsoft Edge's online text-to-speech service from Python WITHOUT needing Microsoft Edge or Windows or…
★ 11.5k
vits
VITS: Conditional Variational Autoencoder with Adversarial Learning for End-to-End Text-to-Speech
★ 7.9k
RealtimeTTS
Converts text to speech in realtime
★ 4k
elevenlabs-python
The official Python SDK for the ElevenLabs API.
★ 3k
vall-e
An unofficial PyTorch implementation of the audio LM VALL-E
★ 3k
MARS5-TTS
MARS5 speech model (TTS) from CAMB.AI
★ 2.8k
marytts
MARY TTS -- an open-source, multilingual text-to-speech synthesis system written in pure java
★ 2.6k
ChatTTS_colab
🚀 一键部署(含离线整合包)!基于 ChatTTS ,支持流式输出、音色抽卡、长音频生…
★ 2.6k
vall-e
PyTorch implementation of VALL-E(Zero-Shot Text-To-Speech), Reproduced Demo…
★ 2.2k
Auto-Synced-Translated-Dubs
Automatically translates the text of a video based on a subtitle file, and then uses AI voice services to…
★ 1.7k
Genie-TTS
GPT-SoVITS ONNX Inference Engine & Model Converter
★ 1.7k
OuteTTS
Interface for OuteTTS models.
★ 1.4k
GPA
[AutoArk] GPA (General Purpose Audio) can do ASR, TTS and voice conversion with one tiny model!
★ 1.3k
soprano
Soprano: Instant, Ultra-Realistic Text-to-Speech
★ 1.2k
XZVoice
Free and open source text-to-speech software
★ 1.2k
dia2
TTS model capable of streaming conversational audio in realtime.
★ 1.2k
botium-speech-processing
Botium Speech Processing
★ 943
bark.cpp
Suno AI's Bark model in C/C++ for fast text-to-speech generation
★ 866
CloneTTS
A lightweight, offline Android Text-to-Speech (TTS) engine enabling seamless system-wide voice cloning and…
★ 738
glow-tts
A Generative Flow for Text-to-Speech via Monotonic Alignment Search
★ 712
bark-voice-cloning-HuBERT-quantizer
The code for the bark-voicecloning model. Training and inference.
★ 711
voicebox-pytorch
Implementation of Voicebox, new SOTA Text-to-speech network from MetaAI, in Pytorch
★ 698
LLaSA_training
LLaSA: Scaling Train-time and Inference-time Compute for LLaMA-based Speech Synthesis
★ 660
f5-tts-mlx
Implementation of F5-TTS in MLX
★ 640
examples
★ 617
tiktok-voice
Simple Python script to interact with the TikTok TTS API
★ 607
vits2_pytorch
unofficial vits2-TTS implementation in pytorch
★ 548
orpheus-tts-local
Run Orpheus 3B Locally With LM Studio
★ 544
e2-tts-pytorch
Implementation of E2-TTS, "Embarrassingly Easy Fully Non-Autoregressive Zero-Shot TTS", in Pytorch
★ 516
vixtts-demo
A Vietnamese Voice Cloning Text-to-Speech Model ✨
★ 515
ComfyUI-OmniVoice-TTS
OmniVoice TTS nodes for ComfyUI - Zero-shot multilingual text-to-speech with voice cloning, voice design, and…
★ 509
google-speech-v2
:speech_balloon: Reverse Engineering Google's Speech To Text API (v2)
★ 469
StyleTTS
Official Implementation of StyleTTS
★ 465
dectalk
Modern builds for the 90s/00s DECtalk text-to-speech application.
★ 451
ProDiff
PyTorch Implementation of ProDiff (ACM-MM'22) with a Extremely-Fast diffusion speech synthesis pipeline
★ 432
FastDiff
PyTorch Implementation of FastDiff (IJCAI'22)
★ 423
Cross-Lingual-Voice-Cloning
Tacotron 2 - PyTorch implementation with faster-than-realtime inference modified to enable cross lingual…
★ 359
easevoice-trainer
EaseVoice Trainer is a simple and user-friendly voice cloning and speech model trainer.
★ 352
StreamingKokoroJS
Unlimited text-to-speech in the Browser using Kokoro-JS, 100% local, 100% open source
★ 346
Speech-and-Text
Speech to text (PocketSphinx, Iflytex API, Baidu API) and text to speech (pyttsx3) |…
★ 342
Whisper-TikTok
From AI tools to TikTok video creation using FFMPEG, Microsoft Edge read aloud and OpenAI Whisper model
★ 335
BlueTTS
Fastest Open Source TTS Model
★ 77
🔗 Famiglie affini

Misurato dai temi di GitHub condivisi da entrambi i progetti, ponderato in base a quanto è raro ciascun tema.