speech-synthesis

29 Projekte teilen dieses GitHub-Topic

speech-synthesis — edge-tts ★11.5kspeech-synthesisvits — ★7.9kespeak-ng — ★6.7kRealtimeTTS — ★4ktacotron — ★3kMARS5-TTS — ★2.8kmarytts — ★2.6kTacotron-2 — ★2.3kVieNeu-TTS — ★2.2kWaveRNN — ★2.2kSpeechT5 — ★1.4kartyom.js — ★1.3kIrene-Voice-Assistant — ★1.1kIrodori-TTS — ★1klocal-talking-llm — ★878glow-tts — ★712vits2 — ★642tiktok-voice — ★607Speech-Backbones — ★604StyleTTS — ★465Ming-UniAudio — ★450speech-recognition-uk — ★439ProDiff — ★432FastDiff — ★423nnmnkwii — ★399MsEdgeTTS — ★335GenerSpeech — ★333VocGAN — ★321manim-voiceover — ★305vits★ 7.9kespeak-ng★ 6.7kRealtimeTTS★ 4ktacotron★ 3kMARS5-TTS★ 2.8kmarytts★ 2.6kTacotron-2★ 2.3kVieNeu-TTS★ 2.2kWaveRNN★ 2.2kSpeechT5★ 1.4kartyom.js★ 1.3kIrene-Voice-Assistant★ 1.1kIrodori-TTS★ 1klocal-talking-llm★ 878glow-tts★ 712vits2★ 642tiktok-voice★ 607Speech-Backbones★ 604StyleTTS★ 465Ming-UniAudio★ 450speech-recognition-uk★ 439ProDiff★ 432FastDiff★ 423nnmnkwii★ 399MsEdgeTTS★ 335GenerSpeech★ 333VocGAN★ 321manim-voiceover★ 305

Linien verbinden Mitglieder, die messbar miteinander verwandt sind. Die Punktgröße spiegelt die Sterne wider.

🧬 Mitglieder
edge-tts
Use Microsoft Edge's online text-to-speech service from Python WITHOUT needing Microsoft Edge or Windows or…
★ 11.5k
vits
VITS: Conditional Variational Autoencoder with Adversarial Learning for End-to-End Text-to-Speech
★ 7.9k
espeak-ng
eSpeak NG is an open source speech synthesizer that supports more than hundred languages and accents.
★ 6.7k
RealtimeTTS
Converts text to speech in realtime
★ 4k
tacotron
A TensorFlow implementation of Google's Tacotron speech synthesis with pre-trained model (unofficial)
★ 3k
MARS5-TTS
MARS5 speech model (TTS) from CAMB.AI
★ 2.8k
marytts
MARY TTS -- an open-source, multilingual text-to-speech synthesis system written in pure java
★ 2.6k
Tacotron-2
DeepMind's Tacotron-2 Tensorflow implementation
★ 2.3k
VieNeu-TTS
Vietnamese TTS with instant voice cloning • On-device • Real-time CPU inference • 24kHz audio quality…
★ 2.2k
WaveRNN
WaveRNN Vocoder + TTS
★ 2.2k
SpeechT5
Unified-Modal Speech-Text Pre-Training for Spoken Language Processing
★ 1.4k
artyom.js
A voice control - voice commands - speech recognition and speech synthesis javascript library. Create your…
★ 1.3k
Irene-Voice-Assistant
Ирина - русский голосовой ассистент для работы оффлайн.…
★ 1.1k
Irodori-TTS
A Flow Matching-based Text-to-Speech Model with Emoji-driven Style Control
★ 1k
local-talking-llm
A talking LLM that runs on your own computer without needing the internet.
★ 878
glow-tts
A Generative Flow for Text-to-Speech via Monotonic Alignment Search
★ 712
vits2
VITS2: Improving Quality and Efficiency of Single-Stage Text-to-Speech with Adversarial Learning and…
★ 642
tiktok-voice
Simple Python script to interact with the TikTok TTS API
★ 607
Speech-Backbones
This is the main repository of open-sourced speech technology by Huawei Noah's Ark Lab.
★ 604
StyleTTS
Official Implementation of StyleTTS
★ 465
Ming-UniAudio
Ming-UniAudio: Speech LLM for Joint Understanding, Generation and Editing with Unified Representation
★ 450
speech-recognition-uk
🇺🇦 Speech Recognition & Synthesis for Ukrainian
★ 439
ProDiff
PyTorch Implementation of ProDiff (ACM-MM'22) with a Extremely-Fast diffusion speech synthesis pipeline
★ 432
FastDiff
PyTorch Implementation of FastDiff (IJCAI'22)
★ 423
nnmnkwii
Library to build speech synthesis systems designed for easy and fast prototyping.
★ 399
MsEdgeTTS
A simple Azure Speech Service module that uses the Microsoft Edge Read Aloud API.…
★ 335
GenerSpeech
PyTorch Implementation of GenerSpeech (NeurIPS'22): a text-to-speech model towards zero-shot style transfer…
★ 333
VocGAN
VocGAN: A High-Fidelity Real-time Vocoder with a Hierarchically-nested Adversarial Network
★ 321
manim-voiceover
Manim plugin for all things voiceover
★ 305
🔗 Verwandte Familien

Gemessen anhand der von beiden Projekten geteilten GitHub-Themen, gewichtet nach der Seltenheit jedes Themas.