speech-to-text

31 Projekte teilen dieses GitHub-Topic

speech-to-text — whisper.cpp ★51.8kspeech-to-textwhisperX — ★23.1kpyvideotrans — ★18.4kspeech_recognition — ★9kannyang — ★6.8kwhisper-jax — ★4.7kstt — ★4.7ktensorflow-speech-recognition — ★2.2kwhisper-ctranslate2 — ★1.3kartyom.js — ★1.3kVoiceStreamAI — ★958botium-speech-processing — ★943react-speech-recognition — ★842whisper-playground — ★833whisper_mic — ★788whisper.unity — ★749curses — ★711speech-demo — ★710expo-speech-recognition — ★652speech-to-text — ★613WhisperS2T — ★577Awesome-Korean-Speech-Recognition — ★534vosk-browser — ★526speech-recognition-uk — ★439whisper-youtube — ★421VRCT — ★406PreenCut — ★405self-supervised-speech-recognition — ★380insanely-fast-whisper-api — ★354Speech-and-Text — ★342MsEdgeTTS — ★335whisperX★ 23.1kpyvideotrans★ 18.4kspeech_recognition★ 9kannyang★ 6.8kwhisper-jax★ 4.7kstt★ 4.7ktensorflow-speech-recogn…★ 2.2kwhisper-ctranslate2★ 1.3kartyom.js★ 1.3kVoiceStreamAI★ 958botium-speech-processing★ 943react-speech-recognition★ 842whisper-playground★ 833whisper_mic★ 788whisper.unity★ 749curses★ 711speech-demo★ 710expo-speech-recognition★ 652speech-to-text★ 613WhisperS2T★ 577Awesome-Korean-Speech-Re…★ 534vosk-browser★ 526speech-recognition-uk★ 439whisper-youtube★ 421VRCT★ 406PreenCut★ 405self-supervised-speech-r…★ 380insanely-fast-whisper-ap…★ 354Speech-and-Text★ 342MsEdgeTTS★ 335

Linien verbinden Mitglieder, die messbar miteinander verwandt sind. Die Punktgröße spiegelt die Sterne wider.

🧬 Mitglieder
whisper.cpp
Port of OpenAI's Whisper model in C/C++
★ 51.8k
whisperX
WhisperX: Automatic Speech Recognition with Word-level Timestamps (& Diarization)
★ 23.1k
pyvideotrans
Translate the video from one language to another and embed dubbing & subtitles.
★ 18.4k
speech_recognition
Speech recognition module for Python, supporting several engines and APIs, online and offline.
★ 9k
annyang
💬 Speech recognition for your site
★ 6.8k
whisper-jax
JAX implementation of OpenAI's Whisper model for up to 70x speed-up on TPU.
★ 4.7k
stt
★ 4.7k
tensorflow-speech-recognition
🎙Speech recognition using the tensorflow deep learning framework, sequence-to-sequence neural networks
★ 2.2k
whisper-ctranslate2
Whisper command line client compatible with original OpenAI client based on CTranslate2.
★ 1.3k
artyom.js
A voice control - voice commands - speech recognition and speech synthesis javascript library. Create your…
★ 1.3k
VoiceStreamAI
Near-Realtime audio transcription using self-hosted Whisper and WebSocket in Python/JS
★ 958
botium-speech-processing
Botium Speech Processing
★ 943
react-speech-recognition
💬Speech recognition for your React app
★ 842
whisper-playground
Build real time speech2text web apps using OpenAI's Whisper https://openai.com/blog/whisper/
★ 833
whisper_mic
Project that allows one to use a microphone with OpenAI whisper.
★ 788
whisper.unity
Running speech to text model (whisper.cpp) in Unity3d on your local machine.
★ 749
curses
Speech to Text and KB input captions for OBS, VRChat, Twitch chat and Discord
★ 711
speech-demo
语音api示例
★ 710
expo-speech-recognition
Speech Recognition for React Native Expo projects
★ 652
speech-to-text
Real-time transcription using faster-whisper
★ 613
WhisperS2T
An Optimized Speech-to-Text Pipeline for the Whisper Model Supporting Multiple Inference Engine
★ 577
Awesome-Korean-Speech-Recognition
한국어 음성인식 STT API 리스트. 각 성능 벤치마크.
★ 534
vosk-browser
A speech recognition library running in the browser thanks to a WebAssembly build of Vosk
★ 526
speech-recognition-uk
🇺🇦 Speech Recognition & Synthesis for Ukrainian
★ 439
whisper-youtube
🔉 Youtube Videos Transcription with OpenAI's Whisper
★ 421
VRCT
VRCT(VRChat Chatbox Translator & Transcription)
★ 406
PreenCut
AI-Powered Video Retrieval & Clipping Tool
★ 405
self-supervised-speech-recognition
speech to text with self-supervised learning based on wav2vec 2.0 framework
★ 380
insanely-fast-whisper-api
An API to transcribe audio with OpenAI's Whisper Large v3!
★ 354
Speech-and-Text
Speech to text (PocketSphinx, Iflytex API, Baidu API) and text to speech (pyttsx3) |…
★ 342
MsEdgeTTS
A simple Azure Speech Service module that uses the Microsoft Edge Read Aloud API.…
★ 335
🔗 Verwandte Familien

Gemessen anhand der von beiden Projekten geteilten GitHub-Themen, gewichtet nach der Seltenheit jedes Themas.