speech-recognition

72 progetti condividono questo topic GitHub

speech-recognition — whisper.cpp ★51.8kspeech-recognitionwhisperX — ★23.1kspeech_recognition — ★9kannyang — ★6.8kwav2letter — ★6.4kwhisper-diarization — ★5.6kwhisper-jax — ★4.7kstt — ★4.7kpocketsphinx — ★4.3kdistil-whisper — ★4.1kChenyme-AAVT — ★3.1kwhishper — ★3kSTT — ★2.6ktensorflow-speech-recognition — ★2.2kmasr — ★2kjulius — ★1.9klip-reading-deeplearning — ★1.9kwer_are_we — ★1.9kwhisper-turbo — ★1.8kSpeechT5 — ★1.4kopen-speech-corpora — ★1.4kwhisper-ctranslate2 — ★1.3kartyom.js — ★1.3kIrene-Voice-Assistant — ★1.1kkaldi-gstreamer-server — ★1.1kVoiceStreamAI — ★958whisper.net — ★935speechpy — ★883local-talking-llm — ★878speech_recognition — ★849react-speech-recognition — ★842whisper-playground — ★833stephanie-va — ★797whisper.rn — ★797whisper_mic — ★788SwiftWhisper — ★784whisper.unity — ★749allosaurus — ★738curses — ★711speech-demo — ★710GigaAM — ★688whisperX★ 23.1kspeech_recognition★ 9kannyang★ 6.8kwav2letter★ 6.4kwhisper-diarization★ 5.6kwhisper-jax★ 4.7kstt★ 4.7kpocketsphinx★ 4.3kdistil-whisper★ 4.1kChenyme-AAVT★ 3.1kwhishper★ 3kSTT★ 2.6ktensorflow-speech-recogn…★ 2.2kmasr★ 2kjulius★ 1.9klip-reading-deeplearning★ 1.9kwer_are_we★ 1.9kwhisper-turbo★ 1.8kSpeechT5★ 1.4kopen-speech-corpora★ 1.4kwhisper-ctranslate2★ 1.3kartyom.js★ 1.3kIrene-Voice-Assistant★ 1.1kkaldi-gstreamer-server★ 1.1kVoiceStreamAI★ 958whisper.net★ 935speechpy★ 883local-talking-llm★ 878speech_recognition★ 849react-speech-recognition★ 842whisper-playground★ 833stephanie-va★ 797whisper.rn★ 797whisper_mic★ 788SwiftWhisper★ 784whisper.unity★ 749allosaurus★ 738curses★ 711speech-demo★ 710GigaAM★ 688

Le linee collegano membri che sono misurabilmente correlati tra loro. La dimensione dei punti riflette le stelle.

🧬 Membri
whisper.cpp
Port of OpenAI's Whisper model in C/C++
★ 51.8k
whisperX
WhisperX: Automatic Speech Recognition with Word-level Timestamps (& Diarization)
★ 23.1k
speech_recognition
Speech recognition module for Python, supporting several engines and APIs, online and offline.
★ 9k
annyang
💬 Speech recognition for your site
★ 6.8k
wav2letter
Facebook AI Research's Automatic Speech Recognition Toolkit
★ 6.4k
whisper-diarization
Automatic Speech Recognition with Speaker Diarization based on OpenAI Whisper
★ 5.6k
whisper-jax
JAX implementation of OpenAI's Whisper model for up to 70x speed-up on TPU.
★ 4.7k
stt
★ 4.7k
pocketsphinx
A small speech recognizer
★ 4.3k
distil-whisper
Distilled variant of Whisper for speech recognition. 6x faster, 50% smaller, within 1% word error rate.
★ 4.1k
Chenyme-AAVT
★ 3.1k
whishper
Transcribe any audio to text, translate and edit subtitles 100% locally with a web UI. Powered by whisper…
★ 3k
STT
🐸STT - The deep learning toolkit for Speech-to-Text. Training and deploying STT models has never been so…
★ 2.6k
tensorflow-speech-recognition
🎙Speech recognition using the tensorflow deep learning framework, sequence-to-sequence neural networks
★ 2.2k
masr
中文语音识别; Mandarin Automatic Speech Recognition;
★ 2k
julius
Open-Source Large Vocabulary Continuous Speech Recognition Engine
★ 1.9k
lip-reading-deeplearning
:unlock: Lip Reading - Cross Audio-Visual Recognition using 3D Architectures
★ 1.9k
wer_are_we
Attempt at tracking states of the arts and recent results (bibliography) on speech recognition.
★ 1.9k
whisper-turbo
Cross-Platform, GPU Accelerated Whisper 🏎️
★ 1.8k
SpeechT5
Unified-Modal Speech-Text Pre-Training for Spoken Language Processing
★ 1.4k
open-speech-corpora
💎 A list of accessible speech corpora for ASR, TTS, and other Speech Technologies
★ 1.4k
whisper-ctranslate2
Whisper command line client compatible with original OpenAI client based on CTranslate2.
★ 1.3k
artyom.js
A voice control - voice commands - speech recognition and speech synthesis javascript library. Create your…
★ 1.3k
Irene-Voice-Assistant
Ирина - русский голосовой ассистент для работы оффлайн.…
★ 1.1k
kaldi-gstreamer-server
Real-time full-duplex speech recognition server, based on the Kaldi toolkit and the GStreamer framwork.
★ 1.1k
VoiceStreamAI
Near-Realtime audio transcription using self-hosted Whisper and WebSocket in Python/JS
★ 958
whisper.net
Whisper.net. Speech to text made simple using Whisper Models
★ 935
speechpy
:speech_balloon: SpeechPy - A Library for Speech Processing and Recognition:…
★ 883
local-talking-llm
A talking LLM that runs on your own computer without needing the internet.
★ 878
speech_recognition
中文语音识别
★ 849
react-speech-recognition
💬Speech recognition for your React app
★ 842
whisper-playground
Build real time speech2text web apps using OpenAI's Whisper https://openai.com/blog/whisper/
★ 833
stephanie-va
Stephanie is an open-source platform built specifically for voice-controlled applications as well as to…
★ 797
whisper.rn
React Native binding of whisper.cpp.
★ 797
whisper_mic
Project that allows one to use a microphone with OpenAI whisper.
★ 788
SwiftWhisper
🎤 The easiest way to transcribe audio in Swift
★ 784
whisper.unity
Running speech to text model (whisper.cpp) in Unity3d on your local machine.
★ 749
allosaurus
Allosaurus is a pretrained universal phone recognizer for more than 2000 languages
★ 738
curses
Speech to Text and KB input captions for OBS, VRChat, Twitch chat and Discord
★ 711
speech-demo
语音api示例
★ 710
GigaAM
Foundational Model for Speech Recognition Tasks
★ 688
cheetah
On-device streaming speech-to-text engine powered by deep learning
★ 669
expo-speech-recognition
Speech Recognition for React Native Expo projects
★ 652
speech-to-text
Real-time transcription using faster-whisper
★ 613
whisperIME
Android Input Method Editor (IME) based on Whisper
★ 606
Speech-Backbones
This is the main repository of open-sourced speech technology by Huawei Noah's Ark Lab.
★ 604
WhisperS2T
An Optimized Speech-to-Text Pipeline for the Whisper Model Supporting Multiple Inference Engine
★ 577
FastASR
★ 553
Awesome-Korean-Speech-Recognition
한국어 음성인식 STT API 리스트. 각 성능 벤치마크.
★ 534
SwiftSpeech
A speech recognition framework designed for SwiftUI.
★ 531
vosk-browser
A speech recognition library running in the browser thanks to a WebAssembly build of Vosk
★ 526
ai-pronunciation-trainer
This tool uses AI to evaluate your pronunciation.
★ 505
spchcat
Speech recognition tool to convert audio to text transcripts, for Linux and Raspberry Pi.
★ 486
leopard
On-device speech-to-text engine powered by deep learning
★ 482
Ming-UniAudio
Ming-UniAudio: Speech LLM for Joint Understanding, Generation and Editing with Unified Representation
★ 450
speech-recognition-uk
🇺🇦 Speech Recognition & Synthesis for Ukrainian
★ 439
whisper-youtube
🔉 Youtube Videos Transcription with OpenAI's Whisper
★ 421
dragonfly
Speech recognition framework allowing powerful Python-based scripting and extension of Dragon…
★ 413
VRCT
VRCT(VRChat Chatbox Translator & Transcription)
★ 406
PreenCut
AI-Powered Video Retrieval & Clipping Tool
★ 405
awesome-russian-speech
Russian speech technology links
★ 404
Freeze-Omni
✨✨Freeze-Omni: A Smart and Low Latency Speech-to-speech Dialogue Model with Frozen LLM
★ 388
xiaoniu
小牛视频翻译 是一款支持本地视频翻译、字幕翻译和 YouTube 视频翻译下载的 AI…
★ 383
self-supervised-speech-recognition
speech to text with self-supervised learning based on wav2vec 2.0 framework
★ 380
wav2vec2-live
A live speech recognition using Facebooks wav2vec 2.0 model.
★ 378
whisper-finetune
Fine-tune and evaluate Whisper models for Automatic Speech Recognition (ASR) on custom datasets or datasets…
★ 365
Speech-and-Text
Speech to text (PocketSphinx, Iflytex API, Baidu API) and text to speech (pyttsx3) |…
★ 342
vakyansh-models
Open source speech to text models for Indic Languages
★ 327
whispering-ui
Native UI for the Whispering Tiger project - https://github.com/Sharrnah/whispering (live transcription /…
★ 323
deepspeech-german
Automatic Speech Recognition (ASR) - German
★ 321
AudioBench
AudioBench: A Universal Benchmark for Audio Large Language Models
★ 319
SmartSpeaker
★ 307
🔗 Famiglie affini

Misurato dai temi di GitHub condivisi da entrambi i progetti, ponderato in base a quanto è raro ciascun tema.