audio-processing

25 projetos partilham este topic do GitHub

audio-processing — mediapipe ★36.1kaudio-processingspleeter — ★28.3kspeechbrain — ★11.7kmlx-audio — ★7.6kpedalboard — ★6.2kDALI — ★5.7kaudioFlux — ★3.3kstemroller — ★3.1kawesome-deep-learning-music — ★3kaudio — ★2.9kailia-models — ★2.4kSnapOtter — ★2kjulius — ★1.9kDeepLearn — ★1.8kSALMONN — ★1.5kvllm-mlx — ★1.4kSoniTranslate — ★1.4kStreamSpeech — ★1.3kSincNet — ★1.2kAPT — ★771awesome-large-audio-models — ★733FoleyCrafter — ★658ltu — ★478whisper-at — ★421WhisperHallu — ★349spleeter★ 28.3kspeechbrain★ 11.7kmlx-audio★ 7.6kpedalboard★ 6.2kDALI★ 5.7kaudioFlux★ 3.3kstemroller★ 3.1kawesome-deep-learning-mu…★ 3kaudio★ 2.9kailia-models★ 2.4kSnapOtter★ 2kjulius★ 1.9kDeepLearn★ 1.8kSALMONN★ 1.5kvllm-mlx★ 1.4kSoniTranslate★ 1.4kStreamSpeech★ 1.3kSincNet★ 1.2kAPT★ 771awesome-large-audio-mode…★ 733FoleyCrafter★ 658ltu★ 478whisper-at★ 421WhisperHallu★ 349

Linhas conectam membros que estão mensuravelmente relacionados entre si. O tamanho do ponto reflete estrelas.

🧬 Membros
mediapipe
Cross-platform, customizable ML solutions for live and streaming media.
★ 36.1k
spleeter
Deezer source separation library including pretrained models.
★ 28.3k
speechbrain
A PyTorch-based Speech Toolkit
★ 11.7k
mlx-audio
A text-to-speech (TTS), speech-to-text (STT) and speech-to-speech (STS) library built on Apple's MLX…
★ 7.6k
pedalboard
🎛 🔊 A Python library for audio.
★ 6.2k
DALI
A GPU-accelerated library containing highly optimized building blocks and an execution engine for data…
★ 5.7k
audioFlux
A library for audio and music analysis, feature extraction.
★ 3.3k
stemroller
Isolate vocals, drums, bass, and other instrumental stems from any song
★ 3.1k
awesome-deep-learning-music
List of articles related to deep learning applied to music
★ 3k
audio
Data manipulation and transformation for audio signal processing, powered by PyTorch
★ 2.9k
ailia-models
The collection of pre-trained, state-of-the-art AI models for ailia SDK
★ 2.4k
SnapOtter
Open-source, self-hosted file-processing tool. Convert, compress, OCR, transcribe & run local AI across…
★ 2k
julius
Open-Source Large Vocabulary Continuous Speech Recognition Engine
★ 1.9k
DeepLearn
Implementation of research papers on Deep Learning+ NLP+ CV in Python using Keras, Tensorflow and Scikit…
★ 1.8k
SALMONN
SALMONN family: A suite of advanced multi-modal LLMs
★ 1.5k
vllm-mlx
OpenAI and Anthropic compatible server for Apple Silicon. Run LLMs and vision-language models (Llama,…
★ 1.4k
SoniTranslate
Synchronized Translation for Videos. Video dubbing
★ 1.4k
StreamSpeech
StreamSpeech is an “All in One” seamless model for offline and simultaneous speech recognition, speech…
★ 1.3k
SincNet
SincNet is a neural architecture for efficiently processing raw audio samples.
★ 1.2k
APT
AI Productivity Tool - Free and open source, improve user productivity, and protect privacy and data…
★ 771
awesome-large-audio-models
Collection of resources on the applications of Large Language Models (LLMs) in Audio AI.
★ 733
FoleyCrafter
[IJCV 2026] FoleyCrafter: Bring Silent Videos to Life with Lifelike and Synchronized Sounds.…
★ 658
ltu
Code, Dataset, and Pretrained Models for Audio and Speech Large Language Model "Listen, Think, and…
★ 478
whisper-at
Code and Pretrained Models for Interspeech 2023 Paper "Whisper-AT: Noise-Robust Automatic Speech Recognizers…
★ 421
WhisperHallu
Experimental code: sound file preprocessing to optimize Whisper transcriptions without hallucinated texts
★ 349
🔗 Familias relacionadas

Medido a partir dos tópicos do GitHub compartilhados por ambos os projetos, ponderado pela raridade de cada tópico.