Cross-platform, customizable ML solutions for live and streaming media.
Audio Processing GitHub Repositories
Explore popular GitHub repositories tagged “audio-processing”.
Compare stars, forks, and programming language using the same GitStar view as GitHub Trending.
Trending Repositories
Deezer source separation library including pretrained models.
A PyTorch-based Speech Toolkit
A text-to-speech (TTS), speech-to-text (STT) and speech-to-speech (STS) library built on Apple's MLX framework, providing efficient speech analysis on Apple Silicon.
macOS System-wide Audio Equalizer & Volume Mixer 🎧
🎛 🔊 A Python library for audio.
A GPU-accelerated library containing highly optimized building blocks and an execution engine for data processing to accelerate deep learning training and inference applications.
An implementation of Shazam's song recognition algorithm.
Effort free video editing!
A library for audio and music analysis, feature extraction.
Isolate vocals, drums, bass, and other instrumental stems from any song
:musical_note: :rainbow: Real-time LED strip music visualization using Python and the ESP8266 or Raspberry Pi
List of articles related to deep learning applied to music
Data manipulation and transformation for audio signal processing, powered by PyTorch
Sampler, Sequencer, Multi-engine synth and effects - in a box! [WIP]
The collection of pre-trained, state-of-the-art AI models for ailia SDK
Open-source, self-hosted file-processing tool. Convert, compress, OCR, transcribe & run local AI across image, video, audio, PDF & documents, via UI, REST API & pipelines. Your files never leave your network.
A little package that brings sound to any Go application. Suitable for playback and audio-processing.
Your Hardcore Loop Machine.
LedFx is a network based LED effect engine designed to deliver advanced real-time audio effects to a wide variety of devices.
Open-Source Large Vocabulary Continuous Speech Recognition Engine
Fast, modern C++ DSP framework, FFT, Sample Rate Conversion, FIR/IIR/Biquad Filters (SSE, AVX, AVX-512, ARM NEON, RISC-V RVV)
Implementation of research papers on Deep Learning+ NLP+ CV in Python using Keras, Tensorflow and Scikit Learn.
Convert Audible aax files to mp3 and m4a/m4b
MLT Multimedia Framework
The owls are not what they seem.
An implementation of the system-wide JamesDSP audio processing engine for non-rooted Android devices
SALMONN family: A suite of advanced multi-modal LLMs
PipeWire Guide. Learn about how PipeWire gives your Linux system a Professional Audio/Video Processing workflow.
Tracktion Engine module