Audio engineering
201 tools and products, each with a live profile and alternatives list.
mediapipeCross-platform, customizable ML solutions for live and streaming media.spleeterDeezer source separation library including pretrained models.eqMacmacOS System-wide Audio Equalizer & Volume Mixer 🎧pedalboard🎛 🔊 A Python library for audio.DALIA GPU-accelerated library containing highly optimized building blocks and an execution engine for data processing to accelerate deep learning training and inferseek-tuneAn implementation of Shazam's song recognition algorithm.auto-editorEffort free video editing!audioFluxA library for audio and music analysis, feature extraction.stemrollerIsolate vocals, drums, bass, and other instrumental stems from any songaudio-reactive-led-strip:musical_note: :rainbow: Real-time LED strip music visualization using Python and the ESP8266 or Raspberry Piawesome-deep-learning-musicList of articles related to deep learning applied to musicaudioData manipulation and transformation for audio signal processing, powered by PyTorchOTTOSampler, Sequencer, Multi-engine synth and effects - in a box! [WIP]ailia-modelsThe collection of pre-trained, state-of-the-art AI models for ailia SDKaudiomentationsA Python library for audio data augmentation. Useful for making audio ML models work well in the real world, not just in the lab.beepA little package that brings sound to any Go application. Suitable for playback and audio-processing.SnapOtterOpen-source, self-hosted file-processing tool. Convert, compress, OCR, transcribe & run local AI across image, video, audio, PDF & documents, via UI, REST API &giadaYour Hardcore Loop Machine.LedFxLedFx is a network based LED effect engine designed to deliver advanced real-time audio effects to a wide variety of devices.kfrFast, modern C++ DSP framework, FFT, Sample Rate Conversion, FIR/IIR/Biquad Filters (SSE, AVX, AVX-512, ARM NEON, RISC-V RVV)DeepLearnImplementation of research papers on Deep Learning+ NLP+ CV in Python using Keras, Tensorflow and Scikit Learn.AaxAudioConverterConvert Audible aax files to mp3 and m4a/m4bmltMLT Multimedia FrameworkarcanThe owls are not what they seem.vocal-removerVocal Remover using Deep Neural NetworksRootlessJamesDSPAn implementation of the system-wide JamesDSP audio processing engine for non-rooted Android devicesPipeWire-GuidePipeWire Guide. Learn about how PipeWire gives your Linux system a Professional Audio/Video Processing workflow.tracktion_engineTracktion Engine moduleawesome-audio-dspMy curated list of audio DSP and plugin development resources (Github fork)chromaprintC library for generating audio fingerprints used by AcoustIDSmartGuitarAmpGuitar plugin made with JUCE that uses neural networks to emulate a tube amplifier.DawDreamerDigital Audio Workstation with Python; VST instruments/effects, parameter automation, FAUST, JAX, Warp Markers, and JUCE processorsffmediaelementFFME: The Advanced WPF MediaElement (based on FFmpeg)spotatuiA fast, standalone terminal music player in Rust: native Spotify streaming plus local, Subsonic, radio, and YouTube sources.torch-audiomentationsFast audio data augmentation in PyTorch. Inspired by audiomentations. Useful for deep learning.audinoOpen source audio annotation tool for humansnnAudioAudio processing by using pytorch 1D convolution networkfritureReal-time audio visualizations (spectrum, spectrogram, etc.)SLAM-LLMA Framework for Speech, Language, Audio, Music Processing with Large Language ModelsoundfingerprintingOpen source audio fingerprinting in .NET. An efficient algorithm for acoustic fingerprinting written purely in C#.kaprekapre: Keras Audio PreprocessorsWave-U-NetImplementation of the Wave-U-Net for audio source separationaudio-visualizer-android🎵 [Android Library] A light-weight and easy-to-use Audio Visualizer for Android.klioSmarter data pipelines for audio.react-native-audio-apiHigh-performance audio engine for react-nativeAPTAI Productivity Tool - Free and open source, improve user productivity, and protect privacy and data security. Including but not limited to: built-in local exclXR3Player🎧 🎼 The MOST ADVANCED JavaFX Media PlayerDTLNTensorflow 2.x implementation of the DTLN real time speech denoising model. With TF-lite, ONNX and real-time audio processing support.r8brain-free-srcHigh-quality pro audio resampler / sample rate conversion header-only C++ library. Very fast, for both audio resampling and time-series interpolation.fast-music-removerA C++ based, lightweight music and noise remover for YouTube and other internet media, using DeepFilterNet for audio enhancement.FoleyCrafter[IJCV 2026] FoleyCrafter: Bring Silent Videos to Life with Lifelike and Synchronized Sounds. AI拟音大师,给你的无声视频添加生动而且同步的音效 😝jumpcutter⏩ Fast-forwards long pauses between sentences — watch lectures ~1.5x faster (browser extension)PESQPESQ (Perceptual Evaluation of Speech Quality) Wrapper for Python Users (narrow band and wide band)spectro🎶 Real-time audio spectrogram generator for the webunsilenceConsole Interface and Library to remove silent parts of a media file 🔈checkrrCheckrr Scans your library files for corrupt media and optionally replaces the files via sonarr and radarrnara_wpeDifferent implementations of "Weighted Prediction Error" for speech dereverberationDplugMake VST2 / VST3 / AU / AAX / CLAP / LV2 / FLP plug-ins for Linux/macOS/Windows, using D.SoundFlowA high-performance, modular audio & MIDI engine for .NET 8+. A complete toolkit for the entire audio lifecycle: Playback, Recording, Multi-track Editing, Pro SyMediaEditorA non-linear editing software that helps you to make nice video.unified-audioAn Open-Source Project to Unify Audio Processing and GenerationComposeMediaPlayerCompose Media Player is a video player library designed for Compose Multiplatform, supporting multiple platforms including Android, macOS, Windows, Linux, iOS aSamplerBoxSamplerBox is a sampler musical instrument based on RaspberryPi.audio-development-toolsAudio Development Tools (ADT) is a project for advancing sound, speech, and music technologies, featuring components for machine learning, sound synthesis, speemusigA shazam like tool to store fingerprints and retrieve thememotion-classification-from-audio-filesUnderstanding emotions from audio files using neural networks and multiple datasets.Retrieval-based-Voice-Conversion-WebUIEasily train a good VC model with voice data <= 10 mins!whisper-atCode and Pretrained Models for Interspeech 2023 Paper "Whisper-AT: Noise-Robust Automatic Speech Recognizers are Also Strong Audio Event Taggers"DSP.jlFilter design, periodograms, window functions, and other digital signal processing functionalityflutter_soloudFlutter low-level audio plugin using SoLoud C++ library and FFIDeepAFx-STDeepAFx-ST - Style transfer of audio effects with differentiable signal processing. Please see https://csteinmetz1.github.io/DeepAFx-ST/TimeSidescalable audio processing framework and server written in PythonmosecaA Streamilt web app for music source separation & karaokeFftSharpA .NET Standard library for computing the Fast Fourier Transform (FFT) of real or complex dataSpectrogram.NET library for creating spectrograms (visual representations of frequency spectrum over time)MornMorn是一个C语言的基础工具和基础算法库,包括数据结构、图像处理、音频处理、机器学习等,具有简单、通用、高效的特点。bungeeA C++ library for time-stretching and pitch-shifting audio with high quality in realtime or offline.mod-desktopMOD Audio for the desktopteensy-junoA Teensy 3.x/4.x based polyphonic synthesizer, modelled after the Juno-106ESP32SynthPolyphonic synthesizer with up to 350 voices/channels for the ESP32 dual core family, offering high-fidelity audio (48kHz @ 32bit).SoundTouchJSA JavaScript library for manipulating WebAudio Contexts, specifically for handling key changeslibopenshot-audioOpenShot Audio Library (libopenshot-audio) is a free, open-source project that enables high-quality editing and playback of audio, and is based on the amazing JLocalText2VoiceA complete local production workflow for clean narration, structured learning content, and podcast-ready audiostemgen🎛 Stemgen is a Stem file generator. Convert any track into a Stem and have fun with Traktor.machinehearingMachine Learning applied to soundcav-maeCode and Pretrained Models for ICLR 2023 Paper "Contrastive Audio-Visual Masked Autoencoder".spectrographicTurn an image into sound whose spectrogram looks like the image.MWEngineAudio engine and DSP library for Android, written in C++ providing low latency performance within a musical context, while providing a Java/Kotlin API. SupportsAmplitudaAudio processing library, which provides waveform dataffmpeg-webA Web and Native UI for ffmpeg-wasm: convert video, audio and images using the power of ffmpeg, directly from your web browser or from your computer.kokoro-iosKokoro TTS for iOS and macOSXPhastFTA high-performance Fast Fourier Transform (FFT) library written in pure and safe Rust.geissThe Geiss screensaver and Winamp music visualization plug-inawesome-audiovisualCurated list of audiovisual projectsFunBoxStereo guitar pedal platform using Daisy Seed.speech-dataset-generator🔊 Create labeled datasets, enhance audio quality, identify speakers, support diverse dataset types. 🎧👥📊 Advanced audio processing.PanakoThe Panako acoustic fingerprinting system.lighthouse[EMNLP2024 Demo], [ICASSP 2025], [ICASSP 2026] A user-friendly library for reproducible video moment retrieval and highlight detection. It also supports audio mprism-mediaEasily transcode media using Node.js 🎶SPTKA suite of speech signal processing toolsweb-voice-processorA library for real-time voice processing in web browserspodcast-transcriberAn open-source tool that turns podcasts into high-quality transcripts and AI-powered summaries.phiolaFast audio player, recorder, converter for Windows, Linux & AndroidfourierFast Fourier transforms (FFTs) in RustaeroThis repo contains the official PyTorch implementation of "Audio Super Resolution in the Spectral Domain" (ICASSP 2023)deepdenoiserAn offline AI Audio Denoiser app built using DeepFilternet3awesome-NLP-resourcesa collection of NLP projects&tools. 自然语言处理方向项目和工具集合。Spleeter-Android-iOSOn-device, Offline Spleeter Solution For MobiledlibAllocators, I/O streams, math, geometry, image and audio processing for Daniraan architecture for neural network inference in real-time audio applicationsaudio-SNRMixing an audio file with a noise file at any Signal-to-Noise Ratio (SNR)VoxWeaveLocal-first high-quality offline RVC voice conversion workstationsynth-plugin-bookSource code for the book Code Your Own Synth Plug-Ins With C++ and JUCEeDSPA cross-platform DSP library written in C++ 11/14. This library harnesses the power of C++ templates to implement a complete set of DSP algorithms.pyaudiodsptoolsNumpy Audio DSP ToolsEasyPulseHQ Easy Effects presets for headphones & earbuds.Speech-Emotion-Classification-with-PyTorchThis repository contains PyTorch implementation of 4 different models for classification of emotions of the speech.audio-plugin-templateA template repository that you can use for creating audio plugins with the JUCE C++ framework. It is based on CMake, uses CPM package manager, the JUCE C++ framvideo-SALMONN-2video-SALMONN 2 is a powerful audio-visual large language model (LLM) that generates high-quality audio-visual video captions, which is developed by the DepartmWebCut🎬 基于 web 端的音视频编辑器。(A web-based audio and video editor.)awesome-sound_event_detectionReading list for research topics in Sound AIProteusGuitar amp and pedal capture plugin using neural networks.openmetersFast and professional audio metering/visualization for Linux.youtube-musical-spectrumAudio visualizer for YouTube, YT Music, Spotify, and SoundCloud with musical notes.stftPitchShiftSTFT based real-time pitch and timbre shifting in C++ and PythonSSRCAn audiophile-grade sample rate converterDDCToolboxCreate and edit DDC headset correction filesPlugalyzerCommand-line VST3, AU and LADSPA plugin host for easier debugging of audio pluginsLocalVQELean neural real-time acoustic echo cancellation with soft delay estimation - GGML and PyTorch inferenceRNNoise_WrapperA simple Python wrapper for audio noise reduction RNNoise. Simplifies work with it, adds new trained models and detailed instructions for training.react-audio-visualizeAn audio visualizer for React. Provides separate components to visualize both live audio and audio blobs.PicoADK-Firmware-Template🎵 🎹 Firmware boilerplate for the RP2040 / RP2350 powered PicoADK Audio Development Boards. Build your own stand alone synthesizers! Includes all nuts and bolts r-audioA library of React components for building Web Audio graphs.encodec-pytorchunofficial implementation of the High Fidelity Neural Audio Compressionffmpeg_kit_flutterFork of the original FFmpeg Kit library to work with Android V2 bindings and Flutter 3+pyACAPython scripts accompanying the book "An Introduction to Audio Content Analysis" (www.AudioContentAnalysis.org)Digital-Signal-Processing-Education-KitEducation kit for teaching digital signal processing on Arm Cortex-M platforms with lectures and lab materials (educational)wscribeez audio transcription tool with flexible processing and post-processing optionsAurioAudio Fingerprinting & Retrieval for .NETyoutube-clips-automatorMARCELO: an AI powered bot to automate the editing and thumbnail creation for your Youtube clips channeltwangLibrary for pure Rust advanced audio synthesis.radioformMusic is meant to sound good. Use an EQ to hear that difference.voxA universal AI toolkit for high-performance Speech-to-Text (STT) and Text-to-Speech (TTS) processing, designed for low-latency and easy model integration.Google-Meet-BotThis project is a Python bot that automates the process of logging into Gmail, joining a Google Meet, recording the audio of the meeting, and then generating a Melody-Forge-EngineAI-Powered Music Production & Live Coding Software 2026chromatone.centerChromatone is a digital garden of visual music theory and a collection of visual music instrumentspaper-listautoupdate paper listfast_bss_evalA fast implementation of bss_eval metrics for blind source separationrustortionA low latency guitar amp simulator.ShaderFlow🔥 Modular shader engine designed for simplicity and speedHushMicReal-time microphone noise suppression for Linux as a system-wide virtual mic (DPDFNet + PipeWire).ultimatevocalremover_apiAPI for a Vocal Remover that uses Deep Neural Networks.pyCrossfadepyCrossfade is the result of a personal project to use beat matching, gradual bpm shift on bars, and EQ modification to provide smooth and tunable transitions bNodeLinkPerformant LavaLink alternative written with Node.Jssee2soundOfficial code for SEE-2-SOUND: Zero-Shot Spatial Environment-to-Spatial Soundwhisper-clipWhisperClip simplifies your life by automatically transcribing audio recordings and saving the text directly to your clipboard. With just a click of a button, ydemucs-rsRust powered waveform source separationFreeAudioPluginListThe ultimate list of free audio processing plugins.spectral_connectivityFrequency domain estimation and functional and directed connectivity analysis tools for electrophysiological dataAudioAuditorA powerful, open-source toolkit for audio analysis and playback. Verify lossless quality, detect AI-generated tracks, and explore your library with a built-in hNeuralSeedNeural networks for guitar amp/pedal emulation on Daisy SeedlibrempegA complete, cross-platform solution to record, convert, filter and stream audio and video.FFaudioConverterGraphical audio convert and filter toolDPDFNetClean up noisy speech in real time with DPDFNet - open-source streaming speech enhancement for research, audio apps, and edge devices. Includes pretrained modelshezem-rsAudio recognition CLI written in Rustlamb-rsA lookahead compressor/limiter that's soft as a lamb.spectralOpen-source self-hosted audio-reactive video editor. Browser preview and server-side export share the same PixiJS runtime.VocalForgeYour one-stop solution for voice dataset creationDISSCOfficial repository for "Speaking Style Conversion With Discrete Self-Supervised Units" (EMNLP 2023). https://arxiv.org/abs/2212.09730HothouseExamplesExample effects code and binaries for the Cleveland Music Co. Hothouse Digital Signal Processing Pedal KitfogpadA VST reverb effect in which the reflections can be frozen, filtered, pitch shifted and ultimately disintegrated.deep-learning-bootcampLauzHack Deep Learning BootcampMicUpReal-time microphone audio processing for AndroidpianolizerAn easy-to-use toolkit for music exploration and visualization, an audio spectrum analyzer helping you turn sounds into piano notesSpeech-SeparationFinal Year Project for Speech SeparationSound-based-bird-species-detectionSound-based Bird Classification - using AI, acoustics and ornithology to classify birds in the environment, an environmental awareness project (Web Application,libvisualLibvisual Audio VisualizationComfyUI-AudioSRComfyUI node for AudioSR - Versatile Audio Super Resolution upscales audio to 48kHz using latent diffusionUnityAudioVisualizerAudio Visualizer in Unity.kaaninhos-mp3A lightweight, retro-styled desktop music player inspired by Winamp. Stream music directly from YouTube, apply real-time audio effects, and play a mini-game whiSpatialFlowSpatialFlow delivers a next-generation music experience on Android, blending online streaming, high-fidelity local playback, intelligent audio processing, and MresonatorsRust implementation of the Resonate algorithm for low-latency spectral analysis, with Python and WebAssembly bindingspydiogment:mega: Python library for audio augmentationlibmidimidi player base on timidity and imguiPartielsPartiels is an audio analysis application that allows you to explore the content and characteristics of soundssdkA powerful cross-platform audio engine, optimized for games.dssspReact Library for Audio Equalizers & Filter VisualizationwarbleRstreamline acoustic analysis in Rspectrogram-threejsA realtime 3d spectrogram visualization of the user's microphone audio. Made with threeJs using shaders.openshot-studio-enhancerOpenshot Studio EnhanceraudeyeCLI tool to visualize the content of an audio filetorchcompDifferentiable dynamic range controller in PyTorch.svelte-audio-waveformGenerate stunning audio waveforms with Svelte 5 and Canvas. Transform an array of peak data into beautifully rendered, customizable waveforms for music players,pafxPython Audio Effectssoundscape_IRTools of soundscape information retrieval, this repository is a developing project. Please go to https://github.com/meil-brcas-org/soundscape_IR for full releasFlipBits文本转无线电声音生成器与双轨 FSK 变声器,带实时二进制可视化 | Text-to-radio audio generator & dual-track FSK voice changer with real-time binary visualizers.doom-audioDoom playable over an audio connectionhms-harmful-brain-activity-classificationKaggle Silver Medal solution archive for HMS harmful brain activity EEG classification.SouPyXSouPyX: An Audio Exploration Space.🪐spectrogram-tutorialA walkthrough of how to make spectrograms in python that are customized for human speech research.soundscape_IRAn open toolbox of soundscape information retrieval