Voice enhancement
78 tools and products, each with a live profile and alternatives list.
KrispAI noise and echo cancellation that sits between your mic and any calling app, plus call transcription features.NVIDIA BroadcastGPU-powered noise removal, room echo removal and virtual background for anyone with an RTX card.Adobe Podcast EnhanceUpload rough voice recordings and get studio-quality speech back — Adobe's free enhancement demo turned essential tool.AuphonicAutomated audio post-production: loudness normalization, leveling, noise reduction and encoding presets for podcasts.Waifu2x-Extension-GUIVideo, Image and GIF upscale/enlarge(Super-Resolution) and Video frame interpolation. Achieved with Waifu2x, Real-ESRGAN, Real-CUGAN, RTX Video Super ResolutiornnoiseRecurrent neural network for audio noise reductionDeepFilterNetNoise supression using deep filteringClearerVoice-StudioAn AI-Powered Speech Processing Toolkit and Open Source SOTA Pretrained Models, Supporting Speech Enhancement, Separation, and Target Speaker Extraction, etc.QualityScalerQualityScaler - image/video AI upscaler appresemble-enhanceAI powered speech denoising and enhancementvoicefixerGeneral Speech RestorationSpeech-Separation-Paper-TutorialA must-read paper for speech separation based on neural networksgtcrnThe official implementation of GTCRN, an ultra-lightweight SE model.FullSubNetPyTorch implementation of "FullSubNet: A Full-Band and Sub-Band Fusion Model for Real-Time Single-Channel Speech Enhancement."sprocketVoice Conversion Tool Kitnoise-repellentA JUCE plugin for broadband noise reductionMP-SENetExplicit Estimation of Magnitude and Phase Spectra in Parallel for High-Quality Speech EnhancementpysepmPython implementation of performance metrics in Loizou's Speech Enhancement bookwaifuExtensionThe waifu2x & Other image-enlargers on MacsetkTools for Speech Enhancement integrated with Kaldiframework-reproducibilityProviding reproducibility in deep learning frameworksRealScalerRealScaler - image/video AI upscaler app (Real-ESRGAN)PercepNetUnofficial implementation of PercepNet: A Perceptually-Motivated Approach for Low-Complexity, Real-Time Enhancement of Fullband SpeechCleanUNetOfficial PyTorch Implementation of CleanUNet (ICASSP 2022)Wave-U-Net-for-Speech-EnhancementImplement Wave-U-Net by PyTorch, and migrate it to the speech enhancement.mayavozPytorch based speech enhancement toolkit.voicefixer_mainGeneral Speech RestorationphasenA unofficial Pytorch implementation of Microsoft's PHASENMTFAA-NetMulti-Scale Temporal Frequency Convolutional Network With Axial Attention for Speech Enhancementtorch-pesqPyTorch implementation of the Perceptual Evaluation of Speech Quality for wideband audioSoundSourceSeparationThe code for multi-channel source separation and dereverberation such as FastMNMF1, FastMNMF2, and AR-FastMNMF2.Noise2Noise-audio_denoising_without_clean_training_dataSource code for the paper titled "Speech Denoising without Clean Training Data: a Noise2Noise Approach". Paper accepted at the INTERSPEECH 2021 conference. ThiAudio-DenoisingNoise removal/ reducer from the audio file in python. De-noising is done using Wavelets and thresholding is done by VISU Shrink thresholding techniquevoicerestoreVoiceRestore: Flow-Matching Transformers for Universal Speech RestorationSEtrainA training code template for DNN-based speech enhancement.clarityClarity Challenge toolkit - software for building Clarity Challenge systemsRealMANA description of "RealMAN: A Real-Recorded and Annotated Microphone Array Dataset for Dynamic Speech Enhancement and Localization" [NeurIPS 2024]looking-to-listenDeep neural network (DNN) for noise reduction, removal of background music, and speech separationRpiANCActive Noise Control on Raspberry PideepvqeAn unofficial implementation of DeepVQE proposed by Microsoft Corp.apsA personal toolkit for single/multi-channel speech recognition & enhancement & separation.spiking-fullsubnetOfficial repository of Spiking-FullSubNet, the Intel N-DNS Challenge Algorithmic Track Winner.McNetThe official repo: "McNet: Fuse Multiple Cues for Multichannel Speech Enhancement", ICASSP 2023open-universeOpen implementation of UNIVERSE and UNIVERSE++ diffusion-based speech enhancement models.NLP-GuideNatural Language Processing (NLP). Covering topics such as Tokenization, Part Of Speech tagging (POS), Machine translation, Named Entity Recognition (NER), ClasFLowHigh_code[ICASSP 2025] "FLowHigh: Towards efficient and high-quality audio super-resolution with single-step flow matching"SpeechDenoiserSpeechDenoiser: Real-Time Speech Denoising with ONNX Welcome to SpeechDenoiser, a simple and effective solution for real-time speech denoising using an ONNX moflowmse(ICASSP 2025, official code)FlowSE: Flow Matching-based Speech EnhancementlibspecbleachC library for audio noise reduction and other spectral effectsDCCRN-with-various-loss-functionsDCCRN with various loss functionsInter-SubNetThe official PyTorch implementation of "Inter-SubNet: Speech Enhancement with Subband Interaction", accepted by ICASSP 2023.ssl_speech_restorationSelfRemaster: SSL Speech RestorationkoalaOn-device noise suppression powered by deep learningTRT-SEAn example of a speech enhancement model deployed with TensorRT.web-noise-suppressorNoise suppressor nodes for Web Audio API.universal-speech-enhancementApply Score diffusion to improve speech signals recorded under various adverse conditions and distortions, including noise, reverberation, clipping, equalizatiopython-speech-enhancementa python library for speech enhancementbLUe_PYSIDEbLUe - A simple and comprehensive image editor featuring automatic contrast enhancement, color correction, 3D LUT creation, raw postprocessing, exposure fusion Python-Sound-ToolSoundPy (alpha stage) is a research-based python package for speech and sound. Applications include deep-learning, filtering, speech-enhancement, audio augmentaebenRepo for source code of EBEN: Extreme Bandwidth Extension NetworkasrpyArtifact Subspace Reconstruction for PythonMANNERMANNER: Multi-view Attention Network for Noise ERasure (Speech enhancement in time-domain)streamfmReal-Time Streamable Generative Speech Restoration with Flow Matchingrapidly-sdkRealtime audio separation running fully on-device across iOS, macOS, Android, Linux and WindowsNested-U-Net-based-Real-time-Speech-Enhancement-Mobile-AppReal-time speech enhancement mobile app using Nested U-Netnoise-xorcistSingle Channel Speech Enhancement Methods and ToolboxMULTI-AUDIODECThis is the official implementation of our multi-channel multi-speaker multi-spatial neural audio codec architecture.vocal-gateFree real-time AI Noise Gate VST3/AU plugin. Removes coughs, sneezes, and other artifacts from your live streams, podcasts, and videos.vibravoxSpeech to Phoneme, Bandwidth Extension and Speaker Verification using the Vibravox dataset.RVAE-EMOfficial PyTorch implementation of "RVAE-EM: Generative speech dereverberation based on recurrent variational auto-encoder and convolutive transfer function" [Ispeexdsp-ns-pythonPython bindings of speexdsp noise suppression libraryRobust-E2E-ASRThis repository contains the code for our upcoming paper An Investigation of End-to-End Models for Robust Speech Recognition at ICASSP 2021.rnnoise_pythonpython wrapper for rnnoise libraryd2geoFramework for computing seismic attributes with Python.Pytorch-Tensor-Train-NetworkJun and Huck's PyTorch-Tensor-Train Network Toolboxremove-noise一个简单的音频降噪工具,提高web UI界面和api接口SignalDecomposition.jlDecompose a signal/timeseries into structure and noise or seasonal and residual componentsdenoisersSimple PyTorch Denoisers for Waveform Audio