A user-friendly toolkit for voice recgonition/transcription/conversion etc. | 简单易用的语音工具箱
-
Updated
Sep 9, 2026 - Python
A user-friendly toolkit for voice recgonition/transcription/conversion etc. | 简单易用的语音工具箱
Source code for the paper titled "Speech Denoising without Clean Training Data: a Noise2Noise Approach". Paper accepted at the INTERSPEECH 2021 conference. This paper tackles the problem of the heavy dependence of clean speech data required by deep learning based audio denoising methods by showing that it is possible to train deep speech denoisi…
Noise removal/ reducer from the audio file in python. De-noising is done using Wavelets and thresholding is done by VISU Shrink thresholding technique
Official pytorch implementation of the paper: "Catch-A-Waveform: Learning to Generate Audio from a Single Short Example" (NeurIPS 2021)
Clean up noisy speech in real time with DPDFNet - open-source streaming speech enhancement for research, audio apps, and edge devices. Includes pretrained models, PyTorch code, ONNX/TFLite inference, 8/16/48 kHz support, and live demos.
基于深度学习的语音增强工具(Speech Enhancement Tools Based on Deep Learning)
Uses machine learning to denoise audio containing speech
logWMSE, an audio quality metric & loss function with support for digital silence target. Useful for training and evaluating audio source separation systems.
Variations of L1 SNR Loss function for training audio source separation machine learning models
Simple PyTorch Denoisers for Waveform Audio
Unofficial PyTorch implementation of "Moises-Light: Resource-efficient Band-split U-Net For Music Source Separation"
Python based audio denoiser 🔉
Code to train a custom time-domain autoencoder to dereverb audio
A re-implementation of the Wavelets package using Cython to improve the speed.
Paper Name: Complex Convolution Neural Network model (Complex DeepLab v3) on STFT time-varying frequency components for audio denoising Creating a Complex Deep Lab v3 model for audio denoising using STFT complex mask Dataset from: https://datashare.is.ed.ac.uk/handle/10283/2791
DeepSuppressor: A deep learning-based approach to speech denoising
Cross‑platform PyQt6 desktop app for video audio denoising using DeepFilterNet3 and FFmpeg
VoxPolish - a TTS dataset cleaner. Remove noise, trim silence, normalize loudness for text-to-speech training data in one command.
Audio denoising in real-time powered by artificial intelligence Python-friendly. Cross-platform. Check ROADMAP!
To associate your repository with the audio-denoising topic, visit your repo's landing page and select "manage topics."