[](https://github.com/FORARTfe/HyMPS#- "AUDIO section") [](https://github.com/FORARTfe/HyMPS/blob/main/Audio/AI-based.md#-- "AI-based category") [](https://github.com/FORARTfe/HyMPS/blob/main/Audio/AI-Voicing.md#--- "Voicing page")

March 10, 2026 · View on GitHub

📁 Cloners - Denoisers - Dubbers - Enhancers - Extractors - Stylers - Speech - TTSers

Warning

SORTING: Language (a>z) > License (openness) > Repository (a>z)\color{orange}\textsf{{SORTING: Language (a>z) > License (openness) > Repository (a>z)}}

Cloners

RepositoryShort descriptionLanguageLicenseWeightsLast commit
ApplioVITS-based Voice Conversion focused on simplicity, quality and performance-
OpenVoiceVersatile Instant Voice Cloning (paper)-
Real-Time Voice CloningClone a voice in 5 seconds to generate arbitrary speech in real-time-
Real-Time Voice CloningAn implementation of Transfer Learning from Speaker Verification to Multispeaker Text-To-Speech Synthesis paper with a vocoder that works in real-time-

Denoisers

RepositoryShort descriptionLanguageLicenseWeightsLast commit
denoise-autoencoderDenoise audio with convolutional autoencoderNA
DenoiserAn AI model to remove noise from the input audio using deep learning model which predicts the type of noise present and filter it out from the audio to give noise-free results1 x h5
Noise2Noise-denoisingSource code for the "Speech Denoising without Clean Training Data: a Noise2Noise Approach" paper24 x pth
Audio-Denoiser-CNN-1 x h5
DNPA PyTorch implementation for the "Speech Denoising by Accumulating Per-Frequency Modeling Fluctuations" paper-
SAB-cnn-audio-denoiserTensorflow 2.0 implementation of the paper A Fully Convolutional Neural Network for Speech EnhancementNA
Clean Sound with AIAims to clean sound recordings from noisy environments using a Convolutional Neural Network (CNN) based on the CleanUNet modelNA
denoiserA PyTorch implementation for the "Real Time Speech Enhancement in the Waveform Domain" paper1 2 3 4 x th
Speech-enhancementDeep learning for audio denoising1 x h5
DeepFilterNetNoise supression using deep filtering3 x ckpt/best 12 x onnx
speech-denoiserdenoiserNA
DTLNTensorflow 2.x implementation of the DTLN real time speech denoising model3 x h5 3 x tflite 2 x onnx

Dubbers

RepositoryShort descriptionLanguageLicenseWeightsLast commit
VoxellaThis app leverages advanced AI algorithms to automatically detect audio and text in the original language, providing effortless translation and native-like voice dubbing-
Voice Craft AIAn AI tool to dub videos into multiple regional languages and lip-sync at the same time-
DubFlowIt lets you effortlessly dub YouTube videos into any language with high-quality translations and synced audio-
Linly-DubbingAn intelligent multi-language AI dubbing and translation tool that offers diverse and high-quality dubbing options by integrating Linly-Talker’s digital human lip-sync technology, creating a more natural multi-language video experience-
voice_ukr_to_engTool to generate English AI Dubbing for a YouTube video-
Emotionally-Intelligent-AI-based-movie-dubbingAI to seamlessly translate and dub content into any language while preserving the original speaker's emotions, characteristics, and authenticity-
Multilingual audio visual system with lip synchronization using GANThis system takes a video in any language and generates a new video with synchronized lip movements speaking in English-
Video Dubbing ToolA fully-featured, multi-language video dubbing tool with a modern Streamlit GUI-
Auto Synced & Translated DubsAutomatically translates the text of a video into chosen languages based on a subtitle file, uses AI voice to dub the video, while keeping it properly synced to the original video using the subtitle's timings-
Kara-AudioGradio web-ui for vocal remover that uses demucs and MDX-Net + automatic subtitle creation using faster Whisper-
Youtube Auto DubbingSimple tool for dubbing youtube videos with AI generatied voice (inspired by Auto Synced & Translated Dubs)-
pyvideotransTranslate the video from one language to another and add dubbing-
Srt-AI-Voice-AssistantSubtitle dubbing with multiple AI projects-
AutoDubAn advanced AI-powered tool that automatically translates and dubs YouTube videos into different languages while dynamically adjusting video speed-
Dubbing AIProject for dubbing a video in many languages and with many different voices with the power of the AI-
Dublaris AI - Video TranslatorAn automated tool for multilingual video dubbing and subtitling-
PollydubleAutomatic Dubbing with Voice Cloning and Speech Recognition using OpenVoice, MeloTTS, Faster Whisper, VoiceFixer, python-audio-separator and FFmpeg-
AI Dubs over SubsA set of python scripts that take a video file as input and try to create a new video dubbed by AI-
AI Video DubbingAutomates translation and dubbing by transcribing audio with Speech-to-Text, translating it with the Translation API, and generating speech with Text-to-Speech using Google APIs-
Google AI DubbingAllows you to create localized videos using the same video base and adding translations using Google AI Powered TextToSpeech API-
Open dubbingAn AI dubbing system which uses machine learning models to automatically translate and synchronize audio dialogue into different languages-
YouDubAn innovative open source tool that focuses on translating and dubbing premium videos from platforms such as YouTube into Chinese-
InstantDubing: AI Video TranslatorIt uses cutting-edge AI technology to transcribe, translate, and then re-voice a video into English in the original speaker's voice-
MultivoiceThis project uses voice cloning and TTS to deliver natural and engaging dubbed dialogue for a seamless viewing adventure-
Open-Source AI Video DubberAutomatically dub any video into English while keeping the original voice styles-
SubdubA command line Python app offering a video-to-dubbed-video workflow with transcription, translation and synchronisation-
WeeaBlindA program to dub non-english media with modern AI speech synthesis, diarization, and voice cloning-
YouTube Auto-DubAutomated voice dubbing (for YouTube videos) that translates and dubs videos with original voice timbre-
AI Voice TranslatorA tool that translates audio into another language with the ability to use your own voice-

Enhancers

RepositoryShort descriptionLanguageLicenseWeightsLast commit
Speech Enhancement / Noise ReductionDemonstrates the process of enhancing and separating mixed audio sources, such as isolating speech from background noise, using a pre-trained model4 x ckpt
AURAL_GAN+predictive_modelAims to transform low-quality phone recordings into professional-quality audio using a Generative Adversarial Network (GAN)NA
Resemble EnhanceAI powered speech denoising and enhancement1 x bin 2 x pth 8 x safetensor 12 x pt
ClearerVoice-StudioAn AI-Powered Speech Processing Toolkit and Open Source SOTA Pretrained Models, Supporting Speech Enhancement, Separation, and Target Speaker Extraction8 x pt
Deep Learning Based Noise Reduction and Speech Enhancement SystemImplements two deep learning models, one can classify the type of noise, the other can retain human voice and reduce environmental noise4 x h5
Audio Enhancement and Denoising using Autoencoders-NA

Extractors

RepositoryShort descriptionLanguageLicenseWeightsLast commit
deeper-wider-melodyCode for the "Enhancing Vocal Melody Extraction with Multilevel Contexts" paper1 x ckpt
Vocal-Extraction-from-Complex-Audio-MixturesA hybrid model combining CNNs and LSTM networks to isolate vocals from complex audio mixturesNA
ultimatevocalremoverguiA GUI for a Vocal Remover that uses Deep Neural Networks1 x pth

Stylers

RepositoryShort descriptionLanguageLicenseWeightsLast commit
AutoPSTGlobal Rhythm Style Transfer Without Text Transcriptions-
AutoVCZero-Shot Voice Style Transfer with Only Autoencoder Loss-
Voice Conversion with Non-Parallel DataDeep neural networks for voice conversion (voice style transfer) in Tensorflow-
StyleSingerCode for Style Transfer for Out-of-Domain Singing Voice Synthesis paper-
Voice style transfer with random CNNAudio style transfer with shallow random parameters CNN-
Robust-Voice-Style-TransferCode for Robust Disentangled Variational Speech Representation Learning for Zero-Shot Voice Conversion paper-

Speech

RepositoryShort descriptionLanguageLicenseWeightsLast commit
ASRTA Deep-Learning-Based Chinese Speech Recognition System-
whisper.cppHigh-performance inference of OpenAI's Whisper automatic speech recognition (ASR) model-

TTSers

RepositoryShort descriptionLanguageLicenseWeightsLast commit
OpenTTSUse Microsoft speech synthesis to generate your own voice package-
TorToiSeA multi-voice TTS system trained with an emphasis on quality-
coquiTTSA deep learning toolkit for Text-to-Speech, battle-tested in research and production-
Fish SpeechSOTA Open Source TTS-