{"repo":"modelscope/ClearerVoice-Studio","free":true,"listed":false,"github":"https://github.com/modelscope/ClearerVoice-Studio","clone":"git clone https://github.com/modelscope/ClearerVoice-Studio.git","description":"An AI-Powered Speech Processing Toolkit and Open Source SOTA Pretrained Models, Supporting Speech Enhancement, Separation, and Target Speaker Extraction, etc.","language":"Python","stars":4416,"topics":["audio","bandwidth-extension","deep-learning","noise-suppression","pytorch","speaker-extraction","speech","speech-enhancement","speech-separation","speech-super-resolution"],"license":"Apache-2.0","category":"machine-learning","readme_excerpt":"ClearerVoice-Studio is an open-source, AI-powered speech processing toolkit designed for researchers, developers, and end-users. It provides capabilities of speech enhancement, speech separation, speech super-resolution, target speaker extraction, and more. The toolkit provides state-of-the-art pre-trained models, along with training and inference scripts, all accessible from this repository. 👉🏻HuggingFace Demo👈🏻 👉🏻ModelScope Demo ｜ 👉🏻SpeechScore Demo👈🏻 ｜ 👉🏻Paper👈🏻 --- Please leave your ⭐ on our GitHub to support this community project！ 记得点击右上角的星星⭐来支持我们一下，您的支持是我们更新模型的最大动力！ News :fire: - Upcoming: More tasks will be added to ClearVoice. - [2025.6] Add an interface for ClearVoice that allows passing a Numpy array into the model and receiving its output as a NumPy array. It allows a more flexible call of the models during a training or inference pipeline. Please check out demo Numpy2Numpy.py . - [2025.5] Updated speechscore with more non-intrusive metrics: NISQA and DISTILL MOS - [2025.4] Updated pip installation for ClearVoice. Now you can simply type pip install clearvoice to use all the pretrained models in ClearVoice, see project description in PyPi link. - [2025.4] Added a training script for speech super-resolution, supporting both retraining and fine-tuning of models. For details, refer to the documentation here. - [2025.4] Added data generation scripts for training/finetuning speech enhancement models. The scripts generate either noisy speech or noisy-rever","default_branch":null,"files":null,"tree":[],"storefront":"/r/modelscope","claimed":false,"request_supported":{"post":"https://gitbuyer.com/r/modelscope/ClearerVoice-Studio/request-supported","requests":0},"note":"indexed from public GitHub; nothing is for sale on this page. Clone it from GitHub. Paid listings live at /search."}