{"repo":"Yuan-ManX/ai-audio-datasets","free":true,"listed":false,"github":"https://github.com/Yuan-ManX/ai-audio-datasets","clone":"git clone https://github.com/Yuan-ManX/ai-audio-datasets.git","description":"AI Audio Datasets (AI-ADS) 🎵, including Speech, Music, and Sound Effects, which can provide training data for Generative AI, AIGC, AI model training, intelligent audio tool development, and audio applications.","language":null,"stars":960,"topics":["aigc","audio","audio-effect","datasets","artificial-intelligence","audio-generation","deep-learning","machine-learning","music-generation"],"license":"MIT","category":"machine-learning","readme_excerpt":"AI Audio Datasets (AI-ADS) 🎵 AI Audio Datasets (AI-ADS) 🎵, including Speech, Music, and Sound Effects, which can provide training data for Generative AI, AIGC, AI model training, intelligent audio tool development, and audio applications. Table of Contents Speech Music Sound Effect Project List Speech AISHELL-1 - AISHELL-1 is a corpus for speech recognition research and building speech recognition systems for Mandarin. AISHELL-3 - AISHELL-3 is a large-scale and high-fidelity multi-speaker Mandarin speech corpus published by Beijing Shell Shell Technology Co.,Ltd. It can be used to train multi-speaker Text-to-Speech (TTS) systems.The corpus contains roughly 85 hours of emotion-neutral recordings spoken by 218 native Chinese mandarin speakers and total 88035 utterances. Arabic speech Corpus - The Arabic Speech Corpus (1.5 GB) is a Modern Standard Arabic (MSA) speech corpus for speech synthesis. The corpus contains phonetic and orthographic transcriptions of more than 3.7 hours of MSA speech aligned with recorded speech on the phoneme level. The annotations include word stress marks on the individual phonemes. Audio-FLAN - Audio-FLAN aims to unify audio-language models that can seamlessly handle both understanding and generation tasks across speech, music, sound. An Instruction-Tuning Dataset for Unified Audio Understanding and Generation Across Speech, Music, and Sound. Audio-FLAN, a large-scale instruction-tuning dataset covering 80 diverse tasks across speech, music, and so","default_branch":null,"files":null,"tree":[],"storefront":"/r/Yuan-ManX","claimed":false,"request_supported":{"post":"https://gitbuyer.com/r/Yuan-ManX/ai-audio-datasets/request-supported","requests":0},"note":"indexed from public GitHub; nothing is for sale on this page. Clone it from GitHub. Paid listings live at /search."}