{"repo":"kigner/audio.cpp-webui","free":true,"listed":false,"github":"https://github.com/kigner/audio.cpp-webui","clone":"git clone https://github.com/kigner/audio.cpp-webui.git","description":"audio.cpp with a full-task WebUI - pure C++ audio-model inference engine powered by ggml. TTS, ASR/STT, VAD, voice conversion, speaker diarization, music generation. No Python dependency.","language":"C++","stars":318,"topics":["asr","audio","cpp","ggml","inference-engine","music-generation","speaker-diarization","speech-recognition","speech-synthesis","speech-to-text"],"license":null,"category":"media-processing","readme_excerpt":"audio.cpp [!NOTE] This repository is a downstream distribution of 0xShug0/audio.cpp (Apache-2.0, Copyright ShugoAI LLC), extended with a full-task WebUI and Windows-friendly local launcher scripts, and periodically merged with upstream. All credit for the core inference framework goes to the upstream project — please star and contribute there. audio.cpp is a high-performance C++ audio inference framework built on top of ggml , designed to make modern local audio models practical, portable, and fast. Tired of juggling a dozen Conda environments, hundreds of Python packages, and dependency conflicts just to try a few audio models? audio.cpp gives those paths a shared native runtime instead. Runs on Windows, Linux, and macOS, with support for NVIDIA, AMD, Apple Silicon, and CPU-only machines. [!IMPORTANT] CUDA performance headline: multiple TTS paths already run 1.8x to up to 8x faster than their Python reference paths while cutting end-to-end latency by 45%-85% . GGUF performance: all released model families support GGUF loading, and tested Q8 packages can run up to 1.53x faster while reducing peak VRAM by up to about 37% on routes such as Higgs Audio, Fish Audio, and Voxtral. See the GGUF guide for support status and the Q8 performance report for 16-bit vs Q8 measurements. Production deployment example: Try Fun-ASR-Nano with audio.cpp on the FunASR platform https://www.funasr.com/en/deploy/audio-cpp.html! VibeVoice 1.5B: generates a 93.9-minute podcast in 18.2 minutes with 10 ","default_branch":null,"files":null,"tree":[],"storefront":"/r/kigner","claimed":false,"request_supported":{"post":"https://gitbuyer.com/r/kigner/audio.cpp-webui/request-supported","requests":0},"note":"indexed from public GitHub; nothing is for sale on this page. Clone it from GitHub. Paid listings live at /search."}