{"repo":"echogarden-project/echogarden","free":true,"listed":false,"github":"https://github.com/echogarden-project/echogarden","clone":"git clone https://github.com/echogarden-project/echogarden.git","description":"Cross-platform speech toolset, used from the command-line or as a Node.js library. Includes a variety of engines for speech synthesis, speech recognition, forced alignment, speech translation, voice isolation, language detection and more.","language":"TypeScript","stars":448,"topics":["forced-alignment","language-identification","speech","speech-alignment","speech-recognition","speech-synthesis","speech-to-text","speech-translation","text-to-speech","language-detection"],"license":null,"category":"cli-tools","readme_excerpt":"Echogarden Echogarden is an easy-to-use speech toolset that includes a variety of speech processing tools. Easy to install, run, and update Written in TypeScript, for the Node.js runtime Can be used either as a command-line utility, or imported as a standard npm package Runs on Windows (x64, ARM64), macOS (x64, ARM64) and Linux (x64, ARM64) Doesn't require Python, Docker, or other system-level dependencies Doesn't rely on essential platform-specific binaries. Engines are either written in pure TypeScript, ported via WebAssembly, or imported using the ONNX runtime Fully open-source (GPL v3) Features Text-to-speech using high-quality Kokoro and VITS offline models for many languages and dialects, and 16 other offline and online engines, including cloud services by Google, Microsoft, Amazon, OpenAI and ElevenLabs Speech-to-text using a custom TypeScript/ONNX port of the OpenAI Whisper speech recognition architecture, whisper.cpp, and several other engines, including cloud services by Google, Microsoft, Amazon and OpenAI Speech-to-transcript alignment using several variants of dynamic time warping (DTW, DTW-RA), including support for multi-pass (hierarchical) processing, or via guided decoding using Whisper recognition models. Supports 100+ languages Speech-to-text translation , translates speech in any of the 98 languages supported by Whisper, to English, with near word-level timing for the translated transcript Speech-to-translated-transcript alignment synchronizes spoken audio","default_branch":null,"files":null,"tree":[],"storefront":"/r/echogarden-project","claimed":false,"request_supported":{"post":"https://gitbuyer.com/r/echogarden-project/echogarden/request-supported","requests":0},"note":"indexed from public GitHub; nothing is for sale on this page. Clone it from GitHub. Paid listings live at /search."}